Patent Searching using Bayesian Statistics
|
|
- Pamela Lyons
- 5 years ago
- Views:
Transcription
1
2 Patent Searching using Bayesian Statistics Willem van Hoorn, Exscientia Ltd Biovia European Forum, London, June 2017
3 Contents Who are we? Searching molecules in patents What can Pipeline Pilot do for you? And what it cannot Proof of concept implementation 3
4 Pioneers of Centaur drug discovery The best of Artificial Intelligence and Human drug discovery expertise combined to deliver radical improvements in productivity Artificial Intelligence AI IA oversee Specialist drug design algorithms evolve, design, and propose what to make Intelligence Augmentation Empowered humans make final decisions and strategy Multiple applications for small molecule discovery Single target compounds Bispecific small molecules Phenotypic drug design 4
5 5
6 Growth Through Collaborations Automated design for single targets Phenotypic Drug Discovery platform Bispecific small molecules immuno-oncology Bispecific small molecule discovery agreement in metabolic disease Company founded First identification of bispecfic small molecules for dual targets Bispecific small molecules targeting 2 distinct GPCRs CNS disease First milestone reached first candidate delivered Global pharma (to be announced Q2 2017) Multi-target discovery agreement Value potential 4M 50:50 240M+ 6
7 Searching structures in patents I have a brilliant new compound idea, is it novel? Patent text searching is easy (Google) Molecular structure searching is not For a quick answer: manual substructure search in SciFinder, SureChEMBL, etc Issue: need a substructure query This is an issue if you want to do this a lot, automatically 7
8 Searching patents - assumptions Structures from patents are available Claimed compounds in patent have common substructure Claimed structures are novel This looks like a set amenable to modelling 8
9 A trivial example Example patent (random): EP A1 Bayesian model: Good : structures from EP A1 Bad : structures from random 200k patents High scoring molecule shown Red atoms: high contribution to score ( novel ) Phenyl is common and has low contribution In a similarity search both would be equally important 9
10 Learn Molecular Categories Training set Set of molecules from multiple categories Typical user case: activity classes Here: category = patent Model prediction Probabilities molecule belongs to category What distinguishes molecule of category A from molecule of category B Prediction made on full structure 10
11 Example with four categories Each column is a category Each row is a fingerprint feature Cell = NormalizedProbability ~ weight of fingerprint Pos: feature associated with being in this class Neg: feature associated with not being in this class Insignificant: < value <
12 A patent model Downloaded SureChembl mapping files (Oct 2015) ~1.7M patents with ~20M Claimed structures (including duplicates) Selected random 1000 patents Random split: 80% training (152k), 20% test set (38k) Multi-category Bayesian model based on ECFP_6 A model with 997 categories was derived. Some patents contained few compounds, none left in 80% Predicted patent for test set 12
13 Result: ~75% ranks in top 3 Frequency (log scale) Rank of true patent in top 997 This looks promising 13
14 Fails: reagents, wrong structures Note these are Claimed structures (really?) Corpus frequency: Bayesian score If these are the failures I am not worried 14
15 The model matrix does not scale Exponential growth: adding a category adds a column, adding a compound adds 0 rows (most likely >0) Building a full patent model runs out of memory 15
16 Redundancy in multi-cat models 8878 rows x 4 categories = 35,512 values But only 1074 unique & significant values ( NP > 0.05) 16
17 Store the model in a database Databases are designed to deal with large data volumes Normalisation removes redundancy Indexing for fast lookup sql query to evaluate model Need the normalised probabilities And building 1.7M models is too slow 17
18 Bayes in PilotScript This is much faster since it skips the publication of the model in xmldb 18
19 Component Does scripted model work? For a random patent (EP A1) Create model by script and Learn Good Molecules Evaluate on SampleDrugs.sd Scores are close (not identical!) but correlate well Script Component Script 19
20 Model stored in database Runtime: ~10.5 hours (Core i7 laptop, 8 CPU, 32 Gb) 20
21 Model scores by sql query* select pt.sc_patent_id, sum(np.normalised_probability) as score from eps_normalised_probabilities np, eps_fingerprints fp, eps_patents pt where np.fingerprint_id = fp. fingerprint_id and np.patent_id = pt. patent_id and fp.feature in ( fingerprints of query structure ) group by pt.number order by score desc Runtime: 2-3 minutes per compound * Need to correct for features not in this patent, see slide in backups 21
22 A test Query: Sildenafil Top-ranked patents: Top ranked patent lists (all?) known PDE5 inhibitors in Claim All other patents also contain Sildenafil However: original Sildenafil patent not found? 22
23 US A Original Sildenafil patent It is not in the 1.7M patents model Yet it is in SureChEMBL, with structures In downloaded copy, all structures are annotated as Description, not Claims. Therefore ignored This is frustrating 23
24 Conclusions Have prototype that shows technique works However it is not yet useful SureChEMBL data not yet accurate enough (Oct 2015) Claims vs Description, etc Multiple structures are wrong 24
25 EXSCIENTIA LTD, LAB 12 - DUNDEE INCUBATOR JAMES LINDSAY PLACE, DD1 5JJ, UNITED KINGDOM CONTACT@EXSCIENTIA.CO.UK
26 SureChEMBL 26
27 Searching virtual libraries Training set: ~100k compounds, ~5k categories 27
28 Evaluate model scores For each compound, fingerprints Keep top 1000 get patent name 1. get sum(np) for known fp 3. get Laplacian, Pactive for patent 2. get array of known FP 4. correct for unknown fingerprints Runtime: 2-3 minutes 28
29 29
Introducing a Bioinformatics Similarity Search Solution
Introducing a Bioinformatics Similarity Search Solution 1 Page About the APU 3 The APU as a Driver of Similarity Search 3 Similarity Search in Bioinformatics 3 POC: GSI Joins Forces with the Weizmann Institute
More informationReaxys Medicinal Chemistry Fact Sheet
R&D SOLUTIONS FOR PHARMA & LIFE SCIENCES Reaxys Medicinal Chemistry Fact Sheet Essential data for lead identification and optimization Reaxys Medicinal Chemistry empowers early discovery in drug development
More informationThe Case for Use Cases
The Case for Use Cases The integration of internal and external chemical information is a vital and complex activity for the pharmaceutical industry. David Walsh, Grail Entropix Ltd Costs of Integrating
More informationThe shortest path to chemistry data and literature
R&D SOLUTIONS Reaxys Fact Sheet The shortest path to chemistry data and literature Designed to support the full range of chemistry research, including pharmaceutical development, environmental health &
More informationhas its own advantages and drawbacks, depending on the questions facing the drug discovery.
2013 First International Conference on Artificial Intelligence, Modelling & Simulation Comparison of Similarity Coefficients for Chemical Database Retrieval Mukhsin Syuib School of Information Technology
More informationPIOTR GOLKIEWICZ LIFE SCIENCES SOLUTIONS CONSULTANT CENTRAL-EASTERN EUROPE
PIOTR GOLKIEWICZ LIFE SCIENCES SOLUTIONS CONSULTANT CENTRAL-EASTERN EUROPE 1 SERVING THE LIFE SCIENCES SPACE ADDRESSING KEY CHALLENGES ACROSS THE R&D VALUE CHAIN Characterize targets & analyze disease
More informationIntroduction. OntoChem
Introduction ntochem Providing drug discovery knowledge & small molecules... Supporting the task of medicinal chemistry Allows selecting best possible small molecule starting point From target to leads
More informationCheminformatics Role in Pharmaceutical Industry. Randal Chen Ph.D. Abbott Laboratories Aug. 23, 2004 ACS
Cheminformatics Role in Pharmaceutical Industry Randal Chen Ph.D. Abbott Laboratories Aug. 23, 2004 ACS Agenda The big picture for pharmaceutical industry Current technological/scientific issues Types
More informationReaxys Pipeline Pilot Components Installation and User Guide
1 1 Reaxys Pipeline Pilot components for Pipeline Pilot 9.5 Reaxys Pipeline Pilot Components Installation and User Guide Version 1.0 2 Introduction The Reaxys and Reaxys Medicinal Chemistry Application
More informationÁkos Tarcsay CHEMAXON SOLUTIONS
Ákos Tarcsay CHEMAXON SOLUTIONS FINDING NOVEL COMPOUNDS WITH IMPROVED OVERALL PROPERTY PROFILES Two faces of one world Structure Footprint Linked Data Reactions Analytical Batch Phys-Chem Assay Project
More informationIntegrated Cheminformatics to Guide Drug Discovery
Integrated Cheminformatics to Guide Drug Discovery Matthew Segall, Ed Champness, Peter Hunt, Tamsin Mansley CINF Drug Discovery Cheminformatics Approaches August 23 rd 2017 Optibrium, StarDrop, Auto-Modeller,
More informationHandling Human Interpreted Analytical Data. Workflows for Pharmaceutical R&D. Presented by Peter Russell
Handling Human Interpreted Analytical Data Workflows for Pharmaceutical R&D Presented by Peter Russell 2011 Survey 88% of R&D organizations lack adequate systems to automatically collect data for reporting,
More informationIntelligent NMR productivity tools
Intelligent NMR productivity tools Till Kühn VP Applications Development Pittsburgh April 2016 Innovation with Integrity A week in the life of Brian Brian Works in a hypothetical pharma company / university
More informationComputational chemical biology to address non-traditional drug targets. John Karanicolas
Computational chemical biology to address non-traditional drug targets John Karanicolas Our computational toolbox Structure-based approaches Ligand-based approaches Detailed MD simulations 2D fingerprints
More informationDe Novo molecular design with Deep Reinforcement Learning
De Novo molecular design with Deep Reinforcement Learning @olexandr Olexandr Isayev, Ph.D. University of North Carolina at Chapel Hill olexandr@unc.edu http://olexandrisayev.com About me Ph.D. in Chemistry
More informationInteractive Feature Selection with
Chapter 6 Interactive Feature Selection with TotalBoost g ν We saw in the experimental section that the generalization performance of the corrective and totally corrective boosting algorithms is comparable.
More informationHow IJC is Adding Value to a Molecular Design Business
How IJC is Adding Value to a Molecular Design Business James Mills Sexis LLP ChemAxon TechTalk Stevenage, ov 2012 james.mills@sexis.co.uk Overview Introduction to Sexis Sexis IJC use cases Data visualisation
More informationEarly Stages of Drug Discovery in the Pharmaceutical Industry
Early Stages of Drug Discovery in the Pharmaceutical Industry Daniel Seeliger / Jan Kriegl, Discovery Research, Boehringer Ingelheim September 29, 2016 Historical Drug Discovery From Accidential Discovery
More informationIntelligent Systems (AI-2)
Intelligent Systems (AI-2) Computer Science cpsc422, Lecture 19 Oct, 24, 2016 Slide Sources Raymond J. Mooney University of Texas at Austin D. Koller, Stanford CS - Probabilistic Graphical Models D. Page,
More informationCheminformatics analysis and learning in a data pipelining environment
Molecular Diversity (2006) 10: 283 299 DOI: 10.1007/s11030-006-9041-5 c Springer 2006 Review Cheminformatics analysis and learning in a data pipelining environment Moises Hassan 1,, Robert D. Brown 1,
More informationFarewell, PipelinePilot Migrating the Exquiron cheminformatics platform to KNIME and the ChemAxon technology
Farewell, PipelinePilot Migrating the Exquiron cheminformatics platform to KNIME and the ChemAxon technology Serge P. Parel, PhD ChemAxon User Group Meeting, Budapest 21 st May, 2014 Outline Exquiron Who
More informationContents 1 Open-Source Tools, Techniques, and Data in Chemoinformatics
Contents 1 Open-Source Tools, Techniques, and Data in Chemoinformatics... 1 1.1 Chemoinformatics... 2 1.1.1 Open-Source Tools... 2 1.1.2 Introduction to Programming Languages... 3 1.2 Chemical Structure
More informationIntelligent Systems (AI-2)
Intelligent Systems (AI-2) Computer Science cpsc422, Lecture 19 Oct, 23, 2015 Slide Sources Raymond J. Mooney University of Texas at Austin D. Koller, Stanford CS - Probabilistic Graphical Models D. Page,
More informationEnvironment (Parallelizing Query Optimization)
Advanced d Query Optimization i i Techniques in a Parallel Computing Environment (Parallelizing Query Optimization) Wook-Shin Han*, Wooseong Kwak, Jinsoo Lee Guy M. Lohman, Volker Markl Kyungpook National
More informationCLRG Biocreative V
CLRG ChemTMiner @ Biocreative V Sobha Lalitha Devi., Sindhuja Gopalan., Vijay Sundar Ram R., Malarkodi C.S., Lakshmi S., Pattabhi RK Rao Computational Linguistics Research Group, AU-KBC Research Centre
More informationReaxys The Highlights
Reaxys The Highlights What is Reaxys? A brand new workflow solution for research chemists and scientists from related disciplines An extensive repository of reaction and substance property data A resource
More informationBayesian Classifiers and Probability Estimation. Vassilis Athitsos CSE 4308/5360: Artificial Intelligence I University of Texas at Arlington
Bayesian Classifiers and Probability Estimation Vassilis Athitsos CSE 4308/5360: Artificial Intelligence I University of Texas at Arlington 1 Data Space Suppose that we have a classification problem The
More informationSciFinder Content Update
SciFinder Content Update 1 Session Agenda New SciFinder is ready to use Current SciFinder data overview Content enhancements in SciFinder SciFinder training resources 2 The new SciFinder is ready to use
More informationApplying Bioisosteric Transformations to Predict Novel, High Quality Compounds
Applying Bioisosteric Transformations to Predict Novel, High Quality Compounds Dr James Chisholm,* Dr John Barnard, Dr Julian Hayward, Dr Matthew Segall*, Mr Edmund Champness*, Dr Chris Leeding,* Mr Hector
More informationStephen McDonald, Mark D. Wrona, Jeff Goshawk Waters Corporation, Milford, MA, USA INTRODUCTION
Combined with a Deep Understanding of Complex Metabolic Routes to Increase Efficiency in the Metabolite Identification Search Stephen McDonald, Mark D. Wrona, Jeff Goshawk Waters Corporation, Milford,
More informationIn Silico Investigation of Off-Target Effects
PHARMA & LIFE SCIENCES WHITEPAPER In Silico Investigation of Off-Target Effects STREAMLINING IN SILICO PROFILING In silico techniques require exhaustive data and sophisticated, well-structured informatics
More informationCS 188: Artificial Intelligence. Outline
CS 188: Artificial Intelligence Lecture 21: Perceptrons Pieter Abbeel UC Berkeley Many slides adapted from Dan Klein. Outline Generative vs. Discriminative Binary Linear Classifiers Perceptron Multi-class
More informationInformation Retrieval and Organisation
Information Retrieval and Organisation Chapter 13 Text Classification and Naïve Bayes Dell Zhang Birkbeck, University of London Motivation Relevance Feedback revisited The user marks a number of documents
More informationStudying the effect of noise on Laplacian-modified Bayesian Analysis and Tanimoto Similarity
Studying the effect of noise on Laplacian-modified Bayesian nalysis and Tanimoto Similarity David Rogers, Ph.D. SciTegic, Inc. (Division of ccelrys, Inc.) drogers@scitegic.com Description of: nalysis methods
More informationData Mining in the Chemical Industry. Overview of presentation
Data Mining in the Chemical Industry Glenn J. Myatt, Ph.D. Partner, Myatt & Johnson, Inc. glenn.myatt@gmail.com verview of presentation verview of the chemical industry Example of the pharmaceutical industry
More informationData Quality Issues That Can Impact Drug Discovery
Data Quality Issues That Can Impact Drug Discovery Sean Ekins 1, Joe Olechno 2 Antony J. Williams 3 1 Collaborations in Chemistry, Fuquay Varina, NC. 2 Labcyte Inc, Sunnyvale, CA. 3 Royal Society of Chemistry,
More informationCOMPARISON OF SIMILARITY METHOD TO IMPROVE RETRIEVAL PERFORMANCE FOR CHEMICAL DATA
http://www.ftsm.ukm.my/apjitm Asia-Pacific Journal of Information Technology and Multimedia Jurnal Teknologi Maklumat dan Multimedia Asia-Pasifik Vol. 7 No. 1, June 2018: 91-98 e-issn: 2289-2192 COMPARISON
More informationCOMBINATORIAL CHEMISTRY IN A HISTORICAL PERSPECTIVE
NUE FEATURE T R A N S F O R M I N G C H A L L E N G E S I N T O M E D I C I N E Nuevolution Feature no. 1 October 2015 Technical Information COMBINATORIAL CHEMISTRY IN A HISTORICAL PERSPECTIVE A PROMISING
More informationLarge scale classification of chemical reactions from patent data
Large scale classification of chemical reactions from patent data Gregory Landrum NIBR Informatics, Basel Novartis Institutes for BioMedical Research 10th International Conference on Chemical Structures/
More informationInformation Retrieval. Lecture 6
Information Retrieval Lecture 6 Recap of the last lecture Parametric and field searches Zones in documents Scoring documents: zone weighting Index support for scoring tf idf and vector spaces This lecture
More informationIntroduction to Google Drive Objectives:
Introduction to Google Drive Objectives: Learn how to access your Google Drive account Learn to create new documents using Google Drive Upload files to store on Google Drive Share files and folders with
More informationProbabilistic Context-free Grammars
Probabilistic Context-free Grammars Computational Linguistics Alexander Koller 24 November 2017 The CKY Recognizer S NP VP NP Det N VP V NP V ate NP John Det a N sandwich i = 1 2 3 4 k = 2 3 4 5 S NP John
More informationFROM MOLECULAR FORMULAS TO MARKUSH STRUCTURES
FROM MOLECULAR FORMULAS TO MARKUSH STRUCTURES DIFFERENT LEVELS OF KNOWLEDGE REPRESENTATION IN CHEMISTRY Michael Braden, PhD ACS / San Diego/ 2016 Overview ChemAxon Who are we? Examples/use cases: Create
More informationIII. Naïve Bayes (pp.70-72) Probability review
III. Naïve Bayes (pp.70-72) This is a short section in our text, but we are presenting more material in these notes. Probability review Definition of probability: The probability of an even E is the ratio
More informationCombinatorial Heterogeneous Catalysis
Combinatorial Heterogeneous Catalysis 650 μm by 650 μm, spaced 100 μm apart Identification of a new blue photoluminescent (PL) composite material, Gd 3 Ga 5 O 12 /SiO 2 Science 13 March 1998: Vol. 279
More informationPerseverance. Experimentation. Knowledge.
2410 Intuition. Perseverance. Experimentation. Knowledge. All are critical elements of the formula leading to breakthroughs in chemical development. Today s process chemists face increasing pressure to
More informationBig Data Analytics. Lucas Rego Drumond
Lucas Rego Drumond Information Systems and Machine Learning Lab (ISMLL) Institute of Computer Science University of Hildesheim, Germany Map Reduce I Map Reduce I 1 / 32 Outline 1. Introduction 2. Parallel
More informationDesign Theory: Functional Dependencies and Normal Forms, Part I Instructor: Shel Finkelstein
Design Theory: Functional Dependencies and Normal Forms, Part I Instructor: Shel Finkelstein Reference: A First Course in Database Systems, 3 rd edition, Chapter 3 Important Notices CMPS 180 Final Exam
More informationChemical Data Retrieval and Management
Chemical Data Retrieval and Management ChEMBL, ChEBI, and the Chemistry Development Kit Stephan A. Beisken What is EMBL-EBI? Part of the European Molecular Biology Laboratory International, non-profit
More informationMetabolite Identification and Characterization by Mining Mass Spectrometry Data with SAS and Python
PharmaSUG 2018 - Paper AD34 Metabolite Identification and Characterization by Mining Mass Spectrometry Data with SAS and Python Kristen Cardinal, Colorado Springs, Colorado, United States Hao Sun, Sun
More informationEnvironmental Chemistry through Intelligent Atmospheric Data Analysis (EnChIlADA): A Platform for Mining ATOFMS and Other Atmospheric Data
Environmental Chemistry through Intelligent Atmospheric Data Analysis (EnChIlADA): A Platform for Mining ATOFMS and Other Atmospheric Data Katie Barton, John Choiniere, Melanie Yuen, and Deborah Gross
More informationBuilding innovative drug discovery alliances. Just in KNIME: Successful Process Driven Drug Discovery
Building innovative drug discovery alliances Just in KIME: Successful Process Driven Drug Discovery Berlin KIME Spring Summit, Feb 2016 Research Informatics @ Evotec Evotec s worldwide operations 2 Pharmaceuticals
More informationGeo Business Gis In The Digital Organization
We have made it easy for you to find a PDF Ebooks without any digging. And by having access to our ebooks online or by storing it on your computer, you have convenient answers with geo business gis in
More informationMathangi Thiagarajan Rice Genome Annotation Workshop May 23rd, 2007
-2 Transcript Alignment Assembly and Automated Gene Structure Improvements Using PASA-2 Mathangi Thiagarajan mathangi@jcvi.org Rice Genome Annotation Workshop May 23rd, 2007 About PASA PASA is an open
More informationMachine learning for ligand-based virtual screening and chemogenomics!
Machine learning for ligand-based virtual screening and chemogenomics! Jean-Philippe Vert Institut Curie - INSERM U900 - Mines ParisTech In silico discovery of molecular probes and drug-like compounds:
More informationNavigating between patents, papers, abstracts and databases using public sources and tools
Navigating between patents, papers, abstracts and databases using public sources and tools Christopher Southan 1 and Sean Ekins 2 TW2Informatics, Göteborg, Sweden, Collaborative Drug Discovery, North Carolina,
More informationApplication Note 12: Fully Automated Compound Screening and Verification Using Spinsolve and MestReNova
Application Note : Fully Automated Compound Screening and Verification Using Spinsolve and MestReNova Paul Bowyer, Magritek, Inc. and Mark Dixon, Mestrelab Sample screening to verify the identity or integrity
More informationPipeline Pilot Integration
Scientific & technical Presentation Pipeline Pilot Integration Szilárd Dóránt July 2009 The Component Collection: Quick facts Provides access to ChemAxon tools from Pipeline Pilot Free of charge Open source
More informationHash-based Indexing: Application, Impact, and Realization Alternatives
: Application, Impact, and Realization Alternatives Benno Stein and Martin Potthast Bauhaus University Weimar Web-Technology and Information Systems Text-based Information Retrieval (TIR) Motivation Consider
More informationCOMPARING PERFORMANCE OF NEURAL NETWORKS RECOGNIZING MACHINE GENERATED CHARACTERS
Proceedings of the First Southern Symposium on Computing The University of Southern Mississippi, December 4-5, 1998 COMPARING PERFORMANCE OF NEURAL NETWORKS RECOGNIZING MACHINE GENERATED CHARACTERS SEAN
More informationReceptor Based Drug Design (1)
Induced Fit Model For more than 100 years, the behaviour of enzymes had been explained by the "lock-and-key" mechanism developed by pioneering German chemist Emil Fischer. Fischer thought that the chemicals
More informationSearch for Substance Data
Search for Substance Data Virtual class structure Part 1 Find property data for a substance Retrieve chemical supplier information for a substance Part 2 Import a structure and conduct an exact structure
More informationLarge Scale Evaluation of Chemical Structure Recognition 4 th Text Mining Symposium in Life Sciences October 10, Dr.
Large Scale Evaluation of Chemical Structure Recognition 4 th Text Mining Symposium in Life Sciences October 10, 2006 Dr. Overview Brief introduction Chemical Structure Recognition (chemocr) Manual conversion
More informationMachine Learning Concepts in Chemoinformatics
Machine Learning Concepts in Chemoinformatics Martin Vogt B-IT Life Science Informatics Rheinische Friedrich-Wilhelms-Universität Bonn BigChem Winter School 2017 25. October Data Mining in Chemoinformatics
More informationThe Rockefeller University Compound Library
The Rockefeller University Compound Library J. Fraser Glickman Rockefeller University High Throughput and Spectroscopy Resource Center April 4 th, 2014 Technologies Resources Available Microplate assay
More informationCross Discipline Analysis made possible with Data Pipelining. J.R. Tozer SciTegic
Cross Discipline Analysis made possible with Data Pipelining J.R. Tozer SciTegic System Genesis Pipelining tool created to automate data processing in cheminformatics Modular system built with generic
More informationCS 188: Artificial Intelligence. Machine Learning
CS 188: Artificial Intelligence Review of Machine Learning (ML) DISCLAIMER: It is insufficient to simply study these slides, they are merely meant as a quick refresher of the high-level ideas covered.
More informationText Mining. Dr. Yanjun Li. Associate Professor. Department of Computer and Information Sciences Fordham University
Text Mining Dr. Yanjun Li Associate Professor Department of Computer and Information Sciences Fordham University Outline Introduction: Data Mining Part One: Text Mining Part Two: Preprocessing Text Data
More informationCS 343: Artificial Intelligence
CS 343: Artificial Intelligence Perceptrons Prof. Scott Niekum The University of Texas at Austin [These slides based on those of Dan Klein and Pieter Abbeel for CS188 Intro to AI at UC Berkeley. All CS188
More informationIntroducing the Morphologi G3 ID The future of particle characterization
Introducing the Morphologi G3 ID The future of particle characterization Dr Anne Virden, Product technical specialist diffraction and analytical imaging What is the Morphologi G3-ID? Advanced R&D particle
More informationDesign and Synthesis of the Comprehensive Fragment Library
YOUR INNOVATIVE CHEMISTRY PARTNER IN DRUG DISCOVERY Design and Synthesis of the Comprehensive Fragment Library A 3D Enabled Library for Medicinal Chemistry Discovery Warren S Wade 1, Kuei-Lin Chang 1,
More informationBiologically Relevant Molecular Comparisons. Mark Mackey
Biologically Relevant Molecular Comparisons Mark Mackey Agenda > Cresset Technology > Cresset Products > FieldStere > FieldScreen > FieldAlign > FieldTemplater > Cresset and Knime About Cresset > Specialist
More informationSpatial Role Labeling CS365 Course Project
Spatial Role Labeling CS365 Course Project Amit Kumar, akkumar@iitk.ac.in Chandra Sekhar, gchandra@iitk.ac.in Supervisor : Dr.Amitabha Mukerjee ABSTRACT In natural language processing one of the important
More informationData Dependencies in the Presence of Difference
Data Dependencies in the Presence of Difference Tsinghua University sxsong@tsinghua.edu.cn Outline Introduction Application Foundation Discovery Conclusion and Future Work Data Dependencies in the Presence
More informationRapid Application Development using InforSense Open Workflow and Daylight Technologies Deliver Discovery Value
Rapid Application Development using InforSense Open Workflow and Daylight Technologies Deliver Discovery Value Anthony Arvanites Daylight User Group Meeting March 10, 2005 Outline 1. Company Introduction
More informationUTAH S STATEWIDE GEOGRAPHIC INFORMATION DATABASE
UTAH S STATEWIDE GEOGRAPHIC INFORMATION DATABASE Data Information and Knowledge Management NASCIO Awards 2009 STATE GEOGRAPHIC INFORMATION DATABASE B. EXECUTIVE SUMMARY Utah has developed one of the most
More informationExpanding the scope of literature data with document to structure tools PatentInformatics applications at Aptuit
Expanding the scope of literature data with document to structure tools PatentInformatics applications at Aptuit Alfonso Pozzan Computational and Analytical Chemistry Drug Design and Discovery Department
More informationFurther information: Basic principles of quantum computing Information on development areas of Volkswagen Group IT
Media information Further information: Basic principles of quantum computing Information on development areas of Volkswagen Group IT Basic principles of quantum computing Development areas of Volkswagen
More informationAlgorithms for NLP. Language Modeling III. Taylor Berg-Kirkpatrick CMU Slides: Dan Klein UC Berkeley
Algorithms for NLP Language Modeling III Taylor Berg-Kirkpatrick CMU Slides: Dan Klein UC Berkeley Announcements Office hours on website but no OH for Taylor until next week. Efficient Hashing Closed address
More informationInstruction to search natural compounds on CH-NMR-NP
Instruction to search natural compounds on CH-NMR-NP The CH-NMR-NP is a charge free service for all users. Please note that required information (name, affiliation, country, email) has to be submitted
More informationDesign and Development of a Large Scale Archaeological Information System A Pilot Study for the City of Sparti
INTERNATIONAL SYMPOSIUM ON APPLICATION OF GEODETIC AND INFORMATION TECHNOLOGIES IN THE PHYSICAL PLANNING OF TERRITORIES Sofia, 09 10 November, 2000 Design and Development of a Large Scale Archaeological
More informationUltra High Throughput Screening using THINK on the Internet
Ultra High Throughput Screening using THINK on the Internet Keith Davies Central Chemistry Laboratory, Oxford University Cathy Davies Treweren Consultants, UK Blue Sky Objectives Reduce Development Failures
More informationSolution Chemical Engineering Kinetics Jm Smith
We have made it easy for you to find a PDF Ebooks without any digging. And by having access to our ebooks online or by storing it on your computer, you have convenient answers with solution chemical engineering
More informationComp 11 Lectures. Mike Shah. July 26, Tufts University. Mike Shah (Tufts University) Comp 11 Lectures July 26, / 40
Comp 11 Lectures Mike Shah Tufts University July 26, 2017 Mike Shah (Tufts University) Comp 11 Lectures July 26, 2017 1 / 40 Please do not distribute or host these slides without prior permission. Mike
More informationECE521 W17 Tutorial 1. Renjie Liao & Min Bai
ECE521 W17 Tutorial 1 Renjie Liao & Min Bai Schedule Linear Algebra Review Matrices, vectors Basic operations Introduction to TensorFlow NumPy Computational Graphs Basic Examples Linear Algebra Review
More informationNEC PerforCache. Influence on M-Series Disk Array Behavior and Performance. Version 1.0
NEC PerforCache Influence on M-Series Disk Array Behavior and Performance. Version 1.0 Preface This document describes L2 (Level 2) Cache Technology which is a feature of NEC M-Series Disk Array implemented
More informationThe Schrödinger KNIME extensions
The Schrödinger KNIME extensions Computational Chemistry and Cheminformatics in a workflow environment Jean-Christophe Mozziconacci Volker Eyrich Topics What are the Schrödinger extensions? Workflow application
More informationSearching Substances in Reaxys
Searching Substances in Reaxys Learning Objectives Understand that substances in Reaxys have different sources (e.g., Reaxys, PubChem) and can be found in Document, Reaction and Substance Records Recognize
More informationACD/Labs Software Impurity Resolution Management. Presented by Peter Russell
ACD/Labs Software Impurity Resolution Management Presented by Peter Russell Impurity Resolution Process Chemists Method Development Specialists Toxicology Groups Stability Groups Analytical Chemists 7/8/2016
More informationEMPIRICAL VS. RATIONAL METHODS OF DISCOVERING NEW DRUGS
EMPIRICAL VS. RATIONAL METHODS OF DISCOVERING NEW DRUGS PETER GUND Pharmacopeia Inc., CN 5350 Princeton, NJ 08543, USA pgund@pharmacop.com Empirical and theoretical approaches to drug discovery have often
More informationConfirmation of In Vitro Nefazodone Metabolites using the Superior Fragmentation of the QTRAP 5500 LC/MS/MS System
Confirmation of In Vitro Nefazodone Metabolites using the Superior Fragmentation of the QTRAP 5500 LC/MS/MS System Claire Bramwell-German, Elliott Jones and Daniel Lebre AB SCIEX, Foster City, California
More informationUsing Artificial Intelligence to Train Your Chatbot.
Using Artificial Intelligence to Train Your Chatbot www.coseer.com praful@coseer.com T h e C o n t e x t 1 A r e c h a t b o t s a f a d o r a d i s r u p t i o n? M i s c e l l a n e o u s G r o w t h
More informationComputational Methods and Drug-Likeness. Benjamin Georgi und Philip Groth Pharmakokinetik WS 2003/2004
Computational Methods and Drug-Likeness Benjamin Georgi und Philip Groth Pharmakokinetik WS 2003/2004 The Problem Drug development in pharmaceutical industry: >8-12 years time ~$800m costs >90% failure
More informationThe Comprehensive Report
High-Throughput Screening 2002: New Strategies and Technologies The Comprehensive Report Presented by HighTech Business Decisions 346 Rheem Blvd., Suite 208, Moraga, CA 94556 Tel: (925) 631-0920 Fax: (925)
More informationAn Optimized Interestingness Hotspot Discovery Framework for Large Gridded Spatio-temporal Datasets
IEEE Big Data 2015 Big Data in Geosciences Workshop An Optimized Interestingness Hotspot Discovery Framework for Large Gridded Spatio-temporal Datasets Fatih Akdag and Christoph F. Eick Department of Computer
More informationCS 5522: Artificial Intelligence II
CS 5522: Artificial Intelligence II Perceptrons Instructor: Alan Ritter Ohio State University [These slides were adapted from CS188 Intro to AI at UC Berkeley. All materials available at http://ai.berkeley.edu.]
More informationThe Quantum Landscape
The Quantum Landscape Computational drug discovery employing machine learning and quantum computing Contact us! lucas@proteinqure.com Or visit our blog to learn more @ www.proteinqure.com 2 Applications
More informationAnomaly Detection for the CERN Large Hadron Collider injection magnets
Anomaly Detection for the CERN Large Hadron Collider injection magnets Armin Halilovic KU Leuven - Department of Computer Science In cooperation with CERN 2018-07-27 0 Outline 1 Context 2 Data 3 Preprocessing
More informationScience of Synthesis Guided Examples
Science of Synthesis Guided Examples 1 Science of Synthesis Guided Examples Compiled by: Dr. Thomas Krimmer Table of Contents 1 Exact Structure Search...2 2 Houben-Weyl...5 3 Reaction Search...6 4 Outbound
More informationIgnasi Belda, PhD CEO. HPC Advisory Council Spain Conference 2015
Ignasi Belda, PhD CEO HPC Advisory Council Spain Conference 2015 Business lines Molecular Modeling Services We carry out computational chemistry projects using our selfdeveloped and third party technologies
More information