IP Library Granted Patent US 12,373,668
Granted Patent B2
US 12,373,668 · App. 16/632,328 · Granted Jul 29, 2025

Methods, systems and non-transitory computer readable media for automated design of molecules with desired properties using artificial intelligence

Inventors: Olexandr Isayev (Chapel Hill, NC); Mariya Popova (Almaty, KZ); Alexander Tropsha (Chapel Hill, NC)
Assignee: THE UNIVERSITY OF NORTH CAROLINA AT CHAPEL HILL
G06N3/045G06N3/044G06N3/088G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,373,668
App. No.
16/632,328
Granted
Jul 29, 2025
Kind
B2
Abstract

The subject matter described herein includes computational methods, systems and non-transitory computer readable media for de-novo drug discovery, which is based on deep learning and reinforcement learning techniques. The subject matter described herein allows generating chemical compounds with desired properties. Two deep neural networks-generative and predictive, represent the general workflow. The process of training consists of two stages. During the first stage, both models are trained separately with supervised learning algorithms, and during the second stage, models are trained jointly with reinforcement learning approach. In this study, we conduct a computational experiment, which demonstrates the efficiency of proposed strategy to maximize, minimize or impose a desired range to a property. We also thoroughly evaluate our models with quantitative approaches and provide visualization and interpretation of internal representation vectors for both predictive and generative models.

Claims (36)

1. A system for automated design of molecules using artificial intelligence, the system comprising:

a computing platform including at least one processor and a memory;

a generative model implemented by the at least one processor for utilizing a first artificial neural network trained in a first training stage to generate representations of valid molecules in a predetermined notation, wherein the first artificial neural network comprises a stack-augmented recurrent neural network trained to model valences of atoms and sequence dependencies including ring openings, ring closures, and bracket sequences to generate the representations of the valid molecules;

a predictive model implemented by the at least one processor and trained in the first training stage separately from the training of the generative model in the first training stage for predicting, from the representations of the valid molecules, numerical properties of the valid molecules; and

a reward function implemented by the at least one processor for generating reward values for the valid molecules, where the reward values are functions of the predicted numerical properties generated by the trained predictive model and a user-specified property, wherein the generative model and the predictive model are trained jointly in a second training stage utilizing the reward values to teach the first artificial neural network to output representations of the valid molecules having the user-specified property,

wherein, in the second training stage, the generative and predictive models are combined into a single reinforcement learning system in which the generative model operates as an agent whose action space is represented by a simplified molecular input line entry system (SMILES) notation alphabet and state space is represented by all possible strings in the SMILES notation alphabet and the predictive model operates as a critic which estimates behavior of the agent by assigning a reward value to each representation of one of the valid molecules generated by the generative model, and

wherein the predictive model includes an embedding layer, a long short term memory layer, and two dense layers.

2. The system of claim 1 wherein the first artificial neural network utilizes a deep learning method.

3. The system of claim 1 wherein the predictive model utilizes a second neural network designed to predict the properties of the valid molecules from the representations of the valid molecules in the predetermined notation.

4. The system of claim 1 wherein at least some of the valid molecules whose representations are generated by the generative model are novel chemical entities.

5. The system of claim 1 wherein the representations of valid molecules generated by the generative model represent molecules that can be synthesized.

6. The system of claim 1 wherein the generative model is configured to generate libraries of the representations of the valid molecules with desired profiles of properties.

7. The system of claim 1 wherein the user-specified property comprises a chemical property, a physical property, or a biological property.

8. The system of claim 7 wherein the user-specified property comprises a biological activity.

9. A method for automated design of molecules using artificial intelligence, the method comprising:

training, in a first training stage, a generative model comprising a first artificial neural network implemented by at least one processor to output representations of valid molecules in a predetermined notation, wherein the first artificial neural network comprises a stack-augmented recurrent neural network trained to model valences of atoms and sequence dependencies including ring openings, ring closures, and bracket sequences to generate the representations of the valid molecules;

training, in the first training stage and separately from the training of the generative model in the first training stage, a predictive model comprising a second neural network to predict properties of the valid molecules in the predetermined notation;

jointly training, in a second training stage, the generative model and the predictive model using reward values generated by a reward function which generates a reward value that is a function of a numerical property predicted by the trained predicted model to teach the first artificial neural network to output representations of the valid molecules having a user-specified property;

utilizing the trained generative model including the first artificial neural network to generate representations of valid molecules in the predetermined notation; and

utilizing the trained predictive model to predict properties of the valid molecules whose representations are output by the trained generative model,

wherein, in the second training stage, the generative and predictive models are combined into a single reinforcement learning system in which the generative model operates as an agent whose action space is represented by a simplified molecular input line entry system (SMILES) notation alphabet and state space is represented by all possible strings in the SMILES notation alphabet and the predictive model operates as a critic which estimates behavior of the agent by assigning a reward value to each representation of one of the valid molecules generated by the generative model, and

wherein the predictive model includes an embedding layer, a long short term memory layer, and two dense layers.

10. The method of claim 9 wherein the first artificial neural network utilizes a deep learning method.

11. The method of claim 9 wherein the predictive model utilizes a second neural network designed to predict the properties of the valid molecules from the representations of the valid molecules in the predetermined notation.

12. The method of claim 9 wherein at least some of the valid molecules whose representations are generated by the generative model are novel chemical entities.

13. The method of claim 9 wherein the user-specified property comprises a chemical property, a physical property, or a biological property.

14. The method of claim 13 wherein the user-specified property comprises a biological activity.

15. The method of claim 9 wherein the representations of the valid molecules generated by the generative model represent molecules that can be synthesized.

16. A non-transitory computer readable medium having stored thereon executable instructions that when executed by a processor of a computer control the computer to perform steps comprising:

training, in a first training stage, a generative model comprising a first artificial neural network to output valid molecules in a predetermined notation, wherein the first artificial neural network comprises a stack-augmented recurrent neural network trained to model valences of atoms and sequence dependencies including ring openings, ring closures, and bracket sequences to generate the representations of the valid molecules;

training, in the first training stage and separately from the training of the generative model in the first training stage, a predictive model comprising a second neural network to predict properties of molecules in the predetermined notation;

jointly training, in a second training stage, the generative model and the predictive model using reward values generated by a reward function which generates a reward value that is a function of a numerical property predicted by the trained predicted model to teach the first artificial neural network to output representations of the valid molecules having a user-specified property;

utilizing the trained generative model including the first artificial neural network to generate valid molecules in the predetermined notation; and

utilizing the trained predictive model to predict properties of the molecules output by the trained generative model,

wherein, in the second training stage, the generative and predictive models are combined into a single reinforcement learning system in which the generative model operates as an agent whose action space is represented by a simplified molecular input line entry system (SMILES) notation alphabet and state space is represented by all possible strings in the SMILES notation alphabet and the predictive model operates as a critic which estimates behavior of the agent by assigning a reward value to each representation of one of the valid molecules generated by the generative model, and

wherein the predictive model includes an embedding layer, a long short term memory layer, and two dense layers.

Assignments (2)
CONFIRMATORY LICENSE Recorded May 10, 2023
From: UNIVERSITY OF NORTH CAROLINA AT CHAPEL HILL
To: NATIONAL SCIENCE FOUNDATION
Reel/Frame 063601/0252 →
CONFIRMATORY LICENSE Recorded Dec 8, 2020
From: UNIVERSITY OF NORTH CAROLINA, CHAPEL HILL
To: NATIONAL SCIENCE FOUNDATION
Reel/Frame 054645/0494 →
Continuity (2)
Provisional Application 62535069 · Jul 20, 2017
Related Publication 20200168302A1 · May 28, 2020
References Cited (89)
US 6587845B1 · Braunheim · 2003 [cited by applicant]
US 20060161407A1 · Lanza et al. · 2006 [cited by applicant]
US 20100010946A1 · De Winter · 2010 [cited by examiner]
US 20120265514A1 · Hopkins et al. · 2012 [cited by applicant]
US 20130252280A1 · Weaver et al. · 2013 [cited by applicant]
Guimaraes GL, Sanchez-Lengeling B, Outeiral C, Farias PL, Aspuru-Guzik A. Objective-reinforced generative adversarial networks (organ) for sequence generation models. arXiv preprint arXiv:1705.10843. May 30, 2017. (Year… [cited by examiner]
Segler MH, Kogej T, Tyrchan C, Waller MP. Generating Focussed Molecule Libraries for Drug Discovery with Recurrent Neural Networks. arXiv preprint arXiv:1701.01329. Jan. 5, 2017. (Year: 2017). [cited by examiner]
Joulin, Armand, and Tomas Mikolov. “Inferring algorithmic patterns with stack-augmented recurrent nets.” Advances in neural information processing systems 28 (2015). (Year: 2015). [cited by examiner]
Yu, Lantao, et al. “Seqgan: Sequence generative adversarial nets with policy gradient.” Proceedings of the AAAI conference on artificial intelligence. vol. 31. No. 1. 2017. (Year: 2017). [cited by examiner]
Jaques, Natasha, et al. “Tuning recurrent neural networks with reinforcement learning.” (2017). (Year: 2017). [cited by examiner]
Olivecrona, Marcus, et al. “Molecular de-novo design through deep reinforcement learning.” Journal of cheminformatics 9.1 (2017): 1-14. (Year: 2017). [cited by examiner]
Jaques, Natasha, et al. “Sequence tutor: Conservative fine-tuning of sequence generation models with kl-control.” International Conference on Machine Learning. PMLR, 2017. (Year: 2017). [cited by examiner]
Yuan, William, et al. “Chemical space mimicry for drug discovery.” Journal of chemical information and modeling 57.4 (2017): 875-882. (Year: 2017). [cited by examiner]
Duvenaud, David K., et al. “Convolutional networks on graphs for learning molecular fingerprints.” Advances in neural information processing systems 28 (2015). (Year: 2015). [cited by examiner]
Lei, Tao, et al. “Deriving neural architectures from sequence and graph kernels.” International Conference on Machine Learning. PMLR, 2017. (Year: 2017). [cited by examiner]
Bjerrum, Esben Jannik, and Richard Threlfall. “Molecular generation with recurrent neural networks (RNNs).” arXiv preprint arXiv:1705.04612 (2017). (Year: 2017). [cited by examiner]
Gómez-Bombarelli, Rafael, et al. “Automatic chemical design using a data-driven continuous representation of molecules.” arXiv:1610.02415, 2016 (Year: 2016). [cited by examiner]
“Scientific Databases,” SRC, Inc., https://www.srcinc.com/what-we-do/environmental/scientific-databases.html, pp. 1-4 (2020). [cited by applicant]
Mullard, “The Drug-Maker's Guide to the Galaxy,” Nature, vol. 549, pp. 445-447 (Sep. 28, 2017). [cited by applicant]
“Introduction to MarvinSketch,” https://docs.chemaxon.com/display/docs/Introduction_to_MarvinSketch.html, pp. 1-7 (Accessed Jul. 20, 2017). [cited by applicant]
“Murcko Scaffolds,” Datagrok, https://datagrok.ai/help/domains/chem/functions/murcko-scaffolds, pp. 1-1 (Accessed Jul. 20, 2017). [cited by applicant]
Stumpfe et al., “Similarity Searching,” Wiley Interdiscip. Rev. Compt. Mol. Sci., vol. 1, pp. 260-282 (2011). [cited by applicant]
“Zinc263823677,” Zinc 15, http://zinc15.docking.org/substances/ZINC000263823677/, pp. 1-10 (2021). [cited by applicant]
“ZinC271402431,” Zinc 15, http://zinc15.docking.org/substances/ZINC271402431/, pp. 1-10 (2021). [cited by applicant]
“Kinase Knowledgebase (KKB),” Eidogen Seranty, pp. 1-2 (2021). [cited by applicant]
“Reinforcement Learning,” Wikipedia, pp. 1-16 (Sep. 16, 2021). [cited by applicant]
Popova et al., “Deep Reinforcement Learning for de-novo Drug Design,” Science Advances, pp. 1-28 (2018). [cited by applicant]
“Chemotext,” Chemotext, http://chemotext.mml.unc.edu/, pp. 1-2 (Nov. 8, 2017). [cited by applicant]
Marblestone et al., “Automatically inferring scientific semantics,” MIT, The Wayback Machine https://web.archive.org/web/20160709144625/http://web.mit.edu/amarbles/www/neuro_word2vec.html, pp. 1-2 (Jul. 2016). [cited by applicant]
McCormick, “Word2Vec Tutorial—The Skip-Gram Model,” Chris McCormick, http://mccormickml.com/2016/04/19/word2vec-tutorial-the-skip-gram-model/, pp. 1-21 (Apr. 19, 2016). [cited by applicant]
De Asis et al., “Multi-Step Reinforcement Learning: A Unifying Algorithm,” The Thirty-Second AAAI Conference on Artificial Intelligence, pp. 2902-2909 (2018). [cited by applicant]
Gomez-Bombarelli et al., “Automatic Chemical Design Using a Data-Driven Continuous Representation of Molecules,” ACS Cent. Sci., vol. 4, pp. 268-276 (2018). [cited by applicant]
Segler et al., “Generating Focused Molecule Libraries for Drug Discovery with Recurrent Neural Networks,” ACS Cent. Sci., vol. 4, pp. 120-131 (2018). [cited by applicant]
Ragoza et al., “Protein-Ligand Scoring with Convolutional Neural Networks,” J Chem Inf Model, vol. 57, No. 4, pp. 1-50 (Apr. 24, 2017). [cited by applicant]
Altae-Tran et al., “Low Data Drug Discovery with One-Shot Learning,” ACS Cent. Sci., vol. 3, pp. 283-293 (Apr. 3, 2017). [cited by applicant]
Segler et al., “Modelling Chemical Reasoning to Predict and Invent Reactions,” Chem. Eur. J., vol. 23, pp. 6118-6128 (2017). [cited by applicant]
Smith et al., “ANI-1: an extensible neural network potential with DFT accuracy at force field computational cost,” Chem. Sci., vol. 8, pp. 3192-3203 (2017). [cited by applicant]
Yu et al., “SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient,” Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence, pp. 2852-2858 (2017). [cited by applicant]
Nguyen et al., “Pharos: Collating protein information to shed light on the druggable genome,” Nucleic Acids Research, vol. 45, pp. D995-D1002 (2017). [cited by applicant]
Santos et al., “A comprehensive map of molecular drug targets,” Nat Rev Drug Discov., vol. 16, No. 1, pp. 1-32 (Jan. 2017). [cited by applicant]
Chockley et al., “The End of Radiology? Three Threats to the Future Practice of Radiology,” J Am Coll Radiol, vol. 13, pp. 1415-1420 (2016). [cited by applicant]
Deleu et al., “Learning Operations on a Stack with Neural Turing Machines,” arXiv:1612.00827v1, pp. 1-6 (Dec. 2, 2016). [cited by applicant]
Jha et al., “Adapting to Artificial Intelligence Radiologists and Pathologists as Information Specialists,” Jama, pp. 1-2 (Nov. 29, 2016). [cited by applicant]
Krakovsky, “Reinforcement Renaissance,” Communications of the ACM, vol. 59, No. 8, pp. 12-14 (Aug. 2016). [cited by applicant]
Morgenstern, “Artificial Intelligence,” The Economist, https://www.economist.com/special-report/2016/06/23/the-return-of-the-machinery-question, pp. 1-5 (Jun. 23, 2016). [cited by applicant]
Longo et al., “Data Sharing,” The New England Journal of Medicine, vol. 374, No. 3, pp. 276-277 (Jan. 21, 2016). [cited by applicant]
Schneider et al., “De Novo Design at the Edge of Chaos,” Journal of Medicinal Chemistry, vol. 59, pp. 4077-4086 (2016). [cited by applicant]
Ursu et al., “DrugCentral: online drug compendium,” Nucleic Acids Research, vol. 25, pp. D932-D939 (2017). [cited by applicant]
Wang et al., “PubChem BioAssay: 2017 Update,” Nucleic Acids Research, vol. 45, pp. D955-D963 (2017). [cited by applicant]
Low et al., “Cheminformatics-aided pharmacovigilance: application to Stevens-Johnson Syndrome,” J Am Med Infom Assoc, vol. 23, pp. 1-11 (2016). [cited by applicant]
Rastegar-Mojarad et al., “Using Social Media Data to Identify Potential Candidates for Drug Repurposing: A Feasibility Study,” JMIR Res Protoc, vol. 5, No. 2, pp. 1-12 (2016). [cited by applicant]
Silver et al., “Mastering the game of Go with deep neural networks and tree search,” Nature, vol. 529, pp. 1-20 (Jan. 28, 2016). [cited by applicant]
Joulin et al., Inferring Algorithmic Patterns with Stack-Augmented Recurrent Nets, arXiv, pp. 1-10 (2015). [cited by applicant]
Oprea et al., “Computational and Practical Aspects of Drug Repositioning,” Assay and Drug Development Technologies, vol. 13, No. 6, pp. 299-306 (Jul./Aug. 2015). [cited by applicant]
Reker et al., “Active-learning strategies in computer-assisted drug discovery,” Drug Discovery Today, vol. 20, No. 4, pp. 458-465 (Apr. 2015). [cited by applicant]
Grefenstette et al., “Learning to Transduce with Unbounded Memory” arXiv, pp. 1-9 (2015). [cited by applicant]
Goodfellow et al., “Generative Adversarial Nets,” Adv. Neural Inf. Process. Syst., No. 27, pp. 1-9 (2014). [cited by applicant]
Spangler et al., “Automated Hypothesis Generation Based on Mining Scientific Literature,” KDD '14, pp. 1877-1886 (Aug. 24-27, 2014). [cited by applicant]
Tetko et al., “How Accurately Can We Predict the Melting Points of Drug-like Compounds?,” Journal of Chemical Information and Modeling, pp. 1-11 (Dec. 2014). [cited by applicant]
Chung et al., “Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling,” arXiv:1412.3555v1, pp. 1-9 (Dec. 11, 2014). [cited by applicant]
Bento et al., “The ChEMBL bioactivity database: an update,” Nucleic Acids Research, vol. 42, pp. D1083-D1090 (Nov. 7, 2013). [cited by applicant]
Yom-Tov et al., “Postmarket Drug Surveillance Without Trial Costs: Discovery of Adverse Drug Reactions Through Large-Scale Analysis of Web Search Queries,” J Med Internet Res, vol. 15, No. 6, pp. 1-12 (2013). [cited by applicant]
Mikolov et al., “Distributed Representations of Words and Phrases and their Compositionality,” Nips, pp. 1-9 (2013). [cited by applicant]
Mikolov et al., “Efficient Estimation of Word Representations in Vector Space,” arXiv:1301.3781v3, pp. 1-12 (Sep. 7, 2013). [cited by applicant]
Polishchuk et al., “Estimation of the size of drug-like chemical space based on GDB-17,” J Comput Aided Mol Des, vol. 27, pp. 675-679 (2013). [cited by applicant]
Scannell, “Diagnosing the decline in pharmaceutical R&D efficiency,” Nature Reviews, Drug Discovery, vol. 11, pp. 191-200 (Mar. 2012). [cited by applicant]
Ruddigkeit et al., “Enumeration of 166 Billion Organic Small Molecules in the Chemical Universe Database GDB-17,” Journal of Chemical Information and Modeling, vol. 52, pp. 2864-2875 (2012). [cited by applicant]
Sutskever et al., “Generating Text with Recurrent Neural Networks,” Proceedings of the 28th International Conference on Machine Learning, pp. 1-8 (2011). [cited by applicant]
Baker et al., “Mining connections between chemicals, proteins, and diseases extracted from Medline annotations,” Journal of Biomedical Informatics, vol. 43, pp. 510-519 (2010). [cited by applicant]
Fourches et al., “Trust, but verify: On the importance of chemical structure curation in cheminformatics and QSAR modeling research,” J Chem Inf Model, vol. 50, No. 7, pp. 1-37 (Jul. 26, 2010). [cited by applicant]
Mauser et al., “Recent developments in de novo design and scaffold hopping,” Current Opinion in Drug Discovery & Development, vol. 11, No. 3, pp. 365-374 (2008). [cited by applicant]
Van der Maaten et al., “Visualizing Data using t-SNE,” Journal of Machine Learning Research, vol. 9, pp. 2579-2605 (2008). [cited by applicant]
Berstel, “Transductions and Context-Free Languages,” Teubner-Verlag, pp. 1-139 (Feb. 19, 2007). [cited by applicant]
Macarron, “Critical review of the role of HTS in drug discovery,” Drug Discovery Today, vol. 11, Nos. 7/8, pp. 277-279 (Apr. 2006). [cited by applicant]
Schnecke et al., “Computational chemistry-driven decision making in lead generation,” DDT, vol. 11, No. 1/2 , pp. 44-50 (Jan. 2006). [cited by applicant]
Irwin et al., “Zinc-A Free Database of Commercially Available Compounds for Virtual Screening,” J Chem Inf Model., vol. 45, No. 1, pp. 1-11 (2005). [cited by applicant]
Schneider et al., “Computer-Based De Novo Design of Drug-Like Molecules,” Nature Reviews, Drug Discovery, vol. 4, pp. 649-663 (Aug. 2005). [cited by applicant]
Kralovics et al., “A Gain-of-Function Mutation of JAK2 in Myeloproliferative Disorders,” The New England Journal of Medicine, vol. 352, pp. 1779-1790 (2005). [cited by applicant]
Lipinski et al., “Navigating chemical space for biology and medicine,” Nature, vol. 432, pp. 1-8 (Dec. 16, 2004). [cited by applicant]
Brown et al., “A Graph-Based Genetic Algorithm and Its Application to the Multiobjective Evolution of Median Molecules,” J. Chem. Inf. Comp. Sci., vol. 44, pp. 1079-1087 (2004). [cited by applicant]
Van den Herik et al., “Games solved: Now and in the future,” Artificial Intelligence, vol. 134, pp. 277-311 (2002). [cited by applicant]
Hochreiter et al., “Long Short-Term Memory,” Neural Computation, vol. 9, No. 8, pp. 1735-1780 (1997). [cited by applicant]
Sakatsume et al., “The Jak Kinases Differentially Associate with the α and β (Accessory Factor) Chains of the Interferon γ Receptor to Form a Functional Receptor Unit Capable of Activating STAT Transcription Factors*,” … [cited by applicant]
Williams, “Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning,” Machine Learning, vol. 8, pp. 229-256 (1992). [cited by applicant]
Swanson et al., “Medical literature as a potential source of new knowledge*,” Bull Med Libr Assoc, vol. 78, No. 1, pp. 29-37 (Jan. 1990). [cited by applicant]
Weininger, “Smiles, a Chemical Language and Information System. 1. Introduction to Methodology and Encoding Rules,” J. Chem. Inf. Comput. Sci., vol. 28, No. 1, pp. 31-36 (1988). [cited by applicant]
Notification of Transmittal of the International Search Report and the Written Opinion of the International Searching Authority, or the Declaration for International Application No. PCT/US2018/043114 (Oct. 15, 2018). [cited by applicant]
Bjerrum, “Smiles Enumeration as Data Augmentation for Neural Network Modeling of Molecules,” Cornell University Library, Computer Science, Machine Learning, arXiv:1703.07076, pp. 1-7 (May 17, 2017). [cited by applicant]
Olivecronia et al., “Molecular De-Novo Design through Deep Reinforcement Learning,” Cornell University Library, Computer Science, Artificial Intelligent, arXiv:1704.07555, pp. 1-16 (Apr. 25, 2017). [cited by applicant]