IP Library › Granted Patent US 12,378,383
Granted Patent B2
US 12,378,383 · App. 17/967,685 · Granted Aug 5, 2025

Molecular structure transformers for property prediction

Inventors: Tusharkumar Gadhiya (Gandhinagar, IN); Falak Shah (Mountain View, CA); Nisarg Vyas (Mountain View, CA); Julia Yang (Berkeley, CA); Vahe Gharakhanyan (Mountain View, CA); Alexander Holiday (Brookfield, WI)
Assignee: X Development LLC
C08J11/16C08J11/10G16C10/00G16C20/10G16C20/20G16C20/40G16C20/70G16C20/80G16C60/00C08J2367/02C08J2367/04C08J2467/02C08J2467/04
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,378,383
App. No.
17/967,685
Filed
Oct 17, 2022
Granted
Aug 5, 2025
Kind
B2
Examiner
LE, JOHN H
Art Unit
2857
USPC
702/27
Abstract

Computer-implemented methods may include accessing a multi-dimensional embedding space that supports relating embeddings of molecules to predicted values of a given property of the molecules. The method may also include identifying one or more points of interest within the embedding space based on the predicted values. Each of the one or more points of interest may include a set of coordinate values within the multi-dimensional embedding space and may be associated with a corresponding predicted value of the given property. The method may further include generating, for each of the one or more points of interest, a structural representation of a molecule by transforming the set of coordinate values included in the point of interest using a decoder network. The method may include outputting a result that identifies, for each of the one or more points of interest, the structural representation of the molecule corresponding to the point of interest.

Claims (48)

1. A computer-implemented method comprising:

accessing a multi-dimensional embedding space that supports relating embeddings of molecules to predicted values of a given property of the molecules;

identifying one or more points of interest within the multi-dimensional embedding space based on the predicted values, wherein each of the one or more points of interest:

includes a set of coordinate values within the multi-dimensional embedding space,

conveys spatial information of atoms or bonds in a molecule, and

is associated with a corresponding predicted value of the given property;

generating, for each of the one or more points of interest, a structural representation of the molecule by transforming the set of coordinate values included in a point of interest using a decoder network, wherein training of the decoder network included learning to transform positions within the embedding space to outputs representing molecular-structure characteristics, wherein the training of the decoder network was performed at least in part concurrently to training an encoder network to transform positions within the embedding space to predictions corresponding to values of the given property; and

outputting a result that identifies, for each of the one or more points of interest, the structural representation of the molecule corresponding to the point of interest.

2. The method of claim 1 , wherein training of the encoder network included learning to transform partial or complete bond string and position (BSP) representations of molecules into positions within the embedding space, and wherein each BSP representation identifies relative positions of atoms connected by a bond in the represented molecule.

3. The method of claim 2 , wherein each BSP representation of the molecules used to train the encoder network includes a set of coordinates for each of the atoms connected by the bond in the represented molecule and further identifies each of the atoms connected by the bond in the represented molecule.

4. The method of claim 2 , wherein the BSP representation of the molecules used to train the encoder network to identify, for each of at least some bonds in a respective molecule, a bond type.

5. The method of claim 2 , wherein a format of the structural representation identified in the result is different than the BSP representation.

6. The method of claim 1 , wherein training of the encoder network included learning to transform partial or complete molecular graph representations of molecules into positions within the embedding space, and wherein each molecular graph representation identifies angles and distances of bonds in the represented molecule.

7. The method of claim 1 , wherein the decoder network and the encoder network were trained by training a transformer model that uses self-attention, wherein the transformer model includes the decoder network and the encoder network.

8. The method of claim 1 , wherein the decoder network and the encoder network were trained by training a transformer model that includes an attention head.

9. The method of claim 1 , further comprising training a machine-learning model that includes the encoder network and the decoder network by:

accessing a set of supplemental training elements, wherein each of the set of training elements includes a representation of a structure of a corresponding given molecule;

masking, for each supplemental training element in the set of supplemental training elements, at least part of the representation to obscure at least part of the structure of the corresponding given molecule; and

training the machine-learning model to predict the obscured at least part of the structure.

10. A system comprising:

one or more data processors; and

a non-transitory computer readable storage medium containing instructions which, when executed on the one or more data processors, cause the one or more data processors to perform a method comprising:

accessing a multi-dimensional embedding space that supports relating embeddings of molecules to predicted values of a given property of the molecules;

identifying one or more points of interest within the multi-dimensional embedding space based on the predicted values, wherein each of the one or more points of interest:

includes a set of coordinate values within the multi-dimensional embedding space,

conveys spatial information of atoms or bonds in a molecule, and

is associated with a corresponding predicted value of the given property;

generating, for each of the one or more points of interest, a structural representation of the molecule by transforming the set of coordinate values included in a point of interest using a decoder network, wherein training of the decoder network included learning to transform positions within the embedding space to outputs representing molecular-structure characteristics, wherein the training of the decoder network was performed at least in part concurrently to training an encoder network to transform positions within the embedding space to predictions corresponding to values of the given property; and

outputting a result that identifies, for each of the one or more points of interest, the structural representation of the molecule corresponding to the point of interest.

11. The system of claim 10 , wherein training of the encoder network included learning to transform partial or complete bond string and position (BSP) representations of molecules into positions within the embedding space, and wherein each BSP representation identifies relative positions of atoms connected by a bond in the represented molecule.

12. The system of claim 11 , wherein each BSP representation of the molecules used to train the encoder network includes a set of coordinates for each of the atoms connected by the bond in the represented molecule and further identifies each of the atoms connected by the bond in the represented molecule.

13. The system of claim 11 , wherein the BSP representation of the molecules used to train the encoder network to identify, for each of at least some bonds in a respective molecule, a bond type.

14. The system of claim 11 , wherein a format of the structural representation identified in the result is different than the BSP representation.

15. The system of claim 10 , wherein training of the encoder network included learning to transform partial or complete molecular graph representations of molecules into positions within the embedding space, and wherein each molecular graph representation identifies angles and distances of bonds in the represented molecule.

16. The system of claim 10 , wherein the decoder network and the encoder network were trained by training a transformer model that uses self-attention, wherein the transformer model includes the decoder network and the encoder network.

17. The system of claim 10 , wherein the decoder network and the encoder network were trained by training a transformer model that includes an attention head.

18. The system of claim 10 , the method further comprising training a machine-learning model that includes the encoder network and the decoder network by:

accessing a set of supplemental training elements, wherein each of the set of training elements includes a representation of a structure of a corresponding given molecule;

masking, for each supplemental training element in the set of supplemental training elements, at least part of the representation to obscure at least part of the structure of the corresponding given molecule; and

training the machine-learning model to predict the obscured at least part of the structure.

19. A computer-program product tangibly embodied in a non-transitory machine-readable storage medium, including instructions configured to cause one or more data processors to perform a method comprising:

accessing a multi-dimensional embedding space that supports relating embeddings of molecules to predicted values of a given property of the molecules;

identifying one or more points of interest within the multi-dimensional embedding space based on the predicted values, wherein each of the one or more points of interest:

includes a set of coordinate values within the multi-dimensional embedding space,

conveys spatial information of atoms or bonds in a molecule, and

is associated with a corresponding predicted value of the given property;

generating, for each of the one or more points of interest, a structural representation of the molecule by transforming the set of coordinate values included in a point of interest using a decoder network, wherein training of the decoder network included learning to transform positions within the embedding space to outputs representing molecular-structure characteristics, wherein the training of the decoder network was performed at least in part concurrently to training an encoder network to transform positions within the embedding space to predictions corresponding to values of the given property; and

outputting a result that identifies, for each of the one or more points of interest, the structural representation of the molecule corresponding to the point of interest.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 23, 2023
From: GADHIYA, TUSHARKUMAR; SHAH, FALAK; VYAS, NISARG; YANG, JULIA; GHARAKHANYAN, VAHE; HOLIDAY, ALEXANDER
To: X DEVELOPMENT LLC
Reel/Frame 062453/0707 →
Continuity (5)
Provisional Application 63264643 · Nov 29, 2021
Provisional Application 63264640 · Nov 29, 2021
Provisional Application 63264641 · Nov 29, 2021
Provisional Application 63264642 · Nov 29, 2021
Related Publication 20230170059A1 · Jun 1, 2023
References Cited (36)
US 5437979A · Rampal et al. · 1995 [cited by applicant]
US 5772855A · Johnson et al. · 1998 [cited by applicant]
US 20020077757A1 · Bunin et al. · 2002 [cited by applicant]
US 20030087334A1 · Bunin et al. · 2003 [cited by applicant]
US 20220108765A1 · Teshima · 2022 [cited by examiner]
CA 1282538C · 1991 [cited by applicant]
CA 2495838C · 2004 [cited by applicant]
CA 2633549C · 2007 [cited by applicant]
CA 2989059C · 2016 [cited by applicant]
KR 20140014641A · 2014 [cited by applicant]
KR 20210123869A · 2021 [cited by applicant]
WO 03044219A1 · 2003 [cited by applicant]
WO 2020113136A1 · 2020 [cited by applicant]
Li et al. , “Training Method of Molecular Understanding Model, Device, Device and Medium”, CN 112786108 A, Date published: May 11, 2021 (Year: 2021). [cited by examiner]
Koge et al., “Embedding of Molecular Structure Using Molecular Hypergraph Variational Autoencoder with Metric Learning”, Molecular Informatics, vol. 40, No. 2, Feb. 2021, pp. 1-7. [cited by applicant]
Lim et al., “Molecular Generative Model Based on Conditional Variational Autoencoder for de Novo Molecular Design”, Journal of Cheminformatics, vol. 10, No. 31, Jul. 11, 2018, pp. 1-9. [cited by applicant]
International Application No. PCT/US2022/046893, International Search Report and Written Opinion, Mailed on Feb. 7, 2023, 10 pages. [cited by applicant]
Yan et al., “Re-balancing Variational Autoencoder Loss for Molecule Sequence Generation”, Available Online at: https://arxiv.org/pdf/1910.00698.pdf, Jan. 15, 2020, 8 pages. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2022/046893, dated Jun. 13, 2024. [cited by applicant]
RDKit, “Getting Started with the RDKit in Python”, Available Online at: https://www.rdkit.org/docs/GettingStartedInPython.html#working-with-3d-molecules, Accessed from Internet on Oct. 11, 2022, 50 pages. [cited by applicant]
Wikipedia, “Ionic Liquid”, Available Online at: https://en.wikipedia.org/wiki/Ionic_liquid, Accessed from Internet on Oct. 11, 2022, 12 pages. [cited by applicant]
Wikipedia, “Simplified Molecular-Input Line-Entry System”, Available Online at: https://en.wikipedia.org/wiki/Simplified_molecular-input_line-entry_system, Accessed from Internet on Oct. 11, 2022, 11 pages. [cited by applicant]
Cereto-Massague et al., “Molecular Fingerprint Similarity Search in Virtual Screening”, Methods, vol. 71, Jan. 2015, 30 pages. [cited by applicant]
Chithrananda et al., “ChemBERTa: Large-Scale Self-Supervised Pretraining for Molecular Property Prediction”, Available Online at: https://arxiv.org/pdf/2010.09885.pdf, Oct. 23, 2020, 7 pages. [cited by applicant]
Coley et al., “Convolutional Embedding of Attributed Molecular Graphs for Physical Property Prediction”, Journal of Chemical Information and Modeling, vol. 57, No. 8, Jul. 11, 2017, 29 pages. [cited by applicant]
Devlin et al., “BERT: Pre-Training of Deep Bidirectional Transformers for Language Understanding”, Available Online at: https://arxiv.org/pdf/1810.04805.pdf, May 24, 2019, 16 pages. [cited by applicant]
Graziano , “Fingerprints of Molecular Reactivity”, Nature Reviews Chemistry, vol. 4, May 2020, p. 227. [cited by applicant]
Krenn et al., “Self-Referencing Embedded Strings (SELFIES): A 100% Robust Molecular String Representation”, Available Online at: https://arxiv.org/pdf/1905.13741.pdf, Mar. 5, 2020, 9 pages. [cited by applicant]
Muegge et al., “An Overview of Molecular Fingerprint Similarity Search in Virtual Screening”, Expert Opinion on Drug Discovery, vol. 11, No. 2, Feb. 2016, pp. 137-148. [cited by applicant]
Schwaller et al., “Molecular Transformer for Chemical Reaction Prediction and Uncertainty Estimation”, Theoretical and Computational Chemistry, Nov. 2018, 11 pages. [cited by applicant]
Wu et al., “MoleculeNet: A Benchmark for Molecular Machine Learning”, Available Online at: https://arxiv.org/pdf/1703.00564.pdf, Oct. 26, 2018, 65 pages. [cited by applicant]
Vasawni et al., “Attention is all you need,” 31st Conference on Neural Information Processing Systems (2017) 15 pages. [cited by applicant]
Office Action for Japanese Patent Application No. 2024-529333 dated Apr. 24, 2024. [cited by applicant]
Office Action for Japanese Patent Application No. 2024-529334, dated Apr. 25, 2025. [cited by applicant]
Office Action for Japanese Patent Application No. 2024-529335, dated Apr. 28, 2025. [cited by applicant]
Office Action for U.S. Appl. No. 17/967,723, dated Jun. 11, 2025. [cited by applicant]