IP Library › Granted Patent US 12,282,851
Granted Patent B2
US 12,282,851 · App. 17/842,247 · Granted Apr 22, 2025

Mutual information adversarial autoencoder

Inventors: Aleksandr Aliper (Moscow, RU); Aleksandrs Zavoronkovs (Rockville, MD); Alexander Zhebrak (Moscow, RU); Artur Kadurin (Taipei, TW); Daniil Polykovskiy (Moscow, RU); Rim Shayakhmetov (Moscow, RU)
Assignee: INSILICO MEDICINE IP LIMITED
G06N3/08G06N3/088G06N7/01G10L19/24
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,282,851
App. No.
17/842,247
Granted
Apr 22, 2025
Kind
B2
Abstract

A method for generating an object includes: providing a dataset having object data and condition data; processing the object data to obtain latent object data and latent object-condition data; processing the condition data to obtain latent condition data and latent condition-object data; processing the latent object data and the latent object-condition data to obtain generated object data; processing the latent condition data and latent condition-object data to obtain generated condition data; comparing the latent object-condition data to the latent condition-object data to determine a difference; processing the latent object data and latent condition data and one of the latent object-condition data or latent condition-object data to obtain a discriminator value; and selecting a selected object based on the generated object data.

Claims (109)

1. A method, comprising:

providing a dataset having object data for an object and condition data for a condition;

providing a computing system having a deep neural network configured with a mutual information adversarial autoencoder;

processing the object data and condition data through the mutual information adversarial autoencoder of the deep neural network to obtain output data;

selecting a selected object based on the output data from the mutual information adversarial autoencoder of the deep neural network that satisfies a given condition;

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

calculating an object loss between generated object data of the output with input object data, for the selected object;

comparing the object loss to a threshold;

providing the selected object with the object loss less than the threshold;

generating the selected object with a decoder of the deep neural network to obtain a generated object that satisfies the given condition; and

providing the generated object that satisfies the given condition.

2. The method of claim 1 , further comprising:

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

calculating a condition loss between generated condition data of the output with input condition data;

comparing the condition loss to a threshold; and

selecting a generated condition that is less than threshold.

3. The method of claim 1 , further comprising:

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

selecting a generated object that is less than a threshold; and

selecting a generated condition that is less than a threshold.

4. The method of claim 1 , further comprising:

obtaining output from the mutual information adversarial autoencoder with the object data and condition data;

obtaining generated object-condition data from latent object data and latent object-condition data of the output;

comparing generated object-condition data with object-condition data; and

selecting generated object-condition data that is less than a threshold.

5. The method of claim 4 , further comprising:

obtaining output from the mutual information adversarial autoencoder with the object data and condition data;

obtaining generated condition-object data from latent condition data and latent condition-object data of the output;

comparing generated condition-object data with condition-object data; and

selecting generated condition-object data that is less than a threshold.

6. The method of claim 1 , further comprising:

processing the object data of the dataset with an object encoder;

processing the condition data of the dataset with a condition encoder;

processing latent object data from the object encoder to obtain generated object data with an object decoder;

processing the latent condition data from the condition encoder to obtain generated condition data with a condition decoder;

comparing the generated object data with the object data;

comparing the generated condition data with the condition data; and

determine a loss from generated data compared to input data.

7. The method of claim 6 , further comprising:

obtaining multidimensional vectors from generated data;

determining a generated data discriminator value;

sampling a reference distribution from a predefined multidimensional distribution;

determining a reference data discriminator value; and

determining a difference between the generated data discriminator value and the reference data discriminator value.

8. The method of claim 7 , further comprising performing an optimization protocol on a loss function based on the difference between the generated data discriminator value and the reference data discriminator value.

9. The method of claim 1 , further comprising:

processing the object data of the dataset with an object encoder;

processing the condition data of the dataset with a condition encoder;

processing latent object data from the object encoder to obtain generated object data with an object decoder;

processing the latent condition data from the condition encoder to obtain generated condition data with a condition decoder;

processing the latent object data and latent condition data to obtain a discriminator value;

selecting a selected object based on outcome in view of the discriminator value for a given condition; and

providing the selected object in a report with a recommendation for validation of a physical form of the object.

10. The method of claim 1 , further comprising obtaining a physical form of the selected object.

11. The method of claim 10 , further comprising validating the physical form of the object.

12. A computer program product comprising:

a non-transient, tangible memory device having computer-executable instructions that when executed by a processor, cause performance of a method comprising:

providing a dataset having object data for an object and condition data for a condition;

providing a computing system having a deep neural network configured with a mutual information adversarial autoencoder;

processing the object data and condition data through the mutual information adversarial autoencoder to obtain output data;

selecting a selected object based on the output data from the mutual information adversarial autoencoder of the deep neural network that satisfies a given condition;

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

calculating an object loss between generated object data of the output with input object data, for the selected object;

comparing the object loss to a threshold;

providing the selected object with the object loss less than the threshold;

generating the selected object with a decoder of the deep neural network; and

providing the generated object that satisfies the given condition.

13. The computer program product of claim 12 , the method further comprising:

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

calculating a condition loss between generated condition data of the output with input condition data;

comparing the condition loss to a threshold; and

selecting a generated condition that is less than threshold.

14. The computer program product of claim 12 , the method further comprising:

comparing the output from the mutual information adversarial autoencoder with the object data and condition data;

selecting a generated object that is less than a threshold; and

selecting a generated condition that is less than a threshold.

15. The computer program product of claim 12 , the method further comprising:

obtaining output from the mutual information adversarial autoencoder with the object data and condition data;

obtaining generated object-condition data from latent object data and latent object-condition data of the output;

comparing generated object-condition data with object-condition data; and

selecting generated object-condition data that is less than a threshold.

16. The computer program product of claim 14 , the method further comprising:

obtaining output from the mutual information adversarial autoencoder with the object data and condition data;

obtaining generated condition-object data from latent condition data and latent condition-object data of the output;

comparing generated condition-object data with condition-object data; and

selecting generated condition-object data that is less than a threshold.

17. The computer program product of claim 12 , the method further comprising:

processing the object data of the dataset with an object encoder;

processing the condition data of the dataset with a condition encoder;

processing latent object data from the object encoder to obtain generated object data with an object decoder;

processing the latent condition data from the condition encoder to obtain generated condition data with a condition decoder;

comparing the generated object data with the object data;

comparing the generated condition data with the condition data; and

determine a loss from generated data compared to input data.

18. The computer program product of claim 17 , the method further comprising:

obtaining multidimensional vectors from generated data;

determining a generated data discriminator value;

sampling a reference distribution from a predefined multidimensional distribution;

determining a reference data discriminator value; and

determining a difference between the generated data discriminator value and the reference data discriminator value.

19. The computer program product of claim 18 , the method further comprising performing an optimization protocol on a loss function based on the difference between the generated data discriminator value and the reference data discriminator value.

20. The computer program product of claim 12 , the method further comprising:

processing the object data of the dataset with an object encoder;

processing the condition data of the dataset with a condition encoder;

processing latent object data from the object encoder to obtain generated object data with an object decoder;

processing the latent condition data from the condition encoder to obtain generated condition data with a condition decoder;

processing the latent object data and latent condition data to obtain a discriminator value;

selecting a selected object based on outcome in view of the discriminator value for a given condition; and

providing the selected object in a report with a recommendation for validation of a physical form of the object.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 16, 2022
From: ALIPER, ALEKSANDR; ZAVORONKOVS, ALEKSANDRS; ZHEBRAK, ALEXANDER; KADURIN, ARTUR; POLYKOVSKIY, DANIIL; SHAYAKHMETOV, RIM
To: INSILICO MEDICINE, INC.
Reel/Frame 060232/0058 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 16, 2022
From: INSILICO MEDICINE, INC.
To: INSILICO MEDICINE IP LIMITED
Reel/Frame 060232/0095 →
Continuity (2)
Continuation 16015990 · Jun 22, 2018
Related Publication 20220391709A1 · Dec 8, 2022
References Cited (58)
US 20150100530A1 · Mnih et al. · 2015 [cited by applicant]
US 20150213361A1 · Gamon et al. · 2015 [cited by applicant]
US 20170228638A1 · Danihelka et al. · 2017 [cited by applicant]
US 20170230675A1 · Wierstra et al. · 2017 [cited by applicant]
US 20180028294A1 · Azernikov et al. · 2018 [cited by applicant]
US 20180039946A1 · Bolte et al. · 2018 [cited by applicant]
US 20180082172A1 · Patel et al. · 2018 [cited by applicant]
US 20180174070A1 · Hoffman et al. · 2018 [cited by applicant]
US 20190147333A1 · Kallur Palli Kumar · 2019 [cited by examiner]
US 20190244609A1 · Olabiyi · 2019 [cited by examiner]
US 20200210842A1 · Baker · 2020 [cited by examiner]
US 20200410361A1 · Okajima · 2020 [cited by examiner]
EP 3188111A1 · 2017 [cited by applicant]
WO 20180013247A2 · 2018 [cited by applicant]
Kadurin A. et al.; “The cornucopia of meaningful leads: Applying deep adversarial autoencoders for new molecule development in oncology”; Oncotarget; Dec. 21, 2016; pp. 10883-10890, XP055471303; DOI: 10.18632/oncotarget… [cited by applicant]
Aliper A. et al.; “Deep Learning Application for Predicting Pharmacological Properties of Druga and Drug Repurposing Using Transcriptomic Data”; Molecular Pharmaceutics, vol. 13, No. 7; Jul. 5, 2016; pp. 2524-2530; XP05… [cited by applicant]
Artemov A.V. et al.; “Integrated Deep Learned Transcriptomic and Structure-based Predictor of Clinical Trials Outcomes”; bioRxiv; Dec. 29, 2016; XP055893302; DOI: 10.1101/095653; 21 pages. [cited by applicant]
Kadurin A. et al.; “druGAN: An Advanced Generative Adversarial Autoencoder Model for de Novo Generation of New Molecules with Desired Molecular Properties in Silico”; Molecular Pharmeceutics, vol. 14, No. 9; Sep. 5, 201… [cited by applicant]
Jing Y. et al.; “Deep Learning for Drug Design: an Artificial Intelligence Paradigm for Drug Discovery in the Big Data Era”; The AAPS Journal, Springer International Publishing, Cham, vol. 20, No. 3; Mar. 30, 2018; pp. … [cited by applicant]
Chen H. et al.; “The rise of deep learning in drug discovery”; Drug Discovery Today, vol. 23, No. 6; Jan. 31, 2018; pp. 1241-1250; XP055664738; Amsterdam, NL ISSN: 1359-6446; DOI: 10.1016/j.drudis.2018.01.039. [cited by applicant]
European Patent Office; Extended European Search Report dated Feb. 28, 2022 issued in Application No. 19822510.4; 10 pages. [cited by applicant]
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. Generative adversarial nets. In Z. Ghahramani, M. Welling, C. Cortes, N. D. Lawrence, and … [cited by applicant]
Antonia Creswell, Tom White, Vincent Dumoulin, Kai Arulkumaran, Biswa Sengupta, and Anil A. Bharath. Generative adversarial networks: An overview. CoRR, abs/1710.07035, 2017. [cited by applicant]
Mehdi Mirza and Simon Osindero. Conditional generative adversarial nets. CoRR, abs/1411.1784, 2014. [cited by applicant]
Alireza Makhzani, Jonathon Shlens, Navdeep Jaitly, and Ian J. Goodfellow. Adversarial autoencoders. CoRR, abs/1511.05644, 2015. [cited by applicant]
Guim Perarnau, Joost van de Weijer, Bogdan Raducanu, and Jose M. Alvarez. Invertible conditional gans for image editing. CoRR, abs/1611.06355, 2016. [cited by applicant]
Navaneeth Bodla, Gang Hua, and Rama Chellappa. Semi-supervised fusedgan for conditional image generation. CORR, abs/1801.05551, 2018. [cited by applicant]
Zhifei Zhang, Yang Song, and Hairong Qi. Age progression/regression by conditional adversarial autoencoder. CoRR, abs/1702.08423, 2017. [cited by applicant]
Grigory Antipov, Moez Baccouche, and Jean-Luc Dugelay. Face aging with conditional generative adversarial networks. CoRR, abs/1702.01983, 2017. [cited by applicant]
Guillaume Lample, Neil Zeghidour, Nicolas Usunier, Antoine Bordes, Ludovic Denoyer, and Marc Aurelio Ranzato. Fader networks:manipulating images by sliding attributes. In I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, … [cited by applicant]
Murat Kocaoglu, Christopher Snyder, Alexandros G. Dimakis, and Sriram Vishwanath. Causalgan: Learning causal Implicit generative models with adversarial training. CoRR, abs/1709.02023, 2017. [cited by applicant]
Diederik P. Kingma and Max Welling. Auto-encoding variational bayes. CoRR, abs/1312.6114, 2013. [cited by applicant]
Kihyuk Sohn, Honglak Lee, and Xinchen Yan. Learning structured output representation using deep conditional generative models. In C. Cortes, N. D. Lawrence, D. D. Lee, M. Sugiyama, and R. Garnett, editors, Advances in N… [cited by applicant]
Zhiting Hu, Zichao Yang, Ruslan Salakhutdinov, and Eric P. Xing. On unifying deep generative models. CoRR, abs/1706.00550, 2017. URL http://arxiv.org/abs/1706.00550. [cited by applicant]
Jianmin Bao, Dong Chen, Fang Wen, Houqiang Li, and Gang Hua. CVAE-GAN: fine-grained image generation through asymmetric training. CoRR, abs/1703.10155, 2017. [cited by applicant]
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros. Unpaired image-to-image translation using cycle-consistent adversarial networks. arXiv preprint arXiv:1703.10593, 2017. [cited by applicant]
Jiquan Ngiam, Aditya Khosla, Mingyu Kim, Juhan Nam, Honglak Lee, and Andrew Y. Ng. Multimodal deep learning. In Proceedings of the 28th International Conference on International Conference on Machine Learning, ICML'11, … [cited by applicant]
Weiran Wang, Raman Arora, Karen Livescu, and Jeff A. Bilmes. On deep multi-view representation learning: Objectives and optimization. CoRR, abs/1602.01024, 2016. URL http://arxiv.org/abs/1602.01024. [cited by applicant]
WeiranWang, Honglak Lee, and Karen Livescu. Deep variational canonical correlation analysis. CoRR, abs/1610.03454, 2016. URL http://arxiv.org/abs/1610.03454. [cited by applicant]
Jimei Yang, Scott E. Reed, Ming-Hsuan Yang, and Honglak Lee. Weakly-supervised disentangling with recurrent transformations for 3d view synthesis. CoRR, abs/1601.00706, 2016. [cited by applicant]
Attila Szabó, Qiyang Hu, Tiziano Portenier, Matthias Zwicker, and Paolo Favaro. Challenges in disentangling independent factors of variation. CoRR, abs/1711.02245, 2017. [cited by applicant]
Antonia Creswell, Anil A. Bharath, and Biswa Sengupta. Conditional autoencoders with adversarial information factorization. CoRR, abs/1711.05175, 2017. [cited by applicant]
Shengjia Zhao, Jiaming Song, and Stefano Ermon. Infovae: Information maximizing variational autoencoders. CoRR, abs/1706.02262, 2017. [cited by applicant]
Yang Li, Quan Pan, Suhang Wang, Haiyun Peng, Tao Yang, and Erik Cambria. Disentangled variational auto-encoder for semi-supervised learning. CoRR, abs/1709.05047, 2017. [cited by applicant]
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel. Infogan: Interpretable representation learning by information maximizing generative adversarial nets. CoRR, abs/1606.03657, 2016 . . .… [cited by applicant]
Qiyang Hu, Attila Szabó, Tiziano Portenier, Matthias Zwicker, and Paolo Favaro. Disentangling factors of variation by mixing them. CoRR, abs/1711.07410, 2017. [cited by applicant]
Marwin H. S. Segler, Thierry Kogej, Christian Tyrchan, and Mark P.Waller. Generating focused molecule libraries for drug discovery with recurrent neural networks. ACS Central Science, 4 (1):120-131, Dec. 2017. [cited by applicant]
Artur Kadurin, Sergey Nikolenko, Kuzma Khrabrov, Alex Aliper, and Alex Zhavoronkov. druGAN: An advanced generative adversarial autoencoder model for de novo generation of new molecules with desired molecular properties … [cited by applicant]
Rafael Gómez-Bombarelli, David K. Duvenaud, JoséMiguel Hernández-Lobato, Jorge Aguilera-Iparraguirre, Timothy D. Hirzel, Ryan P. Adams, and Alán Aspuru-Guzik. Automatic chemical design using a data-driven continuous rep… [cited by applicant]
Mihaela Rosca, Balaji Lakshminarayanan, David Warde-Farley, and Shakir Mohamed. Variational approaches for auto-encoding generative adversarial networks. CoRR, abs/1706.04987, 2017. [cited by applicant]
Florian Schroff, Dmitry Kalenichenko, and James Philbin. Facenet: A unified embedding for face recognition and clustering. CoRR, abs/1503.03832, 2015. [cited by applicant]
Weiran Wang, Raman Arora, Karen Livescu, and Jeff Bilmes. On deep multi-view representation learning. In Proceedings of the 32nd International Conference on International Conference on Machine Learning—vol. 37, ICML'15,… [cited by applicant]
Mohamed Ishmael Diwan Belghazi, Sai Rajeswar, Aristide Baratin, Devon R Hjelm, and Aaron Courville. Mine: Mutual Information neural estimation. arXiv e-prints, 1801.04062, Jan. 2018. URL https://arxiv.org/abs/1801.04062. [cited by applicant]
Qiaonan Duan, Corey Flynn, Mario Niepel, Marc Hafner, Jeremy L. Muhlich, Nicolas F. Fernandez, Andrew D. Rouillard, Christopher M. Tan, Edward Y. Chen, Todd R. Golub, Peter K. Sorger, Aravind Subramanian, and Avi Ma'aya… [cited by applicant]
Kumar Sricharan, Raja Bala, Matthew Shreve, Hui Ding, Kumar Saketh, and Jin Sun. Semi-supervised conditional gans. CoRR, abs/1708.05789, 2017. [cited by applicant]
David S. Wishart, Craig Knox, Anchi Guo, Dean Cheng, Savita Shrivastava, Dan Tzur, Bijaya Gautam, and Murtaza Hassanali. Drugbank: a knowledgebase for drugs, drug actions and drug targets. Nucleic Acids Research, 36 (Da… [cited by applicant]
United States Patent and Trademark Office; International Search Report and Written Opinion dated Aug. 19, 2019, issued in Int'l Application. No. PCT/US2019/033960, 7 pages. [cited by applicant]
Putin, Evgeny et al., “Adversarial Threshold Neural Computer for Molecular De Novo Design”, Molecular Pharmaceutics, Dec. 15, 2017, Manuscript ID: mp-2017-01137w. [cited by applicant]