IP Library Granted Patent US 12,361,542
Granted Patent B2
US 12,361,542 · App. 17/686,131 · Granted Jul 15, 2025

Systems and methods for deep orthogonal fusion for multimodal prognostic biomarker discovery

Inventors: Nathaniel Braman (Cleveland, OH); Jagadish Venkataraman (Menlo Park, CA); Emery T. Goossens (Midvale, UT)
Assignee: TEMPUS AI, INC.
G06T7/0012G06N3/045G16H30/20G06T2207/10116G06T2207/20084G06T2207/30024G06T2207/30096
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,361,542
App. No.
17/686,131
Granted
Jul 15, 2025
Kind
B2
Abstract

A system and method are provided for identifying a multimodal biomarker of a prognostic prediction, using a deep learning framework trained to analyze different modality data, including radiomic image data, pathology image data, and molecular image data to obtain unimodal embedding predictions from those modality data and generate multimodal embedding predictions, through application of a loss minimization and attention-based fusion processes.

Claims (50)

1. A computer-implemented method for identifying a multimodal biomarker of a prognostic prediction for a tumor sample, the method comprising:

a) obtaining, using one or more processors, a radiomic dataset for a tumor sample, the radiomic dataset comprising characteristics of a lesion and/or associated tissue and subsequent analysis and classification of the characteristic corresponding to the tumor sample;

b) obtaining, using the one or more processors, a pathology dataset for the tumor sample, the pathology dataset comprising characteristics of cells, cell types, cell shapes, and/or cell areas corresponding to the tumor sample;

c) obtaining, using the one or more processors, a molecular dataset for the tumor sample, the molecular dataset comprising data derived from sequencing data and/or subsequent analysis and classification of the sequencing data corresponding to the tumor sample;

d) providing, using the more or more processors, the radiomic dataset, the pathology dataset, and the molecular modality dataset to a trained deep learning framework and

e) generating, using the trained deep learning framework, a radiomic embedding prediction, a pathology embedding prediction, and a molecular embedding prediction; and

f) applying, using the trained deep learning framework, the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction to a loss minimization to reduce unimodal embeddings for each prior to an embedding fusion to generate a multimodal embedding prediction as the multimodal biomarker corresponding to the prognostic prediction for the tumor sample,

wherein the trained deep learning framework comprises a unimodal embeddings layer for generating the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction and a fully-connected output layer for generating the multimodal embedding prediction, and

wherein the trained deep learning framework is configured to apply, subsequent to the loss minimization at the unimodal embeddings layer, a second loss function at the fully-connected output layer to generate the multimodal embedding prediction as the multimodal biomarker.

2. The method of claim 1 , wherein the radiomic dataset is selected from the group consisting of a magnetic resonance imaging (MRI) image dataset, a computed tomography (CT) image dataset, a fluorescence image dataset, and an x-ray image dataset.

3. The method of claim 1 , wherein the pathology dataset is selected from the group consisting of a hematoxylin and eosin (H&E) stained slide image dataset, an immunohistochemistry (IHC) stained slide image dataset, and a fluorescence in situ hybridization (FISH) image dataset.

4. The method of claim 1 , wherein the molecular dataset is selected from the group consisting of gene sequencing data, RNA data, DNA data, methylation data, and proteomic data.

5. The method of claim 1 , wherein generating, using the trained deep learning framework, the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction comprises:

i) feeding, within the trained deep learning framework, the radiomic dataset to a radiomic neural network trained to generate the radiomic embedding prediction;

ii) feeding, within the trained deep learning framework, the pathology dataset to a pathology neural network trained to generate the pathology embedding prediction; and

iii) feeding, within the trained deep learning framework, the molecular dataset to a molecular neural network trained to generate the molecular embedding prediction.

6. The method of claim 5 , wherein the radiomic neural network is a convolutional neural network.

7. The method of claim 6 , wherein the convolutional neural network has been trained using multiparametric MRI training images and labeled image features.

8. The method of claim 7 , wherein the convolutional neural network comprises a T1 trained convolutional neural network branch, a T2 trained convolutional neural network branch, and a labeled image features branch.

9. The method of claim 5 , wherein the pathology neural network is a convolutional neural network.

10. The method of claim 5 , wherein the molecular neural network is a self-normalizing neural network.

11. The method of claim 1 , wherein the trained deep learning framework is configured to apply, as the loss minimization, a unimodal loss minimization for the unimodal embeddings layer.

12. The method of claim 1 , wherein the trained deep learning framework is configured to apply, as the loss minimization, a multimodal orthogonalization loss across each of the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction.

13. The method of claim 1 , further comprising:

e) performing, using the trained deep learning framework, a multimodal fusion on the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction;

f) generating a multidimensional fusion matrix containing a plurality of multidimensional embeddings, containing at least one or more bi-modal embeddings or one or more tri-modal embeddings; and

g) determining the multimodal embedding prediction from a comparison of the plurality of multidimensional embeddings.

14. The method of claim 1 , further comprising

h) generating the multimodal embedding prediction is a prediction of overall survival rate corresponding to the tumor sample.

15. The method of claim 1 , further comprising

i) generating the multimodal embedding prediction by:

A) generating a plurality of multimodal embeddings each having a prediction score; and

B) identifying a maximum prediction score as the multimodal embedding prediction.

16. The method of claim 1 , further comprising,

j) using the loss minimization,

k) applying a multimodal orthogonalization across the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction.

17. The method of claim 16 , further comprising

l) applying an attention weighting to the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction and, in response,

m) performing the embedding fusion to generate the multimodal embedding prediction.

18. The method of claim 1 , further comprising:

n) receiving additional features from the radiomic dataset, the pathology dataset, and/or the molecular dataset, the additional features not being used by the deep learning framework in generating the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction; and

o) using, in the trained deep learning framework, the additional features to generate the multimodal embedding prediction.

19. A system for identifying a multimodal biomarker of a prognostic prediction for a tumor sample, the system comprising:

I) one or more processors; and

II) a trained deep learning framework application including computing instructions configured to be executed by the one or more processors to;

A) receive a radiomic image modality dataset for a tumor sample, a pathology image modality dataset for the tumor sample, and a molecular modality dataset for the tumor sample, wherein the radiomic dataset comprises characteristics of a lesion and/or associated tissue and subsequent analysis and classification of the characteristic corresponding to the tumor sample, the pathology dataset comprises characteristics of cells, cell types, cell shapes, and/or cell areas corresponding to the tumor sample, and the molecular dataset comprises data derived from sequencing data and/or subsequent analysis and classification of the sequencing data corresponding to the tumor sample;

B) generate a radiomic embedding prediction, a pathology embedding prediction, and a molecular embedding prediction; and

C) from the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction, applying a loss minimization to reduce unimodal embeddings for each prior to an embedding fusion, generate a multimodal embedding prediction as the multimodal biomarker corresponding to the prognostic prediction for the tumor sample,

wherein the trained deep learning framework comprises a unimodal embeddings layer for generating the radiomic embedding prediction, the pathology embedding prediction, and the molecular embedding prediction and a fully-connected output layer for generating the multimodal embedding prediction, and

wherein the trained deep learning framework is configured to apply, subsequent to the loss minimization at the unimodal embeddings layer, a second loss function at the fully-connected output layer to generate the multimodal embedding prediction as the multimodal biomarker.

Assignments (4)
RELEASE OF SECURITY INTEREST Recorded May 14, 2026
From: ARES CAPITAL CORPORATION, AS COLLATERAL AGENT
To: TEMPUS AI, INC. (F/K/A TEMPUS LABS, INC.)
Reel/Frame 074653/0918 →
CHANGE OF NAME Recorded Feb 9, 2024
From: TEMPUS LABS, INC.
To: TEMPUS AI, INC.
Reel/Frame 066544/0110 →
SECURITY INTEREST Recorded Oct 13, 2023
From: TEMPUS LABS, INC.
To: ARES CAPITAL CORPORATION, AS COLLATERAL AGENT
Reel/Frame 065209/0711 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 7, 2022
From: BRAMAN, NATHANIEL; VENKATARAMAN, JAGADISH; GOOSSENS, EMERY T.
To: TEMPUS LABS, INC.
Reel/Frame 061344/0724 →
Continuity (2)
Provisional Application 63155941 · Mar 3, 2021
Related Publication 20220292674A1 · Sep 15, 2022
References Cited (38)
US 10395772B1 · Lucas et al. · 2019 [cited by applicant]
US 10902952B2 · Lucas et al. · 2021 [cited by applicant]
US 10957041B2 · Yip et al. · 2021 [cited by applicant]
US 10975445B2 · Venkat et al. · 2021 [cited by applicant]
US 11043283B1 · Bell et al. · 2021 [cited by applicant]
US 11043304B2 · Lozac'hmeur et al. · 2021 [cited by applicant]
US 11081210B2 · Perera · 2021 [cited by applicant]
US 11145416B1 · Hafez et al. · 2021 [cited by applicant]
US 11922629B2 · Guida · 2024 [cited by examiner]
US 20200075169A1 · Lau et al. · 2020 [cited by applicant]
US 20200098448A1 · Shah et al. · 2020 [cited by applicant]
US 20200118644A1 · Khan et al. · 2020 [cited by applicant]
US 20200135303A1 · Barber · 2020 [cited by applicant]
US 20200210852A1 · Igartua et al. · 2020 [cited by applicant]
US 20200211716A1 · Lefkofsky et al. · 2020 [cited by applicant]
US 20200258601A1 · Lau · 2020 [cited by applicant]
US 20200335102A1 · Lefkofsky et al. · 2020 [cited by applicant]
US 20200365232A1 · Jaros et al. · 2020 [cited by applicant]
US 20200381087A1 · Ozeran et al. · 2020 [cited by applicant]
US 20200395097A1 · Chang et al. · 2020 [cited by applicant]
US 20210057042A1 · Beaubier et al. · 2021 [cited by applicant]
US 20210057071A1 · Barber et al. · 2021 [cited by applicant]
US 20210090694A1 · Colley et al. · 2021 [cited by applicant]
US 20210098078A1 · Lozac'hmeur et al. · 2021 [cited by applicant]
US 20210115511A1 · Blidner · 2021 [cited by applicant]
US 20210118559A1 · Lefkofsky · 2021 [cited by applicant]
US 20210151192A1 · Lucas et al. · 2021 [cited by applicant]
US 20210155989A1 · Salahudeen et al. · 2021 [cited by applicant]
US 20210172931A1 · Larsen et al. · 2021 [cited by applicant]
US 20220367053A1 · Mahmood · 2022 [cited by examiner]
WO 2020142563A1 · 2020 [cited by applicant]
WO 2021081253A1 · 2021 [cited by applicant]
WO 2021113821A1 · 2021 [cited by applicant]
WO 2021168143A1 · 2021 [cited by applicant]
Lin et al. “Orthogonalization-guided feature fusion network for multimodal 2D+ 3D facial expression recognition.” IEEE Transactions on Multimedia 23 (2020): 1581-1591 (Year: 2020). [cited by examiner]
Singanamalli et al. “Supervised multi-view canonical correlation analysis: Fused multimodal prediction of disease diagnosis and prognosis”; Medical Imaging 2014: Biomedical Applications in Molecular, Structural, and Fun… [cited by examiner]
Lin et al. “Orthogonalization-guided feature fusion network for multimodal 2D+ 3D facial expression recognition.” IEEE Transactions on Multimedia 23: 1581-1591 (Year: 2020). [cited by examiner]
Braman et al., Deep orthogonal fusion: multimodal prognostic biomarker discovery integrating radiology, pathology, genomic and clinical data, 16th European Conference—Computer Vision—ECCV 2020, pp. 667-677 IN: de Bruijn… [cited by applicant]