IP Library › Granted Patent US 12,674,208
Granted Patent B2
US 12,674,208 · App. 19/057,757 · Granted Jul 7, 2026

Methods and systems for determining proportions of distinct cell subsets

Inventors: Aaron M. Newman (Palo Alto, CA); Arash Ash Alizadeh (San Mateo, CA)
Assignee: The Board of Trustees of the Leland Stanford Junior University
C12Q1/6886C12Q1/6809C12Q1/6881G01N33/5005G16B25/00G16B25/10G16B40/10G16C20/20C12Q2600/106C12Q2600/158
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,674,208
App. No.
19/057,757
Filed
Feb 19, 2025
Granted
Jul 7, 2026
Kind
B2
Art Unit
1686
USPC
506/9
Abstract

Methods of deconvolving a feature profile of a physical system are provided herein. The present method may include: optimizing a regression between a) a feature profile of a first plurality of distinct components and b) a reference matrix of feature signatures for a second plurality of distinct components, wherein the feature profile is modeled as a linear combination of the reference matrix, and wherein the optimizing includes solving a set of regression coefficients of the regression, wherein the solution minimizes 1) a linear loss function and 2) an L 2 -norm penalty function; and estimating the fractional representation of one or more distinct components among the second plurality of distinct components present in the sample based on the set of regression coefficients. Systems and computer readable media for performing the subject methods are also provided.

Claims (53)

1 . A method for treating a subject for cancer, comprising:

(a) assaying a biological sample comprising peripheral blood mononuclear cells from a subject having cancer, thereby generating a feature profile m, wherein the biological sample comprises a first plurality of distinct leukocyte cell subsets, wherein the cancer comprises brain cancer, blood cancer, or non-brain solid tumor cancer,

wherein the feature profile m comprises combinations of features associated with the first plurality of distinct leukocyte cell subsets,

wherein the feature profile m comprises a gene expression profile of leukocyte cells in the biological sample, wherein the gene expression profile represents a ribonucleic acid (RNA) transcriptome of the leukocyte cells in the biological sample:

(b) optimizing, by a computer processor, a regression between the feature profile m and a reference matrix B of feature signatures for a second plurality of distinct leukocyte cell subsets in the biological sample, wherein the feature profile m is modeled as a linear combination of the reference matrix B,

wherein the reference matrix B comprises an LM22 signature matrix of relative gene expression values across 22 leukocyte subsets, wherein the LM22 signature matrix comprises ABCB4, ABCB9, ACAP1, ACHE, ACP5, ADAM28, ADAMDEC1, ADAMTS3, ADRB2, AIF1, AIM2, ALOX15, ALOX5, AMPD 1 , ANGPT4, ANKRD55, APOBEC3A, APOBEC3G, APOL3, APOL6, AQP 9 , ARHGAP22, ARRB1, ASGR1, ASGR2, ATHL1, ATP8B4, ATXN80S, AZU1, BACH2, BANK1, BARX2, BCL11B, BCL2A1, BCL7A, BEND5, BFSP1, BHLHE41, BIRC3, BLK, BMP2K, BPI, BRAF, BRSK2, BST1, BTNL8, C11orf80, Clorf54, C3AR1, C5AR1, C5AR2, CA8, CAMP, CASP5, CCDC102B, CCL1, CCL13, CCL14, CCL17, CCL18, CCL19, CCL20, CCL22, CCL23, CCL4, CCL5, CCL7, CCL8, CCND2, CCR10, CCR2, CCR3, CCR5, CCR6, CCR7, CD160, CD180, CD19, CD1A, CD1B, CD1C, CD1D, CD1E, CD2, CD209, CD22, CD244, CD247, CD27, CD28, CD300A, CD33, CD37, CD38, CD3D, CD3E, CD3G, CD4, CD40, CD40LG, CD5, CD6, CD68, CD69, CD7, CD70, CD72, CD79A, CD79B, CD80, CD86, CD8A, CD8B, CD96, CDA, CDC25A, CDH12, CDHR1, CDK6, CEACAM3, CEACAM8, CEMP1, CFP, CHI3L1, CHI3L2, CHST15, CHST7, CLC, CLCA3P, CLEC10A, CLEC2D, CLEC4A, CLEC7A, CLIC2, CMA1, COL8A2, COLO, CPA3, CR2, CREB5, CRISP3, CRTAM, CRYBB1, CSF1, CSF2, CSF3R, CST7, CTLA4, CTSG, CTSW, CXCL10, CXCL11, CXCL13, CXCL3, CXCL5, CXCL9, CXCR1, CXCR2, CXCR5, CXCR6, CXorf57, CYP27A1, CYP27B1, DACH1, DAPK2, DCSTAMP, DEFA4, DENNDSB, DEPDC5, DGKA, DHRS11, DHX58, DPEP2, DPP4, DSC1, DUSP2, EAF2, EBI3, EFNA5, EGR2, ELANE, EMR1, EMR2, EMR3, EPB41, EPHA1, EPN2, ETS1, ETV3, FAIM3, FAM124B, FAM174B, FAM198B, FAM212B, FAM65B, FASLG, FBXL8, FCERIA, FCER2, FCGR2B, FCGR3B, FCN1, FCRL2, FES, FFAR2, FLJ13197, FLT3LG, FLVCR2, FOSB, FOXP3, FPR1, FPR2, FPR3, FRK, FRMD4A, FRMD8, FZD2, FZD3, GAL3ST4, GALR1, GFI1, GGT5, GIPR, GNG7, GNLY, GPC4, GPR1, GPR171, GPR18, GPR183, GPR19, GPR25, GPR65, GPR97, GRAP2, GSTT1, GUSBP11, GYPE, GZMA, GZMB, GZMH, GZMK, GZMM, HAL, HCK, HDC, HESX1, HHEX, HIC1, HISTIH2AE, HIST1H2BG, HK3, HLA-DOB, HLA-DOA1, HMGB3P30, HNMT, HOXA1, HPGDS, HPSE, HRH1, HSPA6, HTR2B, ICA1, ICOS, IDO1, IFI44L, IFNA10, IFNG, IGHD, IGHE, IGHM, IGKC, IGLL3P, IGSF6, IL12B, IL12RB2, IL17A, IL18R 1 , IL18RAP, IL1A, IL1B, IL1RL1, IL21, IL26, IL2RA, IL2RB, IL3, IL4, IL4R, IL5, ILSRA, IL7, IL7R, IL9, IRF8, ITK, KCNA3, KCNG2, KIAA0226L, KIAA0754, KIR2DL1, KIR2DL4, KIR2DS4, KIR3DL2, KIRREL, KLRB1, KLRC3, KLRC4, KLRD1, KLRF1, KLRG1, KLRK1, KRT18P50, KYNU, LAG3, LAIR2, LAMP3, LAT, LCK, LEF1, LHCGR, LILRA2, LILRA3, LILRA4, LILRB2, LIME1, LINC00597, LINC00921, LOC100130100, LOC126987, LRMP, LST1, LTA, LTB, LTC4S, LY86, LY9, MAGEA11, MAK, MANlA1, MANEA, MAP3K13, MAP4K1, MAP4K2, MAP9, MARCH3, MARCO, MAST1, MBL2, MEFV, MEP1A, MGAM, MICAL3, MMP12, MMP25, MMP9, MNDA, MROH7, MS4A1, MS4A2, MS4A3, MS4A6A, MSC, MXD1, MYB, MZB1, NAALADL1, NCF2, NCR3, NFE2, NIPSNAP3B, NKG7, NLRP3, NMBR, NME8, NOD2, NOX3, NPAS1, NPIPB15, NPL, NR4A3, NTN3, NTRK1, ORC1, OSM, P2RX1, P2RX5, P2RY10, P2RY13, P2RY14, P2RY2, PADI4, PAQR5, PASK, PAX7, PBXIP1, PCDHA5, PDCD1, PDCD 1 LG2, PDE6C, PDK1, PGLYRP1, PIK3IP1, PKD2L2, PLAIA, PLA2G 7 , PLCH2, PLEKHF1, PLEKHG3, PMCH, PNOC, PPBP, PPFIBP1, PRF1, PRG2, PRR5L, PSG2, PTGDR, PTGER2, PTG1R, PTPRCAP, PTPRG, PVRIG, QPCT, RAB27B, RALGPS2, RASA3, RASGRP2, RASGRP3, RASSF4, RCAN3, REN, RENBP, REPS2, RGS1, RGS13, RNASE2, RNASE6, RPL10L, RPL3P7, RRP12, RRP9, RSAD2, RYR1, S100A12, S1PR5, SAMSN1, SCN9A, SEC31B, SELL, SEPT5, SEPT8, SERGEF, SH2D1A, SIGLEC1, SIK1, SIRPG, SIT1, SKA1, SKAP1, SLAMF1, SLAMF8, SLC12A1, SLC12A8, SLC15A3, SLC2A6, SLC7A10, SLCO5A1, SMPD3, SMPDL3B, SOCS1, SP140, SPAG4, SPIB, SPOCK2, SSX1, ST3GAL6, ST6GALNAC4, ST8SIA1, STAP1, STEAP4, STXBP6, TARDBPP1, TBX21, TCF7, TCL1A, TEC, TEP1, TGM5, TLR2, TLR7, TLR8, TMEM156, TMEM255A, TNFAIP6, TNFRSF10C, TNFRSF11A, TNFRSF13B, TNFRSF17, TNFRSF4, TNFSF14, TNIP3, TPSAB1, TRAC, TRAF4, TRAT1, TRAV12-2, TRAV13-1, TRAV13-2, TRAV21, TRA V8-6, TRA V9-2, TRBC1, TRDC, TREM1, TREM2, TREML2, TRIB2, TRPM4, TRPM6, TSHR, TTC38, TXK, TYR, UBASH3A, UGT1A8, UGT2B17, UPK3A, VILL, VNN1, VNN2, VNN3, VPREB3, WNT5B, WNT7A, ZAP70, ZBP1, ZBTB10, ZBTB32, ZFP36L2, ZNF135, ZNF165, ZNF204P, ZNF222, ZNF286A, ZNF324, and ZNF442, and

wherein the optimizing comprises solving for a set of regression coefficients f of the regression, wherein the solving minimizes a linear loss function and an L2-norm penalty function:

(c) estimating a relative proportion of one or more distinct leukocyte cell subsets among the second plurality of distinct leukocyte cell subsets present in the biological sample, based at least in part on the set of regression coefficients f;

(d) predicting a clinical outcome of the cancer, based at least in part on a comparison between the estimated relative proportions of the one or more distinct leukocyte cell subsets present in the biological sample and a pre-determined association of the one or more distinct leukocyte cell subsets with clinical outcomes for the cancer, wherein the clinical outcome comprises survival of the subject; and

(e) based on the predicted clinical outcome of the cancer, administering a cancer therapy to the subject thereby treating the cancer of the subject, wherein the cancer therapy comprises a member selected from the group consisting of a chemotherapy, an immunotherapy, and an immunochemotherapy.

2 . The method of claim 1 , wherein the solving for the set of regression coefficients f further comprises selecting a subset of features in the reference matrix B among a plurality of different subsets of feature signatures of the reference matrix B to minimize the linear loss function.

3 . The method of claim 1 , wherein the linear loss function is a linear ε-insensitive loss function.

4 . The method of claim 1 , wherein the optimizing further comprises using support vector regression (SVR).

5 . The method of claim 4 , wherein the support vector regression is ε-SVR.

6 . The method of claim 4 , wherein the support vector regression is v(nu)-SVR.

7 . The method of claim 1 , further comprising determining a significance value for the estimating.

8 . The method of claim 7 , wherein determining the significance value comprises:

(i) generating a random feature profile m* comprising features randomly selected from a parent feature profile, wherein the parent feature profile comprises the feature profile m, and wherein the feature profile m and the random feature profile m* have the same Euclidean norm;

(ii) optimizing a second regression between the random feature profile m* and the reference matrix B, wherein the random feature profile m* is modeled as a linear combination of the reference matrix B,

wherein the optimizing in (ii) comprises solving for a set of regression coefficients f* of the second regression, wherein the solving for the set of regression coefficients f* of the second regression minimizes a linear loss function and an L2-norm penalty function;

(iii) calculating a product of the set of regression coefficients f* and the reference matrix B to generate a reconstituted feature profile;

(iv) determining a difference measurement between the random feature profile m* and the reconstituted feature profile; and

(v) determining the significance value based at least in part on a distribution of difference measurements determined from performing a plurality of i iterations of (i) to (iv).

9 . The method of claim 1 , wherein the reference matrix B comprises at least one distinct feature that is present in the feature profile m of two or more distinct leukocyte cell subsets of the second plurality of distinct leukocyte cell subsets.

10 . The method of claim 1 , wherein the plurality of distinct leukocyte cell subsets comprises two or more distinct immune cell types.

11 . The method of claim 1 , wherein a non-negative regression coefficient of the set of regression coefficients f is indicative of a relative proportion of a corresponding distinct cell subset among the second plurality of distinct cell subsets present in the biological sample.

12 . The method of claim 11 , further comprising setting negative regression coefficients of the set of regression coefficients f to zero values, and normalizing the non-negative regression coefficients of the set of regression coefficients f, thereby generating the estimated relative proportion of the one or more distinct cell subsets.

13 . A method for treating a subject for cancer, comprising:

administering a cancer therapy to the subject thereby treating the cancer of the subject, based on a predicted clinical outcome of the cancer, wherein the cancer comprises brain cancer, blood cancer, or non-brain solid tumor cancer, and wherein the clinical outcome comprises survival of the subject therapy;

wherein the cancer therapy comprises a member selected from the group consisting of a chemotherapy, an immunotherapy, and an immunochemotherapy; and

wherein the predicted clinical outcome of the cancer is determined at least in part by:

(a) assaying a biological sample comprising peripheral blood mononuclear cells from the subject, thereby generating a feature profile m, wherein the biological sample comprises a first plurality of distinct leukocyte cell subsets,

wherein the feature profile m comprises combinations of features associated with the first plurality of distinct leukocyte cell subsets,

wherein the feature profile m comprises a gene expression profile of leukocyte cells in the biological sample, wherein the gene expression profile represents a ribonucleic acid (RNA) transcriptome of the leukocyte cells in the biological sample;

(b) optimizing, by a computer processor, a regression between the feature profile m and a reference matrix B of feature signatures for a second plurality of distinct leukocyte cell subsets in the biological sample, wherein the feature profile m is modeled as a linear combination of the reference matrix B,

wherein the reference matrix B comprises an LM22 signature matrix of relative gene expression values across 22 leukocyte subsets, wherein the LM22 signature matrix comprises ABCB4, ABCB9, ACAP1, ACHE, ACP5, ADAM28, ADAMDEC1, ADAMTS3, ADRB2, AIF1, AIM2, ALOX15, ALOX5, AMPD1, ANGPT4, ANKRD55, APOBEC3A, APOBEC3G, APOL3, APOL6, AQP 9 , ARHGAP22, ARRB1, ASGR1, ASGR2, ATHL1, ATP8B4, ATXN8OS, AZUL, BACH2, BANK1, BARX2, BCL11B, BCL2A1, BCL7A, BENDS, BFSP1, BHLHE41, BIRC3, BLK, BMP2K, BPI, BRAF, BRSK2, BST1, BTNL8, C11orf80, Clorf54, C3AR1, C5AR1, C5AR2, CA8, CAMP, CASP5, CCDCl02B, CCL1, CCL13, CCL14, CCL17, CCL18, CCL19, CCL20, CCL22, CCL23, CCL4, CCL5, CCL7, CCL8, CCND2, CCR10, CCR2, CCR3, CCR5, CCR6, CCR7, CD160, CD180, CD19, CD1A, CD1B, CD1C, CD1D, CD1E, CD2, CD209, CD22, CD244, CD247, CD27, CD28, CD300A, CD33, CD37, CD38, CD3D, CD3E, CD3G, CD4, CD40, CD40LG, CD5, CD6, CD68, CD69, CD7, CD70, CD72, CD79A, CD79B, CD80, CD86, CD8A, CD8B, CD96, CDA, CDC25A, CDH12, CDHR1, CDK6, CEACAM3, CEACAM8, CEMP1, CFP, CHI3L1, CHI3L2, CHST15, CHST7, CLC, CLCA3P, CLEC10A, CLEC2D, CLEC4A, CLEC7A, CLIC2, CMA1, COL8A2, COLQ, CPA3, CR2, CREB5, CRISP3, CRTAM, CRYBB1, CSF1, CSF2, CSF3R, CST7, CTLA4, CTSG, CTSW, CXCL10, CXCL11, CXCL13, CXCL3, CXCL5, CXCL9, CXCR1, CXCR2, CXCR5, CXCR6, CXorf57, CYP27A1, CYP27B1, DACH1, DAPK2, DCSTAMP, DEFA4, DENND5B, DEPDC5, DGKA, DHRS11, DHX58, DPEP2, DPP4, DSC1, DUSP2, EAF2, EBI3, EFNA5, EGR2, ELANE, EMR1, EMR2, EMR3, EPB41, EPHA1, EPN2,ETS1. ETV3. FAIM3. FAM124B. FAM174B. FAM198B. FAM212B, FAM65B. FASLG, FBXL8, FCERIA, FCER2. FCGR2B. FCGR3B. FCN1. FCRL2. FES. FFAR2. FLJ13197. FLT3LG.FLVCR2, FOSB. FOXP3. FPR1. FPR2. FPR3. FRK. FRMD4A. FRMD8. FZD2, FZD3, GAL3ST4. GALR1. GFI1. GGTS. GIPR. GNG7. GNLY. GPC4. GPR1. GPR171. GPR18.GPR183, GPR19. GPR25. GPR65. GPR97. GRAP2. GSTT1, GUSBP11. GYPE. GZMA. GZMB, GZMH. GZMK, GZMM. HAL. HCK, HDC, HESX1. HHEX, HIC1. HIST1H2AE. HIST1H2BG, HK3. HLA-DOB, HLA-DOA1. HMGB3P30. HNMT. HOXAl. HPGDS. HPSE. HRH1, HSPA6.HTR2B. ICA1. ICOS, IDO1. IFI44L, IFNA10. IFNG. IGHD, IGHE, IGHM. IGKC. IGLL3P.IGSF6. IL12B. IL12RB2. IL17A, IL18R1. IL18RAP. IL1A, IL1B, IL1RL1, IL21, IL26. IL2RA.IL2RB. IL3. IL4. IL4R. ILS, ILSRA. IL7, IL7R. IL9, IRF8, ITK. KCNA3, KCNG2. KIAA0226L, KIAA0754. KIR2DL1, KIR2DL4, KIR2DS4, KIR3DL2. KIRREL, KLRB1. KLRC3, KLRC4.KLRD1. KLRF1, KLRG1. KLRK1. KRT18P50. KYNU. LAG3. LAIR2. LAMP3. LAT. LCK, LEFi. LHCGR. LILRA2. LILRA3. LILRA4, LILRB2. LIMEl. LINC00597, LINC00921, LOC100130100. LoC126987. LRMP. LST1, LTA, LTB. LTC4S. LY86, LY9. MAGEAll, MAK, MANlAl. MANEA. MAP3K13, MAP4K1. MAP4K2. MAP9. MARCH3. MARCO. MAST1. MBL2. MEFV, MEP1A. MGAM. MICAL3. MMP12. MMP25. MMP9, MNDA. MROH7. MS4A1.MS4A2. MS4A3. MS4A6A. MSC, MXD1. MYB. MZB1. NAALADL1. NCF2. NCR3. NFE2.NIPSNAP3B, NKG7, NLRP3. NMBR. NME8. NOD2. NOX3. NPAS1. NPIPB15. NPL. NR4A3. NTN3, NTRK1. ORCi. OSM. P2RX1. P2RX5, P2RY10. P2RY13. P2RY14. P2RY2. PADI4. PAQRS. PASK, PAX7, PBXIP1. PCDHAS, PDCD1. PDCD1LG2. PDE6C. PDK1. PGLYRP1. PIK3IP1. PKD2L2. PLAIA, PLA2G7. PLCH2. PLEKHF1. PLEKHG3, PMCH. PNOC. PPBP, PPFIBP1. PRF1. PRG2. PRRSL. PSG2. PTGDR. PTGER2. PTG1R, PTPRCAP. PTPRG. PVRIG, QPCT, RAB27B. RALGPS2. RASA3. RASGRP2, RASGRP3. RASSF4, RCAN3. REN. RENBP.REPS2. RGS1. RGS13. RNASE2. RNASE6, RPLIOL. RPL3P7, RRP12. RRP9. RSAD2, RYRi. S100A12. SiPR5. SAMSN1. SCN9A, SEC31B. SELL. SEPTS, SEPT8. SERGEF, SH2D1A. SIGLECI. SIKi. SIRPG, SITE, SKAL. SKAPI. SLAMFI. SLAMF8. SLC12A1. SLC12A8. SLC15A3. SLC2A6. SLC7A10. SLCO5A1. SMPD3. SMPDL3B. SOCS1. SP140. SPAG4. SPIB. SPOCK2, SSX1. ST3GAL6. ST6GALNAC4. ST8SIA1. STAPl. STEAP4. STXBP6. TARDBPP1. TBX21. TCF7. TCL1A. TEC. TEPi. TGMS. TLR2. TLR7. TLR8. TMEM156. TMEM255A. TNFAIP6, TNFRSF1OC. TNFRSF11A, TNFRSF13B. TNFRSF17. TNFRSF4. TNFSF14. TNIP3. TPSAB1. TRAC, TRAF4. TRATl. TRAV12-2, TRAV13-1, TRAV13-2, TRAV21. TRAV8-6, TRAV9-2, TRBC1, TRDC, TREM1, TREM2, TREML2, TRIB2, TRPM4, TRPM6, TSHR, TTC38, TXK, TYR, UBASH3A, UGT1A8, UGT2B17, UPK3A, VILL, VNN1, VNN2, VNN3, VPREB3, WNT5B, WNT7A, ZAP70, ZBP1, ZBTB10, ZBTB32, ZFP36L2, ZNF135, ZNF165, ZNF204P, ZNF222, ZNF286A, ZNF324, and ZNF442, and

wherein the optimizing comprises solving for a set of regression coefficients f of the regression, wherein the solving minimizes a linear loss function and an L2-norm penalty function;

(c) estimating a relative proportion of one or more distinct leukocyte cell subsets among the second plurality of distinct leukocyte cell subsets present in the biological sample, based at least in part on the set of regression coefficients f; and

(d) predicting a clinical outcome of the cancer, based at least in part on a comparison between the estimated relative proportions of the one or more distinct leukocyte cell subsets present in the biological sample and a pre-determined association of the one or more distinct leukocyte cell subsets with clinical outcomes for the cancer.

14 . The method of claim 13 , wherein the solving for the set of regression coefficients f further comprises selecting a subset of features in the reference matrix B among a plurality of different subsets of feature signatures of the reference matrix B to minimize the linear loss function.

15 . The method of claim 13 , wherein the linear loss function is a linear E-insensitive loss function.

16 . The method of claim 13 , wherein the optimizing further comprises using support vector regression (SVR).

17 . The method of claim 16 , wherein the support vector regression is ε-SVR.

18 . The method of claim 16 , wherein the support vector regression is v(nu)-SVR.

19 . The method of claim 13 , further comprising determining a significance value for the estimating.

20 . The method of claim 19 , wherein determining the significance value comprises:

(i) generating a random feature profile m* comprising features randomly selected from a parent feature profile, wherein the parent feature profile comprises the feature profile m, and wherein the feature profile m and the random feature profile m* have the same Euclidean norm;

(ii) optimizing a second regression between the random feature profile m* and the reference matrix B, wherein the random feature profile m* is modeled as a linear combination of the reference matrix B,

wherein the optimizing in (ii) comprises solving for a set of regression coefficients f* of the second regression, wherein the solving for the set of regression coefficients f* of the second regression minimizes a linear loss function and an L2-norm penalty function;

(iii) calculating a product of the set of regression coefficients f* and the reference matrix B to generate a reconstituted feature profile;

(iv) determining a difference measurement between the random feature profile m* and the reconstituted feature profile; and

(v) determining the significance value based at least in part on a distribution of difference measurements determined from performing a plurality of i iterations of (i) to (iv).

21 . The method of claim 13 , wherein a non-negative regression coefficient of the set of regression coefficients f is indicative of a relative proportion of a corresponding distinct cell subset among the second plurality of distinct cell subsets present in the biological sample.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 20, 2025
From: NEWMAN, AARON M.; ALIZADEH, ARASH ASH
To: THE BOARD OF TRUSTEES OF THE LELAND STANFORD JUNIOR UNIVERSITY
Reel/Frame 070276/0129 →
Continuity (5)
Continuation 18675760 · May 28, 2024
Continuation 16226270 · Dec 19, 2018
Continuation 15004611 · Jan 22, 2016
Provisional Application 62106601 · Jan 22, 2015
Related Publication 20250263800A1 · Aug 21, 2025
References Cited (98)
US 6661004B2 · Aumond et al. · 2003 [cited by applicant]
US 10167514B2 · Newman · 2019 [cited by examiner]
US 11225689B2 · Shekhar · 2022 [cited by applicant]
US 11756651B2 · Givechian · 2023 [cited by applicant]
US 11802314B2 · Newman · 2023 [cited by examiner]
US 12031183B2 · Newman · 2024 [cited by examiner]
US 12249401B2 · Newman · 2025 [cited by examiner]
US 20050108753A1 · Saidi et al. · 2005 [cited by applicant]
US 20130110756A1 · Zhang et al. · 2013 [cited by applicant]
US 20160217253A1 · Newman et al. · 2016 [cited by applicant]
US 20160341731A1 · Sood et al. · 2016 [cited by applicant]
US 20190233898A1 · Newman et al. · 2019 [cited by applicant]
US 20190338364A1 · Newman et al. · 2019 [cited by applicant]
US 20200075169A1 · Lau · 2020 [cited by applicant]
US 20210040442A1 · Rajagopal et al. · 2021 [cited by applicant]
US 20250129425A1 · Newman · 2025 [cited by applicant]
US 20250243551A1 · Newman · 2025 [cited by applicant]
US 20250322909A1 · Newman · 2025 [cited by applicant]
CN 103217411 · 2013 [cited by applicant]
JP T2014517954 · 2014 [cited by applicant]
WO WO2007035613 · 2007 [cited by applicant]
WO WO2012143561 · 2012 [cited by applicant]
WO WO2015109234 · 2015 [cited by applicant]
WO WO2018191558 · 2018 [cited by applicant]
Abraham (2012) Scalable approaches for analysis of human genome-wide expression and genetic variation data. PhD thesis, Dept of Computing and Information Systems, The University of Melbourne, 313 pages. (Year: 2012). [cited by examiner]
Chiu (2008) Using support vector regression to model the correlation between the clinical metastases time and gene expression profile for breast cancer. Artificial Intelligence in Medicine, vol. 44, p. 221-231. (Year: 2… [cited by examiner]
Boardman, M (2006) Extrinsic regularization in parameter optimization for support vector machines. Master of Computer Science, Dalhousie University, 132 pages. (Year: 2008). [cited by examiner]
Wei (2014) RNA-seq accurately identifies cancer biomarker signatures to distinguish tissue of origin. Neoplasia, vol. 16, No. 11, p. 918-927. (Year: 2014). [cited by examiner]
Greene (2014) Big Data Bioinformatics. Journal of Cellular Physiology, vol. 229, p. 1896-1900. (Year: 2014). [cited by examiner]
Mahmoodian (2016) Using support vector regression in gene selection and fuzzy rule generation for relapse time prediction of breast cancer. Biocybernetics and biomedical engineering, vol. 36, p. 468-472. (Year: 2016). [cited by examiner]
Arieshanti (2014) Analysis of SELDI-TOF-MS using ϵ-Support Vector Regression for ovarian cancer identification. The 15th International Conference on Biomedical Engineering, IFMBE proceedings 43, p. 207-210. Springer Int… [cited by examiner]
Karlik (2012) Personalized cancer treatment using naïve bayes classifier. International Journal of Machine Learning and Computing , vol. 2, No. 3, p. 339-344. (Year: 2012). [cited by examiner]
Chuang (2005) dimension reduction with support vector regression for ovarian cancer microarray data. 2005 IEEE international conference on systems, man, and cybernetics. P1-5. DOI: 10.1109/ICSMC.2005.1571284 (Year: 2005… [cited by examiner]
Jimenez (2012) Feasibility of gene expression signature analysis in prostate cancer biopsy specimens to predict outcomes following radiation therapy. Radiation Oncology, vol. 87, Issue 2, supplement, S669 conference abs… [cited by examiner]
Wilson, M.K. et al. (2015) Outcomes and endpoints in trials of cancer treatment: the past, present and future. Lancet Oncology, vol. 16, e32-e42. (Year: 2015). [cited by examiner]
Abbas et al., “Deconvolution of Blood Microarray Data Identifies Cellular Activation Patterns in Systematic Kupus Erythematosus”, PLOS One, Jul. 2009, p. e6098, vol. 4, No. 7. [cited by applicant]
Abbas et al., “Immune response in silico (IRIS): immune-specific genes identified from a compendium of microarray expression data” Genes and Immunity, 2005, pp. 319-331, vol. 6, No. 4. [cited by applicant]
Ahn et al., “DeMix: deconvolution for mixed cancer transcriptomes using raw measured data” Bioinformatics, 2013, pp. 1865-1871, vol. 29, No. 15. [cited by applicant]
Benita et al., “Gene enrichment profiles reveal T-cell development, differentiation, and lineage- specific transcription factors including ZBTB25 as a novel NF-AT repressor”, Blood, 2010, pp. 5376-5384, vol. 115. [cited by applicant]
Burdick and Murray, “Deconvolution of gene expression from cell populations across the C. elegans lineage”, BMC Bioinformatics, Jun. 22, 2012, p. 204, vol. 14. [cited by applicant]
Caicedo, J.C. et al., (2017) Data-analysis strategies for image-based cell profiling. Nature Methods, vol. 14, No. 9, p. 849-863. (Aug. 11, 2017). [cited by applicant]
Chen, C. et al., (2011) Removing batch effects in analysis of expression microarray data: an evaluation of six batch adjustment methods. PLOS ONE vol. 6, issue 2, e17238, 10 pages. [cited by applicant]
Chen et al. (2017) “Inference of immune cell composition on the expression profiles of mouse tissue.” Scientific reports 7: 1-11. [cited by applicant]
Cherlassky et al., “Practical selection of SVM parameters and noise estimation for SVM regression”, Neural Netk, 2004, pp. 113-126, vol. 17. [cited by applicant]
Cobos et al., (2018) “Computational deconvolution of transcriptomics data from mixed cell populations”, Bioinformatics, 34(11), pp. 1969-1779. [cited by applicant]
Coussens et al., “Neutralizing tumor-promoting chronic inflammation: a magic bullet?”, Science, 2013, pp. 286-291, vol. 339. [cited by applicant]
Definition of normal distribution, Wikipedia.com downloaded Jan. 2024 (Year: 2024). [cited by applicant]
Definition of sampling, and random sampling, Wikipedia.com, downloaded Jan. 2024 (Year: 2024). [cited by applicant]
Definition of simple random sampling, Wikipedia.com downloaded Jan. 2024 (Year: 2024). [cited by applicant]
Drucker et al., “Support Vector Regression Machines”, MIT Press, 1997, pp. 155-161, vol. 9. [cited by applicant]
Farrar et al., “Multicollinearity in Regression Analysis: The Problem Revisited”, R. R. Rev. Econ. Stat., 1967, pp. 92-107, vol. 49. [cited by applicant]
Gaiteri et al., (2013) “Beyond modules and hubs: the potential of gene co-expression networks for investigating molecular mechanisms of complex brain disorders.”, Genes, Brain and Behavior, 13: 13-24. [cited by applicant]
Gaujoux and Seoighe, “CellMix: a comprehensive toolbox for gene expression deconvolution”, Bioinformatics, 2013, pp. 2211-2212, vol. 29, No. 17. [cited by applicant]
Goh et al., (2017) Why batch effects matter in Omics data and how to avoid them. Trends in Biotechnology, vol. 25, No. 6, p. 498-507 (Jun. 2017). [cited by applicant]
Gong and Szutakowski, “DeconRNASeq: a statistical framework for deconvolution of heterogeneous tissue samples based on mRNA-Seq data”, Bioinformatics, 2013, pp. 1083-1085, vol. 29, No. 8. [cited by applicant]
Gong et al., “Optimal Deconvolution of Transcriptional Profiling Data Using Quadratic Programming with Application to Complex Clinical Blood Samples”, PLOS One, Nov. 2011, p. e27156. [cited by applicant]
Hanahan et al., “Hallmarks of Cancer: The Next Generation”, Cell, 2011, pp. 646-674, vol. 144. [cited by applicant]
Johnson et al. (2007) “Adjusting batch effects in microarray expression data using empirical Bayes methods” Biostatistics 8(1): 118-127. [cited by applicant]
Ju et al. (2013) “Defining cell-type specificity at the transcriptional level in human disease.” Genome research 23: 1862-1873. [cited by applicant]
Krishnan et al. (2011) “Quantitative Analysis of Sub-Epithelial Connective Tissue Cell Population of Oral Submucous Fibrosis Using Support Vector Machine”, Journal of Medical Imaging and Health Informatics, vol. 1, No. … [cited by applicant]
Kuhn et al., “Population-specific expression analysis (PSEA) reveals molecular changes in diseased brain”, Nat Methods, 2011, pp. 945-947, vol. 8. [cited by applicant]
Le et al. (2020) “A Review of Digital Cytometry Methods: Estimating the Relative Abundance of Cell Types in a Bulk of Cells”, Briefings in Bioinformatics: 1-12. [cited by applicant]
Levy et al., “Active Idiotypic Vaccination Versus Control Immunotherapy for Folicular Lymphoma”, J Clin. Oncol., 2014, pp. 1797-1803, vol. 32. [cited by applicant]
Li et al. (2016) “Comprehensive Analyses of Tumor Immunity: Implications for Cancer Immunotherapy”, Genome Biology, 2016, vol. 17, No. 1: 1-16. [cited by applicant]
Liebner et al., “MMAD: microarray microdissection with analysis of difference is a computational tool for deconvoluting cell type=speficic contributions from tissue samples”, Bioinformatics, 2014, pp. 682-689, vol. 30, … [cited by applicant]
Lu et al., “Expression deconvolution: A reinterpretation of DNA microarray data reveals dynamic changes in cell populations”, PNAS, 2003, pp. 10370-10375, vol. 100, No. 18. [cited by applicant]
Lukk et al., “A global map of human gene expression”, NAt. Biotechnol, 2010, pp. 322-324, vol. 28. [cited by applicant]
Mackey et al. (2011) “Divide-and-Conquer Matrix Factorization” Advances in Neural Information Processing Systems 24, edited by J. Shawe-Taylor et al. Proceedings from the conference, “Neural information Processing Syste… [cited by applicant]
Meng et al., (2013) Scalable simple random sampling and stratified sampling, Proceedings of the 30 [cited by applicant]
Newman et al. (2014) “An ultrasensitive method for quantitating circulating tumor DNA with broad patient coverage” Nature Medicine 20: 548-554. [cited by applicant]
Newman et al., (2014) “Identifying stem cell gene expression patterns and phenotypic networks with AutoSOME.”, Methods in Molecular Biology, 1150: 115-130. [cited by applicant]
Newman et al. (2015) “Robust enumeration of cell subsets from tissue expression profiles”, Nature Methods, vol. 12, No. 5, pp. 453-457. [cited by applicant]
Newman et al. (2017) “Data Normalization Considerations for Digital Tumor Dissection”, Genome Biology, 2017, vol. 18, No. 1: 1-6. [cited by applicant]
Newman et al. (2019) “Determining Cell Type Abundance and Expression From Bulk Tissues with Digital Cytometry”, Nature Biotechnology, vol. 37, No. 7: 773-782. [cited by applicant]
Qiao et al., “PERT: A Method for Expression Deconvolution of Human Blood Samples from Varied Microenvironmental and Developmental Conditions”, PLOS Comput. Biol., 2012, p. e1002838, vol. 8. [cited by applicant]
Rivenbark et al., (2013) “Molecular and cellular heterogeneity in breast cancer: challenges for personalized medicine”, American Journal of Pathology, 183 (4): 1113-1124. [cited by applicant]
Scholkopf et al., “New Support Vector Algorithms”, Neural Comput., 2000, pp. 1207-1245, vol. 12. [cited by applicant]
Shen-Orr and Gajoux, “Computational deconvolution: extracting cell type-specific information from heterogeneous samples”, Curr. Opin. Immunol., 2013, pp. 571-578, vol. 25. [cited by applicant]
Shen-Orr et al., “Cell type-specific gene expression differences in complex tissues”, Nat. Methods., 2010, pp. 287-289, vol. 7. [cited by applicant]
Smola (2004) “A tutorial on support vector regression.”, Statistics and computing, 14: 199-222. [cited by applicant]
Sotiriou et al., (2009) “Gene-expression patterns in Breast Cancer.”, The New England Journal of Medicine, 360 (8): 790-800. [cited by applicant]
Steen et al. (2020) “Profiling Cell Type Abundance and Expression in Bulk Tissues with CIBERSORTx”, Methods Mol Bio, vol. 2117: 135-157. [cited by applicant]
Storey et al., “Statistical significance for genomewide studies”, Proc. Natl. Acad. Sci. U. S. A., 2003, pp. 9440-9445, vol. 100. [cited by applicant]
Sun et al. (2010) “Combined feature selection and cancer prognosis using support vector machine regression.” IEEE/ACM transactions on computational biology and bioinformatics 8(6): 1671-1677. [cited by applicant]
Thiebaut (2002) “Optimization issues in blind deconvolution algorithms”, Proc. SPIE 4847, Astronomical Data Analysis II, 1-9. [cited by applicant]
Tung et al., (2017) Batch effects and the effective design of single cell gene expression studies. Scientific reports, vol. 7, e39921, 15 pages (Jan. 2017). [cited by applicant]
Wagner et al., (2016) Revealing the vectors of cellular identify with single cell genomics. Nature Biotechnology vol. 14, No. 11, p. 1145-1168. [cited by applicant]
Wang et al. (2013) “Non-negative matrix factorization by maximizing correntropy for cancer clustering” BMC Bioinformatics 14(107): 107 (pp. 1-11). [cited by applicant]
Wang et al., “The doubly regularized support vector machine”, Statistica Sinica, 2006, pp. 589-615, vol. 16, No. 2. [cited by applicant]
Wilhelm-Benartzi et al. (2013) “Review of processing and analysis methods for DNA methylation array data” British J of Cancer 109(6): 1394-1402. [cited by applicant]
Yin, (2013) “Identification of differential gene pathways with sparse principal component analysis”, Georgia State Univsrsity, 1-26. [cited by applicant]
Yoshihara et al., “Inferring tumour purity and stromal and immune cell a dmixture from expression data”, Nat. Commun., 2013, p. 2612, vol. 4. [cited by applicant]
Zheng et al., (2014) “Deconvolution of High Dimensional Mixtures via Boosting, with Application to Diffusion-Weighted MRI of Human Brain”, Advances in Neural Information Processing Systems, 27:2699-2707. [cited by applicant]
Zhong and Liu, “Gene expression deconvolution in linear space”, Nat. Methods., 2012, pp. 8-9, vol. 9. [cited by applicant]
Zhong et al., “Digital sorting of complex tissues for cell type-specific gene expression profiles”, BMC Bioinformatics, 2013, p. 89, vol. 14. [cited by applicant]
Zuckerman (2013) PLOS Computational Biology 9:e1003189. [cited by applicant]
Felton, Identification of carcinoma cells in peripheral blood samples of patients with advanced breast carcinoma using RT-PCT amplification of CK7 and MUC1. The Breast. vol. 13, p. 35-41. (2004). [cited by applicant]
Mohammadi et al. (2017) A Critical Survey of Deconvolution Methods for Separating Cell Types in Complex Tissues, Proceedings of the IEEE, 105(2): 340-366. [cited by applicant]