IP Library › Granted Patent US 12,590,326
Granted Patent B2
US 12,590,326 · App. 16/244,966 · Granted Mar 31, 2026

Methods for fragmentome profiling of cell-free nucleic acids

Inventor: Diana Abdueva (Orinda, CA)
Assignee: Guardant Health, Inc.
C12Q1/68G16B20/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,590,326
App. No.
16/244,966
Granted
Mar 31, 2026
Kind
B2
Abstract

The present disclosure contemplates various uses of cell-free DNA. Methods provided herein may use sequence information in a macroscale and global manner, with or without somatic variant information, to assess a fragmentome profile that can be representative of a tissue of origin, disease, progression, etc. In an aspect, disclosed herein is a method for determining a presence or absence of a genetic aberration in deoxyribonucleic acid (DNA) fragments from cell-free DNA obtained from a subject, the method comprising: (a) constructing a multi-parametric distribution of the DNA fragments over a plurality of base positions in a genome; and (b) without taking into account a base identity of each base position in a first locus, using the multi-parametric distribution to determine the presence or absence of the genetic aberration in the first locus in the subject.

Claims (71)

1 . A computer-implemented method for determining a presence or absence of a genetic aberration in deoxyribonucleic acid (DNA) fragments from cell-free DNA obtained from a subject, the method comprising:

(a) obtaining the base sequence of at least 10,000 DNA fragments from the subject, wherein the DNA fragments have been labeled with a unique or non-unique molecular tag, wherein the labeling is performed by ligating sequencing adapters;

(b) constructing, by a computer, a multi-parametric distribution of the DNA fragments over a plurality of base positions in a genome; and

(c) without taking into account a base identity of each base position in a first locus, using the multi-parametric distribution to determine the presence or absence of the genetic aberration in the first locus in the subject.

2 . The method of claim 1 , wherein the genetic aberration comprises a sequence aberration or a copy number variation (CNV), wherein the sequence aberration is selected from the group consisting of:

(i) a single nucleotide variant (SNV),

(ii) an insertion or deletion (indel), and (iii) a gene fusion.

3 . The method of claim 1 , wherein the multi-parametric distribution comprises parameters indicative of one or more of:

(i) a length of the DNA fragments that align with each of the plurality of base positions in the genome,

(ii) a number of the DNA fragments that align with each of the plurality of base positions in the genome, and

(iii) a number of the DNA fragments that start or end at each of the plurality of base positions in the genome.

4 . The method of claim 1 , further comprising using the multi-parametric distribution to determine a distribution score, wherein the distribution score is indicative of a mutation burden of the genetic aberration.

5 . The method of claim 4 , wherein the distribution score comprises values indicating one or more of a number of the DNA fragments with dinucleosomal protection and a number of the DNA fragments with mononucleosomal protection.

6 . A computer-implemented method for analyzing cell-free deoxyribonucleic acid (DNA) fragments derived from a subject, the method comprising:

obtaining sequence information representative of the cell-free DNA fragments, wherein the number of DNA fragments is at least 10,000, wherein the DNA fragments have been labeled with a unique or non-unique molecular tag wherein the labeling is performed by ligating sequencing adapters; and

performing a multi-parametric analysis on a plurality of data sets using the sequence information to generate a multi-parametric model representative of the cell-free DNA fragments, wherein the multi-parametric model comprises three or more dimensions.

7 . The method of claim 6 , wherein the data sets are selected from the group consisting of data comprising:

(a) start position of DNA fragments sequenced,

(b) end position of sequenced DNA fragments,

(c) number of unique sequenced DNA fragments that cover a mappable position,

(d) length of sequenced DNA fragments,

(e) a likelihood that a mappable base-pair position will appear at a terminus of a sequenced DNA fragment,

(f) a likelihood that a mappable base-pair position will appear within a sequenced DNA fragment as a consequence of differential nucleosome occupancy,

(g) a sequence motif of sequenced DNA fragments,

(h) GC content,

(i) sequenced DNA fragment length distribution, and

(j) methylation status.

8 . The method of claim 6 , wherein the multi-parametric analysis comprises mapping to each of a plurality of base positions or regions of a genome, one or more distributions selected from the group consisting of:

(i) a distribution of the number of unique cell-free DNA fragments containing a sequence that covers the mappable position in the genome,

(ii) a distribution of the fragment lengths for each of at least some of the cell-free DNA fragments such that the DNA fragment contains a sequence that covers the mappable position in the genome, and

(iii) a distribution of the likelihoods that a mappable base-pair position will appear at a terminus of a sequenced DNA fragment.

9 . The method of claim 8 , wherein the plurality of base positions or regions of a genome include at least one base position or region associated with one or more of the genes selected from the list consisting of AKT1, ALK, APC, AR, ARAF, ARID1A, ATM, BRAF, BRCA1, BRCA2, CCND1, CCND2, CCNE1, CDH1, CDK4, CDK6, CDKN2A, CDKN2B, CTNNB1, EGFR, ERBB2, ESR1, EZH2, FBXW7, FGFR1, GFR2, FGFR3, GATA3, GNA11, GNAQ, GNAS, HNF1A, HRAS, IDH1, IDH2, JAK2, JAK3, KIT, KRAS, MAP2K1, MAP2K2, MET, MLH1, MPL, MYC, NF1, NFE2L2, NOTCH1, NPM1, NRAS, NTRK1, PDGFRA, PIK3CA, PTEN, PTPN11, RAF1, RB1, RET, RHEB, RHOA, RIT1, ROS1, SMAD4, SMO, SRC, STK11, TERT, TP53, TSC1, and VHL.

10 . The method of claim 8 , wherein the mapping comprises mapping a plurality of values from each of a plurality of the data sets, to each of a plurality of base positions or regions of a genome.

11 . The method of claim 10 , wherein at least one of the plurality of values is a data set selected from the group consisting of

(a) start position of DNA fragments sequenced,

(b) end position of sequenced DNA fragments,

(c) number of unique sequenced DNA fragments that cover a mappable position,

(d) length of sequenced DNA fragments,

(e) a likelihood that a mappable base-pair position will appear at a terminus of a sequenced DNA fragment,

(f) a likelihood that a mappable base-pair position will appear within a sequenced DNA fragment as a consequence of differential nucleosome occupancy, or

(g) a sequence motif of sequenced DNA fragments.

12 . The method of claim 6 , wherein the multi-parametric analysis comprises applying, by a computer, one or more mathematical transforms to generate the multi-parametric model.

13 . The method of claim 6 , wherein the multi-parametric model is a joint distribution model of a plurality of variables selected from the group consisting of:

(a) start position of DNA fragments sequenced,

(b) end position of sequenced DNA fragments,

(c) number of unique sequenced DNA fragments that cover a mappable position,

(d) length of sequenced DNA fragments,

(e) a likelihood that a mappable base-pair position will appear at a terminus of a sequenced DNA fragment,

(f) a likelihood that a mappable base-pair position will appear within a sequenced DNA fragment as a consequence of differential nucleosome occupancy, and

(g) a sequence motif of sequenced DNA fragments.

14 . The method of claim 6 , further comprising identifying in the multi-parametric model, one or more peaks, each peak having a peak distribution width and a peak coverage.

15 . The method of claim 14 , further comprising detecting one or more deviations between the multi-parametric model representative of the cell-free DNA fragments and a reference multi-parametric model.

16 . The method of claim 15 , wherein the deviation is selected from the group consisting of:

(i) an increase in the number of reads outside a nucleosome region,

(ii) an increase in the number of reads within a nucleosome region,

(iii) a broader peak distribution relative to a mappable genomic location,

(iv) a shift in location of a peak,

(v) identification of a new peak,

(vi) a change in depth of coverage of a peak,

(vii) a change in start position around a peak, and

(viii) a change in fragment sizes associated with a peak.

17 . The method of claim 6 , further comprising determining a contribution of the multi-parametric model attributed to

(i) apoptotic processes in cells from which the cell-free DNA originated or

(ii) necrotic processes in cells from which the cell-free DNA originated.

18 . The method of claim 6 , further comprising performing a multi-parametric analysis to

(i) measure RNA expression of the cell-free DNA fragments,

(ii) measure methylation of the cell-free DNA fragments,

(iii) measure a nucleosomal mapping of the cell-free DNA fragments, or

(iv) identify the presence of one or more somatic single nucleotide polymorphisms in the cell-free DNA fragments or one or more germline single nucleotide polymorphisms in the cell-free DNA fragments.

19 . The method of claim 6 , further comprising generating a distribution score comprising values indicating a number of the DNA fragments with dinucleosomal protection or a number of the DNA fragments with mononucleosomal protection.

20 . The method of claim 6 , further comprising estimating a mutation burden of the subject.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 11, 2020
From: ABDUEVA, DIANA
To: GUARDANT HEALTH, INC.
Reel/Frame 052088/0719 →
Continuity (2)
Provisional Application 62615885 · Jan 10, 2018
Related Publication 20190352695A1 · Nov 21, 2019
References Cited (35)
US 9218449B2 · Lo et al. · 2015 [cited by applicant]
US 9598731B2 · Talasaz · 2017 [cited by applicant]
US 9892230B2 · Lo et al. · 2018 [cited by applicant]
US 20130085681A1 · Deciu et al. · 2013 [cited by applicant]
US 20150051087A1 · Rabinowitz et al. · 2015 [cited by applicant]
US 20150368708A1 · Talasaz · 2015 [cited by applicant]
US 20160019338A1 · Chudova et al. · 2016 [cited by applicant]
US 20160040229A1 · Talasaz et al. · 2016 [cited by applicant]
US 20160046986A1 · Eltoukhy et al. · 2016 [cited by applicant]
US 20170024513A1 · Lo et al. · 2017 [cited by applicant]
US 20170298427A1 · Buis · 2017 [cited by examiner]
US 20180251848A1 · Diehn et al. · 2018 [cited by applicant]
WO 2014116598A2 · 2014 [cited by applicant]
WO 2015048535A1 · 2015 [cited by applicant]
WO 2016015058A2 · 2016 [cited by applicant]
WO 2016094853A1 · 2016 [cited by applicant]
Chandrananda, D. et al. “High-resolution characterization of sequence signatures due to non-random cleavage of cell-free DNA” BioMed Central Medical Genomics (Jun. 17, 2015) 8(29):1-19. [cited by applicant]
Clark, T.A. et al. “Analytical Validation of a Hybrid Capture Based Next-Generation Sequencing Clinical Assay for Genomic Profiling of Cell-Free Circulating Tumor DNA,” J. Mol. Diagnostics (2018) 20(5):686-702. [cited by applicant]
Ding, C. et al. “Minimum redundancy feature selection from microarray gene expression data” J Bioinform Comput Biol. Apr. 2005;3(2):185-205. [cited by applicant]
Gaffney, D.J. et al. “Controls of nucleosome positioning in the human genome.” PLoS Genet. (2012) 8:e1003036. [cited by applicant]
He, X. et al. “Different mutations of the human c-mpl gene indicate distinct haematopoietic diseases” J. Hematol & Oncol (2013) 6:11. [cited by applicant]
Hyndman, R.J. (1996). Computing and graphing highest density regions. The American Statistician, 50(2), 120-126. [cited by applicant]
International search report and written opinion dated Sep. 25, 2017 for PCT/US2017/040986. [cited by applicant]
Ivanov, M. et al. “Non-random fragmentation patterns in circulating cell-free DNA reflect epigenetic regulation” BMC Genomics (2015) 16(Suppl. 13):51. [cited by applicant]
Kostyuk, S.V. et al. “An Exposure to the Oxidized DNA Enhances Both Instability of Genome and Survival in Cancer Cells” PLoS (Oct. 17, 2013) 8(10):e77469, pp. 1-23. [cited by applicant]
Liu, F.T. et al. “Isolation forest” Data Mining, 2008. ICDM'08. Eighth IEEE International Conference. [cited by applicant]
Newman, et al. An ultrasensitive method for quantitating circulating tumor DNA with broad patient coverage. Nat Med. May 2014;20(5):548-54. doi: 10.1038/nm.3519. Epub Apr. 6, 2014. [cited by applicant]
Paweletz, C.P. et al. “Bias-corrected targeted next-generation sequencing for rapid, multiplexed detection of actionable alterations in cell-free DNA from advanced lung cancer patients” Clin Canc Res (2016) 22(4):915-92… [cited by applicant]
Phallen, J. et al. “Direct detection of early-stage cancers using circulating tumor DNA” Sci Trans Med (2017) vol. 9, Issue 403, eaan2415DOI: 10.1126/scitranslmed.aan2415. [cited by applicant]
Rousseeuw, P.J. et al. “A Fast Algorithm for the Minimum Covariance Determinant Estimator” Technometrics (1999) 41(3):212-223. [cited by applicant]
Schep, A.N. et al. “Structured nucleosome fingerprints enable high-resolution mapping of chromatin architecture within regulatory regions” Genome Research (2015) 25(11):1757-1770. [cited by applicant]
Scholkopf, B. et al. “Estimating the Support of a High-Dimensional Distribution” Neural Computation (2001) 13(7):1443-1471. [cited by applicant]
Snyder, M.W. et al. “Cell-free DNA Comprises an In Vivo Nucleosome Footprint that Informs Its Tissues-Of-Origin” Cell (2016) 164:57-68 & Supplemental Information. [cited by applicant]
Statham, et al. “Genome-wide nucleosome occupancy and DNA methylation profiling of four human cell lines” Genomics Data (Mar. 2015) 3:94-96. [cited by applicant]
Wand, M.P. “Fast Computation of Multivariate Kernel Estimators.” Journal of Computational and Graphical Statistics (1994) 3:433-445. [cited by applicant]