IP Library Granted Patent US 12,456,541
Granted Patent B2
US 12,456,541 · App. 17/077,967 · Granted Oct 28, 2025

Analysis of genetic variants

Inventors: James Xin Sun (Newton, MA); Roman Yelensky (Newton, MA)
Assignee: Foundation Medicine, Inc.
G16B20/20G16B20/00G16B20/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,456,541
App. No.
17/077,967
Granted
Oct 28, 2025
Kind
B2
Abstract

Methods and systems for analyzing genetic variants are disclosed. The methods and systems can be used to classify a variant in a tumor sample as germline, somatic, subclonal, or ambiguous. The variant can be classified by fitting a genome-wide copy number model to a sequence coverage input and a SNP allele frequency input. The fitted model can be used to determine a tumor purity, and a total copy number and a minor allele copy number for each of a plurality of genomic segments. Using the tumor purity, the total copy number and the minor allele copy number for the genomic segment comprising the variant, and the variant allele frequency for the variant in the tumor sample, the variant can be classified as germline, somatic, subclonal or ambiguous.

Claims (59)

1. A system for classifying a variant in a tumor sample from a subject, comprising:

at least one processor operatively connected to a memory that stores one or more program instructions that, when executed by the at least one processor, are configured to:

cause the at least one processor to receive a sequence coverage input for the tumor sample, a SNP allele frequency input for the tumor sample, and a variant allele frequency for the variant in the tumor sample, wherein the sequence coverage input, the SNP allele frequency input, and the variant allele frequency were determined without using a sample-matched normal control;

fit a genome-wide copy number model for the tumor sample to the sequence coverage input and the SNP allele frequency input;

based on the genome-wide copy number model, determine a tumor purity, a total copy number for each of a plurality of genomic segments, and a minor allele copy number for each of the plurality of genomic segments, wherein at least one genomic segment of the plurality of genomic segments includes the variant; and

classify the variant as germline, somatic, subclonal, or ambiguous using the tumor purity, the total copy number for the genomic segment, the minor allele copy number for the genomic segment, and the variant allele frequency for the variant in the tumor sample.

2. The system of claim 1 , wherein the one or more program instructions, when executed by the at least one processor, are configured to determine the variant allele frequency.

3. The system of claim 1 , wherein the one or more program instructions, when executed by the at least one processor, are configured to determine the sequence coverage input.

4. The system of claim 1 , wherein the one or more program instructions, when executed by the at least one processor, are configured to determine the SNP allele frequency input.

5. The system of claim 1 , wherein the genome-wide copy number model associates:

the sequence coverage input with the total copy number for each genomic segment and the tumor purity, and

the SNP allele frequency input with the total copy number for each genomic segment, the minor allele copy number for each genomic segment, and the tumor purity.

6. The system of claim 1 , wherein the variant is classified as germline, somatic, subclonal, or ambiguous using a process comprising:

computing an expected allele frequency for the variant if the variant were a somatic variant, given the tumor purity, the total copy number for the genomic segment comprising the variant, and the minor allele copy number for the genomic segment,

computing an expected allele frequency for the variant if the variant were a germline variant, given the tumor purity, the total copy number for the genomic segment comprising the variant, and the minor allele copy number for the genomic segment, and

classifying the variant as somatic, germline, subclonal, or ambiguous based on the variant allele frequency for the variant in the tumor sample, the expected allele frequency for the variant if the variant were a somatic variant, and the expected allele frequency for the variant if the variant were a germline variant.

7. The system of claim 6 , wherein the variant is classified somatic, germline, subclonal, or ambiguous with a statistical confidence assessed based on read depth and allele frequency variability within the genomic segment comprising the variant.

8. The system of claim 1 , wherein the variant is classified as germline, somatic, subclonal, or ambiguous using a process comprising:

computing a value for mutation type, comprising fitting a somatic/germline status model that associates the value for mutation type with the tumor purity, the total copy number for the genomic segment comprising the variant, the minor allele copy number for the genomic segment, and the variant allele frequency, and

classifying the variant as germline, somatic, subclonal, or ambiguous based on the value for mutation type.

9. The system of claim 1 , further comprising a sequencer configured to generate sequencing data for the tumor sample.

10. The system of claim 9 , wherein the sequencer is configured to generate the sequencing data by next generation sequencing.

11. The system of claim 9 , wherein the sequence coverage input, the SNP allele frequency input, and the variant allele frequency are computed using the sequencing data.

12. The system of claim 1 , wherein the sequence coverage input is computed as a function of a number of reads for a subgenomic interval of a genome of the tumor sample and a number of reads for a control.

13. The system of claim 12 , wherein the control is a process-matched control.

14. The system of claim 1 , wherein the tumor sample is a tissue sample, a blood sample, or a blood constituent sample.

15. A method for classifying a variant in a tumor sample from a subject, comprising:

receiving, at at least one processor, a sequence coverage input for the tumor sample, a SNP allele frequency input for the tumor sample, and a variant allele frequency for the variant in the tumor sample, wherein the sequence coverage input, the SNP allele frequency input, and the variant allele frequency were determined without using a sample-matched normal control;

fitting, using the at least one processor, a genome-wide copy number model for the tumor sample to the sequence coverage input and the SNP allele frequency input;

based on the genome-wide copy number model, determining, using the at least one processor, a tumor purity, a total copy number for each of a plurality of genomic segments, and a minor allele copy number for each of the plurality of genomic segments, wherein at least one genomic segment of the plurality of genomic segments includes the variant; and

classifying, using the at least one processor, the variant as germline, somatic, subclonal or ambiguous using the tumor purity, the total copy number for the genomic segment comprising the variant, the minor allele copy number for the genomic segment, and the variant allele frequency for the variant in the tumor sample.

16. The method of claim 15 , further comprising determining, using the at least one processor, the variant allele frequency.

17. The method of claim 15 , further comprising determining, using the at least one processor, the sequence coverage input.

18. The method of claim 15 , further comprising determining, using the at least one processor, the SNP allele frequency input.

19. The method of claim 15 , wherein the genome-wide copy number model associates:

the sequence coverage input with the total copy number for each genomic segment and the tumor purity, and

the SNP allele frequency input with the total copy number for each genomic segment, the minor allele copy number for each genomic segment, and the tumor purity.

20. The method of claim 15 , wherein classifying the variant as germline, somatic, subclonal, or ambiguous comprises:

computing, using the at least one processor, an expected allele frequency for the variant if the variant were a somatic variant, given the tumor purity, the total copy number for the genomic segment comprising the variant, and the minor allele copy number for the genomic segment,

computing, using the at least one processor, an expected allele frequency for the variant if the variant were a germline variant, given the tumor purity, the total copy number for the genomic segment comprising the variant, and the minor allele copy number for the genomic segment, and

classifying, using the at least one processor, the variant as somatic, germline, subclonal, or ambiguous based on the variant allele frequency for the variant in the tumor sample, the expected allele frequency for the variant if the variant were a somatic variant, and the expected allele frequency for the variant if the variant were a germline variant.

21. The method of claim 20 , wherein the variant is classified somatic, germline, subclonal, or ambiguous with a statistical confidence assessed based on read depth and allele frequency variability within the genomic segment comprising the variant.

22. The method of claim 15 , wherein classifying the variant as germline, somatic, subclonal, or ambiguous comprises:

computing a value for mutation type, comprising fitting a somatic/germline status model that associates the value for mutation type with the tumor purity, the total copy number for the genomic segment comprising the variant, the minor allele copy number for the genomic segment, and the variant allele frequency, and

classifying the variant as germline, somatic, subclonal, or ambiguous based on the value for mutation type.

23. The method of claim 15 , further comprising sequencing nucleic acid molecules in the tumor sample to generate sequencing data for the tumor sample.

24. The method of claim 23 , wherein the nucleic acid molecules are sequenced by next generation sequencing to generate the sequencing data.

25. The method of claim 23 , wherein the sequence coverage input, the SNP allele frequency input, and the variant allele frequency are computed using the sequencing data.

26. The method of claim 15 , wherein the sequence coverage value is computed as a function of a number of reads for a subgenomic interval of a genome of the tumor sample and a number of reads for a control.

27. The method of claim 26 , wherein the control is a process-matched control.

28. The method of claim 15 , wherein the tumor sample is a tissue sample, a blood sample, or a blood constituent sample.

29. A method of treating a subject having a tumor, comprising:

receiving, at at least one processor, a sequence coverage input for a tumor sample from the subject, a SNP allele frequency input for the tumor sample, and a variant allele frequency for a variant in the tumor sample, wherein the sequence coverage input, the SNP allele frequency input, and the variant allele frequency were determined without using a sample-matched normal control;

fitting, using the at least one processor, a genome-wide copy number model for the tumor sample from the subject to a sequence coverage input and a SNP allele frequency input;

based on the genome-wide copy number model, determining, using the at least one processor, a tumor purity, a total copy number for each of a plurality of genomic segments, and a minor allele copy number for each of the plurality of genomic segments, wherein at least one genomic segment of the plurality of genomic segments includes the variant;

classifying, using the at least one processor, the variant as germline, somatic, or subclonal using the tumor purity, the total copy number for the genomic segment comprising the variant, the minor allele copy number for the genomic segment, and a variant allele frequency for the variant in the tumor sample;

generating a recommendation of a selected treatment modality based on at least a characterization of the variant as germline, somatic, or subclonal; and

treating the subject with the selected treatment modality.

30. The method of claim 29 , wherein the selected treatment modality is a drug combination that treats the tumor.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 6, 2020
From: SUN, JAMES XIN; YELENSKY, ROMAN
To: FOUNDATION MEDICINE, INC.
Reel/Frame 054304/0948 →
Continuity (5)
Continuation 15708475 · Sep 19, 2017
Continuation 14274525 · May 9, 2014
Provisional Application 61939936 · Feb 14, 2014
Provisional Application 61821920 · May 10, 2013
Related Publication 20210043274A1 · Feb 11, 2021
References Cited (102)
US 9340830B2 · Lipson et al. · 2016 [cited by applicant]
US 9792403B2 · Sun et al. · 2017 [cited by applicant]
US 10847249B2 · Sun et al. · 2020 [cited by applicant]
US 20100028873A1 · Belouchi et al. · 2010 [cited by applicant]
US 20110098193A1 · Kingsmore et al. · 2011 [cited by applicant]
US 20120095697A1 · Halpern et al. · 2012 [cited by applicant]
US 20120208706A1 · Downing et al. · 2012 [cited by applicant]
US 20140336996A1 · Sun et al. · 2014 [cited by applicant]
US 20150299795A1 · Lindblad-Toh et al. · 2015 [cited by applicant]
US 20170356053A1 · Otto et al. · 2017 [cited by applicant]
US 20180218113A1 · Sun et al. · 2018 [cited by applicant]
US 20180363066A1 · Chalmers et al. · 2018 [cited by applicant]
US 20190316184A1 · Zimmermann et al. · 2019 [cited by applicant]
US 20210222240A1 · Moshkevich et al. · 2021 [cited by applicant]
US 20230242975A1 · Huang et al. · 2023 [cited by applicant]
US 20240420799A1 · Huang et al. · 2024 [cited by applicant]
WO WO2012092426A1 · 2012 [cited by applicant]
WO WO2014183078A1 · 2014 [cited by applicant]
WO WO2018144782A1 · 2018 [cited by applicant]
WO WO2018213498A1 · 2018 [cited by applicant]
WO WO2019060640A1 · 2019 [cited by applicant]
WO WO2020236941A1 · 2020 [cited by applicant]
WO WO2021247902A2 · 2021 [cited by applicant]
Su et al. “PurityEst: Estimating Purity of Human Tumor Samples Using Next-Generation Sequencing Data,” Bioinformatics (2012) vol. 28:No. 17, pp. 2265-2266. (Year: 2012). [cited by examiner]
Cortes-Ciriano et al., (2017). “A molecular portrait of microsatellite instability across multiple cancers,” Nature Communications, 8:15180, 12 pages. [cited by applicant]
Fabrizio et al., (2018). “Beyond microsatellite testing: assessment of tumor mutational burden identifies subsets of colorectal cancer who may respond to immune checkpoint inhibition,” J. Gastrointestinal Oncology, 9(4)… [cited by applicant]
Girimurugan et al., (2018). “iSeg: an Efficient Algorithm for Segmentation of Genomic and Epigenomic Data,” BMC Bioinformatics, 19:131, 15 pages. [cited by applicant]
Goodman et al., (2017). “Tumor Mutational Burden as an Independent Predictor of Response to Immunotherapy in Diverse Cancers,” Mol. Cancer Ther., 16(11):2598-2608. [cited by applicant]
Hiltemann et al., (2015). “Discriminating somatic and germline mutations in tumor DNA samples without matching normal,” Genome Res., 25(9):1382-1390. [cited by applicant]
International Search Report and Written Opinion received for International Patent Application No. PCT/US2021/035751 mailed on Nov. 23, 2021, 15 pages. [cited by applicant]
Killick et al., (2012). “Optimal detection of changepoints with a linear computational cost,” Journal of the American Statistical Association, 107:500, 25 pages. [cited by applicant]
Kim et al., (2013). “The Landscape of Microsatellite Instability in Colorectal and Endometrial Cancer Genomes,” Cell, 155(4):858-868. [cited by applicant]
Orlandini et al., (2017). “SLMSuite: A Suite of Algorithms for Segmenting Genomic Profiles,” BMC Bioinformatics, 18:321, 5 pages. [cited by applicant]
Richters et al., (2019). “Best practices for bioinformatic characterization of neoantigens for clinical utility,” Genome Medicine, 11(56), 21 pages. [cited by applicant]
Baer et al., (2019). ““Somatic” and “pathogenic”—is the classification strategy applicable in times of large-scale sequencing?” Haematologica, 104(8):1515-1520. [cited by applicant]
Bardia et al., (2018). “Abstract 12097: Landscape of BRCA1 and BRCA2 germline, somatic, and reversion alterations detectable by cell-free DNA testing among patients with metastatic breast, ovarian, pancreatic, or prosta… [cited by applicant]
Fabrizio, (2020). “The Future of TMB as a Predictor of Immunotherapy Response Depends on Harmonization,” available online at <https://www.foundationmedicine.com/blog/the-future-of-tmb-as-a-predictor-of-immunotherapy-res… [cited by applicant]
Huang et al., (2020). “Abstract 5455: Identifying somatic variants with defined confidence level in targeted sequencing of tumor DNA without a matched normal sample,” Cancer Res, 80(16_Suppl):5455, 2 pages. [cited by applicant]
Huang et al., (2020). “Abstract 5455 (E-Poster): Identifying somatic variants with defined confidence level in targeted sequencing of tumor DNA without a matched normal sample,” AACR Annual Meeting 2020, 1 page. [cited by applicant]
Johnson et al., (2016). “Targeted Next Generation Sequencing Identifies Markers of Response to PD-1 Blockade,” Cancer Immunology Research, 4(11):959-967. [cited by applicant]
McKernan et al., (2009). “Sequence and structural variation in a human genome uncovered by short-read, massively parallel ligation sequencing using two-base encoding,” Genome Res, 19(9):1527-41. [cited by applicant]
Popova et al., (2013). “Analysis of Somatic Alterations in Cancer Genome: From SNP Arrays to Next Generation Sequencing,” available online at <https://hal.archives-ouvertes.fr/hal-01108425>, 22 pages. [cited by applicant]
Schrock, (2019). “Chemotherapy or Immunotherapy? TMB May Help Predict Best Treatment Choice for Some Metastatic Colorectal Cancer Patients,” available online at <https://www.foundationmedicine.com/blog/chemotherapy-or-i… [cited by applicant]
Carter, SL et al., “Absolute Quantification of Somatic DNA Alterations in Human Cancer”, Nat Biotechnol., vol. 30, No. 5, pp. 413-421, May 2012. [cited by applicant]
Extended European Search Report for European Application No. EP 14795352.5 dated Mar. 20, 2017. [cited by applicant]
Frampton, G.M. et al., “Development and Validation of a Clinical Cancer Genomic Profiling Test Based on Massively Parallel DNA Sequencing”, Nature Biotechnology, Oct. 20, 2013, pp. 1-11. [cited by applicant]
Greenman, C.D. et al., “PICNIC: An Algorithm to Predict Absolute Allelic Copy Number Variation with Microarray Cancer Data”, Biostatistics, 2010, 11, 1, pp. 164-175. [cited by applicant]
International Search Report and Written Opinion dated Oct. 7, 2014 from International Application No. PCT/US14/37569. [cited by applicant]
Jones, S. et al., “Personalized Genomic Analyses for Cancer Mutation Discovery and Interpretation”, Science Translation Medicine, vol. 7, Issue 283, pp. 1-10, Apr. 15, 2015. [cited by applicant]
Kallioniemi, A. et al., “Comparative Genomic Hybridization for Molecular Cytogenetic Analysis of Solid Tumors”, Science, vol. 258, Oct. 30, 1992, pp. 818-821. [cited by applicant]
Sun et al. “A computational approach to distinguish somatic vs. germline origin of genomic alterations from deep sequencing of cancer specimens without a matched normal” PLOS Computational Biology (2018) vol. 14, No. 2,… [cited by applicant]
Sun et al. Abstract. A computational method for somatic vs. germline variant status determination from targeted NGS of clinical cancer specimens without a matched normal control. HiTSeq 2013: Conference on High Throughp… [cited by applicant]
Sun et al. Abstract. A computational method for somatic versus germline variant status determination from targeted next-generation sequencing of clinical cancer specimens without a matched normal control. American Assoc… [cited by applicant]
Sun et al. Poster presentation. A computational method for somatic vs germ-line variant status determination from targeted next-generation sequencing of clinical cancer specimens without a matched normal control. 15th A… [cited by applicant]
Sun et al. Poster. A computational method for somatic versus germline variant status determination from targeted next-generation sequencing of clinical cancer specimens without a matched normal control. American Associa… [cited by applicant]
Van Loo, P. et al., “Allele-Specific Copy Number Analysis of Tumors”, Proc Natl Acad Sci, vol. 107, No. 39, pp. 16910-16915, Sep. 28, 2010. [cited by applicant]
Yau, Christopher, “OncoSNP-SEQ: A Statistical Approach for the Identification of Somatic Copy Number Alterations from Next-Generation Sequencing of Cancer Genomes”, Bioinformatics, vol. 29, No. 19, 2013, pp. 2482-2484. [cited by applicant]
Yelensky et al. Abstract. A statistical model for detecting gene amplification and homozygous deletion from targeted next-generation sequencing of clinical cancer specimens with significant stromal admixture. HiTSeq 201… [cited by applicant]
Yelensky et al. Presentation. A statistical model for detecting gene amplification and homozygous deletion from targeted next-generation sequencing of clinical cancer specimens with significant stromal admixture. HiTSeq… [cited by applicant]
Fridlyand et al., (2004). “Hidden Markov models approach to the analysis of array CGH data,” Journal of Multivariate Analysis, 90:132-153. [cited by applicant]
Hsu et al., (2005). “Denoising array-based comparative genomic hybridization data using wavelets,” Biostatistics, 6(2):211-226. [cited by applicant]
Metzker, M. (2010). “Sequencing technologies—the next generation,” Nature Rev Genet., 11:31-46. [cited by applicant]
Meyer et al., (2013). “The UCSC Genome Browser database: extensions and updates 2013,” Nucleic Acids Res., 41:D64-69. [cited by applicant]
Olshen et al., (2004). “Circular binary segmentation for the analysis of array-based DNA copy number data,” Biostatistics, 5(4):557-572. [cited by applicant]
Sherry et al., (2001). “dbSNP: the NCBI database of genetic variation,” Nucleic Acids Res., 29(1):308-311. [cited by applicant]
Wang et al., (2005). “A method for calling gains and losses in array CGH data,” Biostatistics, 6(1):45-58. [cited by applicant]
Albers et al., (2011). “Dindel: accurate indel calls from short-read data,” Genome Res., 21(6):961-973. [cited by applicant]
Albert et al., (2007). “Direct selection of human genomic loci by microarray hybridization,” Nat. Methods, 4(11):903-905. [cited by applicant]
Ambion, (2011). “RecoverAll Total Nucleic Acid Isolation Protocol (Ambion, Cat. No. AM1975),” available online at <https://assets.thermofisher.com/TFS-Assets/LSG/manuals/1975MC.pdf>, 29 pages. [cited by applicant]
Batzoglou, et al. (2002), “ARACHNE: A Whole-Genome Shotgun Assembler,” Genome Res., 12:177-189. [cited by applicant]
Braun et al., (1998). “Statistical methods for DNA sequence segmentation,” Statistical Science, 13(2):142-162. [cited by applicant]
Browning et al., (2009). “Simultaneous genotype calling and haplotype phasing improves genotype accuracy and reduces false-positive associations for genome-wide association studies,” Am. J. Hum. Genet., 85(6):847-861. [cited by applicant]
Butler et al., (2008). “ALLPATHS: de novo assembly of whole-genome shotgun microreads,” Genome Res., 18(5):810-820. [cited by applicant]
Cronin et al., (2004). “Measurement of gene expression in archival paraffin-embedded tissues: development and performance of a 92-gene reverse transcriptase-polymerase chain reaction assay,” Am J Pathol., 164(1):35-42. [cited by applicant]
Extended European Search Report received for European Patent Application No. 21819009.8 dated May 23, 2024, 17 pages. [cited by applicant]
Farrar (2007), “Striped Smith-Waterman Speeds Database Searches Six Times Over Other SIMD Implementations,” Bioinformatics, 23(2):156-161. [cited by applicant]
Goya et al., (2010). “SNVMix: predicting single nucleotide variants from next-generation sequencing of tumors,” Bioinformatics, 26(6):730-736. [cited by applicant]
Hodges et al., (2007). “Genome-wide in situ exon capture for selective resequencing,” Nat. Genet., 39(12):1522-1527. [cited by applicant]
International Search Report and Written Opinion received for International Patent Application No. PCT/US2022/080995 mailed on Mar. 9, 2023, 16 pages. [cited by applicant]
Le et al., (2011). “SNP detection and genotyping from low-coverage sequencing data on multiple diploid samples,” Genome Res., 21(6):952-960. [cited by applicant]
Li et al., (2009). “Fast and Accurate Short Read Alignment with Burrows-Wheeler Transform,” Bioinformatics, 25:1754-60. [cited by applicant]
Li et al., (2009). “Genotype imputation,” Annu. Rev. Genomics Hum. Genet., 10:387-406. [cited by applicant]
Li et al., (2009). “The Sequence Alignment/Map format and SAMtools,” Bioinformatics, 25(16):2078-2079. [cited by applicant]
Li et al., (2010). “Fast and accurate long-read alignment with Burrows-Wheeler transform,” Bioinformatics, 26(5):589-595. [cited by applicant]
Lunter et al., (2011). “Stampy: a statistical algorithm for sensitive and fast mapping of Illumina sequence reads,” Genome Res., 21(6):936-939. [cited by applicant]
Masuda et al., (1999). “Analysis of chemical modification of RNA from formalin-fixed samples and optimization of molecular biology applications for such samples,” Nucleic Acids Res., 27(22):4436-4443. [cited by applicant]
McKenna et al., (2010). “The Genome Analysis Toolkit: a MapReduce framework for analyzing next-generation DNA sequencing data,” Genome Res., 20(9):1297-1303. [cited by applicant]
Needleman et al., (1970). “A General Method Applicable to the Search for Similarities in the Amino Acid Sequence of Two Proteins,” J. Molecular Biology, 48(3):443-53. [cited by applicant]
Okou et al., (2007). “Microarray-based genomic selection for high-throughput resequencing,” Nat. Methods, 4(11):907-9. [cited by applicant]
Omega bio-tek, (2009). “E.Z.N.A.® FFPE DNA Kit Quick Guide, product numbers D3399-00, and D3399-01,” Available online at <https://www.omegabiotek.com/wp-content/uploads/2018/01/D3399-WEB.pdf>, 2 pages. [cited by applicant]
Promega, (2007). “Maxwell® 16 Cell Total RNA Purification Kit Technical Bulletin (Promega Literature #TB351),” available online at <http://www.yph-bio.com/equip/app/maxwell/maxwell_16_total_rna_purification_kit.pdf>, 25… [cited by applicant]
Promega, (2017). “Maxwell® 16 FFPE Plus LEV DNA Purification Kit Technical Manual (Promega Literature #TM349),” available online at <https://ch.promega.com/-/media/files/resources/protocols/technical-manuals/101/maxwell… [cited by applicant]
Promega, (2017). “Maxwell® 16 LEV Blood DNA Kit and Maxwell 16 Buccal Swab LEV DNA Purification Kit Technical Manual (Promega Literature #TM333),” available online at <https://www.promega.com/-/media/files/resources/pro… [cited by applicant]
Qiagen, (2020). “QIAamp® DNA FFPE Tissue Handbook (Qiagen, Cat. No. 37625),” available online at <https://www.qiagen.com/nl/resources/download.aspx?id=7d3df4c2-b522-4f6d-b990-0ac3a71799b6&lang=en>, 28 pages. [cited by applicant]
Smith et al., (1981). “Identification of Common Molecular Subsequences,” J. Molecular Biology, 147(1):195-197. [cited by applicant]
Specht et al., (2001). “Quantitative gene expression analysis in microdissected archival formalin-fixed and paraffin-embedded tumor tissue,” Am J Pathol., 158(2):419-429. [cited by applicant]
Tan et al., (2009). “DNA, RNA, and Protein Extraction: The Past and The Present,” J. Biomed. Biotech., 2009:574398, 10 pages. [cited by applicant]
Trapnell et al., (2009). “How to map billions of short reads onto genomes,” Nature Biotech., 27(5):455-457. [cited by applicant]
Van Dijk et al., (2014). “Library preparation methods for next-generation sequencing: tone down the bias,” Exp. Cell Research, 322:12-20. [cited by applicant]
Warren et al., (2007). “Assembling millions of short DNA sequences using SSAKE,” Bioinformatics, 23(4):500-501. [cited by applicant]
Ye et al., (2009). “Pindel: a pattern growth approach to detect break points of large deletions and medium sized insertions from paired-end short reads,” Bioinformatics, 25(21):2865-2871. [cited by applicant]
Zerbino et al., (2008). “Velvet: algorithms for de novo short read assembly using de Bruijn graphs,” Genome Res., 18(5):821-829. [cited by applicant]