IP Library › Granted Patent US 12,674,158
Granted Patent B2
US 12,674,158 · App. 17/438,461 · Granted Jul 7, 2026

Methods for DNA library generation to facilitate the detection and reporting of low frequency variants

Inventors: Morgane Macheret (Saint-Sulpice, CH); Christian Pozzorini (Saint-Sulpice, CH); Adrian Willig (Saint-Sulpice, CH); Jonathan Bieler (Saint-Sulpice, CH); Zhenyu Xu (Saint-Sulpice, CH)
Assignee: Sophia Genetics S.A.
C12N15/1065C12Q1/6869
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,674,158
App. No.
17/438,461
Filed
Sep 12, 2021
Granted
Jul 7, 2026
Kind
B2
Art Unit
1684
USPC
506/4
Abstract

Methods are disclosed for adding adapters to fragmented nucleic acids for next generation sequencing, including providing numerical codes based on variable adapter molecular barcode lengths on both sides of the fragmented nucleic acids and identifying reads from the same fragment based on both barcodes. The methods and products allow for the amplification of the fragmented nucleic acids when there is a low yield of isolated fragmented nucleic acids and also for efficient and reliable detection of low-frequency mutations including in subpopulations of cells within a subject.

Claims (32)

1 . A method for generating a library of DNA-adaptor products from at least two DNA fragments to facilitate the identification of the fragments in a high throughput sequencing data analysis workflow after amplification and sequencing, said method comprising:

generating a pool of adaptors that comprise a plurality of double-stranded or partially double-stranded polynucleotides comprising a spacer sequence on the double-stranded extremity of the adaptors, wherein the adaptors differ from each other by the total length of their spacer sequence of at least 3 and at most L max nucleotides, wherein each spacer sequence-comprises a constant termination subsequence TS of length L TS , wherein L TS comprises at least 3 nucleotides, concatenated with a variable spacer subsequence, and wherein the variable spacer subsequence is truncated from a common constant, predefined nucleotide sequence(S) having a length of L S nucleotides, wherein L S comprises 5-20 nucleotides, and wherein L max corresponds to the sum of L S and L TS ;

ligating, in a reaction mixture, a first and a second adaptor from the pool of adaptors to each end of a first double-stranded DNA fragment to produce a first DNA-adaptor product, wherein the ligation places the constant termination subsequence TS of each of the first and second adaptor between the double-stranded DNA fragment and the variable spacer subsequence of each of the first and second adaptor, respectively, so that the first DNA-adaptor product may be characterized by a numerical code formed by the respective lengths (L 1 , L 2 ) of the first and the second adaptor spacer sequences (SS 1 , SS 2 ); and

ligating, in the same reaction mixture, a third and a fourth adaptor from the pool of adaptors to each end of a second double-stranded DNA fragment to produce a second DNA-adaptor product, wherein the ligation places the constant termination subsequence TS of each of the third and fourth adaptor between the double-stranded DNA fragment and the variable spacer subsequence of each of the third and fourth adaptor, respectively, so that the second DNA-adaptor product may be characterized by a numerical code formed by the respective lengths (L 3 , L 4 ) of the third and the fourth adaptor spacer sequences (SS 3 , SS 4 ),

wherein determining whether the first DNA-adaptor product and second DNA-adaptor product correspond to different fragments relies on the placement of the constant termination subsequence between the double-stranded DNA fragment and the variable spacer subsequence and the characterization of the numerical code of the first and second DNA-adaptor product.

2 . The method of claim 1 , wherein the constant termination subsequence TS differs from the constant, predefined nucleotide sequence S by an edit distance of at least two.

3 . The method of claim 1 , wherein the spacer subsequence is truncated left to right from the start from said constant nucleotide sequence(S).

4 . The method of claim 1 , wherein the spacer subsequence is truncated right to left from the end from said constant nucleotide sequence(S).

5 . The method of claim 1 , wherein the constant termination subsequence TS is a triplet nucleotide ending with a T overhang to facilitate ligation to the DNA fragments.

6 . The method of claim 1 , wherein the constant termination subsequence TS is a quadruplet nucleotide ending with a T overhang to facilitate ligation to the DNA fragments.

7 . The method of claim 1 , further comprising:

amplifying the DNA-adaptor products to produce PCR duplicates suitable for high-throughput sequencing; and

sequencing the PCR duplicates with a high-throughput sequencer to produce raw sequencing reads.

8 . The method of claim 7 , further comprising:

for each sequencing read Rn,

trimming L max nucleotides from the beginning of the read, to produce a trimmed sequencing read;

recording the trimmed sequencing read in a pre-processed sequencing read file; and

aligning to a reference genome the trimmed sequencing reads from the pre-processed sequencing read file, so as to map each trimmed read to a start position and an end position.

9 . The method of claim 7 , further comprising:

for each sequencing read Rn, searching for the constant termination subsequence TS in the first L max nucleotides of the sequencing read;

measuring the length L n of the spacer sequence SS Rn as the distance, in the number of nucleotides, between the start of the sequencing read Rn and the end of the constant termination subsequence TS;

trimming at least L n nucleotides from the beginning of the read, to produce a trimmed sequencing read;

recording the measured length L n and the trimmed sequencing read in a pre-processed sequencing read file; and

aligning to a reference genome the trimmed sequencing reads from the pre-processed sequencing read file, so as to map each trimmed read to a start position and an end position.

10 . The method of claim 9 , wherein sequencing produces pair-end reads, further comprising:

tagging pair-end reads aligned to the same start and end position relative to the reference genome sequence reading direction and having the same numerical code pair of measured spacer sequence lengths (L 1 , L 2 ), as sequencing reads potentially issued from the two strands of the same original double-stranded DNA fragment; and

further subdividing those pair-end reads in two sub-groups according to their strand of origin, where the numerical code pair of measured spacer sequence lengths (L 1 , L 2 ) is given by {L n(forward) , L n(reverse) } in case of pair-end reads with F1R2 orientation and by {L n(reverse) , L n(forward) } in case of pair-end reads with F2R1 orientation.

11 . The method of claim 10 , further comprising collapsing each group of reads sharing the same start, end and numerical code into a consensus sequence for their parent fragment and identifying, with a variant calling method, variants for this parent fragment into the collapsed consensus sequence.

12 . The method of claim 10 , further comprising identifying for each group of reads sharing the same start, end and numerical code, with a statistical variant calling method, the probability of variants for their parent fragment.

13 . The method of claim 1 , wherein the method further includes a multiplex high throughput sequencing genomic analysis method for identifying genomic variants in at least two patient samples from a pool of samples, wherein the library of adaptors are different across samples.

14 . The method of claim 13 , wherein the library of adaptors differs across samples by the termination subsequence TS.

15 . The method of claim 13 , wherein the library of adaptors differs across samples by the predefined nucleotide sequence(S) used for truncating for the variable spacer subsequence.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 29, 2025
From: MACHERET, MORGANE; POZZORINI, CHRISTIAN; WILLIG, ADRIAN; BIELER, JONATHAN; XU, ZHENYU
To: SOPHIA GENETICS S.A.
Reel/Frame 072409/0280 →
SECURITY AGREEMENT Recorded May 3, 2024
From: SOPHIA GENETICS SA
To: PERCEPTIVE CREDIT HOLDINGS IV, LP
Reel/Frame 067307/0266 →
Priority Claims (1)
EP 19198542 · Sep 20, 2019 · regional
Continuity (1)
Related Publication 20220364080A1 · Nov 17, 2022
References Cited (89)
US 6582908B2 · Fodor et al. · 2003 [cited by applicant]
US 8209130B1 · Kennedy et al. · 2012 [cited by applicant]
US 9850523B1 · Chudova et al. · 2017 [cited by applicant]
US 10041127B2 · Talasaz · 2018 [cited by applicant]
US 20010053519A1 · Fodor et al. · 2001 [cited by applicant]
US 20030152490A1 · Trulson et al. · 2003 [cited by applicant]
US 20050122325A1 · Twait · 2005 [cited by applicant]
US 20060199189A1 · Bradford · 2006 [cited by applicant]
US 20110160078A1 · Fodor et al. · 2011 [cited by applicant]
US 20120071331A1 · Casbon et al. · 2012 [cited by applicant]
US 20120100548A1 · Rava et al. · 2012 [cited by applicant]
US 20120136583A1 · Lazar et al. · 2012 [cited by applicant]
US 20120165202A1 · Porreca et al. · 2012 [cited by applicant]
US 20120214678A1 · Rava et al. · 2012 [cited by applicant]
US 20140227705A1 · Vogelstein et al. · 2014 [cited by applicant]
US 20160017419A1 · Chiu et al. · 2016 [cited by applicant]
US 20160032396A1 · Diehn et al. · 2016 [cited by applicant]
US 20160046986A1 · Eltoukhy et al. · 2016 [cited by applicant]
US 20160053301A1 · Raymond et al. · 2016 [cited by applicant]
US 20160319345A1 · Gnerre et al. · 2016 [cited by applicant]
US 20170058332A1 · Kermani et al. · 2017 [cited by applicant]
US 20190085406A1 · Mortimer et al. · 2019 [cited by applicant]
EP 2893040B1 · 2015 [cited by applicant]
EP 3240911B1 · 2020 [cited by applicant]
EP 3766986B1 · 2022 [cited by applicant]
EP 3470533B2 · 2023 [cited by applicant]
EP 3443066B1 · 2024 [cited by applicant]
EP 4123032B1 · 2025 [cited by applicant]
EP 4488686A3 · 2025 [cited by applicant]
GB 2510725B · 2015 [cited by applicant]
JP 6275145B2 · 2018 [cited by applicant]
JP 2022548504A · 2022 [cited by applicant]
WO 2011155833A2 · 2011 [cited by applicant]
WO 2012024543A1 · 2012 [cited by applicant]
WO 2012042374A2 · 2012 [cited by applicant]
WO 2012088348A2 · 2012 [cited by applicant]
WO 2012106559A1 · 2012 [cited by applicant]
WO 2012129363A2 · 2012 [cited by applicant]
WO 2012142213A2 · 2012 [cited by applicant]
WO 2012148477A1 · 2012 [cited by applicant]
WO 2013123442A1 · 2013 [cited by applicant]
WO 2013138510A1 · 2013 [cited by applicant]
WO 2013142389A1 · 2013 [cited by applicant]
WO 2013190441A2 · 2013 [cited by applicant]
WO 2014039556A1 · 2014 [cited by applicant]
WO 2014043763A1 · 2014 [cited by applicant]
WO 2014149134A2 · 2014 [cited by applicant]
WO 2014151117A1 · 2014 [cited by applicant]
WO 2014165549A1 · 2014 [cited by applicant]
WO 2014191938A1 · 2014 [cited by applicant]
WO 2015100427A1 · 2015 [cited by applicant]
WO 2015159293A2 · 2015 [cited by applicant]
WO 2015175705A1 · 2015 [cited by applicant]
WO WO2019180528A1 · 2018 [cited by examiner]
WO WO2018144159A1 · 2018 [cited by examiner]
WO 2019002366A1 · 2019 [cited by applicant]
WO 2019084245A1 · 2019 [cited by applicant]
WO 2019204702A1 · 2019 [cited by applicant]
WO 2020043803A1 · 2020 [cited by applicant]
Notice of Reasons for Refusal of the Japan Patent Office in related Japanese Appl. No. 2022-512862, dated Nov. 19, 2024, 9 pages. [cited by applicant]
Brocks, D. et al., “Intratumor DNA methylation heterogeneity reflects clonal evolution in aggressive prostate cancer”, Cell Rep, vol. 8 Issue No. 3, pp. 798-806, Aug. 7, 2014. [cited by applicant]
Chan, K. C. A. et al., “Noninvasive detection of cancer-associated genome-wide hypomethylation and copy number aberrations by plasma DNA bisulfite sequencing”, PNAS, vol. 110 Issue No. 47, pp. 18761-18768, Nov. 19, 2013. [cited by applicant]
Chiu, R. W. K. et al., “Noninvasive prenatal diagnosis of fetal chromosomal aneuploidy by massively parallel genomic sequencing of DNA in maternal plasma”, Proc Natl Acad Sci U S A, vol. 105 Issue No. 51, pp. 20458-2046… [cited by applicant]
Diehl, F. et al., “Analysis of mutations in DNA isolated from plasma and stool of colorectal cancer patients”, Gastroenterology, vol. 135 Issue No. 2, pp. 489-498, Aug. 2008. [cited by applicant]
Diehl, F. et al., “Detection and quantification of mutations in the plasma of patients with colorectal tumors”, Proc Natl Acad Sci U S A, vol. 102 Issue No. 45, pp. 16368-16373, Nov. 8, 2005. [cited by applicant]
Ding, L. et al., “Clonal evolution in relapsed acute myeloid leukaemia revealed by whole-genome sequencing”, Nature, vol. 481 Issue No. 7382, pp. 506-510, Jan. 11, 2012. [cited by applicant]
Ehrich, M. et al., “Noninvasive detection of fetal trisomy 21 by sequencing of DNA in maternal blood: a study in a clinical setting”, Am J Obstet Gynecol, vol. 204 Issue No. 3, pp. 205.e1-205e.11. [cited by applicant]
Fan, H. C. et al., “Noninvasive diagnosis of fetal aneuploidy by shotgun sequencing DNA from maternal blood”, PNAS, vol. 105 Issue No. 42, pp. 16266-16271, Oct. 21, 2008. [cited by applicant]
Forshew, T. et al., “Noninvasive identification and monitoring of cancer mutations by targeted deep sequencing of plasma DNA”, Sci Transl Med, vol. 4 Issue No. 136, pp. 1-12, May 30, 2012. [cited by applicant]
Fu, G. K. et al., “Counting individual DNA molecules by the stochastic attachment of diverse labels”, PNAS, vol. 108 Issue No. 22, pp. 9026-9031, May 11, 2011. [cited by applicant]
Gerlinger, M. et al., “Intratumor Heterogeneity and Branched Evolution Revealed by Multiregion Sequencing”, N Engl J Med, vol. 366 Issue No. 10, pp. 883-892, Mar. 8, 2012. [cited by applicant]
Hoey, T. “Drug resistance, epigenetics, and tumor cell heterogeneity”, Sci Transl Med, vol. 2 Issue No. 28, pp. 1-3, Apr. 21, 2010. [cited by applicant]
Kinde, I. et al., “Detection and quantification of rare mutations with massively parallel sequencing”, PNAS, vol. 108 Issue No. 23, pp. 9530-9535, May 17, 2011. [cited by applicant]
Kivioja, T. et al., “Counting absolute numbers of molecules using unique molecular identifiers”, Nat Methods, vol. 9 Issue No. 1, pp. 72-74, Nov. 20, 2011. [cited by applicant]
Leary, R. J. et al., “Detection of chromosomal alterations in the circulation of cancer patients with whole-genome sequencing”, Sci Transl Med, vol. 4 Issue No. 162, pp. 1-12, Nov. 28, 2012. [cited by applicant]
Narayan, A. et al., “Ultrasensitive measurement of hotspot mutations in tumor DNA in blood using error-suppressed multiplexed deep sequencing”, Cancer Res, vol. 72 Issue No. 14, pp. 3492-3498, Jul. 15, 2012. [cited by applicant]
Newman, A. M. et al., “Integrated digital error suppression for improved detection of circulating tumor DNA”, Nature Biotechnology, vol. 34, pp. 547-555, Mar. 28, 2016 (Abstract Only). [cited by applicant]
Pan, H. et al., “Epigenomic evolution in diffuse large B-cell lymphomas”, Nat Commun, vol. 6, Issue No. 6921, pp. 1-12, Apr. 20, 2015. [cited by applicant]
Rizzo, J. M. et al., “Key principles and clinical applications of “next-generation” DNA sequencing”, Cancer Prev Res (Phila), vol. 5 Issue No. 7, pp. 887-900, Jul. 2012. [cited by applicant]
Schmitt, M. W. et al., “Detection of ultra-rare mutations by next-generation sequencing”, PNAS, vol. 109 Issue No. 36, pp. 14508-14513, Aug. 1, 2012. [cited by applicant]
Schwarzenbach, H. et al., “Cell-free nucleic acids as biomarkers in cancer patients”, Nat Rev Cancer, vol. 11 Issue No. 6, pp. 426-437, Jun. 2011. [cited by applicant]
Sehnert, A. J. et al., “Optimal detection of fetal chromosomal abnormalities by massively parallel DNA sequencing of cell-free fetal DNA from maternal blood”, Clin Chem, vol. 57 Issue No. 7, pp. 1042-1049, Jul. 2011. [cited by applicant]
Sharma, S. V. et al., “A chromatin-mediated reversible drug-tolerant state in cancer cell subpopulations”, Cell, vol. 141 Issue No. 1, pp. 69-80, Apr. 2, 2010. [cited by applicant]
Shaw, J. A. et al., “Genomic analysis of circulating cell-free DNA infers breast cancer dormancy”, Genome Res, vol. 22 Issue No. 2, pp. 220-231, Feb. 2012. [cited by applicant]
Stefansson, O. A. et al., “Epigenetic Modifications in Breast Cancer and Their Role in Personalized Medicine”, The American Journal of Pathology, vol. 183 Issue No. 4, pp. 1052-1063, Oct. 2013. [cited by applicant]
Tsai, H. et al., “Discovery of rare mutations in populations: TILLING by sequencing”, Plant Physiol, vol. 156 Issue No. 3, pp. 1257-1268, Jul. 2011. [cited by applicant]
Zill, O. A. et al., “Cell-free DNA next-generation sequencing in pancreatobiliary carcinomas”, Cancer Discov., vol. 5 Issue No. 10, pp. 1040-1048, Jun. 24, 2015. [cited by applicant]
Extended European Search Report of the European Patent Office in related European Patent Appl. No. 25160115.9, dated Aug. 21, 2025, 10 pages. [cited by applicant]
Hawkins, J. A. et al., “Error-correcting DNA barcodes for high-throughput sequencing”, bioRxiv, Retrieved from the URL: https://www.biorxiv.org/content/biorxiv/early/2018/05/07/315002.full.pdf, May 7, 2018, 23 pages. [cited by applicant]