IP Library › Granted Patent US 12,338,496
Granted Patent B2
US 12,338,496 · App. 17/389,009 · Granted Jun 24, 2025

Methods, systems, compositions, kits, apparatus and computer-readable media for molecular tagging

Inventors: Kelli Bramlett (Austin, TX); Dumitru Brinza (Montara, CA); Richard Chien (Foster City, CA); Dalia Dhingra (San Francisco, CA); Jian Gu (Austin, TX); Ann Mongan (San Francisco, CA)
Assignee: Life Technologies Corporation
C12Q1/6886C12Q1/6827C12Q1/6874C12Q2600/156
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,338,496
App. No.
17/389,009
Filed
Jul 29, 2021
Granted
Jun 24, 2025
Kind
B2
Art Unit
1681
USPC
506/4
Abstract

In some embodiments, the disclosure relates generally to methods, as well as related systems, compositions, kits, apparatuses and computer-readable media, comprising a multiplex molecular tagging procedure that employs a plurality of tags that are appended to a plurality of polynucleotides. The tags have characteristics, including a sequence, length and/or detectable moiety, or any other characteristic, that uniquely identifies the polynucleotide molecule to which it is appended, and permits tracking individual tagged molecules in a mixture of tagged molecules. For example, the tag having a unique tag sequence, can uniquely identify an individual polynucleotide to which it is appended, and distinguish the individual polynucleotide from other tagged polynucleotides in a mixture. In some embodiments, the multiplex molecular tagging procedure can be used for generating error-corrected sequencing data and for detecting a target polynucleotide which is present at low abundance in a nucleic acid sample.

Claims (30)

1. A method for detecting a genetic variant in a nucleic acid sample, comprising:

a) appending individual polynucleotides of a plurality of polynucleotides of the nucleic acid sample with at least one oligonucleotide tag of a plurality of oligonucleotide tags to generate tagged polynucleotides, wherein the at least one oligonucleotide tag comprises a structure, N 1 N 2 N 3 X 1 X 2 X 3 M 4 M 5 M 6 Y 4 Y 5 Y 6

(i) wherein N represents a random tag sequence wherein each base position in the random tag sequence is independently selected from A, G, C or T;

(ii) wherein X 1 X 2 X 3 represents a first fixed tag sequence that is the same in all of the plurality of tags;

wherein M represents a random tag sequence wherein each base position in the random tag sequence is independently selected from A, G, C or T, wherein the random tag sequence M differs from the random tag sequence N; and

(iv) wherein Y 4 Y 5 Y 6 represents a second fixed tag sequence that is the same in all of the plurality of tags, and the second fixed tag sequence of Y 4 Y 5 Y 6 differs from the first fixed tag sequence of X 1 X 2 X 3 ;

b) amplifying the tagged polynucleotides to generate tagged amplicons;

c) sequencing the tagged amplicons to generate a plurality of candidate sequencing reads, wherein each candidate sequencing read includes at least a portion of a sequence corresponding to the polynucleotide and at least a portion of a sequence corresponding to the at least one oligonucleotide tag that is appended to the polynucleotide, wherein the candidate sequencing reads are stored in a memory in communication with a processor;

d) grouping the candidate sequencing reads into families of grouped candidate sequencing reads having a common tag sequence that is unique to a given family of candidate sequencing reads;

e) removing mistagged sequencing reads from the families of candidate sequencing reads to produce error-corrected families of sequencing reads, wherein a mistagged sequencing read has the common tag sequence for the given family but corresponds to a different polynucleotide than other candidate sequencing reads in the given family due to errors in the appending of step a); and

f) detecting a variant in a plurality of error-corrected families of sequencing reads.

2. The method of claim 1 , further comprising identifying an error in one or more of the candidate sequencing reads.

3. The method of claim 2 , wherein the step of identifying an error further includes applying a culling threshold to a number of nucleotides that differ between the candidate sequencing read and a reference sequence to identify a candidate sequencing read having an error.

4. The method of claim 3 , wherein the reference sequence is a tag-specific reference sequence.

5. The method of claim 3 , wherein the reference sequence is a polynucleotide-specific reference sequence.

6. The method of claim 1 , wherein the removing mistagged sequencing reads of step e) further includes comparing the candidate sequencing read in the given family to a polynucleotide-specific reference sequence to determine a number of nucleotides that differ between the candidate sequencing read and the polynucleotide-specific reference sequence.

7. The method of claim 6 , wherein the removing mistagged sequencing reads of step e) further includes applying a difference counting threshold to identify a mistagged sequencing read.

8. The method of claim 1 , wherein the removing mistagged sequencing reads of step e) includes comparing the candidate sequencing read to one or more other candidate sequencing reads in the given family to identify candidate sequencing reads having a common pattern of variants.

9. The method of claim 8 , wherein the removing mistagged sequencing reads of step e) further includes applying a pattern counting threshold to a number of candidate sequencing reads having the common pattern of variants to identify a group of mistagged sequencing reads.

10. The method of claim 1 , wherein the removing mistagged sequencing reads of step e) includes comparing the candidate sequencing reads in the given family to a polynucleotide-specific reference sequence to identify a candidate mistagged sequencing read.

11. The method of claim 10 , wherein the removing mistagged sequencing reads of step e) further includes comparing the candidate mistagged sequencing read to one or more other candidate mistagged sequencing reads in the given family to identify a common pattern of variants.

12. The method of claim 11 , wherein the removing mistagged sequencing reads of step e) further includes applying a pattern counting threshold to a number of candidate mistagged sequencing reads having the common pattern of variants to determine a group of mistagged sequencing reads.

13. The method of claim 1 , wherein the removing mistagged sequencing reads of step e) includes comparing the candidate sequencing read in the given family to a polynucleotide-specific reference sequence to identify a pattern of differences in a candidate mistagged sequencing read.

14. The method of claim 13 , wherein the removing mistagged sequencing reads of step e) further includes determining a number of matches for the pattern of differences in the candidate mistagged sequencing read compared to a pattern of expected differences between the polynucleotide-specific reference sequence and an expected sequence for a non-target polynucleotide.

15. The method of claim 14 , wherein the removing mistagged sequencing reads of step e) further includes applying a non-target pattern threshold to the number of matches to identify a mistagged sequencing read.

16. The method of claim 1 , wherein the detecting of step f) includes aligning the sequencing reads for the error-corrected family to a polynucleotide-specific reference sequence.

17. The method of claim 16 , wherein the detecting of step f) further includes counting a number of aligned sequencing reads having a particular base difference at a given position with respect to the polynucleotide-specific reference sequence.

18. The method of claim 17 , wherein the detecting of step f) further includes applying a family level threshold to the number of aligned sequencing reads to identify a family-based candidate variant.

19. The method of claim 1 , the detecting of step f) further includes counting a number of error-corrected families having a particular family-based candidate variant.

20. The method of claim 19 , the detecting of step f) further includes applying a multi-family threshold to the number of error-corrected families to identify the variant.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 29, 2021
From: MONGAN, ANN; CHIEN, RICHARD; BRINZA, DUMITRU; BRAMLETT, KELLI; DHINGRA, DALIA; GU, JIAN
To: LIFE TECHNOLOGIES CORPORATION
Reel/Frame 057026/0443 →
Continuity (10)
Continuation 16503262 · Jul 3, 2019
Continuation 15178450 · Jun 9, 2016
Provisional Application 62323142 · Apr 15, 2016
Provisional Application 62311276 · Mar 21, 2016
Provisional Application 62310647 · Mar 18, 2016
Provisional Application 62304530 · Mar 7, 2016
Provisional Application 62248978 · Oct 30, 2015
Provisional Application 62207177 · Aug 19, 2015
Provisional Application 62172836 · Jun 9, 2015
Related Publication 20210363596A1 · Nov 25, 2021
References Cited (103)
US 5333675A · Mullis et al. · 1994 [cited by applicant]
US 5508169A · Deugau et al. · 1996 [cited by applicant]
US 5635400A · Brenner · 1997 [cited by applicant]
US 5846719A · Brenner et al. · 1998 [cited by applicant]
US 5863722A · Brenner · 1999 [cited by applicant]
US 5882856A · Shuber · 1999 [cited by applicant]
US 6172214B1 · Brenner · 2001 [cited by applicant]
US 6207372B1 · Shuber · 2001 [cited by applicant]
US 6521428B1 · Senapathy · 2003 [cited by applicant]
US 6607878B2 · Sorge · 2003 [cited by applicant]
US 6890741B2 · Fan et al. · 2005 [cited by applicant]
US 7393665B2 · Brenner · 2008 [cited by applicant]
US 7537897B2 · Brenner et al. · 2009 [cited by applicant]
US 7579154B2 · Chun · 2009 [cited by applicant]
US 7803550B2 · Makarov et al. · 2010 [cited by applicant]
US 7948015B2 · Rothberg et al. · 2011 [cited by applicant]
US 8053192B2 · Bignell et al. · 2011 [cited by applicant]
US 8148068B2 · Brenner · 2012 [cited by applicant]
US 8168385B2 · Brenner · 2012 [cited by applicant]
US 8182989B2 · Bignell et al. · 2012 [cited by applicant]
US 8318433B2 · Brenner · 2012 [cited by applicant]
US 8470996B2 · Brenner · 2013 [cited by applicant]
US 8476018B2 · Brenner · 2013 [cited by applicant]
US 8481292B2 · Casbon et al. · 2013 [cited by applicant]
US 8666678B2 · Davey et al. · 2014 [cited by applicant]
US 8685678B2 · Casbon et al. · 2014 [cited by applicant]
US 8685889B2 · Van et al. · 2014 [cited by applicant]
US 8715967B2 · Casbon et al. · 2014 [cited by applicant]
US 8722368B2 · Casbon et al. · 2014 [cited by applicant]
US 8728766B2 · Casbon et al. · 2014 [cited by applicant]
US 8741606B2 · Casbon et al. · 2014 [cited by applicant]
US 8822150B2 · Bignell et al. · 2014 [cited by applicant]
US 8835358B2 · Fodor et al. · 2014 [cited by applicant]
US 8865410B2 · Shendure et al. · 2014 [cited by applicant]
US 9018365B2 · Brenner · 2015 [cited by applicant]
US 9085798B2 · Chee · 2015 [cited by applicant]
US 9194001B2 · Brenner · 2015 [cited by applicant]
US 9260753B2 · Xie et al. · 2016 [cited by applicant]
US 9315857B2 · Fu et al. · 2016 [cited by applicant]
US 9340830B2 · Downing et al. · 2016 [cited by applicant]
US 10344336B2 · Bramlett et al. · 2019 [cited by applicant]
US 20040185484A1 · Costa et al. · 2004 [cited by applicant]
US 20060263789A1 · Kincaid et al. · 2006 [cited by applicant]
US 20070020640A1 · McCloskey et al. · 2007 [cited by applicant]
US 20080261204A1 · Lexow · 2008 [cited by applicant]
US 20090318310A1 · Liu et al. · 2009 [cited by applicant]
US 20120283110A1 · Shendure et al. · 2012 [cited by applicant]
US 20130060482A1 · Sikora et al. · 2013 [cited by applicant]
US 20130090860A1 · Sikora et al. · 2013 [cited by applicant]
US 20130302801A1 · Asbury et al. · 2013 [cited by applicant]
US 20140066317A1 · Talasaz · 2014 [cited by examiner]
US 20140227705A1 · Vogelstein et al. · 2014 [cited by applicant]
US 20140255929A1 · Zheng · 2014 [cited by applicant]
US 20140274731A1 · Raymond et al. · 2014 [cited by applicant]
US 20140357499A1 · Gordon et al. · 2014 [cited by applicant]
US 20150044687A1 · Schmitt et al. · 2015 [cited by applicant]
US 20150087535A1 · Patel · 2015 [cited by applicant]
US 20150197798A1 · Xu et al. · 2015 [cited by applicant]
US 20150299812A1 · Talasaz · 2015 [cited by applicant]
US 20150322507A1 · Zimmermann et al. · 2015 [cited by applicant]
US 20160017320A1 · Wang et al. · 2016 [cited by applicant]
US 20160040229A1 · Talasaz et al. · 2016 [cited by applicant]
US 20160046986A1 · Eltoukhy et al. · 2016 [cited by applicant]
US 20160115532A1 · Faham · 2016 [cited by applicant]
CN 103045726A · 2013 [cited by applicant]
CN 103748236A · 2014 [cited by applicant]
CN 104364392A · 2015 [cited by applicant]
WO WO03050304A1 · 2003 [cited by applicant]
WO WO2007037678A2 · 2007 [cited by applicant]
WO WO2012142213A2 · 2012 [cited by applicant]
WO WO2013130512A2 · 2013 [cited by applicant]
WO WO2013142389A1 · 2013 [cited by applicant]
WO WO2013181170A1 · 2013 [cited by applicant]
WO WO2014039556A1 · 2014 [cited by applicant]
WO WO2014130890A1 · 2014 [cited by applicant]
WO WO2015100427A1 · 2015 [cited by applicant]
Binladen, Jonas et al., “The Use of Coded PCR Primers Enables High-Throughput Sequencing of Multiple Homolog Amplification Products by 454 Parallel Sequencing”, PLoS ONE, 2(2):, 2007, e197. [cited by applicant]
Brenner, Sydney et al., “In vitro cloning of complex mixtures of DNA on microbeads: Physical separation of differentially expressed cDNAs”, PNAS vol. 97 No.4 2000, 1665-1670. [cited by applicant]
Brinza, Dumitru et al., “Abstract 2402: A research approach for the detection of somatic mutations at 0.5% frequency from cfDNA and eTc DNA using a multiplex sequencing assay targeting 2000 tumor mutations”, Cancer Rese… [cited by applicant]
Church, George M. et al., “Multiplex DNA sequencing”, Science vol. 240, 1988, pp. 185-188. [cited by applicant]
Craig, et al. Identification of genetic variants using bar-coded multiplexed sequencing. Nat Methods. Oct. 2008;5(10):887-93. doi: 10.1038/nmeth.1251. Epub Sep. 14, 2008. [cited by applicant]
Cronn, Richard et al., “Multiplex sequencing of plant chloroplast genomes using Solexa sequencing-by-synthesis technology”, Nucleic Acids Research, vol. 36, No. 19, e122; doi:10.1093/nar/qkn502, 2008, 1-11. [cited by applicant]
Delhomme, Tiffany, “Needlestack an highly scalable and reproducible pipeline for the detection of ctDNA variants”, International Agency for Research on Cancer IARC (WHO), Jun. 27, 2015, 1-31. [cited by applicant]
Eason, et al., “Characterization of synthetic DNA bar codes in [cited by applicant]
EP19196414.7, Extended Search Report, Mar. 17, 2020, 6 pages. [cited by applicant]
Fernandez-Cuesta, Lynnette et al., “Identification of Circulating Tumor DNA for the Early Detection of Small-cell Lung Cancer”, EBioMedicine, 10, 2016, 117-123. [cited by applicant]
Gray et al., “Selection of Therapeutic H5N1 Monoclonal Antibodies Following IgVH Repertoire Analysis in Mice”, Antiviral Research, vol. 131, Apr. 21, 2016, pp. 100-108, XP029570500. [cited by applicant]
Harper, Diane et al., “Efficacy of a bivalent L1 virus-like particle vaccine in prevention of infection with human papillomavirus types 16 and 18 in young women: a randomised controlled trial”, The Lancet, 364, 2004, 17… [cited by applicant]
Hoffmann, Christian et al., “DNA bar coding and pyrosequencing to identify rare HIV drug resistance mutations”, Nucleic Acids Research, vol. 35, No. 13 e91; doi:10.1093/nar/gkm435, 2007, 1-8. [cited by applicant]
Hug, Hubert et al., “Measurement of the No. of Molecules of a Single mRNA Species in a Complex mRNA Preparation”, Journal of Theoretical Biolog, 221, doi:10.1006/jtbi.2003.3211, 2003, 615-624. [cited by applicant]
Kanagai-Shamanna R, et al., “Next-Generation Sequencing-Based Multi-Gene Mutation Profiling of Solid Tumors Using Fine Needle Aspiration Samples: Promises and Challenges for Routine Clinical Diagnostics”, Modern Patholo… [cited by applicant]
Kennedy et al., “Detecting ultralow-frequency mutations by Duplex Sequencing”, Nature Protocols, 2014, vol. 9, No. 11, pp. 2586-2606. [cited by applicant]
Kinde, Isaac et al., “Detection and quantification of rare mutations with massively parallel sequencing”, Proceedings of the National Academy of Science, vol. 108, No. 23, 2011, 9530-9535. [cited by applicant]
Miner, Brooks et al., “Molecular barcodes detect redundancy and contamination in hairpinbisulfite PCR”, Nucleic Acids Research, vol. 32, No. 17 e135, doi:10.1093/nar/gnh132, 2004, 1-4. [cited by applicant]
Morgenstern B et al., “Multiple Sequence Alignment with User-Defined Anchor Points,” Algorithms for Molecular Biology, Apr. 19, 2006, vol. 1, No. 6, 12 pages. [cited by applicant]
Parameswaran, Poornima et al., “A pyrosequencing-tailored nucleotide barcode design unveils opportunities for large-scale sample multiplexing”, Nucleic Acids Research, vol. 35, No. 19 e130, doi: 10.1093/nar/gkm760, 2007… [cited by applicant]
PCT/US2016/036763, International Search Report and Written Opinion mailed Aug. 12, 2016, 15 pages. [cited by applicant]
Peng, Quan et al., “Reducing amplification artifacts in high multiplex amplicon sequencing by using molecular barcodes”, BMC Genomics, 16:589, DOI 10.1186/s12864-015-1806-8, 2015, 1-12. [cited by applicant]
Pitschi F et al., “Automatic Detection of Anchor Points for Multiple Sequence Alignment,” BMC Bioinformatics, Sep. 2, 2010, vol. 11, No. 445, 11 pages. [cited by applicant]
Rothberg et al., “An integrated semiconductor device enabling non-optical genome sequencing”, Nature, vol. 475, No. 7356, Jul. 21, 2011, pp. 348-352. [cited by applicant]
Schmitt, Michael et al., “Detection of ultra-rare mutations by next-generation sequencing”, Proceedings of the National Academy of Science, vol. 109, No. 36, 2012, 14508-14513. [cited by applicant]
Spencer, D.H. et al., J. Mol. Diagn., vol. 16, pp. 75-88 (2014). [cited by applicant]
Stiller, Mathias et al., “Direct multiplex sequencing (DMPS)—a novel method for targeted high-throughput sequencing of ancient and highly degraded DNA”, Genome Research, vol. 19, 2009, 1843-1848. [cited by applicant]