IP Library › Granted Patent US 12,626,778
Granted Patent B2
US 12,626,778 · App. 18/134,472 · Granted May 12, 2026

Accelerated hidden Markov models for genotype analysis

Inventors: Keith Daniel Noto (San Francisco, CA); James Parker Ferry (Cedar Hills, UT); Bryan Joseph Johnson (American Fork, UT); Alisa Sedghifar (San Francisco, CA); Yong Wang (San Mateo, CA); Shiya Song (San Mateo, CA); Jeffrey Adrion (Salt Lake City, UT)
Assignee: Ancestry.com DNA, LLC
G16B20/00G06N7/01G16B20/40G16B40/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,626,778
App. No.
18/134,472
Granted
May 12, 2026
Kind
B2
Abstract

Disclosed is a configuration for determining a genotyping label composition of a target individual using direct acyclic paths. The configuration includes receiving a phased genotype of the target individual, including a first haplotype and a second haplotype. The configuration initiates a full-ethnicity hidden Markov model (HMM) including nodes with a set of ethnicity labels. The first haplotype is input to determine a first subset of ethnicity labels that match the first haplotype. The second haplotype is input to determine a second subset of ethnicity labels that match the second haplotype. The first and second subsets of ethnicity labels are combined to create a candidate subset of ethnicity labels for the target individual. The configuration initiates a simplified HMM with nodes from the candidate subset of ethnicity labels. The phased genotype of the target individual is input to the simplified HMM to determine genotyping label composition of the target individual.

Claims (62)

1 . A computer-implemented method comprising:

receiving a phased genotype of a target individual, the phased genotype comprising a first haplotype and a second haplotype;

initiating a full-ethnicity hidden Markov model (HMM), the full-ethnicity HMM comprising nodes that have a set of ethnicity labels;

inputting the first haplotype to the full-ethnicity HMM to determine a first subset of ethnicity labels that match the first haplotype;

inputting the second haplotype to the full-ethnicity HMM to determine a second subset of ethnicity labels that match the second haplotype;

combining the first and second subsets of ethnicity labels as a candidate subset of ethnicity labels of the target individual;

initiating a simplified HMM specific to the target individual, the simplified HMM comprising nodes that are simplified from the set of ethnicity labels to the candidate subset of ethnicity labels of the target individual;

inputting the phased genotype of the target individual to the simplified HMM; and

determining an ethnicity composition of the target individual using the simplified HMM.

2 . The computer-implemented method of claim 1 , wherein the nodes in the simplified HMM represent permutations of different first parent ethnicity labels, second parent ethnicity labels, and switch labels.

3 . The computer-implemented method of claim 2 , wherein the switch labels represent a phasing error, the phasing error representative of switching the first and second parent ethnicity labels from one node group to a next node group.

4 . The computer-implemented method of claim 1 , wherein the nodes in the full-ethnicity HMM each represent a haplotype ethnicity from the set of ethnicity labels.

5 . The computer-implemented method of claim 1 , wherein the phased genotype comprises cross-chromosome haplotypes, and wherein the first haplotype and the second haplotype both include a sequence that has span of genetic loci in a plurality of chromosomes.

6 . The computer-implemented method of claim 1 , wherein receiving the phased genotype further comprises:

dividing the phased genotype into a plurality of windows, each window comprising a set of single nucleotide polymorphisms (SNPs).

7 . The computer-implemented method of claim 6 , wherein determining the ethnicity composition further comprises:

determining a path between the nodes in each window of the simplified HMM based on a likelihood of the phased genotype of the target individual traversing nodes along the path;

counting a number of a particular label corresponding to a particular ethnicity label in the path; and

determining an ethnicity composition of the target individual with respect to the particular ethnicity label based on the number of the particular label counted in the path.

8 . The computer-implemented method of claim 7 , wherein determining the likelihood further comprises:

determining a label probability, a label switch probability, and a transition probability, the transition probability associated with a particular edge in the path and representing a likelihood of the first node connected by the path from one window transitioning to the second node connected by the path from another window;

connecting the nodes with edges, each edge corresponding to a determined transition probability.

9 . The computer-implemented method of claim 6 , wherein determining the ethnicity composition further comprises:

displaying the likelihood of the target individual having the particular ethnic origin.

10 . The computer-implemented method of claim 7 , wherein displaying the likelihood of the target individual having the particular ethnic origin further comprises:

determining a minimum label threshold value;

filtering ethnicity labels below the determined minimum label threshold value;

sorting the filtered ethnicity labels in descending order;

determining a delta between the sum of the filtered ethnicity labels and a total number of ethnicity label; and

adding a predetermined value to one of the filtered ethnicity labels at a time until the delta is zero.

11 . The computer-implemented method of claim 1 , wherein the phased genotype is generated with a global phasing algorithm.

12 . The computer-implemented method of claim 1 , wherein the full-ethnicity HMM transition probabilities are determined by reference panels, the reference panel representative of a collection of genotypes from individuals with known ethnicities.

13 . The computer-implemented method of claim 12 , wherein one of the reference panels is an admixed panel, the admixed panels including genetic segments inherited from multiple ethnic origins.

14 . A non-transitory computer readable medium storing computer code comprising instructions that, when executed by one or more processors, causing the one or more processors to perform steps comprising:

receiving a phased genotype of a target individual, the phased genotype comprising a first haplotype and a second haplotype;

initiating a full-ethnicity hidden Markov model (HMM), the full-ethnicity HMM comprising nodes that have a set of ethnicity labels;

inputting the first haplotype to the full-ethnicity HMM to determine a first subset of ethnicity labels that match the first haplotype;

inputting the second haplotype to the full-ethnicity HMM to determine a second subset of ethnicity labels that match the second haplotype;

combining the first and second subsets of ethnicity labels as a candidate subset of ethnicity labels of the target individual;

initiating a simplified HMM specific to the target individual, the simplified HMM comprising nodes that are simplified from the set of ethnicity labels to the candidate subset of ethnicity labels of the target individual;

inputting the phased genotype of the target individual to the simplified HMM; and

determining an ethnicity composition of the target individual using the simplified HMM.

15 . The non-transitory computer readable medium of claim 14 , wherein the nodes in the simplified HMM represent permutations of different first parent ethnicity labels, second parent ethnicity labels, and switch labels.

16 . The non-transitory computer readable medium of claim 15 , wherein the switch labels represent a phasing error, the phasing error representative of switching the first and second parent ethnicity labels from one node group to a next node group.

17 . The non-transitory computer readable medium of claim 14 , wherein the phased genotype comprises cross-chromosome haplotypes, and wherein the first haplotype and the second haplotype both include a sequence that has span of genetic loci in a plurality of chromosomes.

18 . The non-transitory computer readable medium of claim 14 , wherein receiving the phased genotype further comprises:

dividing the phased genotype into a plurality of windows, each window comprising a set of single nucleotide polymorphisms (SNPs).

19 . The non-transitory computer readable medium of claim 18 , wherein determining the ethnicity composition further comprises:

determining a path between the nodes in each window of the simplified HMM based on a likelihood of the phased genotype of the target individual traversing nodes along the path;

counting a number of a particular label corresponding to a particular ethnicity label in the path; and

determining an ethnicity composition of the target individual with respect to the particular ethnicity label based on the number of the particular label counted in the path.

20 . A system comprising:

one or more processors; and

a memory configured to store computer code comprising instructions, the instructions, when executed by one or more processors, cause the one or more processors to perform steps comprising:

receiving a phased genotype of a target individual, the phased genotype comprising a first haplotype and a second haplotype;

initiating a full-ethnicity hidden Markov model (HMM), the full-ethnicity HMM comprising nodes that have a set of ethnicity labels;

inputting the first haplotype to the full-ethnicity HMM to determine a first subset of ethnicity labels that match the first haplotype;

inputting the second haplotype to the full-ethnicity HMM to determine a second subset of ethnicity labels that match the second haplotype;

combining the first and second subsets of ethnicity labels as a candidate subset of ethnicity labels of the target individual;

initiating a simplified HMM specific to the target individual, the simplified HMM comprising nodes that are simplified from the set of ethnicity labels to the candidate subset of ethnicity labels of the target individual;

inputting the phased genotype of the target individual to the simplified HMM; and

determining an ethnicity composition of the target individual using the simplified HMM.

Assignments (3)
PATENT SECURITY AGREEMENT Recorded Aug 3, 2026
From: ANCESTRY.COM OPERATIONS INC.; ANCESTRY.COM DNA, LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS NOTES COLLATERAL AGENT
Reel/Frame 076116/0447 →
PATENT SECURITY AGREEMENT Recorded Aug 3, 2026
From: ANCESTRY.COM OPERATIONS INC.; ANCESTRY.COM DNA, LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS NOTES COLLATERAL AGENT
Reel/Frame 076144/0726 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 10, 2023
From: NOTO, KEITH DANIEL; FERRY, JAMES PARKER; JOHNSON, BRYAN JOSEPH; SEDGHIFAR, ALISA; WANG, YONG; SONG, SHIYA; ADRION, JEFFREY
To: ANCESTRY.COM DNA, LLC
Reel/Frame 064555/0447 →
Continuity (2)
Provisional Application 63330538 · Apr 13, 2022
Related Publication 20230335217A1 · Oct 19, 2023
References Cited (154)
US 8510057B1 · Avey et al. · 2013 [cited by applicant]
US 9213944B1 · Do et al. · 2015 [cited by applicant]
US 9213947B1 · Do et al. · 2015 [cited by applicant]
US 9367800B1 · Do et al. · 2016 [cited by applicant]
US 9836576B1 · Do et al. · 2017 [cited by applicant]
US 9910962B1 · Fakhrai-Rad et al. · 2018 [cited by applicant]
US 9940433B2 · Han et al. · 2018 [cited by applicant]
US 9977708B1 · Do · 2018 [cited by examiner]
US 10114922B2 · Byrnes et al. · 2018 [cited by applicant]
US 10223498B2 · Han et al. · 2019 [cited by applicant]
US 10558930B2 · Noto et al. · 2020 [cited by applicant]
US 10692587B2 · Song et al. · 2020 [cited by applicant]
US 10720229B2 · Barber et al. · 2020 [cited by applicant]
US 11211149B2 · Curtis et al. · 2021 [cited by applicant]
US 11232854B2 · Anderson et al. · 2022 [cited by applicant]
US 20020143578A1 · Cole et al. · 2002 [cited by applicant]
US 20020156596A1 · Caruso et al. · 2002 [cited by applicant]
US 20030113727A1 · Girn et al. · 2003 [cited by applicant]
US 20040267458A1 · Judson et al. · 2004 [cited by applicant]
US 20050025508A1 · Karakama et al. · 2005 [cited by applicant]
US 20050255508A1 · Casey et al. · 2005 [cited by applicant]
US 20080154566A1 · Myres et al. · 2008 [cited by applicant]
US 20080228043A1 · Kenedy et al. · 2008 [cited by applicant]
US 20080255768A1 · Martin · 2008 [cited by examiner]
US 20100256917A1 · McVean et al. · 2010 [cited by applicant]
US 20130085728A1 · Tang · 2013 [cited by examiner]
US 20130149707A1 · Sorenson et al. · 2013 [cited by applicant]
US 20130163860A1 · Suzuki et al. · 2013 [cited by applicant]
US 20130297221A1 · Johnson et al. · 2013 [cited by applicant]
US 20140045705A1 · Bustamante · 2014 [cited by examiner]
US 20140045708A1 · Feng et al. · 2014 [cited by applicant]
US 20140067355A1 · Noto et al. · 2014 [cited by applicant]
US 20140108527A1 · Aravanis et al. · 2014 [cited by applicant]
US 20140194300A1 · Song et al. · 2014 [cited by applicant]
US 20150106115A1 · Hu et al. · 2015 [cited by applicant]
US 20160350479A1 · Han et al. · 2016 [cited by applicant]
US 20170017752A1 · Noto et al. · 2017 [cited by applicant]
US 20170062577A1 · Brewer et al. · 2017 [cited by applicant]
US 20170220738A1 · Barber et al. · 2017 [cited by applicant]
US 20170262577A1 · Ball et al. · 2017 [cited by applicant]
US 20170329904A1 · Naughton et al. · 2017 [cited by applicant]
US 20180329033A1 · Pratt et al. · 2018 [cited by applicant]
US 20180330824A1 · Athey et al. · 2018 [cited by applicant]
US 20190114219A1 · Do et al. · 2019 [cited by applicant]
US 20200005899A1 · Nicula et al. · 2020 [cited by applicant]
US 20200082909A1 · Wang · 2020 [cited by examiner]
US 20200160202A1 · Noto et al. · 2020 [cited by applicant]
US 20200286579A1 · Song et al. · 2020 [cited by applicant]
US 20210134387A1 · McMaster-Schraiber et al. · 2021 [cited by applicant]
US 20220051751A1 · Wilton · 2022 [cited by examiner]
CN 106846029A · 2017 [cited by applicant]
WO WO2014151088A2 · 2014 [cited by applicant]
WO WO2015051006A2 · 2015 [cited by applicant]
WO WO2016061568A1 · 2016 [cited by applicant]
WO WO2018129413A1 · 2018 [cited by applicant]
WO WO2020053789A1 · 2020 [cited by applicant]
23andMe. “Ancestry Composition: 23andMe's State-of-the-Art Geographic Ancestry Analysis.” 23andme.com, Sep. 5, 2015, 10 pages, [Online] [Retrieved Jan. 26, 2024], Retrieved from the Internet Archive <URL:https://web.arc… [cited by applicant]
Alexander, D. H. et al. “Fast Model-Based Estimation of Ancestry in Unrelated Individuals.” Genome Research, vol. 19, Jul. 31, 2009, pp. 1655-1664. [cited by applicant]
Ball, C. A. et al. “AncestryDNA Matching White Paper: Discovering Genetic Matches Across a Massive, Expanding Genetic Database.” AncestryDNA, Mar. 31, 2016, pp. 1-46. [cited by applicant]
Baran, Y. et al. “Fast and Accurate Inference of Local Ancestry in Latino Populations.” Bioinformatics, vol. 28, No. 10, May 2012, pp. 1359-1367. [cited by applicant]
Bastian, M. et al. “Gephi: An Open Source Software for Exploring and Manipulating Networks.” Proceedings of the Third International AAAI Conference on Weblogs and Social Media, vol. 3, No. 1, Mar. 19, 2009, pp. 361-362. [cited by applicant]
Bercovici, S. et al. “Ancestry Inference in Complex Admixtures via Variable-Length Markov Chain Linkage Models.” Proceedings of the 16th Annual International Conference on Research in Computational Molecular Biology, Ap… [cited by applicant]
Brisbin, A. et al. “PCAdmix: Principal Components-Based Assignment of Ancestry along Each Chromosome in Individuals with Admixed Ancestry from Two or More Populations.” Human Biology, vol. 84, No. 4, Aug. 2012, pp. 343-… [cited by applicant]
Brooks, R. R. et al. “Behavior Detection Using Confidence Intervals of Hidden Markov Models.” IEEE Transactions on Systems, Man, and Cybernetics—Part B: Cybernetics, vol. 39, No. 6, Dec. 2009, pp. 1484-1492. [cited by applicant]
Browning, B. L. et al. “A Fast, Powerful Method for Detecting Identity by Descent,” The American Journal of Human Genetics, Feb. 11, 2011, vol. 88, No. 2, pp. 173-182. [cited by applicant]
Browning, B. L. et al. “A Unified Approach to Genotype Imputation and Haplotype Phase Inference for Large Data sets of Trios and Unrelated Individuals,” The American Journal of Human Genetics, Feb. 13, 2009, pp. 210-223… [cited by applicant]
Browning, B. L. et al. “Detecting Identity by Descent and Estimating Genotype Error Rates in Sequence Data.” The American Journal of Human Genetics, vol. 93, Nov. 7, 2013, pp. 840-851. [cited by applicant]
Browning, B. L. et al. “Efficient Multilocus Association Testing for Whole Genome Association Studies Using Localized Haplotype Clustering.” Genetic Epidemiology, vol. 31, Feb. 26, 2007, pp. 365-375. [cited by applicant]
Browning, B. L. et al. “Genotype Imputation with Millions of Reference Samples.” The American Journal of Human Genetics, vol. 98, Jan. 7, 2016, pp. 116-126. [cited by applicant]
Browning, S. R. “Multilocus Association Mapping Using Variable-Length Markov Chains.” American Journal of Human Genetics, Jun. 2006, pp. 903-913, vol. 78, No. 6. [cited by applicant]
Browning, S. R. et al. “Haplotype Phasing: Existing Methods and New Developments.” Nature Reviews Genetics, Author Manuscript, vol. 12, No. 10, Oct. 2011, pp. 703-714. [cited by applicant]
Browning, S. R. et al. “Rapid and Accurate Haplotype Phasing and Missing-Data Inference for Whole-Genome Association Studies by Use of Localized Haplotype Clustering.” American Journal of Human Genetics, vol. 81, Nov. 2… [cited by applicant]
Cann, H. M. et al. “A Human Genome Diversity Cell Line Panel.” Science, vol. 296, No. 5566, Apr. 12, 2002, pp. 261-262. [cited by applicant]
Cavalli-Sforza, L. L. “The Human Genome Diversity Project: Past, Present and Future.” Nature Reviews Genetics, vol. 6, Apr. 1, 2005, pp. 333-340. [cited by applicant]
De Roos, A. P. W. “Genomic Selection in Dairy Cattle.” Dissertation, Wageningen University, Jan. 21, 2011, pp. 1-185. [cited by applicant]
Dilthey, A. et al. “Multi-Population Classical HLA Type Imputation.” PLoS Computational Biology, vol. 9, No. 2, Feb. 2013, pp. 1-13. [cited by applicant]
Dilthey, A. et al. “Multi-Population Classical HLA Type Imputation.” Supporting Text S1, PLoS Computational Biology, vol. 9, No. 2, Feb. 2013, pp. 1-17. [cited by applicant]
Durand, E. Y. et al. “Reducing Pervasive False-Positive Identical-by-Descent Segments Detected by Large-Scale Pedigree Analysis.” Molecular Biology and Evolution, vol. 31, No. 8, Aug. 2014, pp. 2212-2222. [cited by applicant]
Eronen, L. et al. “A Markov Chain Approach to Reconstruction of Long Haplotypes.” Pacific Symposium on Biocomputing, Jan. 1, 2004, pp. 1-12. [cited by applicant]
Falush, D. et al., “Inference of Population Structure Using Multilocus Genotype Data: Linked Loci and Correlated Allele Frequencies,” Genetics, vol. 164, Aug. 2003, pp. 1567-1587. [cited by applicant]
Ghahramani, Z. “An Introduction to Hidden Markov Models and Bayesian Networks.” International Journal of Pattern Recognition and Artificial Intelligence, vol. 15, No. 1, Jun. 2001, pp. 9-42. [cited by applicant]
Gravel, S. “Population Genetics Models of Local Ancestry.” Genetics, vol. 191, No. 2, Jun. 1, 2012, pp. 607-619. [cited by applicant]
Guan, Y. “Detecting Structure of Haplotypes and Local Ancestry.” Genetics, vol. 196, No. 3, Mar. 1, 2014, pp. 625-642. [cited by applicant]
Halperin, E. et al. “Haplotype Reconstruction from Genotype Data Using Imperfect Phylogeny.” Bioinformatics, vol. 20, No. 12, Aug. 2004, pp. 1842-1849. [cited by applicant]
Han, E. et al. “Clustering of 770,000 Genomes Reveals Post-Colonial Population Structure of North America.” Nature Communications, vol. 8, Feb. 7, 2017, pp. 1-12. [cited by applicant]
Harvard. “Plink . . . Whole Genome Association Analysis Toolset.” Harvard.edu, Sep. 9, 2019, 4 pages, [Online] [Retrieved Jan. 29, 2024], Retrieved from the Internet Archive <URL:https://web.archive.org/web/201909091948… [cited by applicant]
Hellenthal, G. et al. “A Genetic Atlas of Human Admixture History.” Science, vol. 343, No. 6172, Feb. 14, 2014, pp. 747-751. [cited by applicant]
Hoggart, C. J. et al. “Design and Analysis of Admixture Mapping Studies.” The American Journal of Human Genetics, vol. 74, No. 5, May 2004, pp. 965-978. [cited by applicant]
Horton, R. et al. “Variation Analysis and Gene Annotation of Eight MHC Haplotypes: The MHC Haplotype Project.” vol. 60, Jan. 10, 2008, pp. 1-18. [cited by applicant]
Howie, B. N. et al. “A Flexible and Accurate Genotype Imputation Method for the Next Generation of Genome-Wide Association Studies.” PLoS Genetics, vol. 5, No. 6, Jun. 2009, pp. 1-15. [cited by applicant]
Hug, N. “Surprise: A Python Scikit for Recommender Systems.” Surpriselib.com, Overview, Sep. 2019, 6 pages, [Online] [Retrieved Jan. 31, 2024], Retrieved from the Internet <URL:https://surpriselib.com/>. [cited by applicant]
Itan, Y. et al. “The Origins of Lactase Persistence in Europe.” PLoS Computational Biology, vol. 5, No. 8, Aug. 2009, pp. 1-13. [cited by applicant]
Jarvis, J.P. et al., “Patterns of Ancestry of Natural Selection and Genetic Association with Stature in Western African Pygmies,” PLoS Genetics, vol. 8, Iss. 4, Apr. 26, 2012, pp. 1-15. [cited by applicant]
Joshi, S. et al. “Identifiable Phenotyping using Constrained Non-Negative Matrix Factorization.” Proceedings of Machine Learning for Healthcare, vol. 56, Aug. 19-20, 2016, pp. 1-24. [cited by applicant]
Ke, X. et al. “Singleton SNPs in the Human Genome and Implications for Genome-Wide Association Studies.” European Journal of Human Genetics, vol. 16, Jan. 16, 2008, pp. 506-515. [cited by applicant]
Khodabandelou, G. et al. “Genome Functional Annotation Across Species Using Deep Convolutional Neural Networks.” bioRxiv: The Preprint Server for Biology, Jun. 7, 2019, pp. 1-11. [cited by applicant]
Krogh, A. S. “Hidden Markov Models for Labeled Sequences.” Proceedings of the 12 [cited by applicant]
Lawson, D. J. et al. “Inference of Population Structure Using Dense Haplotype Data.” PLoS Genetics, vol. 8, No. 1, Jan. 2012, pp. 1-16. [cited by applicant]
Li, J. et al. “Towards Unsupervised Gene Selection: A Matrix Factorization Framework.” IEEE/ACM Transactions on Computational Biology and Bioinformatics, vol. 14, No. 3, May-Jun. 2017, pp. 514-521. [cited by applicant]
Li, N. et al. “Modeling Linkage Disequilibrium and Identifying Recombination Hotspots Using Single-Nucleotide Polymorphism Data.” Genetics, vol. 165, No. 4, Dec. 1, 2003, pp. 2213-2233. [cited by applicant]
Li, Y. et al. “Genotype Imputation.” Annual Review of Genomics and Human Genetics, vol. 10, Sep. 2009, pp. 387-406. [cited by applicant]
Li, Y. et al. “MaCH: Using Sequence and Genotype Data to Estimate Haplotypes and Unobserved Genotypes.” Genetic Epidemiology, vol. 34, No. 8, Dec. 2010, pp. 816-834. [cited by applicant]
Liu, E. Y. et al. “MaCH-Admix: Genotype Imputation for Admixed Populations.” Genetic Epidemiology, vol. 37, No. 1, Jan. 2013, pp. 25-37. [cited by applicant]
Loh, P-R. et al. “Inferring Admixture Histories of Human Populations Using Linkage Disequilibrium.” Genetics, vol. 193, No. 4, Apr. 1, 2013, pp. 1233-1254. [cited by applicant]
Lu, C. et al. “A Normalized Statistical Metric Space for Hidden Markov Models.” IEEE Transactions on Cybernetics, vol. 43, No. 3, Jun. 2013, pp. 806-819. [cited by applicant]
Ma, P. et al. “Comparison of Different Methods for Imputing Genome-Wide Marker Genotypes in Swedish and Finnish Red Cattle.” Journal of Dairy Science, vol. 96, No. 7, Jul. 2013, pp. 4666-4677. [cited by applicant]
Ma, Y. et al. “Accurate Inference of Local Phased Ancestry of Modern Admixed Populations.” Scientific Reports, vol. 4, Jul. 23, 2014, pp. 1-5. [cited by applicant]
Maples, B. K. et al. “RFMix: A Discriminative Modeling Approach for Rapid and Robust Local-Ancestry Inference.” The American Journal of Human Genetics, vol. 93, Aug. 8, 2013, pp. 278-288. [cited by applicant]
McPeek, M. S. et al. “Assessment of Linkage Disequilibrium by the Decay of Haplotype Sharing, with Application to Fine-Scale Genetic Mapping.” American Journal of Human Genetics, vol. 65, No. 3, Sep. 1, 1999, pp. 858-87… [cited by applicant]
Montesinos-López, O. A. et al. “Prediction of Multiple-Trait and Multiple-Environment Genomic Data Using Recommender Systems.” G3 Genes/Genomes/Genetics, vol. 8, No. 1, Jan. 1, 2018, pp. 131-147. [cited by applicant]
Moreno-Estrada, A. et al. “Reconstructing the Population Genetic History of the Caribbean.” PLoS Genetics, vol. 9, No. 11, Nov. 2013, pp. 1-19. [cited by applicant]
Morrison, A.C. et al., “Prediction of Coronary Heart Disease Risk using a Genetic Risk Score: The Atherosclerosis Risk in Communities Study,” American Journal of Epidemiology, vol. 166, No. 1, Apr. 18, 2007, pp. 28-35. [cited by applicant]
Noto, K. et al. “Polly: A Novel Approach for Estimating Local and Global Admixture Proportion Based on Rich Haplotype Models.” ASHG 2015 Abstracts, Abstract 322, The American Society of Human Genetics 65th Annual Meetin… [cited by applicant]
Noto, K. et al. “Polly: A Novel Approach for Estimating Local and Global Admixture Proportion Based on Rich Haplotype Models.” Invited Talk at the American Society of Human Genetics (ASHG) Annual Meeting, Baltimore, MD,… [cited by applicant]
Noto, K. et al. “Underdog: A Fully-Supervised Phasing Algorithm that Learns from Hundreds of Thousands of Samples and Phases in Minutes.” ASHG 2014 Abstracts, Abstract 155, The American Society of Human Genetics 64 [cited by applicant]
Paşaniuc, B et al. “Imputation-Based Local Ancestry Inference in Admixed Populations.” International Symposium on Bioinformatics Research and Applications, May 13-16, 2009, pp. 1-13. [cited by applicant]
Paşaniuc, B. et al. “Inference of Locus-Specific Ancestry in Closely Related Populations.” Bioinformatics, vol. 25, No. 12, Jun. 2009, pp. i213-i221. [cited by applicant]
Patterson, N. et al. “Population Structure and Eigenanalysis.” PLoS Genetics, vol. 2, No. 12, Dec. 2006, pp. 2074-2093. [cited by applicant]
PCT International Search Report and Written Opinion, PCT Application No. PCT/US2023/018531, Jul. 21, 2023, 16 pages. [cited by applicant]
Peck, R. et al. “Introduction to Statistics and Data Analysis.” Section 7.4, 3 [cited by applicant]
Peck, R. et al. “Introduction to Statistics and Data Analysis.” Sections 9.2-9.3, 3 [cited by applicant]
Platt, J.C., “Probabilistic Outputs for Support Vector Machines and Comparisons to Regularized Likelihood Methods,” Mar. 26, 1999, pp. 1-11. [cited by applicant]
Price, A.L. et al., “Sensitive Detection of Chromosomal Segments of Distinct Ancestry in Admixed Populations,” PLoS Genetics, vol. 5, No. 6, Jun. 2009, pp. 1-18. [cited by applicant]
Pritchard, J. K. et al. “Inference of Population Structure Using Multilocus Genotype Data.” Genetics, vol. 155, No. 2, Jun. 1, 2000, pp. 945-959. [cited by applicant]
Purcell, S. et al., “PLINK: A tool set for whole-genome association and population-based linkage analyses,” The American Journal of Human Genetics, vol. 81, Sep. 2007, pp. 559-575. [cited by applicant]
Qian, Y. et al., “Efficient clustering of identity-by-descent between multiple individuals,” Bioinformatics, vol. 30, No. 7, Dec. 19, 2013, pp. 915-922. [cited by applicant]
Rabiner, L.R., “A Tutorial on Hidden Markov Models and Selected Applications in Speech Recognition,” Proceedings of the IEEE, vol. 77, No. 2, Feb. 1989, pp. 257-286. [cited by applicant]
Ranciaro, A. et al. “Genetic Origins of Lactase Persistence and the Spread of Pastoralism in Africa.” The American Journal of Human Genetics, vol. 94, Apr. 3, 2014, pp. 496-510. [cited by applicant]
Roach, J. C. et al. “Analysis of Genetic Inheritance in a Family Quartet by Whole-Genome Sequencing.” Science, vol. 328, No. 5978, Apr. 30, 2010, pp. 636-639. [cited by applicant]
Ron, D. et al., “On the Learnability and Usage of Acyclic Probabilistic Finite Automata,” Journal of Computer and System Sciences, vol. 56, 1998, pp. 133-152. [cited by applicant]
Sankararaman, S. et al. “Estimating Local Ancestry in Admixed Populations.” The American Journal of Human Genetics, vol. 82, Feb. 2008, pp. 290-303. [cited by applicant]
Scheet, P. et al., “A Fast and Flexible Statistical Model for Large-Scale Population Genotype Data: Applications to Inferring Missing Genotypes and Haplotypic Phase,” The American Journal of Human Genetics, vol. 78, Feb… [cited by applicant]
Seligsohn, U. et al., “Genetic Susceptibility to Venous Thrombosis,” The New England Journal of Medicine, vol. 344, No. 16, Apr. 19, 2001, pp. 1222-1231. [cited by applicant]
Silva, M. C. F. et al. “Development of Two Multiplex Mini-Sequencing Panels of Ancestry Informative SNPs for Studies in Latin Americans: An Application to Populations of the State of Minas Gerais (Brazil).” Genetics and… [cited by applicant]
Staples, J. et al., “PRIMUS: Rapid Reconstruction of Pedigrees from Genome-wide Estimates of Identity by Descent,” The American Journal of Human Genetics, vol. 95, Nov. 6, 2014, pp. 553-564. [cited by applicant]
Stephens, M. et al. “Accounting for Decay of Linkage Disequilibrium in Haplotype Inference and Missing-Data Imputation.” The American Journal of Human Genetics, vol. 76, No. 3, Mar. 2005, pp. 449-462. [cited by applicant]
Sturm, R. A. et al. “A Single SNP in an Evolutionary Conserved Region within Intron 86 of the HERC2 Gene Determines Human Blue-Brown Eye Color.” The American Journal of Human Genetics, vol. 82, Feb. 2008, pp. 424-431. [cited by applicant]
Sundquist, A. et al., “Effect of Genetic Divergence in Identifying Ancestral Origin using HAPAA,” Genome Res., vol. 18, Mar. 18, 2008, pp. 676-682. [cited by applicant]
Tang, H. et al. “Reconstructing Genetic Ancestry Blocks in Admixed Individuals.” The American Journal of Human Genetics, vol. 79, Jul. 2006, pp. 1-12. [cited by applicant]
Ter Braak, C. J. F. et al. “Identity-by-Descent Matrix Decomposition Using Latent Ancestral Allele Models.” Genetics, vol. 185, No. 3, Jul. 1, 2010, pp. 1045-1057. [cited by applicant]
THE 1000 Genomes Project Consortium. “A Global Reference for Human Genetic Variation.” Nature, vol. 526, No. 7571, Oct. 1, 2015, pp. 68-74. [cited by applicant]
The International Hapmap 3 Consortium, “Integrating common and rare genetic variation in diverse human populations,” Nature, vol. 467, Sep. 2, 2010, pp. 52-58. [cited by applicant]
The International Hapmap Consortium. “A Second Generation Human Haplotype Map of Over 3.1 Million SNPs.” Nature, Author Manuscript, Oct. 18, 2007, vol. 449, No. 7164, pp. 1-30. [cited by applicant]
The International Hapmap Consortium. “A Haplotype Map of the Human Genome.” Nature, vol. 437, Oct. 27, 2005, pp. 1299-1320. [cited by applicant]
Tipping, M.E., “Sparse Bayesian Learning and the Relevance Vector Machine,” Journal of Machine Learning Research, Jun. 2001, pp. 211-244. [cited by applicant]
Wang, Y. et al. “Ancestry Inference Using Reference Labeled Clusters of Haplotypes.” BMC Bioinformatics, vol. 22, Sep. 25, 2021, pp. 1-14. [cited by applicant]
Weedon, M.N. et al., “Combining Information from Common Type 2 Diabetes Risk Polymorphisms Improves Disease Prediction,” PLoS Med., vol. 3, Iss. 10, Oct. 2006, pp. 1877-1882. [cited by applicant]
Wikipedia. “Ethnicity.” Wikipedia: The Free Encyclopedia, Jul. 30, 2022, 12 pages, [Online] [Retrieved Sep. 27, 2023], Retrieved from the Internet <URL:https://en.wikipedia.org/wiki/Ethnicity>. [cited by applicant]
Wikipedia. “Inverse Distance Weighting.” Wikipedia: The Free Encyclopedia, Dec. 6, 2023, 4 pages, [Online] [Retrieved Jan. 30, 2024], Retrieved from the Internet <URL:https://en.wikipedia.org/wiki/Inverse_distance_weigh… [cited by applicant]
Williams, A. L. et al. “Phasing of Many Thousands of Genotyped Samples.” The American Journal of Human Genetics, vol. 91, Aug. 10, 2012, pp. 238-251. [cited by applicant]
Yang, Q. et al., “Improving the Prediction of Complex Diseases by Testing for Multiple Disease-Susceptibility Genes,” American Journal of Human Genetics, vol. 72, Feb. 14, 2003, pp. 636-649. [cited by applicant]
Yoon, B-J., “Hidden Markov Models and their Applications in Biological Sequence Analysis,” Current Genomics, vol. 10, Sep. 2009, pp. 402-415. [cited by applicant]
Zeng, X. et al. “Probability-Based Collaborative Filtering Model for Predicting Gene-Disease Associations.” BMC Medical Genomics, vol. 10, No. 76, Dec. 28, 2017, pp. 45-53. [cited by applicant]
Zhao, H. et al. “Haplotype Analysis in Population Genetics and Association Studies.” Pharmacogenomics, vol. 4, No. 2, Mar. 1, 2003, pp. 171-178. [cited by applicant]