IP Library Patent Application 17698900
Patent Application
App. No. 17/698,900

DETERMINATION OF COPY NUMBER VARIATIONS USING BINOMIAL PROBABILITY CALCULATIONS

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
17/698,900
Abstract

This invention relates to a binomial calculation of copy number of data obtained from a mixed sample having a first source and a second source.

Claims (31)

1 . (canceled)

2 . (canceled)

3 . A computer-implemented process for calculating copy number variation (CNV) of one or more genomic regions a single source in a mixed sample, wherein at least one processor coupled to a memory executes a software component that performs the process, the process comprising:

accessing by the software component a first data set comprising frequency data for one or more informative loci from a maternal source in the mixed sample;

accessing by the software component a second data set comprising frequency data for one or more informative loci from a fetal source in the mixed sample;

calculating by the software component an estimated fetal source contribution of cell free nucleic based on a binomial distribution of the counts of distinguishing regions from first and second data sets; and

accessing by the software component a third data set comprising frequency data for two or more genomic regions from the combined maternal and fetal sources in the mixed sample; and

calculating by the software component the presence or absence of a CNV for one or more genomic regions in the fetus by comparison of the frequency data from the two or more genomic regions from the combined maternal and fetal sources in the mixed sample and the estimated fetal source contribution of cell free nucleic acids in the mixed sample.

4 . The process of claim 3 , wherein the CNV is calculated based on empirical frequency data for the two or more genomic regions from the combined first and second sources in the mixed sample.

5 . An executable software product stored on a non-transitory computer-readable medium containing program instructions, which when executed by a computer directs performance of steps for estimating copy number variation (CNV) of one or more genomic regions in a mixed sample, the steps comprising:

accessing by the software component a first data set comprising frequency data based on identification of distinguishing regions from copies of one or more informative loci from a first source;

accessing by the software component a second data set comprising frequency data based on identification of distinguishing regions from copies of one or more informative loci from a second source;

calculating by the software component an estimated source contribution of cell free nucleic acids based on a binomial distribution of the first and second data sets;

accessing by the software component a third data set comprising frequency data for two or more genomic regions in the first source and the second source; and

calculating by the software component the estimation of a CNV for one or more genomic regions by comparison of frequency data for two or more genomic regions in the first source and the second source and the estimated source contribution of cell free nucleic acids from the at least one of the first source and the second source.

6 . The process of claim 5 , wherein the CNV is calculated based on empirical frequency data for the two or more genomic regions from the combined first and second sources in the mixed sample.

7 . A system, comprising:

a memory;

a processor coupled to the memory; and

a software component executed by the processor that is configured to:

access a first data set comprising frequency data based on identification of distinguishing regions from copies of one or more informative loci from a first source in a mixed sample;

access a second data set comprising frequency data based on identification of distinguishing regions from copies of one or more informative loci from a second source in the mixed sample;

calculate an estimated contribution of cell free nucleic acids from at least one of the first source and the second source based on a binomial distribution of counts of the distinguishing regions from the first and second data sets;

access a third data set comprising frequency data for two or more genomic regions of the first and second sources in a mixed sample; and

calculate a copy number variation for one or more genomic regions of a single source in the mixed sample by comparison of the frequency data for one or more genomic regions and the estimated source contribution of cell free nucleic acids in the mixed sample.

8 . A computer software product including a non-transitory computer-readable storage medium having fixed therein a sequence of instructions which when executed by a computer directs performance of steps of:

creating a first data set representing a quantity of informative loci from a first source in a mixed sample;

creating a second data set representing a quantity of informative loci from a second source in the mixed sample;

calculating an estimated source contribution of cell free nucleic acids from the first source and the second source in the mixed sample based on a binomial distribution of the quantities of informative loci from the first and second data sets;

accessing a third data set comprising frequency data for two or more genomic regions of the first and second sources in a mixed sample; and

calculating a copy number variation for one or more genomic regions of a single source in the mixed sample by comparison of the frequency data for two or more genomic regions and the estimated source contribution of cell free nucleic acids in the mixed sample.