IP Library Granted Patent US 7,933,912
Granted Patent B2
US 7,933,912 · App. 11/968,998 · Granted Apr 26, 2011

Compiling co-associating bioattributes using expanded bioattribute profiles

Assignee: Expanse Networks, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,933,912
App. No.
11/968,998
Granted
Apr 26, 2011
Kind
B2
Abstract

A bioinformatics method, software, database and system for compiling attribute combinations that co-associate with a query attribute (i.e., an attribute of interest) are presented in which expanded attribute profiles associated with a group of query-attribute-positive individuals and expanded attribute profiles associated with a group of query-attribute-negative individuals are accessed, and combinations of attributes having a higher frequency of occurrence in the set of expanded attribute profiles associated with the group of query-attribute-positive individuals are identified and stored to generate a compilation of attribute combinations that co-associate with the query attribute.

Claims (47)

1. A bioinformatics method for generating a compilation containing combinations of attributes that co-associate with a query attribute associated with a group of individuals, comprising:

a) accessing, within a first computer memory, expanded attribute profiles associated with individuals, wherein the expanded attribute profiles each comprise a set of primary attributes and a set of secondary attributes, and wherein the secondary attributes are derived from the primary attributes and wherein the set of secondary attributes has lower resolution that the set of primary attributes and at least one secondary attribute is a categorical attribute, wherein said categorical attribute includes at least one of a genetic, epigenetic, pangenetic, physical, behavioral, situational, or historical attribute that encompasses at least one primary attribute that is a numerical attribute and is derived through a heuristic rule applied to one or more primary attributes;

b) receiving a query attribute which identifies, within the expanded attribute profiles in the first computer memory, a set of expanded query-attribute-positive attribute profiles associated with a group of query-attribute-positive individuals and a set of expanded query-attribute-negative attribute profiles associated with a group of query-attribute-negative individuals;

c) selecting an expanded query-attribute-positive attribute profile from the set of expanded query-attribute-positive attribute profiles and, a subset of expanded query-attribute-negative attribute profiles from the set of expanded query-attribute-negative attribute profiles;

d) determining a set of candidate attributes by selecting attributes from the expanded query-attribute-positive attribute profile that do not occur in at least a predetermined portion of the subset of expanded query-attribute-negative attribute profiles;

e) computing the frequencies of occurrence of combinations of the candidate attributes in the set of expanded query-attribute-positive attribute profiles and in the set of expanded query-attribute-negative attribute profiles; and

f) storing, within a second computer memory, one or more candidate attribute combinations having higher frequencies of occurrence in the set of expanded query-attribute-positive attribute profiles than in the set of expanded query-attribute-negative attribute profiles to generate a compilation of attribute combinations that co-occur with the query attribute.

2. The bioinformatics method of claim 1 , further comprising storing the frequencies of occurrence of the attribute combinations in the compilation.

3. The bioinformatics method of claim 1 , wherein the query attribute is a combination of two or more attributes.

4. The bioinformatics method of claim 2 , further comprising storing statistical results, generated based on the frequencies of occurrence, which indicate the strength of association of each of the attribute combinations in the compilation with the query attribute.

5. The bioinformatics method of claim 1 , wherein the query-attribute-positive individuals and the query-attribute-negative individuals are derived from a single population of individuals that is preselected based on association or lack of association with one or more user specified attributes.

6. The bioinformatics method of claim 1 , wherein the identity of one or more of the individuals is masked or anonymized.

7. The bioinformatics method of claim 1 , wherein at least a portion of the compilation is transmitted as output to at least one destination selected from the group consisting of a user, a computer readable memory, a computer readable medium, a computer processor, a computer network, a printout device, a visual display, a digital electronic receiver and a wireless receiver.

8. The bioinformatics method of claim 1 , wherein at least one secondary attribute is derived by compounding the values of two or more primary attributes.

9. The bioinformatics method of claim 1 , wherein at least one primary attribute has a continuous value and at least one secondary attribute derived from that primary attribute has a discrete value.

10. The bioinformatics method of claim 1 , wherein at least one secondary attribute comprises an inequality statement containing a quantitative value, wherein the quantitative value is either larger or smaller than that of the primary attribute from which it was derived.

11. The bioinformatics method of claim 1 , wherein two or more of the secondary attributes comprise a sequence of inequality statements containing progressively larger quantitative values.

12. The bioinformatics method of claim 1 , wherein two or more of the secondary attributes comprise a sequence of inequality statements containing progressively smaller quantitative values.

13. A program storage device readable by a machine and containing a set of instructions which, when read by the machine, causes execution of a bioinformatics method for generating a compilation containing attribute combinations that co-associate with a query attribute, comprising:

a) accessing, within a first computer memory, expanded attribute profiles associated with individuals, wherein the expanded attribute profiles each comprise a set of primary attributes and a set of secondary attributes, and wherein the secondary attributes are derived from the primary attributes and wherein said set of secondary attributes has a lower resolution than said set of primary attributes and at least one secondary attribute is a categorical attribute that encompasses at least one primary attribute that is a numerical attribute, wherein said categorical attribute includes at least one of a genetic, epigenetic, pangenetic, physical, behavioral, situational, or historical attribute and said secondary attribute is derived through a heuristic rule applied to one or more primary attributes;

b) receiving a query attribute which identifies, within the expanded attribute profiles in the first computer memory, a set of expanded query-attribute-positive attribute profiles associated with a group of query-attribute-positive individuals and a set of expanded query-attribute-negative attribute profiles associated with a group of query-attribute-negative individuals;

c) selecting an expanded query-attribute-positive attribute profile from the set of expanded query-attribute-positive attribute profiles and, a subset of expanded query-attribute-negative attribute profiles from the set of expanded query-attribute-negative attribute profiles;

d) determining a set of candidate attributes by selecting attributes from the expanded query-attribute-positive attribute profile that do not occur in at least a predetermined portion of the subset of expanded query-attribute-negative attribute profiles;

e) computing the frequencies of occurrence of combinations of the candidate attributes in the set of expanded query-attribute-positive attribute profiles and in the set of expanded query-attribute-negative attribute profiles; and f) storing, within a second computer memory, one or more candidate attribute combinations having higher frequencies of occurrence in the set of expanded query-attribute-positive attribute profiles than in the set of expanded query-attribute-negative attribute profiles to generate a compilation of attribute combinations that co-occur with the query attribute.

14. The program storage device of claim 13 , further comprising:

g) storing statistical results, generated based on the frequencies of occurrence, which indicate the strength of association of each of the attribute combinations in the compilation with the query attribute.

15. A bioinformatics database system for generating a compilation containing attribute combinations that co-associate with a query attribute, comprising:

a) a memory containing:

i) a first data structure containing expanded attribute profiles associated with individuals, wherein the expanded attribute profiles each comprise a set of primary attributes and a set of secondary attributes, and wherein the secondary attributes are derived from the primary attributes and wherein the set of secondary attributes has lower resolution than the set of primary attributes and at least one secondary attribute is a categorical attribute that encompasses at least one primary attribute that is a numerical attribute, wherein said categorical attribute includes at least one of a genetic, epigenetic, pangenetic, physical, behavioral, situational, or historical attribute and said secondary attribute is derived through a heuristic rule applied to one or more primary attributes;

b) a processor for:

i) accessing the first data structure;

ii) receiving a query attribute which identifies, within the expanded attribute profiles in the first data structure, a set of expanded query-attribute-positive attribute profiles associated with a group of query-attribute-positive individuals and a set of expanded query-attribute-negative attribute profiles associated with a group of query-attribute-negative individuals;

iii) selecting an expanded query-attribute-positive attribute profile from the set of expanded query-attribute-positive attribute profiles and, a subset of expanded query-attribute-negative attribute profiles from the set of expanded query-attribute-negative attribute profiles;

iv) determining a set of candidate attributes by selecting attributes from the expanded query-attribute-positive attribute profile that do not occur in at least a predetermined portion of the subset of expanded query-attribute-negative attribute profiles;

v) computing the frequencies of occurrence of combinations of the candidate attributes in the set of expanded query-attribute-positive attribute profiles and in the set of expanded query-attribute-negative attribute profiles; and

vi) storing, within a second data structure in the computer memory, one or more candidate attribute combinations having higher frequencies of occurrence in the set of expanded query-attribute-positive attribute profiles than in the set of expanded query-attribute-negative attribute profiles to generate a compilation of attribute combinations that co-occur with the query attribute.

16. The bioinformatics database system of claim 15 , wherein the processor is also for:

vii storing statistical results, generated based on the frequencies of occurrence, which indicate the strength of association of each of the attribute combinations in the compilation with the query attribute.

17. A bioinformatics computer-based system for generating a compilation containing attribute combinations that co-associate with a query attribute, comprising:

a) a data accessing subsystem for accessing, within a first computer memory, expanded attribute profiles associated with individuals, wherein the expanded attribute profiles each comprise a set of primary attributes and a set of secondary attributes, and wherein the secondary attributes are derived from the primary attributes and wherein the set of secondary attributes has lower resolution than the set of primary attributes and at least one secondary attribute is a categorical attribute that encompasses at least one primary attribute that is a numerical attribute, wherein said categorical attribute includes at least one of a genetic, epigenetic, pangenetic, physical, behavioral, situational, or historical attribute and said secondary attribute is derived through a heuristic rule applied to one or more primary attributes;

b) a data communications subsystem for receiving a query attribute which identifies, within the expanded attribute profiles in the first computer memory, a set of expanded query-attribute-positive attribute profiles associated with a group of query-attribute-positive individuals and a set of expanded query-attribute-negative attribute profiles associated with a group of query-attribute-negative individuals;

c) a data processing subsystem for:

i) selecting an expanded query-attribute-positive attribute profile from the set of expanded query-attribute-positive attribute profiles and, a subset of expanded query-attribute-negative attribute profiles from the set of expanded query-attribute-negative-attribute profiles;

ii) determining a set of candidate attributes by selecting attributes from the expanded query-attribute-positive attribute profile that do not occur in at least a predetermined portion of the subset of expanded query-attribute-negative attribute profiles;

iii) computing the frequencies of occurrence of combinations of the candidate attributes in the set of expanded query-attribute-positive attribute profiles and in the set of expanded query-attribute-negative attribute profiles; and

d) a data storage subsystem for storing, within a second computer memory, one or more candidate attribute combinations having higher frequencies of occurrence in the set of expanded query-attribute-positive attribute profiles than in the set of expanded query-attribute-negative attribute profiles to generate a compilation of attribute combinations that co-occur with the query attribute.

18. The bioinformatics computer-based system of claim 17 , wherein the data processing subsystem is also for generating, based on the frequencies of occurrence, corresponding statistical results which indicate the strength of association of each of the attribute combinations in the compilation with the query attribute, and wherein the data storage subsystem is also for storing the corresponding statistical results within the second computer memory.

Assignments (8)
CORRECTIVE ASSIGNMENT TO CORRECT THE APP. NO. 63806415 TO 63806145 AND APPL NO. 17721779 TO 17731779 PREVIOUSLY RECORDED ON REEL 73168 FRAME 531. ASSIGNOR(S) HEREBY CONFIRMS THE CHANGE OF NAME. Recorded Jan 6, 2026
From: 23ANDME PGS LLC
To: 23ANDME GENOMICS LLC
Reel/Frame 074434/0334 →
CHANGE OF NAME Recorded Oct 22, 2025
From: 23ANDME PGS LLC
To: 23ANDME GENOMICS LLC
Reel/Frame 073168/0531 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 26, 2025
From: 23ANDME, INC.
To: 23ANDME PGS LLC
Reel/Frame 072562/0795 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 4, 2022
From: EXPANSE BIOINFORMATICS, INC.
To: 23ANDME, INC.
Reel/Frame 058537/0913 →
CHANGE OF ADDRESS BY ASSIGNEE Recorded Apr 5, 2021
From: EXPANSE BIOINFORMATICS, INC.
To: EXPANSE BIOINFORMATICS, INC.
Reel/Frame 055823/0714 →
CHANGE OF NAME Recorded Sep 23, 2013
From: EXPANSE NETWORKS, INC.
To: EXPANSE BIOINFORMATICS, INC.
Reel/Frame 031289/0003 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 29, 2008
From: ELDERING, CHARLES A.
To: EXPANSE NETWORKS, INC.
Reel/Frame 020724/0064 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 29, 2008
From: KENEDY, ANDREW A.
To: EXPANSE NETWORKS, INC.
Reel/Frame 020724/0062 →
Continuity (2)
Provisional Application 60895236 · Mar 16, 2007
Related Publication 20080228730A1 · Sep 18, 2008