IP Library Granted Patent US 8,650,190
Granted Patent B2
US 8,650,190 · App. 13/831,565 · Granted Feb 11, 2014

Computer-implemented system and method for generating a display of document clusters

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,650,190
App. No.
13/831,565
Granted
Feb 11, 2014
Kind
B2
Abstract

A computer-implemented system and method for generating a display of document clusters is described. Clusters of documents are presented in a multi-dimensional concept space. At least one document is selected from a collection of documents to be clusters. An angle θ of the document relative to a common origin of the multi-dimensional concept space is computed. The selected document is compared with each of the clusters. An angle σ from the common origin is determined for each cluster. A difference between the angle θ for the document and the angle σ for the cluster is determined. The difference is compared to the variance, and a new cluster is created when the difference exceeds the variance for all the clusters.

Claims (53)

1. A computer-implemented system for generating document clusters, comprising:

clusters of documents in a multi-dimensional concept space;

a selection module select at least one document from a collection of documents to be clustered;

a document angle module to compute an angle θ of the document relative to a common origin of the multi-dimensional concept space;

a comparison to compare the selected document with each of the clusters, comprising:

a cluster angle module to calculate for each cluster, an angle σ from the common origin

a similarity module to determine a difference between the angle θ of the document and the angle σ of the cluster; and

a variance determining module to apply a variance the difference; and

a cluster module to create a new cluster when the difference exceeds the variance for all the clusters.

2. A system according to claim 1 , further comprising:

a display placement module to place the selected document into one of the clusters when the difference between the selected document and that cluster is at least one of greater than or equal to the variance.

3. A system according to claim 1 , further comprising:

a cluster finalizing module to finalize one or more of the clusters comprising at least one of merging two or more of the clusters into a single cluster, splitting at least one of the clusters into two or more clusters, and removing those clusters that are outliers in relation to the remaining clusters.

4. A system according to claim 1 , further comprising:

a concept module to build a concept space over the document collection based on the clusters and any newly generated clusters.

5. A system according to claim 1 , further comprising:

a cluster formation module to form the clusters of documents in the multi-dimensional concept space by creating an initial cluster and adding additional clusters via a k-means clustering technique.

6. A system according to claim 1 , wherein the documents are iteratively compared with the collection to the clusters.

7. A system according to claim 1 , further comprising:

a display module to place one or more of the clusters into the multi-dimensional concept space around the common origin.

8. A system according to claim 1 , wherein the documents within each cluster share related terms and phrases.

9. A system according to claim 1 , wherein the variance is defined as an upper bound on a distance between the documents and the clusters.

10. A system according to claim 1 , further comprising:

an update module to recalculate the angle σ for each cluster when one of the documents is added or one of the documents is removed.

11. A computer-implemented method for generating document clusters, comprising:

presenting clusters of documents in a multi-dimensional concept space and selecting at least one document from a collection of documents to be clustered;

computing an angle θ of the document relative to a common origin of the multi-dimensional concept space;

comparing the selected document with each of the clusters, comprising:

calculating for each cluster, an angle σ from the common origin;

determining a difference between the angle θ for the document and the angle σ for the cluster; and

comparing the difference to a variance; and

creating a new cluster when the difference exceeds the variance for all the clusters.

12. A method according to claim 11 , further comprising:

placing the selected document into one or more of the clusters when the difference between the selected document and that cluster is at least one of greater than or equal to the variance.

13. A method according to claim 11 , further comprising at least one of:

merging two or more of the clusters into a single cluster;

splitting at least one of the clusters into two or more clusters; and

removing those clusters that are outliers in relation to the remaining clusters.

14. A method according to claim 11 , further comprising:

building a concept space over the document collection based on the clusters and any newly generated clusters.

15. A method according to claim 11 , further comprising:

forming the clusters of documents in the multi-dimensional concept space, comprising:

creating an initial cluster; and

adding additional clusters via a k-means clustering technique.

16. A method according to claim 11 , further comprising:

iteratively comparing each document in the collection to the clusters.

17. A method according to claim 11 , further comprising:

placing one or more of the clusters into the multi-dimensional concept space around the common origin.

18. A method according to claim 11 , wherein the documents within each cluster share related terms and phrases.

19. A method according to claim 11 , further comprising:

defining the variance as an upper bound on a distance between the documents and the clusters.

20. A method according to claim 11 , further comprising:

recalculating the angle σ for each cluster when one of the documents is added or one of the documents is removed.

Assignments (6)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 31, 2020
From: GALLIVAN, DAN
To: ATTENEX CORPORATION
Reel/Frame 051679/0324 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 31, 2020
From: ATTENEX CORPORATION
To: FTI TECHNOLOGY LLC
Reel/Frame 051679/0344 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 15, 2018
From: FTI CONSULTING TECHNOLOGY LLC
To: NUIX NORTH AMERICA INC.
Reel/Frame 047237/0019 →
RELEASE OF SECURITY INTEREST IN PATENT RIGHTS AT REEL/FRAME 036031/0637 Recorded Sep 12, 2018
From: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
To: FTI CONSULTING TECHNOLOGY LLC
Reel/Frame 047060/0107 →
CHANGE OF NAME Recorded Apr 20, 2018
From: FTI TECHNOLOGY LLC
To: FTI CONSULTING TECHNOLOGY LLC
Reel/Frame 045785/0645 →
NOTICE OF GRANT OF SECURITY INTEREST IN PATENTS Recorded Jun 29, 2015
From: FTI CONSULTING, INC.; FTI CONSULTING TECHNOLOGY LLC; FTI CONSULTING TECHNOLOGY SOFTWARE CORP
To: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 036031/0637 →