IP Library Granted Patent US 8,713,018
Granted Patent B2
US 8,713,018 · App. 12/833,860 · Granted Apr 29, 2014

System and method for displaying relationships between electronically stored information to provide classification suggestions via inclusion

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,713,018
App. No.
12/833,860
Granted
Apr 29, 2014
Kind
B2
Abstract

A system and for providing reference documents as a suggestion for classifying uncoded documents is provided. A set of reference electronically stored information items, each associated with a classification code, is designated. One or more of the reference electronically stored information items is combined with a set of uncoded electronically stored information items. Clusters of the uncoded electronically stored information items and the one or more reference electronically stored information items are generated. Relationships between the uncoded electronically stored information items and the one or more reference electronically stored information items in at least one cluster are visually depicted as suggestions for classifying the uncoded electronically stored information items in that cluster.

Claims (185)

1. A system for providing reference items as a suggestion for classifying uncoded electronically stored information items, comprising:

a set of reference electronically stored information items each associated with one of a plurality of classification codes and a visual representation of that classification code comprising at least one of a shape and a symbol, wherein the visual representation of each of the classification codes is different from the visual representations of the remaining classification codes;

a set of uncoded electronically stored information items each associated with a visual representation different from the visual representations of the classification codes;

a processor to execute modules, comprising:

a clustering module to combine one or more of the coded reference electronically stored information items with the set of the uncoded electronically stored information items and to group the combined uncoded electronically stored information items and one or more coded reference electronically stored information items into clusters; and

a display to visually depict relationships between the uncoded electronically stored information items and the one or more coded reference electronically stored information items in at least one of the clusters as suggestions for classifying the uncoded electronically stored information items in that cluster by displaying the visual representation associated with each of the coded reference electronically stored information items in that cluster and the visual representation associated with each of the uncoded electronically stored information items in that cluster.

2. A system according to claim 1 , further comprising:

a reference module to generate the set of reference electronically stored information items, comprising at least one of:

a similarity module to identify dissimilar electronically stored information items for a document review project and to assign a classification code to each of the dissimilar electronically stored information items; and

a reference clustering module to cluster electronically stored information items for a document review project, to select one or more of the electronically stored information items in at least one cluster, and to assign a classification code to each of the selected electronically stored information items.

3. A system according to claim 1 , wherein the clusters are generated based on a similarity metric comprising forming a score vector for each uncoded electronically stored information item in the portion and each coded electronically stored information item in the reference set and calculating the similarity metric by comparing the score vectors for one of the uncoded electronically stored information items and one of the coded electronically stored information items in the reference set as an inner product.

4. A system according to claim 3 , wherein the inner product is determined according to the following equation:

cos

σ

AB

=

S

A

·

S

B

S

A

S

B

where cos σ AB comprises a similarity between uncoded electronically stored information item A and coded reference electronically stored information item B, {right arrow over (S)} A comprises a score vector for the uncoded electronically stored information item A, and {right arrow over (S)} B comprises a score vector for the coded reference electronically stored information item B.

5. A system according to claim 1 , further comprising:

a classification module to assign a classification code to one or more of the uncoded electronically stored information items in the at least one cluster.

6. A system according to claim 1 , wherein each uncoded electronically stored information item in the at least one cluster is represented by a symbol in the display and each of the one or more coded reference electronically stored information items is represented by an additional symbol in the display, and further wherein the coded reference electronically stored information items associated with different classification codes are distinguished by assigning a different color to the different symbols.

7. A method for providing reference items as a suggestion for classifying uncoded electronically stored information items, comprising the steps of:

designating a set of reference electronically stored information items each associated with one of a plurality of classification codes and a visual representation of that classification code comprising at least one of a shape and a symbol, wherein the visual representation of each of the classification codes is different from the visual representations of the remaining classification codes;

obtaining a set of uncoded electronically stored information items each associated with a visual representation different from the visual representations of the classification codes;

combining one or more of the coded reference electronically stored information items with the set of the uncoded electronically stored information items;

grouping the combined uncoded electronically stored information items and one or more coded reference electronically stored information items into clusters; and

visually depicting relationships between the uncoded electronically stored information items and one or more coded reference electronically stored information items in at least one of the clusters as suggestions for classifying the uncoded electronically stored information items in that cluster, comprising displaying the visual representation associated with each of the coded reference electronically stored information items in that cluster and the visual representation associated with each of the uncoded electronically stored information items in that cluster,

wherein the steps are performed by a suitably programmed computer.

8. A method according to claim 7 , further comprising:

generating the set of reference electronically stored information items, comprising at least one of:

identifying dissimilar electronically stored information items for a document review project and assigning a classification code to each of the dissimilar electronically stored information items; and

clustering electronically stored information items for a document review project, selecting one or more of the electronically stored information items in at least one cluster and assigning a classification code to each of the selected electronically stored information items.

9. A method according to claim 7 , wherein the clusters are generated based on a similarity metric, comprising:

forming a score vector for each uncoded electronically stored information item in the portion and each coded electronically stored information item in the reference set; and

calculating the similarity metric by comparing the score vectors for one of the uncoded electronically stored information items and one of the electronically stored information items in the reference set as an inner product.

10. A method according to claim 9 , wherein the inner product is determined according to the following equation:

cos

σ

AB

=

S

A

·

S

B

S

A

S

B

where cosσ AB comprises a similarity between uncoded electronically stored information item A and coded reference electronically stored information item B, {right arrow over (S)} A comprises a score vector for the uncoded electronically stored information item A, {right arrow over (S)} B and comprises a score vector for the coded reference electronically stored information item B.

11. A method according to claim 7 , further comprising:

assigning a classification code to one or more of the uncoded electronically stored information items in the at least one cluster.

12. A method according to claim 7 , further comprising:

representing each uncoded electronically stored information item in the at least one cluster with a symbol; and

representing each of the one or more coded reference electronically stored information items with a different symbol; and

distinguishing the coded reference electronically stored information items associated with different classification codes by assigning a different color to the different symbols.

13. A system for clustering reference documents to generate suggestions for classification of uncoded documents, comprising:

a set of reference documents each associated with one of a plurality of classification codes and a visual representation of that classification code comprising at least one of a shape and a symbol, wherein the visual representation of each of the classification codes is different from the visual representations of the remaining classification codes;

a set of uncoded documents each associated with a visual representation different from the visual representations of the classification codes;

a processor to execute modules, comprising:

a clustering module to select one or more of the coded reference documents, to combine the one or more coded reference documents selected with the uncoded documents as a set of combined documents, and to group the combined set of documents into clusters, further comprising:

a cluster similarity module to determine a similarity between each document; and

a grouping module to group the documents into the clusters based on the similarity;

an identification module to identify at least one of the clusters with the coded reference documents; and

a display to visually depict relationships between the uncoded documents and the one or more coded reference documents in the at least one cluster as suggestions for classifying the uncoded documents in that cluster by displaying the visual representation associated with each of the coded reference documents in that cluster and the visual representation associated with each of the uncoded documents in that cluster.

14. A system according to claim 13 , further comprising:

a reference module to generate the set of reference documents, comprising at least one of:

a reference similarity module to identify dissimilar documents for a document review project and assign a classification code to each of the dissimilar documents; and

a reference cluster module to generate clusters of documents for a document review project, selecting one or more of the documents in at least one of the clusters and assigning a classification code to each of the documents.

15. A system according to claim 13 , wherein the one or more coded reference documents are selected from at least one of a predefined, customized, or arbitrary reference document set.

16. A system according to claim 13 , wherein the similarity is determined by forming a score vector for each uncoded document and each coded reference document and calculating a similarity metric between the score vectors for the uncoded documents and coded reference documents as an inner product.

17. A system according to claim 16 , wherein the inner product is determined according to the following equation:

cos

σ

AB

=

S

A

·

S

B

S

A

S

B

where cos σ AB comprises a similarity between uncoded document A and coded reference document B, {right arrow over (S)} A comprises a score vector for the uncoded document A, and {right arrow over (S)} B comprises a score vector for the coded reference document B.

18. A system according to claim 13 , wherein each uncoded document in the at least one cluster is represented by a symbol and each coded reference document is represented by a different symbol, and further wherein the coded reference electronically stored information items associated with different classification codes are distinguished by different color assigned to the different symbols.

19. A method for clustering reference documents to generate suggestions for classification of uncoded documents, comprising the steps of:

designating a set of reference documents each associated with one of a plurality of classification codes and a visual representation of that classification code comprising at least one of a shape and a symbol, wherein the visual representation of each of the classification codes is different from the visual representations of the remaining classification codes;

obtaining a set of uncoded documents each associated with a visual representation different from the visual representations of the classification codes;

selecting one or more of the coded reference documents and combining the one or more coded reference documents selected with the uncoded documents as a set of combined documents;

grouping the combined set of documents into clusters, comprising:

determining a similarity between each document; and

grouping the documents into the clusters based on the similarity;

identifying at least one cluster of the clusters with the coded reference documents; and

visually depicting relationships between the uncoded documents and the one or more coded reference documents in the at least one cluster as suggestions for classifying the uncoded documents in that cluster, comprising displaying the visual representation associated with each of the coded reference documents in that cluster and the visual representation associated with each of the uncoded documents in that cluster,

wherein the steps are performed by a suitably programmed computer.

20. A method according to claim 19 , further comprising:

generating the set of reference documents, comprising at least one of:

identifying dissimilar documents for a document review project and assigning a classification code to each of the dissimilar documents; and

generating clusters of documents for a document review project, selecting one or more of the documents in at least one of the clusters and assigning a classification code to each of the documents.

21. A method according to claim 19 , wherein the one or more coded reference documents are selected from at least one of a predefined, customized, or arbitrary reference document set.

22. A method according to claim 19 , further comprising:

determining the similarity, comprising:

forming a score vector for each uncoded document and each coded reference document; and

calculating a similarity metric between the score vectors for the uncoded documents and coded reference documents as an inner product.

23. A method according to claim 22 , wherein the inner product is determined according to the following equation:

cos

σ

AB

=

S

A

·

S

B

S

A

S

B

where cos σ AB comprises a similarity between uncoded document A and coded reference document B, {right arrow over (S)} A comprises a score vector for the uncoded document A, and {right arrow over (S)} B comprises a score vector for the coded reference document B.

24. A method according to claim 19 , further comprising:

representing each uncoded document in the at least one cluster with a symbol; and

representing each coded reference document with a different symbol; and

distinguishing the coded reference documents with different classification codes with different colors of the different symbols.

Assignments (9)
SECURITY INTEREST Recorded Apr 4, 2024
From: NUIX NORTH AMERICA INC.
To: THE HONGKONG AND SHANGHAI BANKING CORPORATION LIMITED, SYDNEY BRANCH, AS SECURED PARTY
Reel/Frame 067005/0073 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 15, 2018
From: FTI CONSULTING, INC.
To: NUIX NORTH AMERICA INC.
Reel/Frame 047163/0584 →
RELEASE OF SECURITY INTEREST IN PATENT RIGHTS AT REEL/FRAME 036031/0637 Recorded Sep 12, 2018
From: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
To: FTI CONSULTING, INC.
Reel/Frame 047060/0137 →
NOTICE OF GRANT OF SECURITY INTEREST IN PATENTS Recorded Jun 29, 2015
From: FTI CONSULTING, INC.; FTI CONSULTING TECHNOLOGY LLC; FTI CONSULTING TECHNOLOGY SOFTWARE CORP
To: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 036031/0637 →
RELEASE OF SECURITY INTEREST IN PATENT RIGHTS Recorded Jun 29, 2015
From: BANK OF AMERICA, N.A.
To: FTI CONSULTING, INC.; FTI CONSULTING TECHNOLOGY LLC
Reel/Frame 036029/0233 →
RELEASE OF SECURITY INTEREST IN PATENTS Recorded Dec 11, 2012
From: BANK OF AMERICA, N.A.
To: FTI CONSULTING, INC.; FTI TECHNOLOGY LLC; ATTENEX CORPORATION
Reel/Frame 029449/0389 →
NOTICE OF GRANT OF SECURITY INTEREST IN PATENTS Recorded Dec 10, 2012
From: FTI CONSULTING, INC.; FTI CONSULTING TECHNOLOGY LLC
To: BANK OF AMERICA, N.A.
Reel/Frame 029434/0087 →
NOTICE OF GRANT OF SECURITY INTEREST IN PATENTS Recorded Mar 14, 2011
From: FTI CONSULTING, INC.; FTI TECHNOLOGY LLC; ATTENEX CORPORATION
To: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 025943/0038 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 26, 2010
From: KNIGHT, WILLIAM C.; NUSSBAUM, NICHOLAS I.
To: FTI CONSULTING, INC.
Reel/Frame 024742/0704 →