IP Library Granted Patent US 11,048,738
Granted Patent B2
US 11,048,738 · App. 16/263,377 · Granted Jun 29, 2021

Records search and management in compliance platforms

Inventors: Zohar Duchin (Brookline, MA); Ehsan Masud (Leesburg, VA); Michelle Zhong (Marietta, GA)
Assignee: EMC IP Holding Company LLC
G06F16/3347G06F16/313G06F16/319G06F16/338
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,048,738
App. No.
16/263,377
Granted
Jun 29, 2021
Kind
B2
Abstract

A method in one embodiment comprises defining a plurality of fields in a plurality of electronic documents, wherein the plurality of fields respectively correspond to governance, risk and compliance system data structures, identifying a plurality of relationships between the electronic documents based on one or more cross-references between fields of two or more different electronic documents of the plurality of electronic documents, and assigning respective ranks to the plurality of electronic documents based on the relationships. In the method, a query is received from a user device, and a listing of candidate documents of the plurality of electronic documents is retrieved in response to the query. Scores for respective ones of the candidate documents are computed based on at least the assigned ranks, and a response to the query is transmitted to the user device, wherein the response comprises the listing of candidate documents sorted according to the computed scores.

Claims (72)

1. An apparatus comprising:

at least one processing platform comprising a plurality of processing devices;

said at least one processing platform being configured:

to define a plurality of fields in a plurality of electronic documents, wherein the plurality of fields respectively correspond to governance, risk and compliance (GRC) system data structures;

to identify a plurality of relationships between the plurality of electronic documents based on one or more cross-references between fields of two or more different electronic documents of the plurality of electronic documents;

to assign respective ranks to the plurality of electronic documents based on the plurality of relationships;

to receive at least one query from a user device;

to retrieve a listing of candidate documents of the plurality of electronic documents in response to the at least one query;

to compute a plurality of scores for respective ones of the candidate documents based on at least said assigned ranks; and

to transmit a response to the at least one query to the user device, wherein the response comprises the listing of the candidate documents sorted according to the computed plurality of scores;

wherein in defining the plurality of fields, said at least one processing platform is configured:

to define a field of a first one of the two or more different electronic documents comprising a first data element corresponding to a first GRC function; and

to define a field of a second one of the two or more different electronic documents comprising a second data element corresponding to a second GRC function;

wherein the one or more cross-references comprise a reference from the first data element to the second data element; and

wherein, in assigning the respective ranks, said at least one processing platform is configured to determine relative importance of respective ones of the plurality of electronic documents based at least in part on the plurality of relationships, wherein the determination comprises:

comparing document type identifiers of the respective ones of the plurality of electronic documents; and

omitting relationships between two or more electronic documents of the plurality of electronic documents having the same document type identifiers.

2. The apparatus of claim 1 wherein the GRC system data structures comprise data corresponding to one or more GRC functions for an enterprise and one or more correlations between the one or more GRC functions and one or more other GRC functions for the enterprise.

3. The apparatus of claim 2 wherein the two or more different electronic documents are in different applications of the GRC system.

4. The apparatus of claim 1 wherein at least one of the plurality of fields and the plurality of electronic documents correspond to respective unique identifiers.

5. The apparatus of claim 1 wherein said at least one processing platform is further configured to generate an inverted index for the plurality of electronic documents.

6. The apparatus of claim 1 wherein said at least one processing platform is further configured:

to generate a relationship matrix based on the plurality of relationships; and

to compute a vector based on the relationship matrix.

7. The apparatus of claim 6 wherein said at least one processing platform is further configured to modify the vector with a bias factor.

8. The apparatus of claim 1 wherein in assigning the respective ranks to the plurality of electronic documents said at least one processing platform is further configured to compute a vector based on a number of the plurality of documents belonging to a topic.

9. The apparatus of claim 8 wherein the topic is specified in the at least one query.

10. The apparatus of claim 1 wherein the at least one query includes at least one of one or more free text terms and one or more topic terms.

11. The apparatus of claim 10 wherein said at least one processing platform is further configured to identify at least one of a number of the one or more free text terms and a number of the one or more topic terms in a body and a title of respective ones of the candidate documents.

12. The apparatus of claim 10 wherein said at least one processing platform is further configured to compute a similarity between a term frequency-inverse document frequency (TF-IDF) vector corresponding to the one or more free text terms in the at least one query and TF-IDF vectors corresponding to the one or more free text terms in a body and a title of respective ones of the candidate documents.

13. The apparatus of claim 10 wherein said at least one processing platform is further configured to compute a similarity between a term frequency-inverse document frequency (TF-IDF) vector corresponding to one or more topic terms in the at least one query and TF-IDF vectors corresponding to the one or more topic terms in a body and a title of respective ones of the candidate documents.

14. The apparatus of claim 1 wherein said at least one processing platform is further configured to determine one or more features of the respective ones of the candidate documents, wherein the computed plurality of scores are further based on the determined one or more features.

15. The apparatus of claim 14 wherein, in computing the plurality of scores, said at least one processing platform is further configured to weight said assigned ranks higher than the determined one or more features.

16. A method comprising:

defining a plurality of fields in a plurality of electronic documents, wherein the plurality of fields respectively correspond to governance, risk and compliance (GRC) system data structures;

identifying a plurality of relationships between the plurality of electronic documents based on one or more cross-references between fields of two or more different electronic documents of the plurality of electronic documents;

assigning respective ranks to the plurality of electronic documents based on the plurality of relationships;

receiving at least one query from a user device;

retrieving a listing of candidate documents of the plurality of electronic documents in response to the at least one query;

computing a plurality of scores for respective ones of the candidate documents based on at least said assigned ranks; and

transmitting a response to the at least one query to the user device, wherein the response comprises the listing of the candidate documents sorted according to the computed plurality of scores;

wherein defining the plurality of fields comprises:

defining a field of a first one of the two or more different electronic documents comprising a first data element corresponding to a first GRC function; and

defining a field of a second one of the two or more different electronic documents comprising a second data element corresponding to a second GRC function;

wherein the one or more cross-references comprise a reference from the first data element to the second data element;

wherein assigning the respective ranks comprises determining relative importance of respective ones of the plurality of electronic documents based at least in part on the plurality of relationships, wherein the determination comprises:

comparing document type identifiers of the respective ones of the plurality of electronic documents; and

omitting relationships between two or more electronic documents of the plurality of electronic documents having the same document type identifiers; and

wherein the method is performed by at least one processing platform comprising at least one processing device comprising a processor coupled to a memory.

17. The method of claim 16 further comprising:

generating a relationship matrix based on the plurality of relationships; and

computing a vector based on the relationship matrix.

18. The method of claim 16 wherein assigning the respective ranks to the plurality of electronic documents further comprises computing a vector based on a number of the plurality of documents belonging to a topic.

19. A computer program product comprising a non-transitory processor-readable storage medium having stored therein program code of one or more software programs, wherein the program code when executed by at least one processing platform causes said at least one processing platform:

to define a plurality of fields in a plurality of electronic documents, wherein the plurality of fields respectively correspond to governance, risk and compliance (GRC) system data structures;

to identify a plurality of relationships between the plurality of electronic documents based on one or more cross-references between fields of two or more different electronic documents of the plurality of electronic documents;

to assign respective ranks to the plurality of electronic documents based on the plurality of relationships;

to receive at least one query from a user device;

to retrieve a listing of candidate documents of the plurality of electronic documents in response to the at least one query;

to compute a plurality of scores for respective ones of the candidate documents based on at least said assigned ranks; and

to transmit a response to the at least one query to the user device, wherein the response comprises the listing of the candidate documents sorted according to the computed plurality of scores;

wherein in defining the plurality of fields, the program code further causes said at least one processing platform:

to define a field of a first one of the two or more different electronic documents comprising a first data element corresponding to a first GRC function; and

to define a field of a second one of the two or more different electronic documents comprising a second data element corresponding to a second GRC function;

wherein the one or more cross-references comprise a reference from the first data element to the second data element; and

wherein, in assigning the respective ranks, the program code further causes said at least one processing platform to determine relative importance of respective ones of the plurality of electronic documents based at least in part on the plurality of relationships, wherein the determination comprises:

comparing document type identifiers of the respective ones of the plurality of electronic documents; and

omitting relationships between two or more electronic documents of the plurality of electronic documents having the same document type identifiers.

20. The computer program product according to claim 19 wherein, in assigning the respective ranks to the plurality of electronic documents the program code further causes said at least one processing platform to compute a vector based on a number of the plurality of documents belonging to a topic.

21. The computer program product according to claim 19 wherein the program code further causes said at least one processing platform:

to generate a relationship matrix based on the plurality of relationships; and

to compute a vector based on the relationship matrix.

Assignments (4)
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (053546/0001) Recorded Jun 23, 2022
From: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
To: DELL MARKETING L.P. (ON BEHALF OF ITSELF AND AS SUCCESSOR-IN-INTEREST TO CREDANT TECHNOLOGIES, INC.); DELL INTERNATIONAL L.L.C.; DELL PRODUCTS L.P.; DELL USA L.P.; EMC CORPORATION; DELL MARKETING CORPORATION (SUCCESSOR-IN-INTEREST TO FORCE10 NETWORKS, INC. AND WYSE TECHNOLOGY L.L.C.); EMC IP HOLDING COMPANY LLC
Reel/Frame 071642/0001 →
SECURITY AGREEMENT Recorded Apr 22, 2020
From: CREDANT TECHNOLOGIES INC.; DELL INTERNATIONAL L.L.C.; DELL MARKETING L.P.; DELL PRODUCTS L.P.; DELL USA L.P.; EMC CORPORATION; FORCE10 NETWORKS, INC.; WYSE TECHNOLOGY L.L.C.; EMC IP HOLDING COMPANY LLC
To: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A.
Reel/Frame 053546/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 25, 2019
From: DUCHIN, ZOHAR; MASUD, EHSAN; ZHONG, MICHELLE
To: EMC IP HOLDING COMPANY LLC
Reel/Frame 048685/0819 →
SECURITY AGREEMENT Recorded Mar 21, 2019
From: CREDANT TECHNOLOGIES, INC.; DELL INTERNATIONAL L.L.C.; DELL MARKETING L.P.; DELL PRODUCTS L.P.; DELL USA L.P.; EMC CORPORATION; FORCE10 NETWORKS, INC.; WYSE TECHNOLOGY L.L.C.; EMC IP HOLDING COMPANY LLC
To: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A.
Reel/Frame 049452/0223 →
Continuity (1)
Related Publication 20200250213A1 · Aug 6, 2020