IP Library Granted Patent US 9,846,740
Granted Patent B2
US 9,846,740 · App. 14/021,391 · Granted Dec 19, 2017

Associative search systems and methods

Inventors: Christopher David Bamford (Farnham, GB); Clive Nicholas Jordan (St. Albans, GB)
Assignee: MIMECAST SERVICES LTD.
G06F17/30675G06F17/30011G06F17/3053G06F17/30395G06F17/30554G06F17/30864
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,846,740
App. No.
14/021,391
Granted
Dec 19, 2017
Kind
B2
Abstract

A computer system including a memory, at least one processor coupled to the memory, and a search component executable by the at least one processor is provided. The search component is configured to receive information descriptive of at least one search term; execute a first query against a plurality of documents that identifies at least one first document of the plurality of documents responsive to the at least one search term; identify one or more secondary terms associated with the at least one first document based on occurrence of the one or more secondary terms within the at least one first document; and provide a search result including at least one of the one or more secondary terms and one or more identifiers of one or more documents including the one or more secondary terms. The search result may also include one or more identifiers of bookmarked documents.

Claims (78)

1. A computer system comprising:

a memory;

at least one processor coupled to the memory; and

a search component executable by the at least one processor and configured to:

receive information descriptive of at least one search term;

execute a first query against a plurality of documents that identifies at least two documents of the plurality of documents responsive to the at least one search term, wherein the at least two documents include the at least one search term;

automatically identify one or more secondary terms included within the at least two documents based on a correlation between the at least one search term and the one or more secondary terms included within the at least two documents, wherein the correlation is based on any one of: the parts of speech embodied by the at least one search term and the one or more secondary terms, the presence of the one or more secondary terms as recorded in a search history or a location of the at least one search term within an email;

rank the one or more secondary terms across the at least two documents based on the correlation;

record a subset of secondary terms, including one or more of the ranked secondary terms;

execute a second query against the plurality of documents using the recorded subset of secondary terms; and

provide, within a search result interface, a search result including the at least two documents from the first query including the recorded subset of secondary terms and one or more documents from the second query including the recorded subset of secondary terms.

2. The computer system of claim 1 , wherein the at least two documents include a plurality of documents, the memory is configured to store a configurable parameter specifying a maximum number of documents, and the search component is further configured to limit the plurality of documents to a number of documents having a predefined relationship with the configurable parameter.

3. The computer system of claim 1 , wherein the one or more secondary terms include a plurality of secondary terms, the memory is configured to store a configurable parameter specifying a maximum number of secondary terms, and the search component is further configured to limit the plurality of secondary terms to a number of secondary terms having a predefined relationship with the configurable parameter.

4. The computer system of claim 1 , wherein the search result includes at least one identifier of a bookmarked document.

5. The computer system of claim 1 , wherein the recorded subset of secondary terms includes a plurality of secondary terms and the search component is further configured to assign a frequency score to each secondary term of the plurality of secondary terms according to a frequency with which each secondary term occurs within the at least two documents.

6. The computer system of claim 5 , wherein the search component is configured to assign a frequency score to each secondary term using at least one of a term frequency-inverse document frequency process and an Okapi BM25 process.

7. The computer system of claim 5 , further comprising an interface component configured to display each secondary term of the plurality of secondary terms sorted by the frequency score within the search result.

8. The computer system of claim 7 , wherein the interface component is further configured to:

receive information identifying at least one secondary term of the plurality of secondary terms; and

identify, responsive to receiving the information identifying the at least one secondary term, one or more additional documents including the at least one secondary term.

9. The computer system of claim 1 , wherein the one or more documents include one or more seed documents and the search component is further configured to:

identify at least one seed document of the one or more seed documents;

identify at least one additional secondary term associated with the at least one seed document based on occurrence of the at least one additional secondary term within the at least one seed document;

identify one or more additional documents including the at least one additional secondary term; and

provide the one or more additional documents within the search result.

10. The computer system of claim 9 , further comprising an interface component configured to receive information selecting the at least one seed document.

11. The computer system of claim 1 , wherein the correlation is based on the location of the at least one search term within an email.

12. The computer system of claim 1 , wherein the secondary terms are recorded based on evaluating a rule.

13. A method of executing an associative search using a computer system including memory and at least one processor coupled to the memory, the method comprising:

receiving, by the computer system, information describing at least one search term;

executing, by the computer system, a first query against a plurality of documents that identifies at least two documents of the plurality of documents responsive to the at least one search term, wherein the at least two documents include the at least one search term;

automatically identifying, by the computer system, one or more secondary terms included within the at least two documents based on a correlation between the at least one search term and the one or more secondary terms included within the at least two documents, wherein the correlation is based on any one of: the parts of speech embodied by the at least one search term and the one or more secondary terms, the presence of the one or more secondary terms as recorded in a search history or a location of the at least one search term within an email;

rank the one or more secondary terms across the at least two documents based on the correlation;

recording a subset of secondary terms, including one or more of the ranked secondary terms;

executing a second query against the plurality of documents using the recorded subset of secondary terms; and

providing, by the computer system within a search result interface, a search result including the at least two documents from the first query including the recorded subset of secondary terms and one or more documents from the second query including the recorded subset of secondary terms provided by the second query.

14. The method of claim 13 , wherein the at least two documents include a plurality of documents, the memory is configured to store a configurable parameter specifying a maximum number of documents, and the method further comprises limiting the plurality of documents to a number of documents having a predefined relationship with the configurable parameter.

15. The method of claim 13 , wherein the one or more secondary terms include a plurality of secondary terms, the memory is configured to store a configurable parameter specifying a maximum number of secondary terms, and the method further comprises limiting the plurality of secondary terms to a number of secondary terms having a predefined relationship with the configurable parameter.

16. The method of claim 13 , wherein providing the search result includes providing at least one identifier of a bookmarked document.

17. The method of claim 13 , wherein the recorded subset of secondary terms includes a plurality of secondary terms and the method further comprises assigning a frequency score to each secondary term of the plurality of secondary terms according to a frequency with which each secondary term occurs within the at least two documents.

18. The method of claim 17 , wherein scoring each secondary term comprises assigning a frequency score to each secondary term using at least one of a term frequency-inverse document frequency process and an Okapi BM25 process.

19. The method of claim 17 , further comprising displaying the plurality of secondary terms sorted by the frequency score within the search result.

20. The method of claim 19 , further comprising:

receiving information identifying at least one secondary term of the plurality of secondary terms; and

identifying, responsive to receiving the information identifying the at least one secondary term, one or more additional documents including the at least one secondary term.

21. The method of claim 13 , wherein the one or more documents include one or more seed documents and the method further comprises:

identifying at least one seed document of the one or more seed documents;

identifying at least one additional secondary term associated with the at least one seed document based on occurrence of the at least one additional secondary term within the at least one seed document;

identifying one or more additional documents including the at least one additional secondary term; and

providing the one or more additional documents within the search result.

22. The method of claim 21 , wherein identifying the at least one seed document includes identifying the at least one seed document in response to receiving information selecting the at least one seed document.

23. The method of claim 13 , wherein the correlation is based on the location of the at least one search term within an email.

24. The method of claim 13 , wherein the recording the subset of secondary terms includes evaluating a rule to determine whether to record a secondary term.

25. A non-transitory computer readable medium storing instructions executable by at least one processor to execute an associative search method, the instructions being coded to instruct the at least one processor to:

receive information describing at least one search term;

execute a first query against a plurality of documents that identifies at least two documents of the plurality of documents responsive to the at least one search term, wherein the at least two documents include the at least one search term;

automatically identify one or more secondary terms included within the at least two documents based on a correlation between the at least one search term and the one or more secondary terms included within the at least two documents, wherein the correlation is based on any one of: the parts of speech embodied by the at least one search term and the one or more secondary terms, the presence of the one or more secondary terms as recorded in a search history or a location of the at least one search term within an email;

rank the one or more secondary terms across the at least two documents based on the correlation;

record a subset of secondary terms, including one or more of the ranked secondary terms;

execute a second query against the plurality of documents using the recorded subset of secondary items; and

provide, within a search result interface, a search result including the at least two documents from the first query including the recorded subset of secondary terms and one or more documents from the second query including the recorded subset of secondary terms.

26. The computer readable medium of claim 25 , wherein the at least two documents include a plurality of documents and the instructions further instruct the at least one processor to limit the plurality of documents to a number of documents having a predefined relationship with a configurable parameter specifying a maximum number of documents.

27. The computer readable medium of claim 25 , wherein the one or more secondary terms include a plurality of secondary terms and the instructions further instruct the at least one processor to limit the plurality of secondary terms to a number of secondary terms having a predefined relationship with a configurable parameter specifying a maximum number of secondary terms.

28. The computer readable medium of claim 25 , wherein the instructions to provide the search result include instructions that instruct the at least one processor to provide at least one identifier of a bookmarked document.

29. The computer readable medium of claim 25 , wherein the recorded subset of secondary terms includes a plurality of secondary terms and the instructions further instruct the at least one processor to assign a frequency score to each secondary term of the plurality of secondary terms according to a frequency with which each secondary term occurs within the at least two documents.

30. The computer readable medium of claim 29 , wherein the instructions to score each secondary term include instructions that instruct the at least one processor to assign a frequency score to each secondary term using at least one of a term frequency-inverse document frequency process and an Okapi BM25 process.

31. The computer readable medium of claim 29 , wherein the instructions further instruct the at least one processor to display the plurality of secondary terms sorted by the frequency score within the search result.

32. The computer readable medium of claim 31 , wherein the instructions further instruct the at least one processor to:

receive information identifying at least one secondary term of the plurality of secondary terms; and

identify, responsive to receiving the information identifying the at least one secondary term, one or more additional documents including the at least one secondary term.

33. The computer readable medium of claim 25 , wherein the one or more documents include one or more seed documents and the instructions further instruct the at least one processor to:

identify at least one seed document of the one or more seed documents;

identify at least one additional secondary term associated with the at least one seed document based on occurrence of the at least one additional secondary term within the at least one seed document;

identify one or more additional documents including the at least one additional secondary term; and

provide the one or more additional documents within the search result.

34. The computer readable medium of claim 33 , wherein the instructions further instruct the at least one processor to receive information identifying the at least one seed document and the instructions to identify the at least one seed document include instructions that instruct the at least one processor to identify the at least one seed document in response to receiving the information identifying the at least one seed document.

35. The computer readable medium of claim 25 , wherein the correlation is based on the location of the at least one search term within an email.

36. The computer readable medium of claim 25 , wherein the secondary terms are recorded based on evaluating a rule.

Assignments (5)
SECURITY INTEREST Recorded May 20, 2022
From: MIMECAST NORTH AMERICA, INC.; MIMECAST SERVICES LIMITED
To: ARES CAPITAL CORPORATION
Reel/Frame 060132/0429 →
RELEASE OF SECURITY INTEREST Recorded May 19, 2022
From: JPMORGAN CHASE BANK, N.A.
To: MIMECAST SERVICES LTD.; ETORCH INC.
Reel/Frame 059962/0294 →
SECURITY AGREEMENT Recorded Jul 23, 2018
From: MIMECAST SERVICES LIMITED
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 046616/0242 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 27, 2017
From: MIMECAST NORTH AMERICA, INC.
To: MIMECAST SERVICES LTD.
Reel/Frame 042821/0798 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 5, 2014
From: BAMFORD, CHRISTOPHER D.; JORDAN, CLIVE N.
To: MIMECAST NORTH AMERICA, INC.
Reel/Frame 032360/0012 →
Continuity (1)
Related Publication 20150074085A1 · Mar 12, 2015