IP Library Granted Patent US 8,069,174
Granted Patent B2
US 8,069,174 · App. 12/945,710 · Granted Nov 29, 2011

System and method for automatic anthology creation using document aspects

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,069,174
App. No.
12/945,710
Granted
Nov 29, 2011
Kind
B2
Abstract

A generic and expandable document aspect system and method for searching, browsing, presenting, and interacting with data assembled from document contents and related external data is provided. New varieties of document aspects are added to existing installations and can be accessed by users without requiring upgrades to server or clients, for example by using plug-in technology.

Claims (49)

1. A computer implemented method for searching, browsing, presenting, and interacting with data assembled from contents of at least one document and any related external data, comprising the steps of:

for a document aspect comprising a portion of said document's contents in user viewable form, where said portion is all and only those parts of said at least one document that have something semantically useful in common, and for one or more predetermined document aspect types within said at least one document, performing the following steps:

extracting from said at least one document one or more user viewable objects related to said predetermined type in accordance with said document aspect to yield a document aspect instance;

for any related external data, extracting from a related external data source one or more user viewable objects related to said predetermined type in accordance with said document aspect;

storing said document aspect instance extracted from said at least one document and any user viewable objects extracted from said related external data source, as a user viewable assembled collection in a repository for a subsequent use; and

when said document aspect instance is accessed, providing said assembled collection to a user.

2. The method of claim 1 , further comprising the step of:

post-processing said assembled collection by any of or any combination of searching, browsing, presenting, and interacting.

3. The method of claim 1 , wherein an extracted object comprises either of or the combination of:

a portion of said document's contents and relevant metadata obtained from sources other than the document's contents.

4. The method of claim 1 , said extracting step further comprising the step of:

formatting said one or more extracted objects and storing said one or more formatted extracted objects.

5. The method of claim 1 , wherein a new predetermined type and said extracting and storing steps can be added as a piece of server plug-in software, such that a client recognizes and presents said assembled collection without need for client upgrades or plug-ins.

6. The method of claim 1 , wherein said extracting and storing steps are performed responsive to a trigger event.

7. The method of claim 6 , wherein a trigger event is any of:

document submission, periodic spidering, administrator request, and user request via a servlet.

8. The method of claim 1 , wherein said repository is a server's collection of per-document data stored in a hierarchy of files on a hard disk.

9. The method of claim 1 , wherein said stored collection is a cross-document collection, formed from a union of collections of multiple documents.

10. The method of claim 1 , wherein said stored collection of extracted objects is associated with a user.

11. The method of claim 1 , wherein said extracting step uses heuristic techniques to identify objects of said predetermined type.

12. The method of claim 1 , wherein said extracting and storing steps are encapsulated in a self-contained jar.

13. The method of claim 1 , wherein said assembled collection comprises one or more collections hierarchically.

14. The method of claim 1 , wherein when said document is accessed, further comprising the step of:

requesting available collections for presenting.

15. The method of claim 1 , wherein said stored collection of extracted objects is associated with a user.

16. The method of claim 1 , wherein said extracting step uses heuristic techniques to identify objects of said predetermined type.

17. A system for searching, browsing, presenting, and interacting with data assembled from contents of at least one document and any related external data, comprising:

for a document aspect comprising a portion of said document's contents in user viewable form, where said portion is all and only those parts of said at least one document that have something semantically useful in common, and for one or more predetermined types within said at least one document, means for performing the following:

extracting from said at least one document one or more user viewable objects related to said predetermined type in accordance with said document aspect to yield a document aspect instance;

for any related external data, extracting from a related external data source one or more user viewable objects related to said predetermined type in accordance with said document aspect;

storing said document aspect instance extracted from said at least one document and any user viewable objects extracted from said related external data source, as a user viewable assembled collection in a repository for a subsequent use; and

when said document aspect instance is accessed, means for providing said assembled collection to a user.

18. The system of claim 17 , further comprising:

means for post-processing said assembled collection by any of or any combination of searching, browsing, presenting, and interacting.

19. The system of claim 17 , wherein an extracted object comprises either of or the combination of: a portion of said document's contents and relevant metadata obtained from sources other than the document's contents.

20. The system of claim 17 , said means for extracting further comprising:

means for formatting said one or more extracted objects and means for storing said one or more formatted extracted objects.

21. The system of claim 17 , wherein a new predetermined type and said associated extracting and storing steps can be added as a piece of server plug-in software, such that a client recognizes and presents said associated collection without need for client upgrades or plug-ins.

22. The system of claim 17 , wherein said extracting and storing steps are performed responsive to a trigger event.

23. The system of claim 22 , wherein a trigger event is any of: document submission, periodic spidering, administrator request, and user request via a servlet.

24. The system of claim 17 , wherein said repository is a server's collection of per-document data stored in a hierarchy of files on a hard disk.

25. The system of claim 17 , wherein said stored collection is a cross-document collection, formed from a union of collections of multiple documents.

26. The system of claim 17 , wherein said stored collection of extracted objects is associated with a user.

27. The system of claim 17 , wherein said means for extracting uses heuristic techniques to identify objects of said predetermined type.

28. The system of claim 17 , wherein said means for extracting and storing are encapsulated in a self-contained jar.

29. The system of claim 17 , wherein said one or more collections comprises one or more collections hierarchically.

30. The system of claim 17 , wherein when said document is accessed, further comprising:

means for requesting available collections for presenting.

31. The system of claim 17 , wherein said stored collection of extracted objects is associated with a user.

Assignments (9)
PATENT SECURITY AGREEMENT Recorded Apr 9, 2025
From: PROQUEST LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS COLLATERAL AGENT
Reel/Frame 070793/0579 →
PATENT SECURITY AGREEMENT Recorded Apr 9, 2025
From: PROQUEST LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS COLLATERAL AGENT
Reel/Frame 070793/0587 →
PATENT SECURITY AGREEMENT Recorded Apr 9, 2025
From: PROQUEST LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 070793/0595 →
TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENT COLLATERAL Recorded Dec 1, 2021
From: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
To: PROQUEST LLC; EBRARY
Reel/Frame 058294/0036 →
SECURITY INTEREST Recorded Dec 17, 2015
From: PROQUEST LLC; EBRARY
To: BANK OF AMERICA, N.A. AS COLLATERAL AGENT
Reel/Frame 037318/0946 →
RELEASE OF SECURITY INTEREST Recorded Oct 30, 2014
From: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
To: PROQUEST LLC; EBRARY; DIALOG LLC; CAMBRIDGE SCIENTIFIC ABSTRACTS, LIMITED PARTNERSHIP; PROQUEST INFORMATION AND LEARNING LLC
Reel/Frame 034076/0672 →
SECURITY INTEREST Recorded Oct 24, 2014
From: PROQUEST LLC; EBRARY
To: BANK OF AMERICA, N.A. AS COLLATERAL AGENT
Reel/Frame 034033/0293 →
AMENDED AND RESTATED INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Apr 24, 2012
From: CAMBRIDGE SCIENTIFIC ABSTRACTS, LIMITED PARTNERSHIP; PROQUEST LLC; PROQUEST INFORMATION AND LEARNING LLC; DIALOG LLC; EBRARY
To: MORGAN STANLEY SENIOR FUNDING, INC.
Reel/Frame 028101/0914 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT SUPPLEMENT Recorded Feb 9, 2011
From: EBRARY
To: MORGAN STANLEY & CO. INCORPORATED
Reel/Frame 025777/0332 →