IP Library Granted Patent US 9,075,847
Granted Patent B2
US 9,075,847 · App. 13/687,829 · Granted Jul 7, 2015

Methods, apparatus and system for identifying a document

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,075,847
App. No.
13/687,829
Granted
Jul 7, 2015
Kind
B2
Abstract

A computerized method for identifying a document. A signature may be determined for a first document and compared with a signature for each of one or more additional documents. A document similarity score may be determined and one or more similar documents may be identified based on the document similarity score.

Claims (32)

1. A computerized method for document processing, the method comprising:

obtaining one or more counts of keypoints of a first document and one or more counts of remarkable characteristics of a second document;

comparing at least one of the one or more counts of keypoints of the first document and at least one of one or more corresponding counts of remarkable characteristics of the second document, wherein the comparing step further comprises selecting a cell of a plurality of cells of the first document as a base cell and determining a value of a difference between a keypoint count of the base cell associated with the first document and a corresponding keypoint count of a base cell of the second document to determine a base cell similarity score; and

determining a document similarity score between the first and second document.

2. The computerized method of claim 1 , wherein the second document is one of a plurality of stored documents and further comprising repeating the comparing and determining steps for each of one or more of the plurality of stored documents.

3. The computerized method of claim 2 , further comprising identifying one or more of the plurality of stored documents based on one or more lowest document similarity scores, a lower similarity score corresponding to a higher similarity.

4. The computerized method of claim 1 , further comprising determining the one or more counts of keypoints of the first document for each of a plurality of cells of the first document.

5. The computerized method of claim 1 , wherein the one or more counts of keypoints of the first document and the one or more corresponding counts of keypoints of a second document are normalized values.

6. The computerized method of claim 1 , further comprising determining and persistently storing the one or more counts of keypoints for each of a plurality of cells of one or more of the plurality of stored documents.

7. The computerized method of claim 1 , further comprising normalizing and persistently storing the one or more counts of keypoints for each of one or more of the plurality of stored documents.

8. The computerized method of claim 1 , further comprising repeating the selecting and determining the value steps for each of a plurality of cells of the first document.

9. The computerized method of claim 8 , wherein the determining the document similarity score step further comprises summing the determined values.

10. The computerized method of claim 8 , wherein the determining the value step further comprises: selecting a neighboring cell of the base cell associated with the first document; and determining a value of a difference between a remarkable characteristic count of the selected neighboring cell associated with the first document and a keypoint count of a corresponding neighboring cell of the second document to determine a neighbor cell similarity score.

11. The computerized method of claim 10 , wherein the determining the document similarity score step further comprises summing the determined base cell similarity scores and one or more neighbor cell similarity scores.

12. The computerized method of claim 10 , further comprising weighting each neighbor cell similarity score and wherein the determining the document similarity score step further comprises summing the determined base cell similarity scores and one or more weighted neighbor cell similarity scores.

13. A computerized method for document processing, the method comprising:

capturing an image of a printed document;

determining one or more counts of keypoints of the printed document;

comparing at least one of the one or more counts of keypoints of the first document and at least one of one or more corresponding counts of remarkable characteristics of a second document of a plurality of stored documents; determining a document similarity score, wherein the comparing step further comprises selecting a cell of a plurality of cells of the first document as a base cell and determining a value of a difference between a keypoint count of the base cell associated with the first document and a corresponding keypoint count of a base cell of the second document to determine a base cell similarity score;

repeating the comparing and determining steps for each of the plurality of stored documents; and

identifying one or more of the plurality of stored documents corresponding to one or more lowest document similarity scores.

14. The computerized method of claim 13 , further comprising obtaining one or more of the identified stored documents.

15. The computerized method of claim 14 , further comprising providing a link to one or more of the identified stored documents.

16. Apparatus for document processing, the apparatus comprising:

a processor; memory to store instructions that, when executed by the processor cause the processor to:

obtain one or more counts of keypoints of a first document and one or more counts of keypoints of a second document;

compare at least one of the one or more counts of keypoints of the first document and at least one of one or more corresponding counts of keypoints of the second document, wherein the comparing step further comprises selecting a cell of a plurality of cells of the first document as a base cell and determining a value of a difference between a keypoint count of the base cell associated with the first document and a corresponding keypoint count of a base cell of the second document to determine a base cell similarity score; and

determine a document similarity score between the first and second document.

17. Non-transitory a computer-readable medium embodying instructions that, when executed by a processor perform operations comprising:

obtaining one or more counts of keypoints of a first document and one or more counts of keypoints of a second document;

comparing at least one of the one or more counts of keypoints of the first document and at least one of one or more corresponding counts of keypoints of the second document, wherein the comparing step further comprises selecting a cell of a plurality of cells of the first document as a base cell and determining a value of a difference between a keypoint count of the base cell associated with the first document and a corresponding keypoint count of a base cell of the second document to determine a base cell similarity score; and

determining a document similarity score between the first and second document.

Assignments (3)
CHANGE OF NAME Recorded Aug 26, 2014
From: SAP AG
To: SAP SE
Reel/Frame 033625/0223 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 11, 2013
From: HOBBS, GODFREY
To: SAP AG
Reel/Frame 031577/0037 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 28, 2012
From: RUPP, STEFANIE; GUSTAV, AXEL
To: SAP AG
Reel/Frame 029367/0053 →