IP Library › Granted Patent US 11,367,297
Granted Patent B2
US 11,367,297 · App. 16/907,935 · Granted Jun 21, 2022

Method of automatically extracting information of a predefined type from a document

Inventors: Sebastian Andreas Bildner (Munich, DE); Paul Krion (Munich, DE); Thomas Stark (Munich, DE); Martin Christopher Stämmler (Munich, DE); Martin Von Schledorn (Munich, DE); Jürgen Oesterle (Munich, DE); Renjith Karimattathil Sasidharan (Bangalore, IN)
Assignee: Amadeus S.A.S.
G06V30/414G06K9/6256G06V10/22G06V10/44G06V30/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,367,297
App. No.
16/907,935
Granted
Jun 21, 2022
Kind
B2
Abstract

Method and system of automatically extracting information of a predefined type from a document is provided. The method comprises using an object detection algorithm to identify at least one segment of the document that is likely to comprise the information of the predefined type. The method further comprises building at least one bounding box corresponding to the at least one segment and if the bounding box is likely to comprise the information of the predefined type extracting the information comprised by the bounding box from the at least one bounding box.

Claims (50)

1. A method comprising:

identifying, by an object detection algorithm, at least one segment of a document that is likely to comprise information of a predefined type;

building at least one bounding box corresponding to the at least one segment;

identifying that the at least one bounding box likely comprises the information of the predefined type; and

extracting, by a character identification algorithm, the information of the predefined type from the at least one bounding box based on identifying, by a multilayer neural network, the information of the predefined type based on characteristics of the information of the predefined type, wherein the neural network includes:

a first layer configured to differentiate between empty regions and non-empty regions of the document and to identify basic patterns present on the document, and

a second layer configured to identify shapes that are more complex compared to the basic patterns present on the document.

2. The method of claim 1 wherein the characteristics of the information of the predefined type comprise a number format and at least one of a comma or a decimal point.

3. The method of claim 1 wherein the multilayer neural network is compatible with a decision layer, and the decision layer is configured to detect at least one of (i) a location of the bounding box, (ii) a height and a width of a bounding box, and (iii) a classification score indicating a classification of a detected character.

4. The method of claim 1 wherein a convolutional multilayer neural network is used by the object detection algorithm.

5. The method of claim 1 wherein a fully-convolutional neural network is used by the object detection algorithm and/or the character identification algorithm.

6. The method of claim 1 further comprising:

training the neural network with a plurality of documents in a training activity to correctly extract the information of the predefined type.

7. The method of claim 1 wherein a probability value is assigned to the at least one bounding box, and the probability value is indicative of the probability that a certain bounding box contains the information of the predefined type.

8. The method of claim 1 further comprising:

identifying a character by the character identification algorithm,

wherein a probability value is assigned to the character, and the probability value is indicative of the probability that the identified character is identical with a character actually comprised by the information of the predefined type.

9. The method of claim 1 further comprising:

assigning a probability value assigned to the at least one bounding box; and

assigning probability values to characters within the at least one bounding box in order to provide a combined confidence score.

10. The method of claim 1 wherein the document is digitally scanned from a paper-based document, and the information of the predefined type is at least one of a creation date, a total amount, an arrival/departure date, a VAT-ID, a receipt id, and an invoice number.

11. The method of claim 1 wherein the document is a paper-based receipt or a paper-based invoice.

12. A system comprising:

a computing device; and

a computer-readable storage medium comprising a set of instructions that upon execution by the computing device cause the system to:

identify, by an object detection algorithm, at least one segment of a document that is likely to comprise information of a predefined type;

build at least one bounding box corresponding to the at least one segment;

identify that the at least one bounding box likely comprises the information of the predefined type; and

extract, by a character identification algorithm, the information of the predefined type from the at least one bounding box based on identifying, by a multilayer neural network, the information of the predefined type based on characteristics of the information of the predefined type, wherein the neural network includes:

a first layer configured to differentiate between empty regions and non-empty regions of the document and to identify basic patterns present on the document, and

a second layer configured to identify shapes that are more complex compared to the basic patterns present on the document.

13. The system of claim 12 wherein the characteristics of the information of the predefined type comprise a number format and at least one of a comma or a decimal point.

14. The system of claim 12 wherein the multilayer neural network is compatible with a decision layer, and the decision layer is configured to detect at least one of (i) a location of the bounding box, (ii) a height and a width of a bounding box, and (iii) a classification score indicating a classification of a detected character.

15. The system of claim 12 wherein the set of instructions, upon execution by the computing device, further cause the system to:

train the neural network with a plurality of documents in a training activity to correctly extract the information of the predefined type.

16. The system of claim 12 wherein a probability value is assigned to the at least one bounding box, and the probability value is indicative of the probability that a certain bounding box contains the information of the predefined type.

17. The system of claim 12 wherein the set of instructions, upon execution by the computing device, further cause the system to:

identify a character by the character identification algorithm,

wherein a probability value is assigned to the character, and the probability value is indicative of the probability that the identified character is identical with a character actually comprised by the information of the predefined type.

18. The system of claim 12 wherein the set of instructions, upon execution by the computing device, further cause the system to:

assign a probability value assigned to the at least one bounding box; and

assign probability values to characters within the at least one bounding box in order to provide a combined confidence score.

19. The system of claim 12 wherein the document is a paper-based receipt or a paper-based invoice.

20. A non-transitory computer-readable storage medium comprising computer-readable instructions that upon execution by a processor of a computing device cause the computing device to:

identify, by an object detection algorithm, at least one segment of a document that is likely to comprise information of a predefined type;

build at least one bounding box corresponding to the at least one segment;

identify that the at least one bounding box likely comprises the information of the predefined type; and

extract, by a character identification algorithm, the information of the predefined type from the at least one bounding box based on identifying, by a multilayer neural network, the information of the predefined type based on characteristics of the information of the predefined type, wherein the neural network includes:

a first layer configured to differentiate between empty regions and non-empty regions of the document and to identify basic patterns present on the document, and

a second layer configured to identify shapes that are more complex compared to the basic patterns present on the document.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 10, 2020
From: BILDNER, SEBASTIAN ANDREAS; KRION, PAUL; STARK, THOMAS; STÄMMLER, MARTIN CHRISTOPHER; VON SCHLEDORN, MARTIN; OESTERLE, JÜRGEN; SASIDHARAN, RENJITH KARIMATTATHIL
To: AMADEUS S.A.S.
Reel/Frame 053441/0609 →
Priority Claims (1)
FR 1907252 · Jul 1, 2019 · national
Continuity (1)
Related Publication 20210004584A1 · Jan 7, 2021