IP Library Granted Patent US 10,956,786
Granted Patent B2
US 10,956,786 · App. 16/790,945 · Granted Mar 23, 2021

Machine learning based extraction of partition objects from electronic documents

Inventors: Dan G. Tecuci (Austin, TX); Ravi Kiran Reddy Palla (Sunnyvale, CA); Hamid Reza Motahari Nezhad (Los Altos, CA); Vincent Poon (Millbrae, CA); Nigel Paul Duffy (San Francisco, CA); Joseph Nipko (Austin, TX)
G06K9/6257G06F3/0482G06K9/00442G06K9/6263G06N3/02G06N20/20G06K2009/00489
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,956,786
App. No.
16/790,945
Granted
Mar 23, 2021
Kind
B2
Abstract

An object-extraction method includes generating multiple partition objects based on an electronic document, and receiving a first user selection of a data element via a user interface of a compute device. In response to the first user selection, and using a machine learning model, a first subset of partition objects from the multiple partition objects is detected and displayed via the user interface. A user interaction, via the user interface, with one of the partition objects is detected, and in response, a weight of the machine learning model is modified, to produce a modified machine learning model. A second user selection of the data element is received via the user interface, and in response and using the modified machine learning model, a second subset of partition objects from the multiple partition objects is detected and displayed via the user interface, the second subset different from the first subset.

Claims (41)

1. A method, comprising:

automatically generating a modified electronic document, via a processor of a compute device and in response to one of: (1) detecting an error in a first electronic document having an associated domain, or (2) detecting a first user interaction, via a user interface of the compute device, with a representation of the first electronic document;

identifying, via a domain-specific machine learning model for the associated domain, and in response to generating the modified electronic document, a set of objects associated with the modified electronic document;

detecting a second user interaction, via the user interface of the compute device, with a representation of an object from the set of objects;

in response to detecting the second user interaction, modifying the domain-specific machine learning model based on the second user interaction to produce a modified machine learning model;

detecting, based on a signal representing a user-selected data element for a second electronic document having the associated domain, and using the modified machine learning model, a set of objects associated with the second electronic document; and

displaying, via the user interface, a representation of each object from the set of objects associated with the second electronic document.

2. The method of claim 1 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the error is an optical character recognition (OCR) induced error.

3. The method of claim 1 , wherein the domain-specific machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

4. The method of claim 1 , wherein the domain-specific machine learning model operates within an artificial neural network (ANN).

5. The method of claim 1 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the detecting the error in the first electronic document includes receiving an error signal representing a non-extracted data string, associated with the first electronic document, from a named-entity recognition (NER) model.

6. The method of claim 1 , wherein the associated domain is a document ype of the first document.

7. The method of claim 1 , wherein the second user interaction includes one of clicking on or highlighting the representation of the object from the set of objects.

8. An apparatus, comprising:

a processor; and

a memory storing instructions, executable by the processor, to:

automatically generate a modified electronic document, in response to one of: (1) detecting an error in a first electronic document having an associated domain, or (2) detecting a first user interaction, via a user interface of a compute device, with a representation of the first electronic document;

identify, via a domain-specific machine learning model for the associated domain, and in response to generating the modified electronic document, a set of objects associated with the modified electronic document;

detect a second user interaction, via the user interface of the compute device, with a representation of an object from the set of objects;

in response to detecting the second user interaction, modify the domain-specific machine learning model based on the second user interaction to produce a modified machine learning model;

detect, based on a signal representing a user-selected data element a second electronic document having the associated domain, and using the modified machine learning model, a set of objects associated with the second electronic document; and

display, via the user interface, a representation of each object from the set of objects associated with the second electronic document.

9. The apparatus of claim 8 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the error is an optical character recognition (OCR) induced error.

10. The apparatus of claim 8 , wherein the domain-specific machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

11. The apparatus of claim 8 , wherein the domain-specific machine learning model operates within an artificial neural network (ANN).

12. The apparatus of claim 8 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the detecting the error in the first electronic document includes receiving an error signal representing a non-extracted data string; associated with the first electronic document, from a named-entity recognition (NER) model.

13. The apparatus of claim 8 , wherein the associated domain is a document type of the first document.

14. The apparatus of claim 8 , wherein the second user interaction includes one of clicking on or highlighting the representation of the object from the set of objects.

15. A non-transitory, processor-readable medium storing instructions to cause a processor to:

automatically generate a modified electronic document, in response to one of: (1) detecting an error in a first electronic document having an associated domain, or (2) detecting a first user interaction, via a user interface of a compute device, with a representation of the first electronic document;

identify, via a domain-specific machine learning model for the associated domain, and in response to generating the modified electronic document, a set of objects associated with the modified electronic document;

detect a second user interaction, via the user interface of the compute device, with a representation of an object from the set of objects;

in response to detecting the second user interaction, modify the domain-specific machine learning model based on the second user interaction to produce a modified machine learning model;

detect, based on a signal representing a user-selected data element for a second electronic document having the associated domain, and using the modified machine learning model, a set of objects associated with the second electronic document; and

display, via the user interface, a representation of each object from the set of objects associated with the second electronic document.

16. The non-transitory, processor-readable medium of claim 15 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the error is an optical character recognition (OCR) induced error.

17. The non-transitory, processor-readable medium of claim 15 , wherein the domain-specific machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

18. The non-transitory, processor-readable medium of claim 15 , wherein the domain-specific machine learning model operates within an artificial neural network (ANN).

19. The non-transitory, processor-readable medium of claim 15 , wherein the generating the modified electronic document is in response to detecting the error in the first electronic document, and the detecting the error in the first electronic document includes receiving an error signal representing a non-extracted data string, associated with the first electronic document, from a named-entity recognition (NER) model.

20. The non-transitory, processor-readable medium of claim 15 , wherein the associated domain is a document type of the first document.

21. The method of claim 1 , wherein the automatically generating the modified electronic document is in response to detecting the error in the first electronic document, and the error is one of an automated language translation induced error or an automated spell-checking induced error.

Continuity (2)
Division 16382707 · Apr 12, 2019
Related Publication 20200327373A1 · Oct 15, 2020
Cited By (2)
US 12,373,083 US 12,444,163