IP Library › Granted Patent US 11,170,055
Granted Patent B2
US 11,170,055 · App. 16/235,112 · Granted Nov 9, 2021

Artificial intelligence augmented document capture and processing systems and methods

Inventor: Gareth Edward Hutchins (Farnham, GB)
Assignee: OPEN TEXT SA ULC
G06F16/93G06F16/285G06F40/20G06K9/00449G06K9/00463G06K9/00469G06N5/02G06N20/00G06K2209/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,170,055
App. No.
16/235,112
Granted
Nov 9, 2021
Kind
B2
Abstract

A document capture server receives a document image from a document capture client and processes the image into an electronic document containing textual content. During capture, the document capture server determines a graphical layout of the document, extracts keywords from the document, classifies the document accordingly, and calls an artificial intelligence (AI) platform to gain insights on the textual content. The AI platform analyzes the textual content and returns additional, insightful data such as a sentiment of the textual content. The document capture server can validate the additional data, integrate the additional data in a process or workflow, and/or provide the textual content and the additional data to a content repository or a computing facility operating in an enterprise computing environment. The document capture server can provide validated data to the AI platform to improve future analyses by the AI platform.

Claims (57)

1. A method, comprising:

receiving, by a document capture server computer from a document capture module running on a client device, an image of a paper document;

processing, by the document capture server computer, the image received from the document capture module on the client device into an electronic document, the processing producing textual content of the electronic document;

extracting, by the document capture server computer, keywords from the textual content of the electronic document;

classifying the electronic document, the classifying performed by the document capture server computer based at least in part on the keywords extracted from the textual content of the electronic document;

making, by the document capture server computer, a call to an artificial intelligence (Al) platform server computer, the call containing the textual content of the electronic document and a class of the electronic document, wherein the Al platform server computer analyzes the textual content of the electronic document utilizing the class and returns additional data including a sentiment of the textual content of the electronic document;

validating, by the document capture server computer, the additional data returned by the Al platform server computer; and

providing the electronic document and the additional data to a content repository or a computing facility operating in an enterprise computing environment.

2. The method according to claim 1 , wherein the call includes an identification of a knowledge base accessible by the Al platform server computer, the knowledge base specific to the class of the electronic document.

3. The method according to claim 1 , wherein the additional data includes at least an entity, a summary, a category, or a concept that the Al platform server computer learned from the textual content.

4. The method according to claim 1 , further comprising:

providing a validated concept or entity from the validating to the Al platform server computer to improve future analyses by the Al platform server computer.

5. The method according to claim 1 , wherein the processing comprises performing at least a format conversion on the image, an image enhancement on the image, or an optical character recognition procedure on the image.

6. The method according to claim 1 , further comprising:

determining a graphical layout of the electronic document; and

classifying the electronic document based at least in part on the graphical layout of the electronic document.

7. The method according to claim 1 , further comprising:

performing a zonal extraction of a form in the image; or

performing a freeform extraction for a regular expression in the image.

8. A system, comprising:

a processor;

a non-transitory computer-readable medium; and

stored instructions translatable by the processor to perform:

receiving, from a document capture module running on a client device, an image of a paper document;

processing the image received from the document capture module on the client device into an electronic document, the processing producing textual content of the electronic document;

extracting keywords from the textual content of the electronic document;

classifying the electronic document, the classifying based at least in part on the keywords extracted from the textual content of the electronic document;

making a call to an artificial intelligence (Al) platform server computer, the call containing the textual content of the electronic document and a class of the electronic document, wherein the Al platform server computer analyzes the textual content of the electronic document utilizing the class and returns additional data including a sentiment of the textual content of the electronic document;

validating the additional data returned by the Al platform server computer; and

providing the electronic document and the additional data to a content repository or a computing facility operating in an enterprise computing environment.

9. The system of claim 8 , wherein the call includes an identification of a knowledge base accessible by the Al platform server computer, the knowledge base specific to the class of the electronic document.

10. The system of claim 8 , wherein the additional data includes at least an entity, a summary, a category, or a concept that the Al platform server computer learned from the textual content.

11. The system of claim 8 , wherein the stored instructions are further translatable by the processor to perform:

providing a validated concept or entity from the validating to the Al platform server computer to improve future analyses by the Al platform server computer.

12. The system of claim 8 , wherein the processing comprises performing at least a format conversion on the image, an image enhancement on the image, or an optical character recognition procedure on the image.

13. The system of claim 8 , wherein the stored instructions are further translatable by the processor to perform:

determining a graphical layout of the electronic document; and

classifying the electronic document based at least in part on the graphical layout of the electronic document.

14. The system of claim 8 , wherein the stored instructions are further translatable by the processor to perform:

performing a zonal extraction of a form in the image; or

performing a freeform extraction for a regular expression in the image.

15. A computer program product comprising a non-transitory computer readable medium storing instructions translatable by a processor to perform:

receiving, from a document capture module running on a client device, an image of a paper document;

processing the image received from the document capture module on the client device into an electronic document, the processing producing textual content of the electronic document;

extracting keywords from the textual content of the electronic document;

classifying the electronic document, the classifying based at least in part on the keywords extracted from the textual content of the electronic document;

making a call to an artificial intelligence (Al) platform server computer, the call containing the textual content of the electronic document and a class of the electronic document, wherein the Al platform server computer analyzes the textual content of the electronic document utilizing the class and returns additional data including a sentiment of the textual content of the electronic document;

validating the additional data returned by the Al platform server computer; and

providing the electronic document and the additional data to a content repository or a computing facility operating in an enterprise computing environment.

16. The computer program product of claim 15 , wherein the call includes an identification of a knowledge base accessible by the Al platform server computer, the knowledge base specific to the class of the electronic document.

17. The computer program product of claim 15 , wherein the additional data includes at least an entity, a summary, a category, or a concept that the Al platform server computer learned from the textual content.

18. The computer program product of claim 15 , wherein the instructions are further translatable by the processor to perform:

providing a validated concept or entity from the validating to the Al platform server computer to improve future analyses by the Al platform server computer.

19. The computer program product of claim 15 , wherein the processing comprises performing at least a format conversion on the image, an image enhancement on the image, or an optical character recognition procedure on the image.

20. The computer program product of claim 15 , wherein the instructions are further translatable by the processor to perform:

determining a graphical layout of the electronic document; and

classifying the electronic document based at least in part on the graphical layout of the electronic document.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 28, 2018
From: HUTCHINS, GARETH EDWARD
To: OPEN TEXT SA ULC
Reel/Frame 047994/0571 →
Continuity (1)
Related Publication 20200210490A1 · Jul 2, 2020
Cited By (1)
US 12,596,681