IP Library Granted Patent US 11,715,310
Granted Patent B1
US 11,715,310 · App. 17/062,414 · Granted Aug 1, 2023

Using neural network models to classify image objects

Inventors: Daniel Sammons (Alameda, CA); Andy Mahdavi (San Francisco, CA)
Assignee: States Title, LLC
G06V10/22G06F18/214G06F18/2163G06N3/08G06V30/412G06V30/413G06V30/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,715,310
App. No.
17/062,414
Granted
Aug 1, 2023
Kind
B1
Abstract

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for machine learning. One of the methods includes receiving an image; providing the image to a neural network model, wherein the neural network model is trained to output predictions of one or more locations within the image and corresponding classifications; extracting text content within one or more of the one or more locations; analyzing the extracted text content using the corresponding classifications to evaluate one or more of external consistency with other data records or internal consistency with content from one or more of the particular locations; and generating one or more outputs based on the analyzing.

Claims (36)

1 . A method comprising:

receiving an image;

providing the image to a neural network model, wherein the neural network model is trained to output predictions of one or more locations within the image, each location defined by a respective bounding box, and corresponding classifications for each location;

extracting text content within one or more of the bounding boxes output by the neural network model;

analyzing the extracted text content using the corresponding classifications to evaluate one or more of external consistency with other data records or internal consistency with content from at least one other location of the one or more locations; and

generating one or more outputs based on results of the analyzing the extracted text content to evaluate one or more of external consistency or internal consistency.

2 . The method of claim 1 , wherein each location within the image is defined by bounding box coordinates defining a particular region of the image.

3 . The method of claim 2 , wherein the extracting text content comprises performing character recognition within the boundaries of each bounding box.

4 . The method of claim 1 , wherein the neural network model is a convolutional neural network and wherein the convolutional neural network is trained based on a collection of labeled images defining coordinates of regions of interest in the respective documents along with labeled classifications for each region of interest.

5 . The method of claim 4 , wherein the collection of labeled images comprise both human-labeled images and augmented images.

6 . The method of claim 1 , wherein the generating one or more outputs comprises flagging the image as requiring additional evaluation.

7 . The method of claim 1 , wherein each image corresponds to a document, wherein the document is a form having a plurality of fields, each field having a form label and wherein one or more of the fields have been filled out with information responsive to the corresponding form labels.

8 . One or more non-transitory computer storage media encoded with instructions that, when executed by one or more computers, cause the one or more computers to perform operations comprising:

receiving an image;

providing the image to a neural network model, wherein the neural network model is trained to output predictions of one or more locations within the image, each location defined by a respective bounding box, and corresponding classifications for each location;

extracting text content within one or more of the bounding boxes output by the neural network model;

analyzing the extracted text content using the corresponding classifications to evaluate one or more of external consistency with other data records or internal consistency with content from at least one other location of the one or more locations; and

generating one or more outputs based on results of the analyzing the extracted text content to evaluate one or more of external consistency or internal consistency.

9 . The non-transitory computer storage media of claim 8 , wherein each location within the image is defined by bounding box coordinates defining a particular region of the image.

10 . The non-transitory computer storage media of claim 9 , wherein the extracting text content comprises performing character recognition within the boundaries of each bounding box.

11 . The non-transitory computer storage media of claim 8 , wherein the neural network model is a convolutional neural network and wherein the convolutional neural network is trained based on a collection of labeled images defining coordinates of regions of interest in the respective documents along with labeled classifications for each region of interest.

12 . The non-transitory computer storage media of claim 11 , wherein the collection of labeled images comprise both human-labeled images and augmented images.

13 . The non-transitory computer storage media of claim 8 , wherein the generating one or more outputs comprises flagging the image as requiring additional evaluation.

14 . The non-transitory computer storage media of claim 8 , wherein each image corresponds to a document, wherein the document is a form having a plurality of fields, each field having a form label and wherein one or more of the fields have been filled out with information responsive to the corresponding form labels.

15 . A system comprising one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving an image;

providing the image to a neural network model, wherein the neural network model is trained to output predictions of one or more locations within the image, each location defined by a respective bounding box, and corresponding classifications for each location;

extracting text content within one or more of the bounding boxes output by the neural network model;

analyzing the extracted text content using the corresponding classifications to evaluate one or more of external consistency with other data records or internal consistency with content from at least one other location of the one or more locations; and

generating one or more outputs based on results of the analyzing the extracted text content to evaluate one or more of external consistency or internal consistency.

16 . The system of claim 15 , wherein each location within the image is defined by bounding box coordinates defining a particular region of the image.

17 . The system of claim 16 , wherein the extracting text content comprises performing character recognition within the boundaries of each bounding box.

18 . The system of claim 15 , wherein the neural network model is a convolutional neural network and wherein the convolutional neural network is trained based on a collection of labeled images defining coordinates of regions of interest in the respective documents along with labeled classifications for each region of interest.

19 . The system of claim 18 , wherein the collection of labeled images comprise both human-labeled images and augmented images.

20 . The system of claim 15 , wherein the generating one or more outputs comprises flagging the image as requiring additional evaluation.

21 . The system of claim 15 , wherein each image corresponds to a document, wherein the document is a form having a plurality of fields, each field having a form label and wherein one or more of the fields have been filled out with information responsive to the corresponding form labels.

Assignments (9)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 21, 2024
From: STATES TITLE, LLC
To: DOMA TECHNOLOGY LLC
Reel/Frame 069659/0492 →
RELEASE OF SECURITY INTEREST Recorded Sep 30, 2024
From: ALTER DOMUS (US) LLC
To: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
Reel/Frame 068742/0050 →
RELEASE OF SECURITY INTEREST Recorded Sep 30, 2024
From: HUDSON STRUCTURED CAPITAL MANAGEMENT LTD.
To: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
Reel/Frame 068742/0095 →
SECURITY INTEREST Recorded May 3, 2024
From: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
To: ALTER DOMUS (US) LLC
Reel/Frame 067312/0857 →
CHANGE OF NAME Recorded Oct 1, 2021
From: STATES TITLE, INC.
To: STATES TITLE, LLC
Reel/Frame 057684/0602 →
CORRECTIVE ASSIGNMENT TO CORRECT THE APPLICATION NUMBERS 10255550, 10510009 AND 10755184 TO PATENT NUMBERS 10255550, 10510009 AND 10755184 PREVIOUSLY RECORDED ON REEL 054804 FRAME 0211. ASSIGNOR(S) HEREBY CONFIRMS THE SECURITY INTEREST. Recorded Jan 8, 2021
From: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
To: HUDSON STRUCTURED CAPITAL MANAGEMENT LTD
Reel/Frame 056322/0310 →
SECURITY INTEREST Recorded Jan 5, 2021
From: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
To: HUDSON STRUCTURED CAPITAL MANAGEMENT LTD.
Reel/Frame 054812/0286 →
SECURITY INTEREST Recorded Jan 4, 2021
From: STATES TITLE HOLDING, INC.; TITLE AGENCY HOLDCO, LLC
To: HUDSON STRUCTURED CAPITAL MANAGEMENT LTD.
Reel/Frame 054804/0211 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 16, 2020
From: SAMMONS, DANIEL; MAHDAVI, ANDY
To: STATES TITLE, INC.
Reel/Frame 054083/0155 →
Cited By (3)
US 12,525,046 US 12,602,791 US 12,608,967