IP Library Granted Patent US 11,562,586
Granted Patent B2
US 11,562,586 · App. 17/808,157 · Granted Jan 24, 2023

Systems and methods for generating search results based on optical character recognition techniques and machine-encoded text

Inventors: Vinod Balakrishnan (Fremont, CA); Xiaoyu Guo (Sunnyvale, CA)
Assignee: Yahoo Assets LLC
G06V30/153G06F16/332G06F16/338G06N20/00G06V30/414G06V30/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,562,586
App. No.
17/808,157
Granted
Jan 24, 2023
Kind
B2
Abstract

Disclosed are systems and methods for generating search result data based on machine-encoded text generated by computer vision optical character recognition machine learning techniques performed on digital media. The disclosed systems and methods provide a novel framework for performing machine learning visual search or machine learning text extraction techniques on digital media in order to extract and analyze the data therein and further conduct search queries based on the extracted and analyzed data. The disclosed framework may leverage the aforementioned computer vision machine learning techniques in order to provide a user with relevant search results regarding objects and text detect in digital media captured on a user device.

Claims (63)

1. A computer-implemented method comprising:

transmitting an image from a user device to an OCR machine learning module;

receiving machine-encoded text from the OCR machine learning module;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving a selection via a graphical user interface of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

receiving, from the user device, a search request comprising a query of the portion of the machine-encoded text;

retrieving, from a search engine module, media and text data corresponding to the selected portion of the machine-encoded text; and

transmitting, to the user device, search results corresponding to the search request.

2. The computer-implemented method of claim 1 , further comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

3. The computer-implemented method of claim 1 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

4. The computer-implemented method of claim 1 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

5. The computer-implemented method of claim 1 , further comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

6. The computer-implemented method of claim 1 , further comprising:

authenticating the received search request via an API gateway; and

determining whether the search request comprises an authentication token.

7. The computer-implemented method of claim 1 , wherein transmitting to the user device, search results corresponding to the search request further comprises:

displaying search results comprising content ranked according to relevance associated with subject matter or text corresponding to the selected portion of the machine-encoded text.

8. A system comprising:

a memory storage device storing instructions for generating search result data based on machine-encoded text data; and

one or more processors configured to execute the instructions for generating search result data based on machine-encoded text data by performing the steps of:

transmitting an image from a user device to an OCR machine learning module;

receiving machine-encoded text from the OCR machine learning module;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving a selection via a graphical user interface of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

receiving, from the user device, a search request comprising a query of the portion of the machine-encoded text;

retrieving, from a search engine module, media and text data corresponding to the selected portion of the machine-encoded text; and

transmitting, to the user device, search results corresponding to the search request.

9. The system of claim 8 , wherein the processor is further configured to execute the instructions to perform one more additional steps comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

10. The system of claim 8 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

11. The system of claim 8 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

12. The system of claim 8 , wherein the processor is further configured to execute the instructions to perform one more additional steps comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

13. The system of claim 8 , wherein the processor is further configured for authenticating the received search request via an API gateway by determining whether the search request comprises an authentication token.

14. The system of claim 8 , wherein receiving at the user device search results corresponding to the search request further comprises:

displaying search results comprising content ranked according to relevance associated with subject matter or text corresponding to the selected portion of the machine-encoded text.

15. A non-transitory computer-readable medium for generating search result data based on machine-encoded text data, the method comprising:

at least one storage medium with instructions thereon for generating search result data based on machine-encoded text data; and

at least one processor that executes the instructions to perform a method comprising:

transmitting an image from a user device to an OCR machine learning module;

receiving machine-encoded text from the OCR machine learning module;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving a selection via a graphical user interface of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

receiving, from the user device, a search request comprising a query of the portion of the machine-encoded text;

retrieving, from a search engine module, media and text data corresponding to the selected portion of the machine-encoded text; and

transmitting, to the user device, search results corresponding to the search request.

16. The non-transitory computer-readable medium of claim 15 , further comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

17. The non-transitory computer-readable medium of claim 15 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

18. The non-transitory computer-readable medium of claim 15 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

19. The non-transitory computer-readable medium of claim 15 , further comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

20. The non-transitory computer-readable medium of claim 15 , further comprising:

authenticating the received search request via an API gateway; and

determining whether the search request comprises an authentication token.

Assignments (4)
SUPPLEMENTAL PATENT SECURITY AGREEMENT Recorded Sep 17, 2025
From: YAHOO ASSETS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 072915/0540 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 17, 2022
From: BALAKRISHNAN, VINOD; GUO, XIAOYU
To: OATH INC.
Reel/Frame 060826/0977 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 17, 2022
From: YAHOO AD TECH LLC (FORMERLY VERIZON MEDIA INC.)
To: YAHOO ASSETS LLC
Reel/Frame 061201/0818 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 17, 2022
From: OATH INC.
To: VERIZON MEDIA INC.
Reel/Frame 061202/0001 →
Continuity (2)
Continuation 16657201 · Oct 18, 2019
Related Publication 20220319212A1 · Oct 6, 2022