IP Library Granted Patent US 11,893,815
Granted Patent B2
US 11,893,815 · App. 18/146,024 · Granted Feb 6, 2024

Systems and methods for generating search results based on optical character recognition techniques and machine-encoded text

Inventors: Vinod Balakrishnan (Fremont, CA); Xiaoyu Guo (Sunnyvale, CA)
Assignee: Yahoo Assets LLC
G06V30/153G06F16/332G06F16/338G06N20/00G06V30/414G06V30/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,893,815
App. No.
18/146,024
Granted
Feb 6, 2024
Kind
B2
Abstract

Disclosed are systems and methods for generating search result data based on machine-encoded text generated by computer vision optical character recognition machine learning techniques performed on digital media. The disclosed systems and methods provide a novel framework for performing machine learning visual search or machine learning text extraction techniques on digital media in order to extract and analyze the data therein and further conduct search queries based on the extracted and analyzed data. The disclosed framework may leverage the aforementioned computer vision machine learning techniques in order to provide a user with relevant search results regarding objects and text detect in digital media captured on a user device.

Claims (60)

1. A computer-implemented method comprising:

receiving, from a user device an image comprising text;

converting the text into machine-encoded text;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving, from the user device, a search request comprising a selection of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

identifying media and text data corresponding to the selected portion of the machine encoded text; and

transmitting, to the user device, the identified media and text data to the user device.

2. The computer-implemented method of claim 1 , further comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

3. The computer-implemented method of claim 1 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

4. The computer-implemented method of claim 1 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

5. The computer-implemented method of claim 1 , further comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

6. The computer-implemented method of claim 1 , further comprising:

authenticating the received search request via an API gateway; and

determining whether the search request comprises an authentication token.

7. The computer-implemented method of claim 1 , wherein transmitting to the user device, search results corresponding to the search request further comprises:

displaying search results comprising content ranked according to relevance associated with subject matter or text corresponding to the selected portion of the machine-encoded text.

8. A system comprising:

a memory storage device storing instructions for generating search result data based on machine-encoded text data; and

one or more processors configured to execute the instructions for generating search result data based on machine-encoded text data by performing the steps of:

receiving, from a user device an image comprising text;

converting the text into machine-encoded text;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving, from the user device, a search request comprising a selection of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

identifying media and text data corresponding to the selected portion of the machine encoded text; and

transmitting, to the user device, the identified media and text data to the user device.

9. The system of claim 8 , wherein the processor is further configured to execute the instructions to perform one more additional steps comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

10. The system of claim 8 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

11. The system of claim 8 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

12. The system of claim 8 , wherein the processor is further configured to execute the instructions to perform one more additional steps comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

13. The system of claim 8 , wherein the processor is further configured for authenticating the received search request via an API gateway by determining whether the search request comprises an authentication token.

14. The system of claim 8 , wherein receiving at the user device search results corresponding to the search request further comprises:

displaying search results comprising content ranked according to relevance associated with subject matter or text corresponding to the selected portion of the machine-encoded text.

15. A non-transitory computer-readable medium for generating search result data based on machine-encoded text data, the method comprising:

at least one storage medium with instructions thereon for generating search result data based on machine-encoded text data; and

at least one processor that executes the instructions to perform a method comprising:

receiving, from a user device an image comprising text;

converting the text into machine-encoded text;

rendering a geometrical bounding element comprising the machine-encoded text;

receiving, from the user device, a search request comprising a selection of at least a portion of the machine-encoded text corresponding to the geometrical bounding element;

identifying media and text data corresponding to the selected portion of the machine encoded text; and

transmitting, to the user device, the identified media and text data to the user device.

16. The non-transitory computer-readable medium of claim 15 , further comprising:

automatically determining a location of the user device based on GPS coordinates or identifying a name of the location from object data or text data in the image.

17. The non-transitory computer-readable medium of claim 15 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

analyzing text in the image by conducting natural language processing techniques on the analyzed text.

18. The non-transitory computer-readable medium of claim 15 , wherein rendering a geometrical bounding element comprising the machine-encoded text, further comprises:

generating an enhanced location profile based on a determined location and search results corresponding to the search request.

19. The non-transitory computer-readable medium of claim 15 , further comprising:

determining whether the geometrical bounding element is rendered visible or transparent based on a determination of an amount of text in the image.

20. The non-transitory computer-readable medium of claim 15 , further comprising:

authenticating the received search request via an API gateway; and

determining whether the search request comprises an authentication token.

Assignments (4)
SUPPLEMENTAL PATENT SECURITY AGREEMENT Recorded Sep 17, 2025
From: YAHOO ASSETS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 072915/0540 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 23, 2022
From: BALAKRISHNAN, VINOD; GUO, XIAOYU
To: OATH INC.
Reel/Frame 062196/0020 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 23, 2022
From: OATH INC.
To: VERIZON MEDIA INC.
Reel/Frame 062211/0683 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 23, 2022
From: YAHOO AD TECH LLC (FORMERLY VERIZON MEDIA INC.)
To: YAHOO ASSETS LLC
Reel/Frame 062212/0001 →
Continuity (3)
Continuation 17808157 · Jun 22, 2022
Continuation 16657201 · Oct 18, 2019
Related Publication 20230126412A1 · Apr 27, 2023
Cited By (1)
US 12,405,125