IP Library Granted Patent US 12,380,117
Granted Patent B2
US 12,380,117 · App. 18/807,305 · Granted Aug 5, 2025

Voice query refinement to embed context in a voice query

Inventors: Rajendran Pichaimurthy (Bangalore, IN); Madhusudhan Seetharam (Bangalore, IN); Harshith Kumar Gejjegondanahally Sreekanth (Bangalore, IN)
Assignee: Adeia Guides Inc.
G06F16/24575G06F16/24522G06F40/253G06F40/289G06F40/30G06V40/20G10L15/22G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,380,117
App. No.
18/807,305
Granted
Aug 5, 2025
Kind
B2
Abstract

Systems and methods are described for providing contextual search results. The system may receive a search query during presentation of a video. If the query is ambiguous, the system accesses some of the frames of the video. The frames are analyzed to identify a performed action depicted in the frames. The system retrieves a keyword related to the identified action. The ambiguous query is augmented with the keyword. The augmented search query is used to search for and output relevant search results.

Claims (54)

1. A method comprising:

identifying a character in a plurality of frames presented during playback of a video while an ambiguous search query is received;

generating a model of at least one movement of the character that occurred during presentation of at least two frames of the plurality of frames;

comparing the model to at least one movement template in a database;

determining, based on the comparing, the model is comparable to the at least one movement template;

in response to determining the model is comparable to the at least one movement template, extracting a keyword from the at least one movement template; and

retrieving search results for display based on the keyword and the ambiguous search query.

2. The method of claim 1 , wherein extracting the keyword comprises retrieving metadata of the at least one movement template.

3. The method of claim 1 , wherein generating the model comprises:

identifying at least two body parts of the character, wherein each of the at least two body parts are visible in at least two frames of the plurality of frames;

computing at least one angle formed between the at least two body parts; and

storing the at least one angle as part of the model.

4. The method of claim 3 , wherein the at least two body parts are identified as a body part combination.

5. The method of claim 4 , further comprising:

searching the database for the body part combination; and

identifying the at least one movement template based on the body part combination.

6. The method of claim 1 , wherein determining the model is comparable to the at least one movement template comprises determining the model comprises a body part combination angle that matches a corresponding body part combination angle of the at least one movement template.

7. The method of claim 1 , further comprising:

generating context data characterizing objects present in the plurality of frames; and

retrieving the search results based on the context data and the ambiguous search query.

8. The method of claim 1 , further comprising:

receiving a spoken search query;

performing speech to text analysis to derive words that comprise the spoken search query; and

determining at least one word of the spoken search query is ambiguous based on analysis of the derived words.

9. The method of claim 8 , wherein determining the at least one word of the spoken search query is ambiguous comprises determining the at least one word is at least one of a pronoun or an auxiliary verb.

10. The method of claim 1 , wherein the plurality of frames comprises a predetermined number of displayed frames after receiving the ambiguous search query.

11. A system comprising:

input circuitry to receive a query during presentation of a video; and

control circuitry configured to:

identify a character in a plurality of frames presented during playback of a video while an ambiguous search query is received;

generate a model of at least one movement of the character that occurred during presentation of at least two frames of the plurality of frames;

compare the model to at least one movement template in a database;

determine, based on the comparing, the model is comparable to the at least one movement template;

in response to determining the model is comparable to the at least one movement template, extract a keyword from the at least one movement template; and

retrieve search results for display based on the keyword and the ambiguous search query.

12. The system of claim 11 , wherein the control circuitry configured to extract the keyword is further configured to retrieve metadata of the at least one movement template.

13. The system of claim 11 , wherein the control circuitry configured to generate the model is further configured to:

identify at least two body parts of the character, wherein each of the at least two body parts are visible in at least two frames of the plurality of frames;

compute at least one angle formed between the at least two body parts; and

store the at least one angle as part of the model.

14. The system of claim 13 , wherein the at least two body parts are identified as a body part combination.

15. The system of claim 14 , wherein the control circuitry is further configured to:

searching the database for the body part combination; and

identifying the at least one movement template based on the body part combination.

16. The system of claim 11 , wherein the control circuitry configured to determine the model is comparable to the at least one movement template is further configured to determine the model comprises a body part combination angle that matches a corresponding body part combination angle of the at least one movement template.

17. The system of claim 11 , wherein the control circuitry is further configured to:

generate context data characterizing objects present in the plurality of frames; and

retrieve the search results based on the context data and the ambiguous search query.

18. The system of claim 11 , wherein the control circuitry is further configured to:

receive a spoken search query;

perform speech to text analysis to derive words that comprise the spoken search query; and

determine at least one word of the spoken search query is ambiguous based on analysis of the derived words.

19. The system of claim 18 , wherein the control circuitry configured to determine the at least one word of the spoken search query is ambiguous is further configured to determine the at least one word is at least one of a pronoun or an auxiliary verb.

20. The system of claim 11 , wherein the plurality of frames comprises a predetermined number of displayed frames after receiving the ambiguous search query.

Assignments (3)
SECURITY INTEREST Recorded May 28, 2025
From: ADEIA INC. (F/K/A XPERI HOLDING CORPORATION); ADEIA HOLDINGS INC.; ADEIA MEDIA HOLDINGS INC.; ADEIA IMAGING LLC; ADEIA MEDIA LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA TECHNOLOGIES INC.; ADEIA GUIDES INC.; ADEIA SOLUTIONS LLC; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR INTELLECTUAL PROPERTY LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA PUBLISHING INC.
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 071454/0343 →
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0171 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2024
From: PICHAIMURTHY, RAJENDRAN; SEETHARAM, MADHUSUDHAN; SREEKANTH, HARSHITH KUMAR GEJJEGONDANAHALLY
To: ROVI GUIDES, INC.
Reel/Frame 068687/0001 →
Continuity (4)
Continuation 18136642 · Apr 19, 2023
Continuation 17866852 · Jul 18, 2022
Continuation 16206385 · Nov 30, 2018
Related Publication 20250053566A1 · Feb 13, 2025
References Cited (22)
US 9830391B1 · Bakir et al. · 2017 [cited by applicant]
US 10437833B1 · Nguyen · 2019 [cited by applicant]
US 11468071B2 · Pichaimurthy et al. · 2022 [cited by applicant]
US 11663222B2 · Pichaimurthy et al. · 2023 [cited by applicant]
US 12093267B2 · Pichaimurthy et al. · 2024 [cited by applicant]
US 20130086105A1 · Hammontree et al. · 2013 [cited by applicant]
US 20140188925A1 · Skolicki · 2014 [cited by applicant]
US 20160195856A1 · Spero · 2016 [cited by applicant]
US 20170242857A1 · Kim · 2017 [cited by applicant]
US 20180107748A1 · Skolicki · 2018 [cited by applicant]
US 20180204111A1 · Zadeh et al. · 2018 [cited by applicant]
US 20190244270A1 · Kim et al. · 2019 [cited by applicant]
US 20190258851A1 · Rajan et al. · 2019 [cited by applicant]
US 20190294668A1 · Goel et al. · 2019 [cited by applicant]
US 20200175019A1 · Pichaimurthy et al. · 2020 [cited by applicant]
US 20230004567A1 · Pichaimurthy et al. · 2023 [cited by applicant]
US 20230376490A1 · Pichaimurthy et al. · 2023 [cited by applicant]
EP 3557441A1 · 2019 [cited by applicant]
WO 2017044260A1 · 2017 [cited by applicant]
WO 2018043990A1 · 2018 [cited by applicant]
International Preliminary Report on Patentability received for PCT Patent Application No. PCT/US2019/063181, mailed on Jun. 10, 2021, 08 pages. [cited by applicant]
International Search Report and Written Opinion of PCT/US2019/063181 dated Feb. 21, 2020. [cited by applicant]