IP Library Granted Patent US 11,468,071
Granted Patent B2
US 11,468,071 · App. 16/206,385 · Granted Oct 11, 2022

Voice query refinement to embed context in a voice query

Inventors: Rajendran Pichaimurthy (Bangalore, IN); Madhusudhan Seetharam (Bangalore, IN); Harshith Kumar Gejjegondanahally Sreekanth (Bangalore, IN)
Assignee: Rovi Guides, Inc.
G06F16/24575G06F16/24522G06F40/253G06F40/289G06F40/30G06V40/20G10L15/22G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,468,071
App. No.
16/206,385
Granted
Oct 11, 2022
Kind
B2
Abstract

Systems and methods are described for providing contextual search results. The system may receive a search query during presentation of a video. If the query is ambiguous, the system accesses some of the frames of the video. The frames are analyzed to identify a performed action depicted in the frames. The system retrieves a keyword related to the identified action. The ambiguous query is augmented with the keyword. The augmented search query is used to search for and output relevant search results.

Claims (72)

1. A method for providing contextual search results to ambiguous queries, the method comprising:

receiving a search query during a presentation of a video;

determining whether at least one word in the search query is ambiguous;

in response to determining that at least one word in the search query is ambiguous:

(i) accessing a plurality of frames from the video that were presented concurrently with receiving the search query;

(ii) analyzing the plurality of frames to identify a performed action;

(iii) retrieving a keyword associated with the identified action; and

performing a text-based search based on the search query and the keyword; and

outputting results of the search.

2. The method of claim 1 ,

wherein identifying the performed action comprises:

identifying a character in each of the plurality of frames;

generating a model of the identified character's movements;

determining that the generated model matches a movement template; and

wherein retrieving the keyword associated with the identified action comprises retrieving metadata of the movement template.

3. The method of claim 2 , wherein generating a model of the identified character's movements comprises:

identifying body parts of the identified character; and

calculating an angle between two body parts of the identified character.

4. The method of claim 3 , wherein determining that the generated model matches the movement template comprises:

comparing the calculated angle with a reference angle of the movement template; and

in response to determining that the calculated angle matches a reference angle determining that the generated model matches the movement template.

5. The method of claim 1 , wherein receiving the search query comprises:

detecting user voice input; and

performing speech to text analysis of the user voice input to derive the search query.

6. The method of claim 1 , wherein determining that at least one word in the search query is ambiguous comprises:

determining that the search query comprises at least one of: a pronoun and an auxiliary verb.

7. The method of claim 1 , wherein accessing the plurality of frames from the video that were presented concurrently with receiving the search query comprises:

receiving an audio sample of the video that were presented concurrently with receiving the search query;

identifying a time location in the video where the audio sample occurred; and

extracting frames corresponding to the time location.

8. The method of claim 7 , wherein extracting frames corresponding to the time location comprises:

extracting frames from a predetermined time period prior to the time location; and

extracting frames from a predetermined time period after the time location.

9. The method of claim 1 , wherein accessing the plurality of frames from the video that were presented concurrently with receiving the search query comprises capturing displayed frames of the video for a predetermined time period after receiving the search query.

10. The method of claim 1 , wherein accessing the plurality of frames from the video that were presented concurrently with receiving the search query comprises capturing a predetermined number of displayed frames after receiving the search query.

11. A system for providing contextual search results to ambiguous queries, the system comprising;

input circuitry of a device configured to:

receive a search query during a presentation of a video; and

control circuitry of the device configured to:

determine whether at least one word in the search query is ambiguous;

in response to determining that at least one word in the search query is ambiguous:

(i) access a plurality of frames from the video that were presented concurrently with receiving the search query;

(ii) analyze the plurality of frames to identify a performed action;

(iii) retrieve a keyword associated with the identified action; and

perform a text-based search based on the augmented search query and the keyword; and

output results of the search.

12. The system of claim 11 ,

wherein, when identifying the performed action, the control circuitry is configured to:

identify a character in each of the plurality of frames;

generate a model of the identified character's movements;

determine that the generated model matches a movement template; and

wherein, when retrieving the keyword associated with the identified action, the control circuitry is configured to retrieve metadata of the movement template.

13. The system of claim 12 , wherein, when generating a model of the identified character's movements, the control circuitry is configured to:

identify body parts of the identified character; and

calculate an angle between two body parts of the identified character.

14. The system of claim 13 , wherein, when determining that the generated model matches the movement template, the control circuitry is configured to:

compare the calculated angle with a reference angle of the movement template; and

in response to determining that the calculated angle matches a reference angle, determine that the generated model matches the movement template.

15. The system of claim 11 , wherein, when receiving the search query, the control circuitry is configured to:

detect user voice input; and

perform speech to text analysis of the user voice input to derive the search query.

16. The system of claim 11 , wherein, when determining that at least one word in the search query is ambiguous, the control circuitry is configured to:

determine that the search query comprises at least one of: a pronoun and an auxiliary verb.

17. The system of claim 11 , wherein, when accessing the plurality of frames from the video that were presented concurrently with receiving the search query, the control circuitry is configured to:

receive an audio sample of the video that were presented concurrently with receiving the search query;

identify a time location in the video where the audio sample occurred; and

extract frames corresponding to the time location.

18. The system of claim 17 , wherein, when extracting frames corresponding to the time location, the control circuitry is configured to:

extract frames from a predetermined time period prior to the time location; and

extract frames from a predetermined time period after the time location.

19. The system of claim 11 , wherein, when accessing the plurality of frames from the video that were presented concurrently with receiving the search query, the control circuitry is configured to capture displayed frames of the video for a predetermined time period after receiving the search query.

20. The system of claim 11 , wherein, when accessing the plurality of frames from the video that were presented concurrently with receiving the search query, the control circuitry is configured to capture a predetermined number of displayed frames after receiving the search query.

Assignments (7)
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0171 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: MORGAN STANLEY SENIOR FUNDING, INC.
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053481/0790 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: HPS INVESTMENT PARTNERS, LLC
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053458/0749 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
PATENT SECURITY AGREEMENT Recorded Nov 25, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
Reel/Frame 051110/0006 →
SECURITY INTEREST Recorded Nov 22, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: HPS INVESTMENT PARTNERS, LLC, AS COLLATERAL AGENT
Reel/Frame 051143/0468 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 22, 2019
From: PICHAIMURTHY, RAJENDRAN; SEETHARAM, MADHUSUDHAN; SREEKANTH, HARSHITH KUMAR GEJJEGONDANAHALLY
To: ROVI GUIDES, INC.
Reel/Frame 048404/0393 →
Continuity (1)
Related Publication 20200175019A1 · Jun 4, 2020
Cited By (1)
US 12,380,117