IP Library Granted Patent US 11,520,821
Granted Patent B2
US 11,520,821 · App. 16/201,352 · Granted Dec 6, 2022

Systems and methods for providing search query responses having contextually relevant voice output

Inventors: Ankur Anil Aher (Bangalore, IN); Harish Ashok Kumar (Bangalore, IN)
Assignee: ROVI GUIDES, INC.
G06F16/637G06F16/907G06F16/90332G10L13/027
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,520,821
App. No.
16/201,352
Granted
Dec 6, 2022
Kind
B2
Abstract

Systems and methods are described for responding to a search query with a contextually relevant voice output. An illustrative method receives a search query, determines an answer to the search query, identifies a media content reference included in the search query, determines, based on the media content reference, a personality associated with the media content reference, identifies a voice profile of the personality, and generates audio output using the voice profile of the personality, the audio output including the answer to the search query.

Claims (80)

1. A method for responding to a search query with a contextually relevant voice output, the method comprising:

receiving a search query;

determining, using control circuitry, an answer to the search query;

determining, using the control circuitry, that the search query includes a reference to a video content item;

in response to determining that the search query includes the reference to the video content item:

accessing a voice profile database that comprises a plurality of voice profiles of a plurality of voice performers and an indication, for each respective voice performer, of a video content item that the respective voice performer is included in a cast of;

querying the voice profile database to identify a voice performer included in a cast of the video content item referenced in the search query; and

determining that the voice profile database indicates that a voice profile of the identified voice performer is available; and

in response to (a) identifying the voice performer included in the cast of the video content item referenced in the search query, and (b) accessing the indication that the voice profile of the identified voice performer is available at the voice profile database:

retrieving the voice profile from the voice profile database; and

generating, using the control circuitry, audio output by synthesizing audio of the answer to the search query in a voice of the voice performer using the retrieved voice profile of the voice performer identified as being included in the cast of the video content item referenced in the search query, the audio output including the answer to the search query.

2. The method of claim 1 , wherein generating the audio output using the retrieved voice profile of the voice performer comprises:

synthesizing, using the retrieved voice profile of the voice performer, audio matching a plurality of words included in the answer; and

generating an audio response including the synthesized audio.

3. The method of claim 2 , wherein synthesizing, using the retrieved voice profile of the voice performer, audio matching the plurality of words included in the answer comprises:

determining, based on the voice profile of the voice performer, a characteristic of the voice performer's voice;

retrieving, from a database, a plurality of audio templates, wherein each audio template of the plurality of audio templates corresponds to a respective word of the plurality of words;

modifying, using the control circuitry, each audio template of the plurality of audio templates based on the characteristic of the voice performer's voice; and

generating, using the control circuitry, audio corresponding to each modified audio template.

4. The method of claim 2 , further comprising:

retrieving an audio recording of the voice performer's voice,

wherein the audio recording includes at least a portion of the answer, and

wherein generating the audio response including the synthesized audio comprises generating audio output including the synthesized audio and the audio recording.

5. The method of claim 4 , further comprising:

identifying a subset of the plurality of words included in the answer that is not included in the audio recording,

wherein synthesizing audio matching a plurality of words included in the answer comprises synthesizing audio corresponding to the subset of the plurality of words included in the answer that is not included in the audio recording.

6. The method of claim 1 , wherein generating the audio output using the voice profile of the voice performer comprises:

retrieving an audio recording corresponding to the voice profile, the audio recording including at least a portion of the answer; and

generating an audio response including the audio recording.

7. The method of claim 1 , wherein determining that the voice profile database indicates that a voice profile of the identified voice performer is available comprises:

retrieving, from the voice profile database, an indication of a character associated with the video content item identified by the reference to the video content item; and

identifying a voice performer providing a voice of the character.

8. The method of claim 7 , wherein identifying the voice profile of the voice performer comprises retrieving, from the voice profile database, an indication of a characteristic of the performer's voice.

9. The method of claim 1 , further comprising:

identifying a phrase associated with the video content item identified by the reference to the video content item,

wherein generating audio output using the retrieved voice profile of the voice performer comprises generating audio output including the phrase associated with the video content item identified by the reference to the video content item.

10. The method of claim 9 , wherein identifying the phrase associated with the video content item identified by the reference to the video content item comprises:

identifying a plurality of words included in the search query;

identifying, based on the plurality of words, the video content item identified by the reference to the video content item; and

retrieving, from a database, the phrase associated with the video content item identified by the reference to the video content item.

11. A system for responding to a search query with a contextually relevant voice output, the system comprising:

control circuitry configured to:

receive a search query;

determine an answer to the search query;

determine that the search query includes a reference to a video content item;

in response to determining that the search query includes the reference to the video content item:

access a voice profile database that comprises a plurality of voice profiles of a plurality of voice performers and an indication, for each respective voice performer, of a video content item that the respective voice performer is included in a cast of;

a query the voice profile database to identify a voice performer included in a cast of the video content item referenced in the search query; and

determine that the voice profile database indicates that a voice profile of the identified voice performer is available; and

in response to (a) identifying the voice performer included in the cast of the video content item referenced in the search query, and (b) accessing the indication that the voice profile of the identified voice performer is available at the voice profile database:

retrieve the voice profile from the voice profile database; and

generate audio output by synthesizing audio of the answer to the search query in a voice of the voice performer using the retrieved voice profile of the voice performer identified as being included in the cast of the video content item referenced in the search query, the audio output including the reply to the user input.

12. The system of claim 11 , wherein the control circuitry is further configured to:

synthesize, using the retrieved voice profile of the voice performer, audio matching a plurality of words included in the answer; and

generate an audio response including the synthesized audio.

13. The system of claim 12 , wherein the control circuitry is further configured to:

determine, based on the voice profile of the voice performer, a characteristic of the voice performer's voice;

retrieve, from a database, a plurality of audio templates, wherein each audio template of the plurality of audio templates corresponds to a respective word of the plurality of words;

modify each audio template of the plurality of audio templates based on the characteristic of the voice performer's voice; and

generate audio corresponding to each modified audio template.

14. The system of claim 12 , wherein the control circuitry is further configured to:

retrieve an audio recording of the voice performer's voice, wherein the audio recording includes at least a portion of the answer; and

generate audio output including the synthesized audio and the audio recording.

15. The system of claim 14 , wherein the control circuitry is further configured to:

identify a subset of the plurality of words included in the answer that is not included in the audio recording; and

synthesize audio corresponding to the subset of the plurality of words included in the answer that is not included in the audio recording.

16. The system of claim 11 , wherein the control circuitry is further configured to:

retrieve an audio recording corresponding to the voice profile, the audio recording including at least a portion of the answer; and

generate an audio response including the audio recording.

17. The system of claim 11 , wherein the control circuitry is further configured to:

retrieve, from the voice profile database, an indication of a character associated with the video content item identified by the reference to the video content item;

identify a voice performer providing a voice of the character; and

retrieve, from the voice profile database, an indication of a characteristic of the voice of the voice performer.

18. The system of claim 11 , wherein the control circuitry is further configured to:

identify a phrase associated with the video content item identified by the reference to the video content item; and

generate audio output including the phrase associated with the video content item identified by the reference to the video content item.

19. The system of claim 18 , wherein the control circuitry is further configured to:

identify a plurality of words included in the search query;

identify, based on the plurality of words, the video content item identified by the reference to the video content item; and

retrieve, from a database, the phrase associated with the video content item identified by the reference to the video content item.

Assignments (7)
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0171 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: MORGAN STANLEY SENIOR FUNDING, INC.
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053481/0790 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: HPS INVESTMENT PARTNERS, LLC
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053458/0749 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
PATENT SECURITY AGREEMENT Recorded Nov 25, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
Reel/Frame 051110/0006 →
SECURITY INTEREST Recorded Nov 22, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: HPS INVESTMENT PARTNERS, LLC, AS COLLATERAL AGENT
Reel/Frame 051143/0468 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 6, 2018
From: AHER, ANKUR ANIL; KUMAR, HARISH ASHOK
To: ROVI GUIDES, INC.
Reel/Frame 047689/0091 →
Continuity (1)
Related Publication 20200167384A1 · May 28, 2020