IP Library Granted Patent US 11,024,296
Granted Patent B2
US 11,024,296 · App. 16/815,609 · Granted Jun 1, 2021

Systems and methods for conversations with devices about media using interruptions and changes of subjects

Inventors: Charles Dawes (Ryton, GB); Walter R. Klappert (North Hollywood, CA)
Assignee: Rovi Guides, Inc.
G10L15/1815G10L15/222
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,024,296
App. No.
16/815,609
Granted
Jun 1, 2021
Kind
B2
Abstract

Systems and methods are described herein for providing media guidance. Control circuitry may receive a first voice input and access a database of topics to identify a first topic associated with the first voice input. A user interface may generate a first response to the first voice input, and subsequent to generating the first response, the control circuitry may receive a second voice input. The control circuitry may determine a match between the second voice input and an interruption input such as a period of silence or a keyword or a phrase, such as “Ahh,”, “Umm,”, or “Hmm.” The user interface may generate a second response that is associated with a second topic related to the first topic. By interrupting the conversation and changing the subject from time to time, media guidance systems can appear to be more intelligent and human.

Claims (91)

1. A method for providing a follow-up response to a conversational input based on detecting a user hesitation to an initial response, the method comprising:

receiving, using control circuitry, a first input;

generating, using a user interface, a first response to the first input;

subsequent to generating the first response, receiving, using the control circuitry, a voice input followed by a period of silence;

in response to receiving the voice input followed by the period of silence, comparing the voice input to a plurality of verbal cues indicative of hesitations;

determining, based on the comparing, that the voice input matches a verbal cue of the plurality of verbal cues; and

in response to determining that the voice input matches the verbal cue, generating, using the user interface, a second response to the first input.

2. The method of claim 1 , further comprising:

accessing, using the control circuitry, a database of topics comprising a semantic network indicating relationships between a plurality of topics;

identifying, using the control circuitry, a first topic that is associated with the first input from the database of topics; and

generating, using the user interface, the second response to the first input comprising a media asset recommendation associated with a second topic related to the first topic in the database of topics.

3. The method of claim 2 , further comprising:

extracting, from the database, relationships between the first topic and a remainder of the plurality of topics;

comparing each of the relationships between the first topic and the remainder of the plurality of topics to a relationship threshold; and

storing to memory a list indicating a subset of the relationships between the first topic and the remainder of the plurality of topics that do not exceed the relationship threshold and a list of topics of the plurality of topics that correspond to the subset of the relationships;

wherein the second topic is selected from the list of topics.

4. The method of claim 2 , wherein the semantic network comprises numerical relationships between the plurality of topics, the numerical relationships indicating a statistical likelihood that the second topic is related to the first topic.

5. The method of claim 4 , further comprising:

receiving a plurality of voice inputs from a plurality of users;

identifying a first subset of the plurality of voice inputs that relate to the first topic;

identifying, from the first subset of the plurality of voice inputs, a second subset of the plurality of voice inputs relating to the second topic; and

calculating a statistical likelihood, based on the identified first subset of the plurality of voice inputs and second subset of the plurality of voice inputs, that the second topic follows the first topic.

6. The method of claim 2 , further comprising:

accessing, using the control circuitry, a user profile indicating media preferences of a user;

retrieving a genre preference from the user profile;

identifying a subset of the plurality of topics that are associated with the retrieved genre; and

selecting the second topic from the subset of the plurality of topics.

7. The method of claim 2 , wherein the database of topics indicates a respective genre associated with each respective topic of the plurality of topics, further comprising:

extracting, from the database of topics, a genre associated with the first topic;

identifying a subset of the plurality of topics that are associated with the extracted genre; and

selecting the second topic from the subset of the plurality of topics.

8. The method of claim 1 , further comprising:

receiving, using the control circuitry, a third voice input;

comparing the third voice input to the plurality of verbal cues to determine a match between the third voice input and a second verbal cue from the plurality of verbal cues;

determining whether a threshold period of time has elapsed between a current time and the second response; and

in response to determining that the threshold period of time has elapsed between the current time and the second response, generating, using the user interface, a third response to the first input.

9. The method of claim 1 , further comprising:

receiving an identifier of a user associated with the first input;

accessing, using the control circuitry, a plurality of voice personality profiles, each voice personality profile corresponding to a respective user and comprising indications of a respective plurality of verbal cues; and

selecting one of the plurality of voice personality profiles based on the identifier of the user associated with the first input;

wherein comparing the voice input to the plurality of verbal cues comprises comparing the voice input to the indications of the respective plurality of verbal cues associated with the selected voice personality profiles.

10. The method of claim 9 , further comprising:

extracting a threshold period of time from the selected voice personality profile; and

calculating a time elapsed since the first input by comparing a current time to a receipt time associated with the first input;

wherein generating the second response is performed in response to determining that the time elapsed has exceeded the threshold period of time.

11. A system for providing a follow-up response to a conversational input based on detecting a user hesitation to an initial response, the system comprising:

a user interface; and

control circuitry configured to:

receive, using control circuitry, a first input;

generate, using a user interface, a first response to the first input;

subsequent to generating the first response, receive, a voice input followed by a period of silence;

in response to receiving the voice input followed by the period of silence, compare the voice input to a plurality of verbal cues indicative of hesitations;

determine, based on the comparing, that the voice input matches a verbal cue of the plurality of verbal cues; and

in response to determining that the voice input matches the verbal cue, generate, using the user interface, a second response to the first input.

12. The system of claim 11 , wherein the control circuitry is further configured to:

access a database of topics comprising a semantic network indicating relationships between a plurality of topics;

identify a first topic that is associated with the first input from the database of topics; and

generate, using the user interface, the second response to the first input comprising a media asset recommendation associated with a second topic related to the first topic in the database of topics.

13. The system of claim 12 , further comprising a memory, wherein the control circuitry is further configured to:

extract, from the database, relationships between the first topic and a remainder of the plurality of topics;

compare each of the relationships between the first topic and the remainder of the plurality of topics to a relationship threshold; and

store to the memory a list indicating a subset of the relationships between the first topic and the remainder of the plurality of topics that do not exceed the relationship threshold and a list of topics of the plurality of topics that correspond to the subset of the relationships;

select the second topic from the list of topics.

14. The system of claim 12 , wherein the semantic network comprises numerical relationships between the plurality of topics, the numerical relationships indicating a statistical likelihood that the second topic is related to the first topic.

15. The system of claim 14 , wherein the control circuitry is further configured to:

receive a plurality of voice inputs from a plurality of users;

identify a first subset of the plurality of voice inputs that relate to the first topic;

identify, from the first subset of the plurality of voice inputs, a second subset of the plurality of voice inputs relating to the second topic; and

calculate a statistical likelihood, based on the identified first subset of the plurality of voice inputs and second subset of the plurality of voice inputs, that the second topic follows the first topic.

16. The system of claim 12 , wherein the control circuitry is further configured to:

access a user profile indicating media preferences of a user;

retrieve a genre preference from the user profile;

identify a subset of the plurality of topics that are associated with the retrieved genre; and

select the second topic from the subset of the plurality of topics.

17. The system of claim 12 , wherein the database of topics indicates a respective genre associated with each respective topic of the plurality of topics, wherein the control circuitry is further configured to:

extract, from the database of topics, a genre associated with the first topic;

identify a subset of the plurality of topics that are associated with the extracted genre; and

select the second topic from the subset of the plurality of topics.

18. The system of claim 11 , wherein the control circuitry is further configured to: receive a third voice input;

compare the third voice input to the plurality of verbal cues to determine a match between the third voice input and a second verbal cue from the plurality of verbal cues;

determine whether a threshold period of time has elapsed between a current time and the second response; and

in response to determining that the threshold period of time has elapsed between the current time and the second response, generate, using the user interface, a third response to the first input.

19. The system of claim 11 , wherein the control circuitry is further configured to:

receive an identifier of a user associated with the first input;

access a plurality of voice personality profiles, each voice personality profile corresponding to a respective user and comprising indications of a respective plurality of verbal cues;

select one of the plurality of voice personality profiles based on the identifier of the user associated with the first input; and

compare the voice input to the indications of the respective plurality of verbal cues associated with the selected voice personality profiles.

20. The system of claim 19 , wherein the control circuitry is further configured to:

extract a threshold period of time from the selected voice personality profile;

calculate a time elapsed since the first input by comparing a current time to a receipt time associated with the first input; and

generate the second response is performed in response to determining that the time elapsed has exceeded the threshold period of time.

Assignments (3)
CHANGE OF NAME Recorded Sep 25, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069049/0212 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 11, 2020
From: DAWES, CHARLES; KLAPPERT, WALTER R.
To: ROVI GUIDES, INC.
Reel/Frame 052086/0833 →
Continuity (3)
Continuation 16379312 · Apr 9, 2019
Continuation 14757910 · Dec 23, 2015
Related Publication 20200302920A1 · Sep 24, 2020