IP Library Granted Patent US 10,140,982
Granted Patent B2
US 10,140,982 · App. 15/693,162 · Granted Nov 27, 2018

Method for using pauses detected in speech input to assist in interpreting the input during conversational interaction for information retrieval

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,140,982
App. No.
15/693,162
Granted
Nov 27, 2018
Kind
B2
Abstract

A method for using speech disfluencies detected in speech input to assist in interpreting the input is provided. The method includes providing access to a set of content items with metadata describing the content items, and receiving a speech input intended to identify a desired content item. The method further includes detecting a speech disfluency in the speech input and determining a measure of confidence of a user in a portion of the speech input following the speech disfluency. If the confidence measure is lower than a threshold value, the method includes determining an alternative query input based on replacing the portion of the speech input following the speech disfluency with another word or phrase. The method further includes selecting content items based on comparing the speech input, the alternative query input (when the confidence measure is low), and the metadata associated with the content items.

Claims (37)

1. A method for using speech disfluencies detected in speech input to assist in interpreting the input, the method comprising:

providing access to a set of content items, each of the content items being associated with metadata that describes the corresponding content item;

receiving a speech input from a user, the input intended by the user to identify at least one desired content item;

detecting a speech disfluency in the speech input;

determining a confidence measure of the user in a portion of the speech input following the speech disfluency based on a manner by which the user utters the portion of the speech input following the speech disfluency;

upon a condition in which the confidence measure does not exceed a threshold value:

retrieving from memory preferences of the user for particular content items of the set of content items;

comparing the portion of the speech input to metadata associated with the particular content items of the set of content items;

determining, based on the comparing, whether metadata associated with a content item of the particular content items at least partially matches the portion of the speech input;

in response to determining that metadata associated with the content item of the particular content items at least partially matches the portion of the speech input;

replacing the portion of the speech input with the metadata associated with the content item to generate a query comprising a modified speech input for searching a database associated with the set of content items;

selecting from the database a subset of content items from the set of content items based on comparing the query and metadata associated with the subset of content items;

upon a condition in which the confidence measure exceeds a threshold value, selecting the subset of content items from the set of content items based on comparing the speech input and the metadata associated with the subset of content items; and

presenting the subset of content items to the user.

2. The method of claim 1 , further comprising measuring a duration of the speech disfluency, wherein the determination of the confidence measure is based on the duration of the speech disfluency.

3. The method of claim 1 , wherein the subset of content items is a first subset of content items, and wherein the determination of the confidence measure is based on a second subset of the content items.

4. The method of claim 1 , further comprising offering assistance when the user engages in a speech disfluency.

5. The method of claim 4 , wherein the assistance is inferring a word or phrase following the speech disfluency and presenting the word or phrase to the user.

6. The method of claim 1 , wherein the speech disfluency is a pause or an auditory time filler.

7. A system for using speech disfluencies detected in speech input to assist in interpreting the input, the system comprising control circuitry configured to:

provide access to a set of content items, each of the content items being associated with metadata that describes the corresponding content item;

receive a speech input from a user, the input intended by the user to identify at least one desired content item;

detect a speech disfluency in the speech input;

determine a confidence measure of the user in a portion of the speech input following the speech disfluency based on a manner by which the user utters the portion of the speech input following the speech disfluency;

upon a condition in which the confidence measure does not exceed a threshold value:

retrieve from memory preferences of the user for particular content items of the set of content items;

compare the portion of the speech input to metadata associated with a content item of the particular content items at least partially matches the portion of the speech input;

in response to determining that the metadata associated with the content item of the particular content items at least partially matches the portion of the speech input:

replacing the portion of the speech input with the metadata associated with the content item to generate a query comprising a modified speech input for searching a database associated with the set of content items;

select from the database a subset of content items from the set of content items based on comparing the query and metadata associated with the subset of content items;

upon a condition in which the confidence measure exceeds a threshold value, select the subset of content items from the set of content items based on comparing the speech input and the metadata associated with the subset of content items; and

present the subset of content items to the user.

8. The system of claim 7 , wherein the control circuitry is further configured to measure a duration of the speech disfluency and wherein the determination of the confidence measure is based on the duration of the speech disfluency.

9. The system of claim 7 , wherein the subset of content items is a first subset of content items, and wherein the determination of the confidence measure is based on a second subset of the content items.

10. The system of claim 7 , wherein the control circuitry is further configured to offer assistance when the user engages in a speech disfluency.

11. The system of claim 10 , wherein the assistance is inferring a word or phrase following the speech disfluency and presenting the word or phrase to the user.

12. The system of claim 7 , wherein the speech disfluency is a pause or an auditory time filler.

Assignments (10)
CHANGE OF NAME Recorded Sep 24, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069036/0209 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 24, 2024
From: VEVEO LLC
To: ROVI GUIDES, INC.
Reel/Frame 069036/0130 →
CHANGE OF NAME Recorded Sep 24, 2024
From: VEVEO, INC.
To: VEVEO LLC
Reel/Frame 069036/0068 →
PARTIAL RELEASE OF SECURITY INTEREST IN PATENTS Recorded Oct 27, 2022
From: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
To: VEVEO LLC (F.K.A. VEVEO, INC.); DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
Reel/Frame 061786/0675 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: MORGAN STANLEY SENIOR FUNDING, INC.
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053481/0790 →
RELEASE OF SECURITY INTEREST Recorded Jun 5, 2020
From: HPS INVESTMENT PARTNERS, LLC
To: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
Reel/Frame 053458/0749 →
SECURITY INTEREST Recorded Jun 1, 2020
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS INC.; VEVEO, INC.; INVENSAS CORPORATION; INVENSAS BONDING TECHNOLOGIES, INC.; TESSERA, INC.; TESSERA ADVANCED TECHNOLOGIES, INC.; DTS, INC.; PHORUS, INC.; IBIQUITY DIGITAL CORPORATION
To: BANK OF AMERICA, N.A.
Reel/Frame 053468/0001 →
PATENT SECURITY AGREEMENT Recorded Nov 25, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
Reel/Frame 051110/0006 →
SECURITY INTEREST Recorded Nov 22, 2019
From: ROVI SOLUTIONS CORPORATION; ROVI TECHNOLOGIES CORPORATION; ROVI GUIDES, INC.; TIVO SOLUTIONS, INC.; VEVEO, INC.
To: HPS INVESTMENT PARTNERS, LLC, AS COLLATERAL AGENT
Reel/Frame 051143/0468 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 11, 2018
From: ARAVAMUDAN, MURALI; GILL, DAREN; VENKATARAMAN, SASHIKUMAR; AGARWAL, VINEET; RAMAMOORTHY, GANESH
To: VEVEO, INC.
Reel/Frame 044594/0543 →