IP Library Granted Patent US 11,854,529
Granted Patent B2
US 11,854,529 · App. 17/984,479 · Granted Dec 26, 2023

Method and apparatus for generating hint words for automated speech recognition

Inventors: Ankur Aher (Maharashtra, IN); Jeffry Copps Robert Jose (Tamil Nadu, IN)
Assignee: Rovi Guides, Inc.
G10L15/02G10L15/22G10L2015/025G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,854,529
App. No.
17/984,479
Granted
Dec 26, 2023
Kind
B2
Abstract

Systems and methods for determining hint words that improve the accuracy of automated speech recognition (ASR) systems. Hint words are determined in the context of a user issuing voice commands in connection with a voice interface system. Terms are initially taken from most frequently occurring terms in operation of a voice interface system. For example, most frequently occurring terms that arise in electronic search queries or received commands are selected. Certain of these terms are selected as hint words, and the selected hint words are then transmitted to an ASR system to assist in translation of speech to text.

Claims (32)

1. A method of determining hint words for automated speech recognition, the method comprising:

determining, using control circuitry, a first set of terms comprising terms that are most frequently occurring terms from speech input to a voice interface system;

determining, using the control circuitry, a second set of terms that are most frequently occurring terms arising during a predetermined time period of operation of the voice interface system;

selecting as a first set of hint words facilitating operation of an automated speech recognition application, the first set of hint words comprising common terms of the first set of terms and the second set of terms;

selecting a second set of hint words comprising a plurality of terms that are not among the common terms from the first set of terms or the second set of terms; and

transmitting, using the control circuitry, the hint words to the automated speech recognition application, the hint words comprising the first set of hint words and the second set of hint words.

2. The method of claim 1 , wherein selecting the second set of hint words further comprises selecting a predetermined number of most frequently occurring ones of the terms as the second set of hint words.

3. The method of claim 1 , wherein the second set of hint words further comprises a first predetermined number of the first set of terms and a second predetermined number of the second set of terms.

4. The method of claim 1 , wherein a sum of the number of the first set of hint words and a number of the second set of hint words is equal to a predetermined number.

5. The method of claim 1 , wherein the second set of hint words further comprises one or more terms from the second set of terms that are not among the common terms; and

wherein a sum of the number of the first set of hint words and a number of the second set of hint words is equal to a predetermined number.

6. The method of claim 1 , wherein the most frequently occurring terms are selected from one or more of terms of most recent speech input to the voice interface system.

7. The method of claim 1 , wherein the most frequently occurring terms are selected from one or more of terms of speech input to the voice interface system, or phonemes thereof.

8. The method of claim 1 , wherein the most frequently occurring terms are selected from one or more of terms of speech input to the voice interface system, or phonetic neighbors thereof.

9. The method of claim 1 , wherein at least one of the terms of speech input comprise one or more of names of consumer goods, tasks, reminders, calendar items, dates, or items of a list of items.

10. A system for determining hint words for automated speech recognition, the system comprising:

processing circuitry configured to:

determine a first set of terms comprising terms that are most frequently occurring terms from operation of a voice interface system, the most frequently occurring terms selected from one or more of terms of speech input to the voice interface system;

determine a second set of terms that are most frequently occurring terms arising during a predetermined time period of operation of the voice interface system;

select as a first set of hint words facilitating operation of an automated speech recognition application, the first set of hint words comprising common terms of the first set of terms and the second set of terms; and

select, if less than a predetermined number of common terms are selected, a second set of hint words comprising a plurality of terms from the first set of terms that are not among the common terms; and

input/output circuitry configured to:

transmit the first set of hint words and the second set of hint words to the automated speech recognition application.

11. The system of claim 10 , wherein the selecting further comprises selecting a predetermined number of the most frequently occurring ones of the terms as the hint words.

12. The system of claim 10 , wherein the second set of hint words further comprises a first predetermined number of the first set of terms and a second predetermined number of the second set of terms.

13. The system of claim 10 , wherein a sum of the number of the first set of hint words and a number of the second set of hint words is equal to the predetermined number of common terms.

14. The system of claim 10 , wherein the second set of hint words further comprises one or more terms from the second set of terms that are not among the common terms; and

wherein a sum of the number of the first set of hint words and a number of the second set of hint words is equal to the predetermined number of common terms.

15. The system of claim 10 , wherein the most frequently occurring terms are selected from one or more of terms of most recent speech input to the voice interface system.

16. The system of claim 10 , wherein the most frequently occurring terms are selected from one or more of terms of speech input to the voice interface system, or phonemes thereof.

17. The system of claim 10 , wherein the most frequently occurring terms are selected from one or more of terms of speech input to the voice interface system, or phonetic neighbors thereof.

18. The system of claim 10 , wherein at least one of the terms of speech input comprise one or more of names of consumer goods, tasks, reminders, calendar items, dates, or items of a list of items.

Assignments (3)
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0238 →
SECURITY INTEREST Recorded May 3, 2023
From: ADEIA GUIDES INC.; ADEIA IMAGING LLC; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA SOLUTIONS LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063529/0272 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2022
From: AHER, ANKUR; ROBERT JOSE, JEFFRY COPPS
To: ROVI GUIDES, INC.
Reel/Frame 061721/0676 →