IP Library Granted Patent US 8,532,992
Granted Patent B2
US 8,532,992 · App. 13/762,978 · Granted Sep 10, 2013

System and method for standardized speech recognition infrastructure

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,532,992
App. No.
13/762,978
Granted
Sep 10, 2013
Kind
B2
Abstract

Disclosed herein are systems, methods, and computer-readable storage media for selecting a speech recognition model in a standardized speech recognition infrastructure. The system receives speech from a user, and if a user-specific supervised speech model associated with the user is available, retrieves the supervised speech model. If the user-specific supervised speech model is unavailable and if an unsupervised speech model is available, the system retrieves the unsupervised speech model. If the user-specific supervised speech model and the unsupervised speech model are unavailable, the system retrieves a generic speech model associated with the user. Next the system recognizes the received speech from the user with the retrieved model. In one embodiment, the system trains a speech recognition model in a standardized speech recognition infrastructure. In another embodiment, the system handshakes with a remote application in a standardized speech recognition infrastructure.

Claims (50)

1. A method comprising:

receiving speech from a user;

determining, via a processor, to apply one of supervised training and unsupervised training; and

when supervised training is selected:

determining whether available data are sufficient to build a new speech recognition model;

when the available data is sufficient to build the new speech recognition model, building the new speech recognition model using the available data; and

when the available data is not sufficient to build the new speech recognition model:

selecting an existing speech recognition model; and

generating an adapted speech recognition model based on transformations generated from the existing speech recognition model based on the speech and associated transcriptions.

2. The method of claim 1 , wherein the new speech recognition model, the existing speech recognition model and the adapted speech recognition model are standardized speech models.

3. The method of claim 1 , wherein one of the new speech recognition model, the existing speech recognition model and the adapted speech recognition model is publicly available.

4. The method of claim 1 , wherein the existing speech recognition model is a general model and wherein the adapted speech recognition model is a result of applying transformations to the existing speech recognition model.

5. The method of claim 1 , further comprising recognizing additional speech using the adapted speech recognition model.

6. The method of claim 5 , wherein recognizing the additional speech is performed off-line at a later time.

7. The method of claim 1 , further comprising reusing speech models for additional received speech.

8. The method of claim 1 , further comprising:

recognizing voice commands in the speech; and

controlling elements of a game based on the voice commands.

9. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored, which, when executed by the processor, result in the processor performing operations comprising:

receiving speech from a user;

determining, via a processor, to apply one of supervised training and unsupervised training; and

when supervised training is selected:

determining whether available data are sufficient to build a new speech recognition model;

when the available data is sufficient to build the new speech recognition model, building the new speech recognition model using the available data; and

when the available data is not sufficient to build the new speech recognition model:

selecting an existing speech recognition model; and

generating an adapted speech recognition model based on transformations generated from the existing speech recognition model based on the speech and associated transcriptions.

10. The system of claim 9 , wherein the new speech recognition model, the existing speech recognition model and the adapted speech recognition model are standardized speech models.

11. The system of claim 9 , wherein one of the new speech recognition model, the existing speech recognition model and the adapted speech recognition model is publicly available.

12. The system of claim 9 , wherein the existing speech recognition model is a general model and wherein the adapted speech recognition model is a result of applying transformations to the existing speech recognition model.

13. The system of claim 9 , the computer-readable storage medium having additional instructions stored which result in the operations further comprising recognizing additional speech using the adapted speech recognition model.

14. The system of claim 13 , wherein recognizing the additional speech is performed off-line at a later time.

15. The system of claim 9 , the computer-readable storage medium having additional instructions stored which result in the operations further comprising reusing speech models for additional received speech.

16. The system of claim 9 , the computer-readable storage medium having additional instructions stored which result in the operations further comprising:

recognizing voice commands in the speech; and

controlling elements of a game based on the voice commands.

17. A non-transitory computer-readable storage device having instructions stored, which, when executed by a computing device, result in the computing device performing operations comprising:

receiving speech from a user;

determining, via a processor, to apply one of supervised training and unsupervised training; and

when supervised training is selected:

determining whether available data are sufficient to build a new speech recognition model;

when the available data is sufficient to build the new speech recognition model, building the new speech recognition model using the available data; and

when the available data is not sufficient to build the new speech recognition model:

selecting an existing speech recognition model; and

generating an adapted speech recognition model based on transformations generated from the existing speech recognition model based on the speech and associated transcriptions.

18. The non-transitory computer-readable storage device of claim 17 , wherein the new speech recognition model, the existing speech recognition model and the adapted speech recognition model are standardized speech models.

19. The non-transitory computer-readable storage device of claim 17 , wherein one of the new speech recognition model, the existing speech recognition model and the adapted speech recognition model is publicly available.

20. The computer-readable storage medium of claim 17 , wherein the existing speech recognition model is a general model and wherein the adapted speech recognition model is a result of applying transformations to the existing speech recognition model.

Assignments (16)
RELEASE OF SECURITY INTEREST Recorded Sep 4, 2025
From: RUNWAY GROWTH FINANCE CORP., AS AGENT
To: INTERACTIONS CORPORATION; INTERACTIONS LLC
Reel/Frame 072802/0931 →
CORRECTIVE ASSIGNMENT TO CORRECT THE THE APPLICATION NUMBER PREVIOUSLY RECORDED AT REEL: 060445 FRAME: 0733. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Feb 1, 2023
From: INTERACTIONS LLC; INTERACTIONS CORPORATION
To: RUNWAY GROWTH FINANCE CORP.
Reel/Frame 062919/0063 →
RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY RECORDED AT REEL/FRAME: 036100/0925 Recorded Jul 1, 2022
From: SILICON VALLEY BANK
To: INTERACTIONS LLC
Reel/Frame 060559/0576 →
RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY RECORDED AT REEL/FRAME: 049388/0082 Recorded Jun 30, 2022
From: SILICON VALLEY BANK
To: INTERACTIONS LLC
Reel/Frame 060558/0474 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jun 27, 2022
From: INTERACTIONS LLC; INTERACTIONS CORPORATION
To: RUNWAY GROWTH FINANCE CORP.
Reel/Frame 060445/0733 →
TERMINATION AND RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY Recorded May 23, 2022
From: ORIX GROWTH CAPITAL, LLC
To: INTERACTIONS CORPORATION; INTERACTIONS LLC
Reel/Frame 061749/0825 →
RELEASE OF SECURITY INTEREST Recorded May 18, 2020
From: BEARCUB ACQUISITIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 052693/0866 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jun 5, 2019
From: INTERACTIONS LLC
To: SILICON VALLEY BANK
Reel/Frame 049388/0082 →
ASSIGNMENT OF IP SECURITY AGREEMENT Recorded Nov 17, 2017
From: ARES VENTURE FINANCE, L.P.
To: BEARCUB ACQUISITIONS LLC
Reel/Frame 044481/0034 →
CORRECTIVE ASSIGNMENT TO CORRECT THE CHANGE PATENT 7146987 TO 7149687 PREVIOUSLY RECORDED ON REEL 036009 FRAME 0349. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Nov 17, 2015
From: INTERACTIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 037134/0712 →
FIRST AMENDMENT TO INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jul 13, 2015
From: INTERACTIONS LLC
To: SILICON VALLEY BANK
Reel/Frame 036100/0925 →
SECURITY INTEREST Recorded Jun 23, 2015
From: INTERACTIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 036009/0349 →
SECURITY INTEREST Recorded Dec 19, 2014
From: INTERACTIONS LLC
To: ORIX VENTURES, LLC
Reel/Frame 034677/0768 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 16, 2014
From: AT&T ALEX HOLDINGS, LLC
To: INTERACTIONS LLC
Reel/Frame 034642/0640 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 10, 2014
From: AT&T INTELLECTUAL PROPERTY I, L.P.
To: AT&T ALEX HOLDINGS, LLC
Reel/Frame 034462/0764 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 8, 2013
From: LJOLJE, ANDREJ; RENGER, BERNARD S.; TISCHER, STEVEN NEIL
To: AT&T INTELLECTUAL PROPERTY I, L.P.
Reel/Frame 029783/0019 →