IP Library Granted Patent US 8,463,608
Granted Patent B2
US 8,463,608 · App. 13/417,824 · Granted Jun 11, 2013

Interactive speech recognition model

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,463,608
App. No.
13/417,824
Granted
Jun 11, 2013
Kind
B2
Abstract

A method and apparatus for updating a speech model on a multi-user speech recognition system with a personal speech model for a single user. A speech recognition system, for instance in a car, can include a generic speech model for comparison with the user speech input. A way of identifying a personal speech model, for instance in a mobile phone, is connected to the system. A mechanism is included for receiving personal speech model components, for instance a BLUETOOTH connection. The generic speech model is updated using the received personal speech model components. Speech recognition can then be performed on user speech using the updated generic speech model.

Claims (30)

1. A method for updating a first speech model in a speech recognition system, comprising:

receiving from a user device of a user, the user device comprising a personal speech model trained for the user through previous speech recognition operations, one or more personal speech model components of the personal speech model trained for the user through previous speech recognition operations, the one or more personal speech model components describing personal speech characteristics of the user;

updating the first speech model using at least some of the one or more personal speech model components, by modifying at least one speech model component of the first speech model and/or adding at least one speech model component to the first speech model; and

performing speech recognition on user speech using the first speech model updated with the at least some of the one or more personal speech model components.

2. The method of claim 1 , wherein the user device is a mobile device.

3. The method of claim 2 , wherein the network connection includes a BLUETOOTH® connection between the mobile device and the speech recognition system.

4. The method of claim 3 , wherein the speech recognition system is part of a navigation system located in a vehicle.

5. The method of claim 1 , wherein the personal speech model components are personal language model components.

6. The method of claim 1 , wherein the personal speech model components are personal acoustic model components.

7. At least one non-transitory computer readable medium encoded with instructions that, when executed on at least one computer, performs a method for updating a first speech model in a speech recognition system, comprising:

receiving, from a user device of a user, the user device comprising a personal speech model trained for the user through previous speech recognition operations, one or more personal speech model components of the personal speech model trained for the user through previous speech recognition operations, the one or more personal speech model components describing personal voice characteristics of the user;

updating the first speech model using at least some of the one or more personal speech model components, by modifying at least one speech model component of the first speech model and/or adding at least one speech model component to the first speech model; and

performing speech recognition on user speech using the first speech model updated with the at least some of the one or more personal speech model components.

8. The at least one non-transitory computer readable medium of claim 7 , wherein the user device is a mobile device.

9. The at least one non-transitory computer readable medium of claim 8 , wherein the network connection includes a BLUETOOTH® connection between the mobile device and the speech recognition system.

10. The at least one non-transitory computer readable medium of claim 9 , wherein the speech recognition system is part of a navigation system located in a vehicle.

11. The at least one non-transitory computer readable medium of claim 7 , wherein the personal speech model components are personal language model components.

12. The at least one non-transitory computer readable medium of claim 7 , wherein the personal speech model components are personal acoustic model components.

13. A speech recognition system accessible over a network, comprising:

a first speech model;

at least one processor programmed to implement a speech recognition engine capable of recognizing speech data based, at least in part, on the first speech model; and

a speech model controller to;

receive, from a user device of a user, the user device comprising a personal speech model trained for the user through previous speech recognition operations, one or more personal speech model components of the personal speech model trained for the user through previous speech recognition operations, the one or more personal speech model components describing personal voice characteristics of the user;

update the first speech model using at least some of the one or more personal speech model components, by modifying at least one speech model component of the first speech model and/or adding at least one speech model component to the first speech model; and

transmit user speech to be recognized by the speech recognition engine using the first speech model updated with the at least some of the one or more personal speech model components.

14. The speech recognition system of claim 13 , wherein the user device is a mobile device.

15. The speech recognition system of claim 14 , wherein the network connection includes a BLUETOOTH® connection between the mobile device and the speech recognition system.

16. The speech recognition system of claim 15 , wherein the speech recognition system is part of a navigation system located in a vehicle.

17. The speech recognition system of claim 13 , wherein the personal speech model components are personal language model components.

18. The speech recognition system of claim 13 , wherein the personal speech model components are personal acoustic model components.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 13, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065552/0934 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 13, 2013
From: DOW, BARRY NEIL; JANKE, ERIC WILLIAM; CHEUNG, DANIEL LEE YUK; STANIFORD, BENJAMIN TERRICK
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 030402/0638 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 13, 2013
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 030402/0804 →