System and method for mobile automatic speech recognition
View Patent ↗A system and method of updating automatic speech recognition parameters on a mobile device are disclosed. The method comprises storing user account-specific adaptation data associated with ASR on a computing device associated with a wireless network, generating new ASR adaptation parameters based on transmitted information from the mobile device when a communication channel between the computing device and the mobile device becomes available and transmitting the new ASR adaptation data to the mobile device when a communication channel between the computing device and the mobile device becomes available. The new ASR adaptation data on the mobile device more accurately recognizes user utterances.
1. A method comprising:
receiving, at a first device, automatic speech recognition data from a second device used for speech recognition, wherein the automatic speech recognition data comprises information associated with at least one of a speaker, an environment, an utterance, and a transducer;
analyzing the automatic speech recognition data to yield automatic speech recognition adaptation parameters; and
transmitting the automatic speech recognition adaptation parameters to the second device for use in recognizing speech on the second device.
2. The method of claim 1 , wherein the second device is a mobile device.
3. The method of claim 1 , wherein the automatic speech recognition data comprises at least one of:
a representation of an environment of the second device and multi-modal data associated with a multi-modal input from a user.
4. The method of claim 1 , wherein the automatic speech recognition data comprises an automatic speech recognition output from the second device.
5. The method of claim 1 , wherein the first device is remote from the second device.
6. The method of claim 1 , wherein the automatic speech recognition adaptation parameters are based on the automatic speech recognition data and stored user specific adaptation data.
7. The method of claim 1 , wherein the automatic speech recognition data comprises audio data gathered by the second device during automatic speech recognition with a user.
8. A system comprising:
a processor; and
a memory storing instructions for controlling the processor to perform steps comprising:
receiving, at a first device, automatic speech recognition data from a second device used for speech recognition, wherein the automatic speech recognition data comprises information associated with at least one of a speaker, an environment, an utterance, and a transducer;
analyzing the automatic speech recognition data to yield automatic speech recognition adaptation parameters; and
transmitting the automatic speech recognition adaptation parameters to the second device for use in recognizing speech on the second device.
9. The system of claim 8 , wherein the automatic speech recognition data comprises at least one of:
a representation of an environment of the second device and multi-modal data associated with a multi-modal input from a user.
10. The system of claim 8 , wherein the first device is remote from the second device.
11. The system of claim 8 , wherein the second device is a mobile device.
12. The system of claim 8 , wherein the automatic speech recognition data comprises an automatic speech recognition output from the second device.
13. The system of claim 8 , wherein the automatic speech recognition adaptation parameters are based on the automatic speech recognition data and stored user specific adaptation data.
14. The system of claim 8 , wherein the automatic speech recognition data comprises audio data gathered by the second device during automatic speech recognition with a user.
15. A non-transitory computer-readable storage medium storing instructions which, when executed by a computing device, cause the computing device to perform steps comprising:
receiving, at a first device, automatic speech recognition data from a second device used for speech recognition, wherein the automatic speech recognition data comprises a representation of at least one of a speaker, an environment, an utterance, and a transducer;
analyzing the automatic speech recognition data to yield automatic speech recognition adaptation parameters; and
transmitting the automatic speech recognition adaptation parameters to the second device for use in recognizing speech on the second device.
16. The non-transitory computer-readable storage medium of claim 15 , wherein the automatic speech recognition data comprises at least one of:
a representation of an environment of the second device and multi-modal data associated with a multi-modal input from a user.
17. The non-transitory computer-readable storage medium of claim 15 , wherein the first device is remote from the second device.
18. The non-transitory computer-readable storage medium of claim 15 , wherein the second device is a mobile device.
19. The non-transitory computer-readable storage medium of claim 15 , wherein the automatic speech recognition data comprises an automatic speech recognition output from the second device.
20. The non-transitory computer-readable storage medium of claim 15 , wherein the automatic speech recognition adaptation parameters are based on the automatic speech recognition data and stored user specific adaptation data.