IP Library Granted Patent US 8,433,574
Granted Patent B2
US 8,433,574 · App. 13/372,241 · Granted Apr 30, 2013

Hosted voice recognition system for wireless devices

Inventors: Victor R. Jablokov (Charlotte, NC); Igor R. Jablokov (Charlotte, NC); Marc White (Boca Raton, FL)
Assignee: Canyon IP Holdings, LLC
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,433,574
App. No.
13/372,241
Granted
Apr 30, 2013
Kind
B2
Abstract

Methods, systems, and software for converting the audio input of a user of a hand-held client device or mobile phone into a textual representation by means of a backend server accessed by the device through a communications network. The text is then inserted into or used by an application of the client device to send a text message, instant message, email, or to insert a request into a web-based application or service. In one embodiment, the method includes the steps of initializing or launching the application on the device; recording and transmitting the recorded audio message from the client device to the backend server through a client-server communication protocol; converting the transmitted audio message into the textual representation in the backend server; and sending the converted text message back to the client device or forwarding it on to an alternate destination directly from the server.

Claims (58)

1. A method comprising:

receiving a selection of an application at a device;

receiving an audio input at the device;

transmitting, to a server:

the audio input; and

an identifier of the selected application;

receiving an identifier of the audio input from the server;

transmitting, to the server:

the identifier of the audio input; and

a request for speech recognition results for the audio input;

receiving, from the server, the requested speech recognition results for the audio input, wherein the requested speech recognition results were generated using a speech recognition technique associated with the selected application; and

processing at least a portion of the requested speech recognition results with the selected application.

2. The method of claim 1 , wherein the identifier of the audio input comprises a receipt.

3. The method of claim 1 , wherein the selected application comprises at least one of an instant messaging application, a text messaging application, or an email application.

4. The method of claim 3 , wherein processing at least a portion of the requested speech recognition results with the selected application comprises sending a message comprising at least a portion of the requested speech recognition results, wherein the message is sent as at least one of an instant message, a text message, or an email message.

5. The method of claim 1 , wherein the selection of the application at the device comprises a user pressing a button of the device, the button being associated with the application.

6. The method of claim 1 , wherein the device comprises at least one of a mobile phone or a personal digital assistant.

7. The method of claim 1 , wherein the audio input is transmitted to the server as a streaming transmission.

8. The method of claim 1 further comprising causing the device to present an indication that the requested speech recognition results have been received by the device.

9. The method of claim 1 , wherein the requested speech recognition results comprise at least one of a transcription of the audio input or a web service result corresponding to the audio input.

10. The method of claim 1 further comprising:

generating audio corresponding to the requested speech recognition results; and

causing the device to play the audio corresponding to the requested speech recognition results.

11. The method of claim 1 , wherein the speech recognition technique associated with the selected application comprises a grammar associated with the selected application.

12. Non-transitory computer storage having stored thereon a computer-executable module configured to execute in one or more processors of a device, the computer-executable module being further configured to:

receive a selection of an application;

receive an audio input;

transmit, to a server:

the audio input;

an identifier of the selected application; and

a request for speech recognition results for the audio input;

receive the speech recognition results from the server, wherein the requested speech recognition results were generated using a speech recognition technique associated with the selected application; and

process at least a portion of the requested speech recognition results with the selected application.

13. The non-transitory computer storage of claim 12 , wherein processing at least a portion of the requested speech recognition results comprises causing the device to display an advertisement corresponding to the requested speech recognition results.

14. The non-transitory computer storage of claim 12 , wherein the computer-executable module is further configured to:

determine a position of the device using a global positioning system;

identify, based at least in part on the requested speech recognition results, a target of interest near the position of the device; and

cause the device to display a location of the target of interest.

15. The non-transitory computer storage of claim 12 , wherein the computer-executable module is configured to process the requested speech recognition results by causing the device to display at least a portion of the requested speech recognition results.

16. The non-transitory computer storage of claim 12 , wherein the computer-executable module is further configured to cause the device to play audio corresponding to at least a portion of the requested speech recognition results.

17. The non-transitory computer storage of claim 12 , wherein the computer-executable module is further configured to receive a correction to the requested speech recognition results from a user.

18. The non-transitory computer storage of claim 12 , wherein the requested speech recognition results comprise at least one of a transcription of the audio input or a web service result corresponding to the audio input.

19. The non-transitory computer storage of claim 12 , wherein the speech recognition technique associated with the selected application comprises a grammar associated with the selected application.

20. A system comprising:

an electronic data store configured to store one or more speech recognition implementations, each speech recognition implementation being associated with at least one application; and

a computing device in communication with the electronic data store, the computing device configured to:

receive, from a device:

an audio input;

an identifier of a selected application; and

a request for speech recognition results;

select a speech recognition implementation, wherein the selected speech recognition implementation is associated with the selected application;

process at least a portion of the audio input with the selected speech recognition implementation to generate the requested speech recognition results; and

transmit the requested speech recognition results to the device.

21. The system of claim 20 , wherein the selected speech recognition implementation comprises a grammar associated with the selected application.

22. The system of claim 20 , wherein the computing device is further configured to associate a receipt with the audio input and to transmit the receipt to the device.

23. The system of claim 20 , wherein the selected application comprises at least one of an instant messaging application, a text messaging application, or an email application.

24. The system of claim 20 , wherein the computing device is configured using one or more servlets.

25. The system of claim 20 , wherein the requested speech recognition results comprise at least one of a transcription of the audio input or a web service result corresponding to the audio input.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2015
From: CANYON IP HOLDINGS LLC
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 037083/0914 →
Continuity (3)
Continuation 11697074 · Apr 5, 2007
Provisional Application 60789837 · Apr 5, 2006
Related Publication 20120166199A1 · Jun 28, 2012