IP Library Granted Patent US 8,682,663
Granted Patent B2
US 8,682,663 · App. 13/924,112 · Granted Mar 25, 2014

Performing speech recognition over a network and using speech recognition results based on determining that a network connection exists

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,682,663
App. No.
13/924,112
Granted
Mar 25, 2014
Kind
B2
Abstract

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for generating, distributing, and using speech recognition models are disclosed. The methods, systems, and apparatus include actions of receiving a speech data representation of an utterance and obtaining feature data extracted from the speech data representation of the utterance. Additional actions include obtaining a transcription of the utterance using the feature data and identifying contact information associated with the transcription. Further actions include determining that a network connection exists and initiating communication using the contact information based on determining that a network connection exists.

Claims (44)

1. A computer-implemented method comprising:

receiving, by a mobile device, a speech data representation of an utterance;

obtaining, by the mobile device, feature data extracted from the speech data representation of the utterance;

obtaining, by the mobile device, a transcription of the utterance using the feature data;

identifying, by the mobile device, contact information associated with the transcription;

determining, by the mobile device, that a network connection exists; and

initiating, by the mobile device, communication using the contact information based on determining that a network connection exists.

2. The method of claim 1 , comprising determining, by the mobile device, that the mobile device supports local feature extraction, wherein the feature data is extracted from the speech data representation of the mobile device after the mobile device determines that local feature extraction is supported.

3. The method of claim 1 , comprising determining, by the mobile device, that the mobile device supports local speech recognition, wherein the transcription is obtained based on determining that local speech recognition is supported.

4. The method of claim 1 , wherein obtaining the transcription comprises generating the transcription using one or more speech recognition models that are locally stored by mobile device.

5. The method of claim 1 , wherein identifying contact information associated with the transcription comprises:

determining that a portion of the transcription matches a name of a contact; and

determining that the contact information is associated with the contact.

6. The method of claim 1 , wherein determining that a network connection exists comprises determining whether a computer-to-telephone connection exists.

7. A non-transitory computer-readable medium storing software comprising

instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:

receiving, by a mobile device, a speech data representation of an utterance;

obtaining, by the mobile device, feature data extracted from the speech data representation of the utterance;

obtaining, by the mobile device, a transcription of the utterance using the feature data;

identifying, by the mobile device, contact information associated with the transcription;

determining, by the mobile device, that a network connection exists; and

initiating, by the mobile device, communication using the contact information based on determining that a network connection exists.

8. The medium of claim 7 , the operations further comprising determining, by the mobile device, that the mobile device supports local feature extraction, wherein the feature data is extracted from the speech data representation of the mobile device after the mobile device determines that local feature extraction is supported.

9. The medium of claim 7 , the operations further comprising determining, by the mobile device, that the mobile device supports local speech recognition, wherein the transcription is obtained based on determining that local speech recognition is supported.

10. The medium of claim 7 , wherein obtaining the transcription comprises generating the transcription using one or more speech recognition models that are locally stored by mobile device.

11. The medium of claim 7 , wherein identifying contact information associated with the transcription comprises:

determining that a portion of the transcription matches a name of a contact; and

determining that the contact information is associated with the contact.

12. The medium of claim 7 , wherein determining that a network connection exists comprises determining whether a computer-to-telephone connection exists.

13. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving, by a mobile device, a speech data representation of an utterance;

obtaining, by the mobile device, feature data extracted from the speech data representation of the utterance;

obtaining, by the mobile device, a transcription of the utterance using the feature data;

identifying, by the mobile device, contact information associated with the transcription;

determining, by the mobile device, that a network connection exists; and

initiating, by the mobile device, communication using the contact information based on determining that a network connection exists.

14. The system of claim 13 , the operations further comprising determining, by the mobile device, that the mobile device supports local feature extraction, wherein the feature data is extracted from the speech data representation of the mobile device after the mobile device determines that local feature extraction is supported.

15. The system of claim 13 , the operations further comprising determining, by the mobile device, that the mobile device supports local speech recognition, wherein the transcription is obtained based on determining that local speech recognition is supported.

16. The system of claim 13 , wherein obtaining the transcription comprises generating the transcription using one or more speech recognition models that are locally stored by mobile device.

17. The system of claim 13 , wherein identifying contact information associated with the transcription comprises:

determining that a portion of the transcription matches a name of a contact; and

determining that the contact information is associated with the contact.

18. The system of claim 13 , wherein determining that a network connection exists comprises determining whether a computer-to-telephone connection exists.

Assignments (4)
CHANGE OF NAME Recorded Oct 6, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044141/0827 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2013
From: REDING, CRAIG L.; LEVAS, SUZI
To: TELESECTOR RESOURCES GROUP, INC.
Reel/Frame 031107/0973 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2013
From: TELESECTOR RESOURCES GROUP, INC.
To: VERIZON PATENT AND LICENSING INC.
Reel/Frame 031108/0024 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2013
From: VERIZON PATENT AND LICENSING INC.
To: GOOGLE INC.
Reel/Frame 031108/0038 →