IP Library Granted Patent US 8,301,454
Granted Patent B2
US 8,301,454 · App. 12/546,636 · Granted Oct 30, 2012

Methods, apparatuses, and systems for providing timely user cues pertaining to speech recognition

Assignee: Canyon IP Holdings LLC
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,301,454
App. No.
12/546,636
Granted
Oct 30, 2012
Kind
B2
Abstract

A method is provided of providing cues from am electronic communication device to a user while capturing an utterance. A plurality of cues associated with the user utterance are provided by the device to the user in at least near real-time. For each of a plurality of portions of the utterance, data representative of the respective portion of the user utterance is communicated from the electronic communication device to a remote electronic device. In response to this communication, data, representative of at least one parameter associated with the respective portion of the user utterance, is received at the electronic communication device. The electronic communication device provides one or more cues to the user based on the at least parameter. At least one of the cues is provided by the electronic communication device to the user prior to completion of the step of capturing the user utterance.

Claims (83)

1. A system for providing cues from a device to a user while capturing an utterance, the system comprising:

a backend server; and

a hand-held mobile communication device in communication with the backend server over a network, the hand-held communication device configured to:

capture a first portion of a user utterance; and

communicate to the backend server, data representative of the first portion of the user utterance;

the backend server being configured to:

receive the data representative of the first portion of the user utterance;

process the data representative of the first portion of the user utterance, said processing including determining a metric associated with the first portion of the user utterance; and

communicate to the hand-held mobile communication device, data representative of the metric associated with the first portion of the user utterance;

the hand-held mobile communication device being further configured to:

receive from the backend server, the data representative of the metric associated with the first portion of the user utterance;

provide to the user, a cue based at least in part on the received data representative of a metric associated with the first portion of the user utterance;

capture a second portion of the user utterance; and

communicate to the backend server, data representative of the second portion of the user utterance;

the backend server being further configured to:

receive the data representative of the second portion of the user utterance;

process the data representative of the second portion of the user utterance, said processing including determining a metric associated with the second portion of the user utterance; and

communicate to the hand-held mobile communication device, data representative of the metric associated with the second portion of the user utterance;

the hand-held mobile communication device being further configured to:

receive from the backend server, the data representative of the metric associated with the second portion of the user utterance; and

provide to the user, a cue based on the received data representative of the metric associated with the second portion of the user utterance;

wherein said providing to the user, a cue based at least in part on the received data representative of a metric associated with the first portion of the user utterance, occurs prior to said capturing, by the hand-held mobile communication device, a second portion of the user utterance.

2. The system of claim 1 , wherein:

the hand-held mobile communication device is further configured to:

capture a third portion of the user utterance; and

communicate to the backend server, data representative of the third portion of the user utterance;

the backend server is further configured to:

receive the data representative of the third portion of the user utterance;

process the data representative of the third portion of the user utterance, said processing including determining a metric associated with the third portion of the user utterance; and

communicate to the hand-held mobile communication device, data representative of the metric associated with the third portion of the user utterance; and

the hand-held mobile communication device is further configured to:

receive from the backend server, the data representative of the metric associated with the third portion of the user utterance; and

provide to the user, a cue based at least in part on the received data representative of the metric associated with the third portion of the user utterance;

wherein said providing to the user, a cue based at least in part on the received data representative of a metric associated with the second portion of the user utterance, occurs prior to said capturing, by the hand-held mobile communication device, a third portion of the user utterance.

3. A computer-implemented method of providing cues to a user while capturing an utterance, the computer-implemented method comprising:

capturing, by an electronic communication device, a user utterance; and

providing, by the electronic communication device to the user in at least near real-time, one or more cues associated with the user utterance, wherein said providing includes, for each portion of a plurality of portions of the user utterance:

communicating, from the electronic communication device, data representative of the respective portion of the user utterance to a remote electronic device,

in response to the communication of data representative of the respective portion of the user utterance, receiving, at the electronic communication device, data representative of at least one parameter associated with the respective portion of the user utterance, and

providing, by the electronic communication device to the user, at least one cue based at least in part on the at least one parameter associated with the respective portion of the user utterance;

wherein at least one cue of the one or more cues is provided by the electronic communication device to the user prior to completion of capturing the user utterance.

4. The computer-implemented method of claim 3 , wherein said communicating, from the electronic communication device, data representative of each respective portion of the user utterance comprises streaming, from the electronic communication device, data representative of the user utterance.

5. The computer-implemented method of claim 3 , wherein said receiving, at the electronic communication device, data representative of at least one parameter associated with the respective portion of the user utterance comprises receiving, at the electronic communication device, a token comprising data representative of at least one parameter associated with the respective portion of the user utterance.

6. The computer-implemented method of claim 3 , wherein each respective portion of the user utterance consists of a word, each respective portion of the user utterance consists of a syllable, each respective portion of the user utterance consists of a phrase, or each respective portion of the user utterance consists of a sentence.

7. The computer-implemented method of claim 3 , wherein each respective portion of the user utterance comprises a word, each respective portion of the user utterance comprises a syllable, each respective portion of the user utterance comprises a phrase, or each respective portion of the user utterance comprises a sentence.

8. The computer-implemented method of claim 3 , wherein the at least one parameter associated with each respective portion of the user utterance comprises at least one metric associated with each respective portion of the user utterance.

9. The computer-implemented method of claim 3 , wherein the at least one parameter associated with each respective portion of the user utterance comprises a confidence level corresponding to a transcription result of the respective portion of the user utterance.

10. The computer-implemented method of claim 3 , further comprising receiving, for each respective portion of the plurality of portions, at the electronic communication device, together with the data representative of at the least one parameter associated with the respective portion of the user utterance, data representative of a transcription result of the respective portion of the user utterance.

11. The computer-implemented method of claim 3 , wherein the at least one parameter associated with each respective portion of the user utterance comprises a volume level of the utterance or a background noise level of the user utterance.

12. The computer-implemented method of claim 3 , wherein said providing, by the electronic communication device to the user, at least one cue, comprises displaying, via the electronic communication device, a graphical cue.

13. The computer-implemented method of claim 12 , wherein the graphical cue comprises an emoticon.

14. The computer-implemented method of claim 3 , wherein said providing, by the electronic communication device to the user, at least one cue, comprises outputting, via the electronic communication device, an auditory cue, verbal cue, or optical cue.

15. The computer-implemented method of claim 3 , wherein said providing, by the electronic communication device to the user, at least one cue, comprises providing, by the electronic communication device to the user, a plurality of cues, each cue being based at least in part on a different parameter associated with the respective portion of the user utterance.

16. The computer-implemented method of claim 3 , wherein said providing, by the electronic communication device to the user, at least one cue, comprises providing, by the electronic communication device to the user, a combination cue based at least in part on a plurality of parameters associated with the respective portion of the user utterance.

17. The computer-implemented method of claim 16 , wherein the combination cue is configured to be perceivable as representative of at least two different parameters associated with the respective portion of the user utterance.

18. The computer-implemented method of claim 3 , further comprising adjusting, by the user, his or her speech pattern based at least in part on at least one provided cue.

19. A system for providing cues to a user while capturing an utterance, the system comprising:

a local computing device, the local computing device configured to:

capture a user utterance;

provide to the user in at least near real-time one or more cues associated with the user utterance, wherein said providing includes, for each portion of a plurality of portions of the user utterance:

transmitting, to a remote computing device over a network, data representative of the respective portion of the user utterance;

receiving, from the remote computing device, data representative of at least one parameter associated with the respective portion of the user utterance; and

providing to the user at least one cue based at least in part on the at least one parameter associated with the respective portion of the user utterance;

wherein at least one cue of the one or more cues is provided to the user prior to completion of capturing the user utterance; and

a data storage device in communication with the local computing device, the data storage device configured to store data representative of one or more parameters received from the remote computing device.

20. The system of claim 19 , wherein each respective portion of the user utterance comprises at least one of a syllable, a word, a phrase, or a sentence.

21. The system of claim 19 , wherein the at least one parameter associated with each respective portion of the user utterance comprises a confidence level corresponding to a transcription result of the respective portion of the user utterance.

22. The system of claim 19 , wherein, for each respective portion of the plurality of portions of the user utterance, the local computing device is further configured to receive from the remote computing device, together with the data representative of the at least one parameter associated with the respective portion of the user utterance, data representative of a transcription result of the respective portion of the user utterance.

23. The system of claim 19 , wherein the at least one parameter associated with each respective portion of the user utterance comprises at least one of a one of a volume level of the respective portion of the user utterance or a background noise level of the respective portion of the user utterance.

24. The system of claim 19 , wherein the provided at least one cue comprises at least one of a graphical cue, an optical cue, a verbal cue, or an auditory cue.

25. A non-transitory computer-readable medium having a computer-executable component for providing cues to a user while capturing an utterance, the computer-executable component comprising:

a cue-providing component configured to:

cause an electronic communication device to capture a user utterance; and

cause the electronic communication device to provide to the user in at least near real-time one or more cues associated with the user utterance, wherein said providing includes, for each portion of a plurality of portions of the user utterance:

causing the electronic communication device to communicate data representative of the respective portion of the user utterance to a remote electronic device;

in response to the communication of data representative of the respective portion of the user utterance, causing the electronic communication device to receive data representative of at least one parameter associated with the respective portion of the user utterance; and

causing the electronic communication device to provide at least one cue based at least in part on the at least one parameter associated with the respective portion of the user utterance;

wherein the electronic communication device is caused to provide at least one cue of the one or more cues prior to completion of capturing the user utterance.

26. The non-transitory computer-readable medium of claim 25 , wherein each respective portion of the user utterance comprises at least one of a syllable, a word, a phrase, or a sentence.

27. The non-transitory computer-readable medium of claim 25 , wherein the at least one parameter associated with each respective portion of the user utterance comprises a confidence level corresponding to a transcription result of the respective portion of the user utterance.

28. The non-transitory computer-readable medium of claim 25 , wherein the cue-providing component is further configured to cause the electronic communication device to receive, together with the data representative of the at least one parameter associated with the respective portion of the user utterance, data representative of a transcription result of the respective portion of the user utterance.

29. The non-transitory computer-readable medium of claim 25 , wherein the at least one parameter associated with each respective portion of the user utterance comprises at least one of a volume level of the respective portion of the user utterance or a background noise level of the respective portion of the user utterance.

30. The non-transitory computer-readable medium of claim 25 , wherein the provided at least one cue comprises at least one of a graphical cue, an optical cue, a verbal cue, or an auditory cue.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2015
From: CANYON IP HOLDINGS LLC
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 037083/0914 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 27, 2012
From: YAP LLC
To: CANYON IP HOLDINGS LLC
Reel/Frame 027770/0733 →
RELEASE OF SECURITY INTEREST Recorded Oct 1, 2011
From: VENTIRE LENDING & LEASING V, INC. AND VENTURE LENDING & LEASING VI, INC.
To: YAP INC.
Reel/Frame 027001/0859 →
SECURITY AGREEMENT Recorded Dec 21, 2010
From: YAP INC.
To: VENTURE LENDING & LEASING VI, INC.; VENTURE LENDING & LEASING V, INC.
Reel/Frame 025521/0513 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 12, 2010
From: PADEN, SCOTT EDWARD
To: YAP, INC.
Reel/Frame 024230/0209 →
Continuity (2)
Provisional Application 61091330 · Aug 22, 2008
Related Publication 20100049525A1 · Feb 25, 2010