IP Library Granted Patent US 8,005,679
Granted Patent B2
US 8,005,679 · App. 11/933,191 · Granted Aug 23, 2011

Global speech user interface

Assignee: Promptu Systems Corporation
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,005,679
App. No.
11/933,191
Granted
Aug 23, 2011
Kind
B2
Abstract

A global speech user interface (GSUI) comprises an input system to receive a user's spoken command, a feedback system along with a set of feedback overlays to give the user information on the progress of his spoken requests, a set of visual cues on the television screen to help the user understand what he can say, a help system, and a model for navigation among applications. The interface is extensible to make it easy to add new applications.

Claims (54)

1. A method for operating a global speech user interface (GSUI), comprising:

providing a speech input device having a switch, where the GSUI is activated by user activation of the switch;

performing speech recognition to transcribe spoken commands into commands acceptable by a communications system;

using the transcribed spoken commands to navigate among applications hosted on said communications system;

displaying a set of immediate speech feedback overlays including visual cues to guide a user in issuing proper spoken commands where each immediate speech feedback overlay provides non-textual feedback information about a state of said communications system, comprising:

(a) checking if a current screen is speech-enabled when said switch is activated;

(b) if the current screen is speech-enabled, displaying a first tab signaling that a speech input system is activated;

(c) if the current screen is not speech-enabled, displaying a second tab signaling a non speech-enabled alert, said second tab staying on screen for a first interval;

(d) if said switch is re-activated, repeating Step(a);

(e) if said switch is not deactivated within a second interval, interrupting recognition;

(f) if said switch is deactivated after a third interval lapsed but before said second interval in Step (e) lapsed, displaying a third tab signaling that speech recognition is in processing; and

(g) if said switch was deactivated before said third interval in Step (f) lapsed, removing any tab on the screen.

2. The method of claim 1 , wherein said first tab includes a solid image of a predetermined logo.

3. The method of claim 1 , wherein said second tab comprises a prohibiting sign overlaid on a logo.

4. The method of claim 3 , wherein said second tab further comprises a text box for textual message.

5. The method of claim 1 , wherein said first interval in Step (c) is approximately ten seconds.

6. The method of claim 1 , wherein said second interval in Step (e) is approximately ten seconds and said third interval in Step (f) is approximately 0.1 second.

7. The method of claim 1 , wherein said third tab is a flashing predetermined logo which is approximately 40% transparent.

8. The method of claim 1 , where said GSUI is implemented in a communications including a set top box and a head-end, and wherein said Step (f) further comprises the steps of:

(h) if said set top box takes longer than a fourth interval measured from the time that the user releases said switch to the time that the last speech data is sent to said head-end, interrupting speech recognition processing and displaying a fourth tab signaling an application alert, said fourth tab staying on the screen for a fifth interval; and

(i) if a remote control button other than said switch is pressed while a spoken command is being processed, interrupting speech recognition processing and removing any tab on the screen.

9. The method of claim 8 , wherein said fourth interval is approximately five seconds and said fifth interval is approximately ten seconds.

10. The method of claim 8 , wherein said fourth tab comprises an exclamation point overlaid on said logo.

11. The method of claim 10 , wherein said fourth tab further comprises a text box for textual messages.

12. The method of claim 8 , wherein said Step (h) further comprises the steps of:

(j) if said switch is re-activated while said fourth tab on the screen, removing the fourth tab and repeating Step (a); and

(k) when said fifth interval lapses or if a remote control button other than said switch is activated while said fourth tab is on the screen, removing said fourth tab.

13. The method of claim 1 , wherein said Step (f), upon a complete recognition, further comprises the steps of:

(l) checking whether the speech recognition is successful;

(m) if the speech recognition is successful, displaying a fifth tab signaling a positive speech recognition, said fifth tab staying on the screen for a predetermined period; and

(n) if said switch is re-activated before said fifth tab disappears, repeating Step (a).

14. The method of claim 13 , wherein said fifth tab comprises a check mark overlaid on said logo.

15. The method of claim 13 , further comprising automatically counting unsuccessful recognitions, including resetting a count to zero after each successful recognition or when any button of a remote control device is pressed;

wherein said Step (I) further comprises the steps of:

(o) if the speech recognition is unsuccessful, checking the count of unsuccessful recognitions;

(p) if the complete recognition is the first unsuccessful recognition, displaying a sixth tab signaling a misrecognition speech, said sixth tab staying on the screen for about one second; and

(q) if said switch is repressed before said sixth tab disappears, repeating Step (a).

16. The method of claim 15 , wherein said sixth tab in Step (p) is a question mark overlaid on said logo.

17. The method of claim 15 , wherein said Step (o) further comprises the steps of:

(r) if the complete recognition is the second unsuccessful recognition, displaying a first variant of said sixth tab signaling a misrecognition speech and displaying a short textual message, said first variant of said sixth tab staying on the screen for about ten seconds; and

(s) if said switch is repressed before said first variant of said sixth tab disappears, repeating Step (a).

18. The method of claim 17 , wherein said first variant of said sixth tab comprises:

a question mark overlaid on said logo; and

a short text box displaying a short textual message.

19. The method of claim 15 , wherein said Step (o) further comprises the steps of:

(t) if the complete recognition is the third unsuccessful recognition, displaying a second variant of said sixth tab signaling a misrecognition speech and displaying a long textual message, said second variant of said sixth tab staying on the screen for about ten seconds; and

(u) if said switch is re-activated before said second variant of said sixth tab disappears, repeating Step (a).

20. The method of claim 1 , wherein said Step (e) further comprises the steps of:

(v) displaying a first variant of said fourth tab, said first variant staying on the screen for a sixth interval;

(w) removing said first variant of said fourth tab from the screen if said switch is deactivated after said sixth interval lapsed; and

(x) displaying a second variant of said fourth tab, said second variant staying on the screen until said switch is deactivated.

21. The method of claim 20 , wherein said first variant comprises an exclamation point and a first textual message.

22. The method of claim 20 , wherein said sixth interval is approximately ten seconds.

23. The method of claim 20 , wherein said second variant comprises an exclamation point and a second textual message.

Assignments (1)
CHANGE OF NAME Recorded Nov 5, 2010
From: AGILETV CORPORATION
To: PROMPTU SYSTEMS CORPORATION
Reel/Frame 025326/0067 →
Continuity (3)
Division 10260906 · Sep 30, 2002
Provisional Application 60327207 · Oct 3, 2001
Related Publication 20080120112A1 · May 22, 2008