IP Library Granted Patent US 8,818,804
Granted Patent B2
US 8,818,804 · App. 13/786,998 · Granted Aug 26, 2014

Global speech user interface

Inventors: Adam Jordan (El Cerrito, CA); Scott Lynn Maddux (San Francisco, CA); Tim Plowman (Berkeley, CA); Victoria Stanbach (Santa Cruz, CA); Jody Williams (San Carlos, CA)
Assignee: Promptu Systems Corporation
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,818,804
App. No.
13/786,998
Granted
Aug 26, 2014
Kind
B2
Abstract

A global speech user interface (GSUI) comprises an input system to receive a user's spoken command, a feedback system along with a set of feedback overlays to give the user information on the progress of his spoken requests, a set of visual cues on the television screen to help the user understand what he can say, a help system, and a model for navigation among applications. The interface is extensible to make it easy to add new applications.

Claims (93)

1. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command;

wherein said main menu overlay comprises:

a first sub-menu overlay specifically for access to an interactive program guide system;

a second sub-menu overlay specifically for access to a video on demand system which provides access to a library of video content; and

a third sub-menu overlay specifically for access to a walled garden system which provides browser-based Internet service;

wherein each of said sub-menus provides a set of speech-activated virtual buttons.

2. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising a speaker personalization and identification mechanism for training said communications system to identify respective users by voice;

wherein said speaker personalization and identification mechanism is programmed to perform any of presenting different custom interfaces and personalized television content for different respective trained speakers, and accessing blocked content response to one or more utterances from a given speaker.

3. A system for establishing a global speech user interface (GSUI) comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising said processor configured for allowing the user to navigate television programs by spoken command.

4. The system of claim 3 , further comprising:

said processor configured for allowing the user to initiate via spoken command an automatic scan search for television programs pursuant to a search category, wherein each matching program remains on screen for a predetermined period of time before advancing to next matching program.

5. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising said processor configured for allowing the user to search, via spoken command, for particular television programs by specific attributes.

6. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising said processor configured for allowing the user to perform any of:

adding television programs to categories;

editing television programs in categories; and

deleting television programs from categories.

7. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising said processor configured for allowing the user to set parental control, with which children are blocked from accessing controlled television channels or television programs.

8. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising said processor configured for allowing the user to filter groups of television programs by specific attributes.

9. The system of claim 8 , wherein said interactive program guide comprises any of:

allowing the user to, via spoken command, sort television programs by category;

allowing the user to set parental controls, with which children are blocked from accessing controlled television channels or television programs;

allowing the user to, via spoken command, set reminders for television programs to play in the future;

allowing the user to, via spoken command, search television programs based on a specific criteria;

processing pay per view purchases; and

allowing the user to, via spoken command, access and upgrade premium television services.

10. A system for establishing a global speech user interface (GSUI), comprising:

a processor configured for performing speech recognition to transcribe spoken commands into commands acceptable by said communications system;

said processor configured for using the transcribed spoken commands to navigate among applications hosted on said communications system; and

said processor configured for displaying a set of visual cues to guide a user in issuing proper spoken commands, said visual cues comprising:

a set of immediate speech feedback overlays, each of which provides non-textual feedback information about a state of said communications system;

a set of help overlays, each of which provides a context-sensitive list of frequently used speech-activated commands for each screen of every speech-activated application;

a set of feedback overlays, each of which provides information about a problem that said communications system is experiencing; and

a main menu overlay that shows a list of services available to the user, each of said services being accessible by spoken command; and

further comprising, based upon voice identification, any of:

targeting television advertisement or banner advertisement contained in an application screen to the user;

targeting television programming recommendations to the user;

delivering personalized information to the user; and

automatically configuring the user's interface preferences.

Continuity (5)
Division 13179294 · Jul 8, 2011
Continuation 11933191 · Oct 31, 2007
Division 10260906 · Sep 30, 2002
Provisional Application 60327207 · Oct 3, 2001
Related Publication 20130211836A1 · Aug 15, 2013