IP Library Granted Patent US 9,848,243
Granted Patent B2
US 9,848,243 · App. 14/572,596 · Granted Dec 19, 2017

Global speech user interface

Inventors: Adam Jordan (El Cerrito, CA); Scott Lynn Maddux (San Francisco, CA); Tim Plowman (Berkeley, CA); Victoria Stanbach (Santa Cruz, CA); Jody Williams (San Carlos, CA)
Assignee: PROMPTU SYSTEMS CORPORATION
H04N21/47G06Q30/0271G06Q30/0631G10L15/22G10L21/06H04N5/44543H04N21/42203H04N21/4316H04N21/4622H04N21/472H04N21/475H04N21/478H04N21/4781H04N21/4782H04N21/4788H04N21/47202H04N21/47211H04N21/47214H04N21/482H04N21/4826H04N21/4828H04N21/4852H04N21/812H04N21/8173G06F3/16G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,848,243
App. No.
14/572,596
Granted
Dec 19, 2017
Kind
B2
Abstract

A global speech user interface (GSUI) comprises an input system to receive a user's spoken command, a feedback system along with a set of feedback overlays to give the user information on the progress of his spoken requests, a set of visual cues on the television screen to help the user understand what he can say, a help system, and a model for navigation among applications. The interface is extensible to make it easy to add new applications.

Claims (58)

1. A global speech user interface (GSUI) device, comprising:

a processor device configured for performing speech recognition to transcribe spoken commands for use in navigating among one or more applications hosted on a communications system;

said processor device configured for displaying a set of visual cues to guide a user in issuing spoken commands, said visual cues comprising an overlay that shows a list of services available to the user;

wherein said overlay comprises:

a first sub-menu overlay specifically for access, by spoken command, to an interactive program guide system;

a second sub-menu overlay specifically for access, by spoken command, to a content on demand system; and

a third sub-menu overlay specifically for access, by spoken command, to a browser-based service;

wherein each of said sub-menus provides a set of speech-activated virtual buttons; and

a speaker personalization and identification mechanism for training said communications system to identify respective users by voice;

wherein said speaker personalization and identification mechanism presents different custom interfaces and personalized content for different respective speakers.

2. The interface device of claim 1 , further comprising:

said processor device configured for initiating, via spoken command, an automatic scan search for content comprising one or more programs pursuant to a search category.

3. The interface device of claim 2 , wherein each matching program remains on screen for a predetermined period of time before advancing to next matching program.

4. The interface device of claim 1 , further comprising:

said processor device configured for, via spoken command, performing any of:

adding content to categories;

editing content in categories; and

deleting content from categories.

5. The interface device of claim 1 , further comprising:

said processor device configured for, via spoken command, filtering groups of content by specific attributes.

6. The interface device of claim 1 , said interactive program guide comprising:

responsive to a spoken command, sorting content by category;

responsive to a spoken command, setting parental controls, wherein children are blocked from accessing controlled channels or content;

responsive to a spoken command, setting reminders for content to play in the future;

responsive to a spoken command, searching content based on a specific criteria;

processing pay per view purchases;

responsive to a spoken command, any of accessing and upgrading premium content services; or

a combination thereof.

7. The interface device of claim 1 , further comprising, based upon voice identification, any of:

targeting one or more advertisements contained in an application screen to the user;

targeting content recommendations to the user;

delivering personalized information to the user; and

automatically configuring the user's interface preferences.

8. The interface device of claim 1 , further comprising:

varying display of content comprising any of:

displaying advertisements personalized to the specific speaker;

displaying programming recommendations personalized to the specific speaker;

displaying video-on-demand purchase recommendations personalized to the specific speaker;

expediting an on-screen purchase transaction based upon recognition of the specific speaker;

displaying an interface and content personalized to the specific speaker;

automatically configuring the display or user interface according to preferences of the specific speaker; and

implementing a designated scheme for parental control of content by automatically blocking or allowing content according to identity of the specific speaker.

9. The interface device of claim 1 , further comprising:

responsive to using the transcribed spoken commands to navigate among one or more applications, initiating instant messaging communication.

10. The interface device of claim 1 , further comprising:

analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and

responsive to voiceprint recognition of a specific speaker, initiating instant messaging communication.

11. The interface device of claim 1 , further comprising:

responsive to using the transcribed spoken commands to navigate among one or more applications, accessing one or more games and allowing said specific speaker to engage in game play.

12. The interface device of claim 1 , further comprising:

analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and

responsive to voiceprint recognition of a specific speaker, accessing one or more games and allowing said specific speaker to engage in said game play.

13. The interface device of claim 1 , further comprising:

using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device; and

responsive thereto, varying display of the content presentation device according to identity of the specific speaker.

14. The interface device of claim 1 , further comprising:

analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and

responsive to voiceprint recognition of a specific speaker, varying display of the content presentation device according to identity of the specific speaker.

Continuity (7)
Continuation 14029729 · Sep 17, 2013
Division 13786998 · Mar 6, 2013
Division 13179294 · Jul 8, 2011
Continuation 11933191 · Oct 31, 2007
Division 10260906 · Sep 30, 2002
Provisional Application 60327207 · Oct 3, 2001
Related Publication 20150106836A1 · Apr 16, 2015