Global speech user interface
A global speech user interface (GSUI) comprises an input system to receive a user's spoken command, a feedback system along with a set of feedback overlays to give the user information on the progress of his spoken requests, a set of visual cues on the television screen to help the user understand what he can say, a help system, and a model for navigation among applications. The interface is extensible to make it easy to add new applications.
1. A method of providing a speech user interface for television, comprising operations of:
analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers;
performing speech recognition to transcribe spoken input into transcribed spoken commands satisfying predetermined criteria;
using the transcribed spoken commands to navigate among predetermined applications concerning operation of a television; and
responsive to voiceprint recognition of a specific speaker, varying display of the television according to identity of the specific speaker;
the operation of varying display of the television comprising each of:
causing the television to display advertisements personalized to the specific speaker;
causing the television to display programming recommendations personalized to the specific speaker;
causing the television to display video-on-demand purchase recommendations personalized to the specific speaker;
expediting an on-screen purchase transaction based upon voiceprint;
causing the television to display an interface and television content personalized to the specific speaker; and
automatically configuring the television or speech user interface according to preferences of the specific speaker.
2. The method of claim 1 , the operation of varying display of the television further comprising
implementing a designated scheme for parental control of television content by automatically blocking or allowing television content according to identity of the specific speaker.
3. A method of providing a speech user interface for content presentation devices, comprising operations of:
performing speech recognition to transcribe spoken input into transcribed spoken commands satisfying predetermined criteria;
using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device;
responsive to recognition of a specific speaker by voice identification, varying display of the content presentation device according to identity of the specific speaker;
the operation of varying display of the content presentation device comprising each of:
causing the content presentation device to display advertisements personalized to the specific speaker;
causing the content presentation device to display programming recommendations personalized to the specific speaker;
causing the content presentation device to display video-on-demand purchase recommendations personalized to the specific speaker;
expediting an on-screen purchase transaction based upon recognition of the specific speaker;
causing the content presentation device to display an interface and television content personalized to the specific speaker; and
automatically configuring the content presentation device or speech user interface according to preferences of the specific speaker.
4. The method of claim 3 , the operation of varying display of the content presentation device further comprising:
implementing a designated scheme for parental control of television content by automatically blocking or allowing content presentation device content according to identity of the specific speaker.
5. The method of claim 3 , further comprising:
responsive to using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device, initiating instant messaging communication.
6. The method of claim 3 , further comprising:
analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and
responsive to voiceprint recognition of a specific speaker, initiating instant messaging communication.
7. The method of claim 3 , further comprising:
responsive to using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device, accessing one or more games and allowing said specific speaker to engage in game play.
8. The method of claim 3 , further comprising:
analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and
responsive to voiceprint recognition of a specific speaker, accessing one or more games and allowing said specific speaker to engage in said game play.
9. A method of providing a speech user interface for content presentation devices, comprising operations of:
performing speech recognition to transcribe spoken input into transcribed spoken commands satisfying predetermined criteria;
using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device;
responsive thereto, varying display of the content presentation device according to identity of the specific speaker;
analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers; and
responsive to voiceprint recognition of a specific speaker, varying display of the content presentation device according to identity of the specific speaker;
the operation of varying display of the content presentation device comprising each of:
causing the content presentation device to display advertisements personalized to the specific speaker;
causing the content presentation device to display programming recommendations personalized to the specific speaker;
causing the content presentation device to display video-on-demand purchase recommendations personalized to the specific speaker;
expediting a purchase transaction based upon voiceprint;
causing the content presentation device to display an interface and content personalized to the specific speaker; and
automatically configuring the content presentation device or speech user interface according to preferences of the specific speaker.
10. The method of claim 9 , the operation of varying display of the content presentation device further comprising:
implementing a designated scheme for parental control of content by automatically blocking or allowing content according to identity of the specific speaker.
11. A method of providing a speech user interface for content presentation devices, comprising operations of:
analyzing utterances of different speakers to facilitate voiceprint identification of the different speakers;
performing speech recognition to transcribe spoken input into transcribed spoken commands satisfying predetermined criteria;
using the transcribed spoken commands to navigate among predetermined applications concerning operation of a content presentation device; and
responsive to voiceprint recognition of a specific speaker, varying display of the content presentation device according to identity of the specific speaker;
the operation of varying display of the content presentation device comprising causing the television to display advertisements personalized to the specific speaker.