IP Library Granted Patent US 9,661,381
Granted Patent B2
US 9,661,381 · App. 14/720,546 · Granted May 23, 2017

Using an audio stream to identify metadata associated with a currently playing television program

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,661,381
App. No.
14/720,546
Granted
May 23, 2017
Kind
B2
Abstract

Systems and methods for using an audio stream to identify metadata associated with a currently playing television program are disclosed. A video stream including audio description data is received. A set of information is determined from the audio description data. A request including the set of information is sent to a server remotely located from the client for additional processing. A set of instructions is received from the server. The set of instructions is determined based on the additional processing of the set of information. One or more applications are executed in accordance with the set of instructions in response to receiving the set of instructions.

Claims (45)

1. A method, comprising:

at a computing device having one or more processors and memory storing one or more programs to be executed by the one or more processors:

obtaining a video stream, an audio stream, and audio description data for a media program, the audio description data comprising a synchronized audio narrative describing what is happening in the media program during one or more of: natural pauses in primary audio content included in the audio stream of the media program, and the primary audio content;

determining a set of information from the audio description data, wherein the set of information includes one or more symbols or words derived from the audio description data;

sending a request including the set of information to a server remotely located from the computing device for processing;

receiving from the server a set of instructions, wherein the set of instructions includes instructions to display content information relating to the set of information; and

in response to receiving, and in accordance with, the set of instructions,

executing one or more applications based on a type of the content information, wherein the one or more applications include at least one of: a web browser, a music application, a feed reader application, a coupon application, and a content viewer.

2. The method of claim 1 , further comprising formatting for display, adjacent to a display of the video stream, output from the one or more applications.

3. The method of claim 1 , wherein the set of information includes non-speech information derived from the audio description data.

4. The method of claim 1 , wherein the set of information includes at least one symbol extracted from the audio description data.

5. The method of claim 1 , wherein the set of information includes at least one sentence extracted from the audio description data.

6. The method of claim 1 , wherein the output from the one or more applications is displayed concurrently with the video stream.

7. The method of claim 1 , wherein the output from the one or more applications is displayed concurrently on a second device synchronized with the computing device.

8. The method of claim 1 , wherein determining the set of information from the audio description data includes applying a speech recognition technique to convert the audio description data to text.

9. The method of claim 1 , wherein determining the set of information from the audio description data includes converting the audio description data into text without playing the audio description data.

10. The method of claim 1 , wherein the set of information includes at least some text extracted from the audio description data.

11. The method of claim 1 , further comprising:

transmitting a code to initiate playing of audio content included in the audio description data;

recording at least a portion of the audio content; and

extracting text from the recorded audio content.

12. A computing system, comprising:

one or more processors;

memory; and

one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for:

obtaining a video stream, an audio stream, and audio description data for a media program, the audio description data comprising a synchronized audio narrative describing what is happening in the media program during one or more of: natural pauses in primary audio content included in the audio stream of the media program, and the primary audio content;

determining a set of information from the audio description data, wherein the set of information includes one or more symbols or words derived from the audio description data;

sending a request including the set of information to a server remotely located from the computing device for processing;

receiving from the server a set of instructions, wherein the set of instructions includes instructions to display content information relating to the set of information; and

in response to receiving, and in accordance with, the set of instructions,

executing one or more applications based on a type of the content information, wherein the one or more applications include at least one of: a web browser, a music application, a feed reader application, a coupon application, and a content viewer.

13. The system of claim 12 , wherein the one or more programs further comprise instructions for: formatting for display, adjacent to a display of the video stream, output from the one or more applications.

14. The system of claim 12 , wherein the set of information includes non-speech information derived from the audio description data.

15. The system of claim 12 , wherein the set of information includes at least one symbol extracted from the audio description data.

16. The system of claim 12 , wherein the set of information includes at least one sentence extracted from the audio description data.

17. A non-transitory computer readable storage medium storing one or more programs configured for execution by one or more processors of a server system, the one or more programs comprising instructions, to be executed by the one or more processors, for:

obtaining a video stream, an audio stream, and audio description data for a media program, the audio description data comprising a synchronized audio narrative describing what is happening in the media program during one or more of: natural pauses in primary audio content included in the audio stream of the media program, and the primary audio content;

determining a set of information from the audio description data, wherein the set of information includes one or more symbols or words derived from the audio description data;

sending a request including the set of information to a server remotely located from the computing device for processing;

receiving from the server a set of instructions, wherein the set of instructions includes instructions to display content information relating to the set of information; and

in response to receiving, and in accordance with, the set of instructions,

executing one or more applications based on a type of the content information, wherein the one or more applications include at least one of: a web browser, a music application, a feed reader application, a coupon application, and a content viewer.

18. The non-transitory computer readable storage medium of claim 17 , wherein the one or more programs further comprise instructions for: formatting for display, adjacent to a display of the video stream, output from the one or more applications.

19. The non-transitory computer readable storage medium of claim 17 , wherein the set of information includes non-speech information derived from the audio description data.

20. The non-transitory computer readable storage medium of claim 17 , wherein the set of information includes at least one symbol extracted from the audio description data.

Assignments (1)
CHANGE OF NAME Recorded Dec 5, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044695/0115 →