IP Library Granted Patent US 9,042,703
Granted Patent B2
US 9,042,703 · App. 11/263,059 · Granted May 26, 2015

System and method for content-based navigation of live and recorded TV and video programs

Inventors: David Crawford Gibbon (Lincroft, NJ); Zhu Liu (Marlboro, NJ); Behzad Shahraray (Freehold, NJ)
Assignee: AT&T Intellectual Property II, L.P.
H04N5/91H04N21/4828
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,042,703
App. No.
11/263,059
Granted
May 26, 2015
Kind
B2
Abstract

A system, method and computer-readable medium are presented for providing real-time content-based navigation of live video programming. The video programming is received and a searchable database is generated. The method aspect of the invention comprises receiving a live video program, generating an index to the received live video program by extracting images and/or text from the video program, recording the live video program, presenting at least a portion of data associated with the generated index to a user, receiving user input regarding a search to a portion of the recorded video program to which the user desires to navigate and playing back the recorded video program starting at the searched portion identified by the user input. The search may be of an image and/or text portion of the presentation.

Claims (47)

1. A method comprising:

receiving a broadcast video program from a first source;

recording the broadcast video program, to yield a recorded video program;

performing an analysis of images and text associated with the recorded video program, to generate a content-based index, wherein the content-based index:

is generated independent of index information transmitted with the broadcast video program; and

is generated by combining extracted images and the text from the recorded video program with network-based content provided by an external source, the network-based content being transmitted separately from the broadcast video program;

automatically, via a processor and without user input, supplementing the content-based index with the network-based content as the content-based index is generated, wherein the network-based content is from a second source that is distinct from the first source and the external source;

presenting a search field, wherein presenting of the search field occurs on a first computing device different from a second computing device which displays the broadcast video program, wherein the first computing device is a hand-held device comprising a touch-sensitive screen;

receiving user input in the search field, the user input describing one of an item and an activity;

searching the content-based index to identify a location in the recorded video program associated with the user input; and

playing back the recorded video program starting at the location on the second device.

2. The method of claim 1 , wherein the network-based content accompanies the broadcast video program.

3. The method of claim 1 , wherein:

when a plurality of portions of the broadcast video program match the user input, then presenting the location to the user further comprises presenting the plurality of portions; and

playing back the recorded video program starting at the location is in response to a user selection from the plurality of portions.

4. The method of claim 1 , wherein generating the content-based index further comprises generating a searchable database of the images and the text from the broadcast video program and one of: a content-based sampling module to extract images, a textual module to extract text, and an automatic speech recognition module to extract text.

5. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

receiving a broadcast video program from a first source;

recording the broadcast video program, to yield a recorded video program;

performing an analysis of images and text associated with the recorded video program, to generate a content-based index, wherein the content-based index:

is generated independent of index information transmitted with the broadcast video program; and

is generated by combining extracted images and the text from the recorded video program with network-based content provided by an external source, the network-based content being transmitted separately from the broadcast video program;

automatically, via a processor and without user input, supplementing the content-based index with the network-based content as the content-based index is generated, wherein the network-based content is from a second source that is distinct from the first source and the external source;

presenting a search field, wherein presenting of the search field occurs on a first computing device different from a second computing device which displays the broadcast video program, wherein the first computing device is a hand-held device comprising a touch-sensitive screen;

receiving user input in the search field, the user input describing one of an item and an activity;

searching the content-based index to identify a location in the recorded video program associated with the user input; and

playing back the recorded video program on the second computing device starting at the location.

6. The system of claim 5 , the computer-readable storage medium having additional instructions stored which, when executed by the processor, result in the operations further comprising:

when a plurality of portions of the broadcast video program match the user input, presenting of the location further comprises presenting the plurality of portions.

7. The system of claim 5 , wherein performing the analysis further comprises using one of: a content-based sampling module to extract images, a textual module to extract text, and an automatic speech recognition module to extract text.

8. A computer-readable device having instructions stored which, when executed by a computing device, cause the computing device to perform operations comprising:

receiving a broadcast video program from a first source;

recording the broadcast video program, to yield a recorded video program;

performing an analysis of images and text associated with the recorded video program, to generate a content-based index, wherein the content-based index:

is generated independent of index information transmitted with the broadcast video program; and

is generated by combining extracted images and the text from the recorded video program with network-based content provided by an external source, the network-based content being transmitted separately from the broadcast video program;

automatically, via a processor and without user input, supplementing the content-based index with the network-based content as the content-based index is generated, wherein the network-based content is from a second source that is distinct from the first source and the external source;

presenting a search field, wherein presenting of the search field occurs on a first computing device different from a second computing device which displays the broadcast video program, wherein the first computing device is a hand-held device comprising a touch-sensitive screen;

receiving user input in the search field, the user input describing one of an item and an activity;

searching the content-based index to identify a location in the recorded video program associated with the user input; and

playing back the recorded video program starting at the location on the second device.

9. The computer-readable device of claim 8 , wherein:

when a plurality of portions of the broadcast video program match the user input, then presenting the location to the user further comprises presenting the plurality of portions; and

playing back the recorded video program starting at the location is in response to a user selection from the plurality of portions.

10. The computer-readable device of claim 8 , wherein generating the content-based index further comprises generating a searchable database of images and text from the broadcast video program and one of: a content-based sampling module to extract images, a textual module to extract text and an automatic speech recognition module to extract text.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2015
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 035024/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2015
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 035024/0114 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 3, 2006
From: GIBBON, DAVID CRAWFORD; LIU, ZHU; SHAHRARAY, BEHZAD
To: AT&T CORP.
Reel/Frame 017119/0205 →
Continuity (1)
Related Publication 20070098350A1 · May 3, 2007