IP Library Granted Patent US 8,793,583
Granted Patent B2
US 8,793,583 · App. 13/654,327 · Granted Jul 29, 2014

Method and apparatus for annotating video content with metadata generated using speech recognition technology

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,793,583
App. No.
13/654,327
Granted
Jul 29, 2014
Kind
B2
Abstract

A method and apparatus is provided for annotating video content with metadata generated using speech recognition technology. The method begins by rendering video content on a display device. A segment of speech is received from a user such that the speech segment annotates a portion of the video content currently being rendered. The speech segment is converted to a text-segment and the text-segment is associated with the rendered portion of the video content. The text segment is stored in a selectively retrievable manner so that it is associated with the rendered portion of the video content.

Claims (55)

1. A method for annotation of video content in a device communicatively coupled to a network, the method comprising:

receiving, in the device, a captured speech segment comprising speech from a user of a second device, wherein the captured speech segment annotates a portion of the video content streamed to the second device for being played to the user contemporaneously with the speech from the user;

converting the captured speech segment to a text-segment;

associating the text-segment with the portion of the video content contemporaneously played to the user; and

storing in a selectively retrievable manner the text-segment so that the text-segment is associated with the portion of the video content.

2. An apparatus for annotation of a video content, the apparatus comprising:

a memory; and

a processor communicatively coupled to the memory and to a network interface,

the processor configured to be communicatively coupled via the network interface to a network;

the processor further configured to receive, via the network interface, a captured speech segment comprising speech from a user of a second device coupled to the network, wherein the captured speech segment annotates a portion of the video content streamed to the second device for being played to the user contemporaneously with the speech from the user;

the processor further configured to convert the captured speech segment to a text-segment, to associate the text-segment with the portion of the video content contemporaneously played to the user; and to store in a selectively retrievable manner the text-segment so that the text-segment is associated with the portion of the video content.

3. The method of claim 1 further comprising:

streaming the video content via the network to the second device.

4. The method of claim 1 further comprising:

receiving, in the device, a timestamp associated with the captured speech segment;

wherein associating the text-segment with the portion of the video content further comprises using the timestamp.

5. The method of claim 1 further comprising:

generating metadata based on the text-segment.

6. The method of claim 1 further comprising:

generating metadata based on an identified speaker associated with the speech segment.

7. The method of claim 1 further comprising:

generating metadata based on specific words of the text-segment.

8. The method of claim 1 further comprising:

receiving, in the device, before receiving the captured speech segment, a message comprising a user input selecting an operational state.

9. The method of claim 1 wherein the operational state is selected from the group consisting of an annotate state, a narrate state, a commentary state, an analyze state, and a review/edit state.

10. The method of claim 1 wherein storing the text-segment further comprises:

storing the text-segment and metadata in a database of a storage device communicatively coupled to the network.

11. The method of claim 1 wherein storing the text-segment further comprises:

storing the text-segment in a database of a storage device communicatively coupled to the network.

12. The method of claim 1 further comprising:

storing, in a database of a storage device communicatively coupled to the network, metadata comprising a timestamp for associating the text-segment with the portion of the video content.

13. The method of claim 1 further comprising:

storing, in a storage device communicatively coupled to the network, a modified version of the video content comprising metadata for associating the text-segment with the portion of the video content.

14. The method of claim 1 wherein storing the text-segment further comprises:

storing, in a storage device communicatively coupled to the network, a modified version of the video content comprising the text-segment and metadata for associating the text-segment with the portion of the video content.

15. A method for annotation of video content in a device communicatively coupled to a network, the method comprising:

receiving, in the device, a text-segment of recognized speech comprising recognized speech from a user of a second device coupled to the network, wherein the text-segment annotates a portion of the video content streamed to the second device for being played to the user contemporaneously with the speech from the user;

associating the text-segment with the portion of the video content; and

storing in a selectively retrievable manner the text-segment so that it is associated with the portion of the video content.

16. The method of claim 15 further comprising:

streaming the video content via the network to the second device.

17. The method of claim 15 further comprising:

receiving, in the device, a timestamp associated with the captured speech segment;

wherein associating the text-segment with the portion of the video content further comprises using the timestamp.

18. A non-transitory computer-readable medium having computer-executable instructions embodied thereon for annotation of video content in a device communicatively coupled to a network, wherein the instructions, when executed by at least one processor of the device, cause the at least one processor to perform the method of claim 15 .

19. A method for annotation of video content in a device communicatively coupled to a network, the method comprising:

receiving, in the device, a text-segment of recognized speech comprising recognized speech from a user of a second device, wherein the text-segment annotates a portion of the video content streamed to the second device for being played to the user contemporaneously with the speech from the user;

receiving, in the device, metadata comprising a timestamp for associating the text-segment with the portion of the video content; and

storing in a selectively retrievable manner the text-segment so that it is associated with the portion of the video content.

20. The method of claim 19 further comprising:

streaming the video content via the network to the second device.

21. The method of claim 19 further comprising:

receiving, in the device, a timestamp associated with the captured speech segment;

wherein associating the text-segment with the portion of the video content further comprises using the timestamp.

22. A non-transitory computer-readable medium having computer-executable instructions embodied thereon for annotation of video content in a device communicatively coupled to a network, wherein the instructions, when executed by at least one processor of the device, cause the at least one processor to perform the method of claim 19 .

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2014
From: MOTOROLA MOBILITY LLC
To: GOOGLE TECHNOLOGY HOLDINGS LLC
Reel/Frame 034244/0014 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 12, 2014
From: MCKOEN, KEVIN M.; GROSSMAN, MICHAEL A.
To: GENERAL INSTRUMENT CORPORATION
Reel/Frame 033091/0252 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 9, 2013
From: GENERAL INSTRUMENT CORPORATION
To: GENERAL INSTRUMENT HOLDINGS, INC.
Reel/Frame 030764/0575 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 9, 2013
From: GENERAL INSTRUMENT HOLDINGS, INC.
To: MOTOROLA MOBILITY LLC
Reel/Frame 030866/0113 →