IP Library Granted Patent US 9,769,524
Granted Patent B2
US 9,769,524 · App. 15/203,336 · Granted Sep 19, 2017

Method and apparatus for augmenting media content

Inventors: Raghuraman Gopalan (Union City, CA); Lee Begeja (Gillette, NJ); David Crawford Gibbon (Lincroft, NJ); Zhu Liu (Marlboro, NJ); Amy Ruth Reibman (Chatham, NJ); Bernard S. Renger (New Providence, NJ); Behzad Shahraray (Holmdel, NJ); Eric Zavesky (Austin, TX)
Assignee: AT&T Intellectual Property I, L.P.
H04N21/44218G06K9/00302G06K9/00711G06K9/00718H04N21/231H04N21/233H04N21/2353H04N21/23418H04N21/242H04N21/4394H04N21/44008H04N21/44016H04N21/454H04N21/458H04N21/4532H04N21/4542H04N21/4667H04N21/4725H04N21/8106H04N21/8133H04N21/84H04N21/8405H04N21/858H04N21/8547
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,769,524
App. No.
15/203,336
Granted
Sep 19, 2017
Kind
B2
Abstract

Aspects of the subject disclosure may include, for example, generating narrative descriptions corresponding to visual features, visual events, and interactions there between for media content, where the narrative descriptions are associated with time stamps of the media content, and presenting the media content and an audio reproduction of the narrative descriptions, wherein the audio reproduction is synchronized to video of the media content according to the time stamps. Other embodiments are disclosed.

Claims (46)

1. A non-transitory machine-readable storage medium, comprising executable instructions that, when executed by a processing system including a processor, facilitate performance of operations, comprising:

determining visual capabilities of a user viewing a media content item at a device;

filtering a plurality of narrative descriptions associated with the media content item according to the visual capabilities of the user viewing the media content item to generate a plurality of filtered descriptions;

determining if a feature in the plurality of narrative descriptions has been provided in a first portion of the filtered description; and

replacing a non-shortened description of the feature with a shortened description of the feature in a second portion of the filtered description responsive to determining that the feature has been provided in the first portion, wherein the shortened description occupies a smaller memory space than the non-shortened description of the feature provided in the first portion of the filtered description; and

transmitting the media content item, the first portion of the plurality of filtered descriptions, and the second portion of the plurality of filtered descriptions to the device for presentation via an application.

2. The non-transitory machine-readable storage medium of claim 1 , wherein the narrative description is further filtered according to requirements of the application.

3. The non-transitory machine-readable storage medium of claim 1 , wherein the plurality of narrative descriptions are associated with a plurality of time stamps of the media content item.

4. The non-transitory machine-readable storage medium of claim 1 , wherein the operations further comprise:

scanning a plurality of images from the media content item to detect a plurality of visual features present within the plurality of images;

analyzing interactions between the plurality of visual features to determine a plurality of visual events that are depicted in the media content item; and

generating the plurality of narrative descriptions corresponding to the plurality of visual features and the plurality of visual events.

5. The non-transitory machine-readable storage medium of claim 4 , wherein the filtered descriptions comprise emotional states in the media content item, and wherein the operations further comprise:

recognizing facial expressions of persons in the visual features in the media content item; and

determining the emotional states in the media content according to the facial expressions that are recognized.

6. The non-transitory machine-readable storage medium of claim 4 , wherein the operations further comprise filtering the plurality of visual features according to second requirements of the application.

7. The non-transitory machine-readable storage medium of claim 4 , wherein the plurality of narrative descriptions comprise natural language descriptions of the plurality of visual features and the plurality of visual events.

8. The non-transitory machine-readable storage medium of claim 4 , wherein the operations further comprise analyzing audio content of the media content item to extract audio information associated with the plurality of visual features, wherein the step of analyzing interactions further includes the audio information.

9. The non-transitory machine-readable storage medium of claim 4 , wherein the operations further comprise analyzing lighting characteristics of the plurality of visual features to extract mood information, wherein the step of analyzing interactions further includes the mood information.

10. The non-transitory machine-readable storage medium of claim 1 , wherein the operations further comprise:

searching the plurality of filtered descriptions based on a search term;

selecting a video clip from the media content item, wherein a filtered description of the plurality of filtered descriptions substantially matches the search term;

generating a speech rendering corresponding to the filtered descriptions; and

presenting the video clip and the speech rendering via the application.

11. The non-transitory machine-readable storage medium of claim 1 , wherein the filtering of the narrative descriptions according to second requirements of a profile of user associated with the application.

12. The non-transitory machine-readable storage medium of claim 1 , wherein the operations further comprise:

comparing the plurality of narrative descriptions to a plurality of objectionable event descriptions; and

censoring a portion of the media content item corresponding to a narrative description of the plurality of narrative descriptions that substantially matches an objectionable event description of the plurality of objectionable event descriptions.

13. The non-transitory machine-readable storage medium of claim 12 , wherein the censoring comprises removing the portion of the media content item.

14. The non-transitory machine-readable storage medium of claim 12 , wherein the censoring comprises generating an advisory of the objectionable event.

15. A communication device, comprising:

a processing system including a processor; and

a memory that stores executable instructions that, when executed by the processing system, facilitate performance of operations, comprising:

providing, to a server, visual capabilities of a user viewing a media content item;

receiving, from the server, the media content item, a first portion of a plurality of filtered descriptions associated with the media content item, and a second portion of the plurality of filtered descriptions associated with the media content item, a wherein the filter descriptions are generated from a plurality of narrative descriptions associated with the media content according to the visual capabilities of the user of the media content item, and wherein the first portion includes a non-shortened description of a feature and the second portion includes a shortened description of the feature that occupies a smaller memory space than the non-shortened description; and

presenting the media content item, the first portion of the plurality of filtered descriptions associated with the media content item, and the second portion of the plurality of filtered descriptions associated with the media content item via an application.

16. The communication device of claim 15 , wherein the operations further comprise:

receiving a portion of the media content item from the server, wherein the portion is extracted from the media content item corresponding to a narrative description of the plurality of narrative descriptions that substantially matches a target event description of a plurality of target event descriptions; and

presenting the portion of the media content item via the application.

17. The communication device of claim 15 , wherein the operations further comprise receiving a graphical representation of a portion of the plurality of narrative descriptions.

18. The communication device of claim 15 , wherein the narrative description is further filtered according to requirements of the application.

19. The communication device of claim 15 , wherein the plurality of narrative descriptions are associated with a plurality of time stamps of the media content item.

20. A method, comprising:

providing, by a processing system including a processor, to a server, visual capabilities of a user viewing a media content item;

receiving, by the processing system, from the server, the media content item, a first portion of a plurality of filtered descriptions associated with the media content item, and a second portion of the plurality of filtered descriptions associated with the media content item, a wherein the filter descriptions are generated from a plurality of narrative descriptions associated with the media content according to the visual capabilities of the user of the media content item, and wherein the first portion includes a non-shortened description of a feature and the second portion includes a shortened description of the feature that occupies a smaller memory space than the non-shortened description; and

presenting, by the processing system, the media content item, the first portion of the plurality of filtered descriptions associated with the media content item, and the second portion of the plurality of filtered descriptions associated with the media content item via an application.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 26, 2016
From: GOPALAN, RAGHURAMAN; BEGEJA, LEE; GIBBON, DAVID CRAWFORD; LIU, ZHU; REIBMAN, AMY RUTH; RENGER, BERNARD S.; SHAHRARAY, BEHZAD; ZAVESKY, ERIC
To: AT&T INTELLECTUAL PROPERTY I, LP
Reel/Frame 039255/0863 →
Continuity (2)
Continuation 14264580 · Apr 29, 2014
Related Publication 20160316265A1 · Oct 27, 2016