IP Library Granted Patent US 10,210,907
Granted Patent B2
US 10,210,907 · App. 15/627,980 · Granted Feb 19, 2019

Systems and methods for adding content to video/multimedia based on metadata

Inventors: Atul Puri (Redmond, CA); Hari Kalva (Delray Beach, FL)
Assignee: INTEL CORPORATION
G11B27/3081G06F3/0481G06T5/50G06T11/60H04N5/45H04N19/115H04N19/167H04N19/61H04N21/23412H04N21/234318H04N21/234345H04N21/44012H04N21/440245H04N21/4722H04N21/4725H04N21/4728H04N21/47205H04N21/8543H04N21/8545G06T2200/24G06T2200/32H04L47/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,210,907
App. No.
15/627,980
Granted
Feb 19, 2019
Kind
B2
Abstract

An interactive video/multimedia application (IVM application) may specify one or more media assets for playback. The IVM application may define the rendering, composition, and interactivity of one or more the assets, such as video. Video multimedia application data (IVMA data may) be used to define the behavior of the IVM application. The IVMA data may be embodied as a standalone file in a text or binary, compressed format. Alternatively, the IVMA data may be embedded within other media content. A video asset used in the IVM application may include embedded, content-aware metadata that is tightly coupled to the asset. The IVM application may reference the content-aware metadata embedded within the asset to define the rendering and composition of application display elements and user-interactivity features. The interactive video/multimedia application (defined by the video and multimedia application data) may be presented to a viewer in a player application.

Claims (40)

1. An apparatus comprising:

processing logic to automatically detect an object and its location in a plurality of frames of video while the video is playing, and to associate content with the detected object using metadata, wherein the metadata associates the location of the object in the frames to content to provide for metadata-to-content synchronization for said frames, and upon rendering of the frames, said processing logic to add the content to the frames, based on the metadata, at the locations in the frames, in association with the detected object, and track movement of the object and to display the content on the object as it moves; and

memory coupled to the processing logic, the memory to store the frames.

2. The apparatus of claim 1 , said processing logic to automatically detect an object by detecting at least one of a shape or texture.

3. The apparatus of claim 1 , wherein the metadata comprises at least one of a region of interest (ROI) bounding the detected object, and a location of the detected object in the frame.

4. The apparatus of claim 1 , wherein said content is an image.

5. The apparatus of claim 4 , wherein said image is an advertisement.

6. The apparatus of claim 1 , wherein said content is an advertisement.

7. The apparatus of claim 1 , wherein said content is text.

8. A method comprising:

automatically detecting an object and its location in a plurality of frames of video while the video is playing;

associating content with the detected object using metadata, wherein the metadata associates the location of the object in the frames to content to provide for metadata-to-content synchronization for said frames;

upon rendering of the frames, adding the content to the frames, based on the metadata, at the locations in the frames, in association with the detected object; and

tracking movement of the object and to display the content on the object as it moves.

9. The method of claim 8 , including detecting an object by detecting at least one of a shape or texture.

10. The method of claim 8 , wherein the metadata comprises at least one of a region of interest (ROI) bounding the detected object, and a location of the detected object in the frame.

11. The method of claim 8 , wherein said content is an image.

12. The method of claim 11 , wherein said image is an advertisement.

13. The method of claim 8 , wherein said content is an advertisement.

14. The method of claim 8 , wherein said content is text.

15. One or more non-transitory computer readable media storing instructions to perform a sequence comprising:

automatically detecting an object and its location in a plurality of frames of video while the video is playing;

associating content with the detected object using metadata, wherein the metadata associates the location of the object in the frames to content to provide for metadata-to-content synchronization for said frames,

upon rendering of the frames, adding the content to the frames, based on the metadata, at the locations in the frames, in association with the detected object; and

track movement of the object and to display the content on the object as it moves.

16. The media of claim 15 , further storing instructions to automatically detect an object by detecting at least one of a shape or texture.

17. The media of claim 15 , wherein the metadata comprises at least one of a region of interest (ROI) bounding the detected object, and a location of the detected object in the frame.

18. The media of claim 15 , wherein said content is an image.

19. The media of claim 18 , wherein said image is an advertisement.

20. The media of claim 15 , wherein said content is an advertisement.

21. The media of claim 15 , wherein said content is text.

22. An apparatus comprising:

means for automatically detecting an object and its location in a plurality of frames of video while the video is playing;

means for associating content with the detected object using metadata, wherein the metadata associates the location of the object in the frames to content to provide for metadata-to-content synchronization for said frames;

upon rendering of the frames, means for adding the content to the frames, based on the metadata, at the locations in the frames, in association with the detected object; and

means for tracking movement of the object and to display the content on the object as it moves.

23. The apparatus of claim 22 , including means for detecting an object by detecting at least one of a shape or texture.

24. The apparatus of claim 22 , wherein the metadata comprises at least one of a region of interest (ROI) bounding the detected object, and a location of the detected object in the frame.

25. The apparatus of claim 22 , wherein said content is an image.

26. The apparatus of claim 22 wherein said image is an advertisement.

Continuity (7)
Continuation 14962563 · Dec 8, 2015
Continuation 14524565 · Oct 27, 2014
Continuation 13972013 · Aug 21, 2013
Continuation 13742523 · Jan 16, 2013
Continuation 12586057 · Sep 16, 2009
Provisional Application 61192136 · Sep 16, 2008
Related Publication 20170287525A1 · Oct 5, 2017
Cited By (1)
US 12,289,483