IP Library Granted Patent US 7,809,760
Granted Patent B2
US 7,809,760 · App. 12/559,817 · Granted Oct 5, 2010

Multimedia integration description scheme, method and system for MPEG-7

Assignees: AT&T Intellectual Property II, L.P.; Columbia University in the City of New York
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,809,760
App. No.
12/559,817
Granted
Oct 5, 2010
Kind
B2
Abstract

The invention provides a system and method for integrating multimedia descriptions in a way that allows humans, software components or devices to easily identify, represent, manage, retrieve, and categorize the multimedia content. In this manner, a user who may be interested in locating a specific piece of multimedia content from a database, Internet, or broadcast media, for example, may search for and find the multimedia content. In this regard, the invention provides a system and method that receives multimedia content and separates the multimedia content into separate components which are assigned to multimedia categories, such as image, video, audio, synthetic and text. Within each of the multimedia categories, the multimedia content is classified and descriptions of the multimedia content are generated. The descriptions are then formatted, integrated, using a multimedia integration description scheme, and the multimedia integration description is generated for the multimedia content. The multimedia description is then stored into a database. As a result, a user may query a search engine which then retrieves the multimedia content from the database whose integration description matches the query criteria specified by the user. The search engine can then provide the user a useful search result based on the multimedia integration description.

Claims (46)

1. A method for utilizing description records from multimedia content, the method causing a computing device to perform steps comprising:

receiving multimedia content having description records generated by steps comprising:

generating multimedia object descriptions for at least one identified multimedia type based on multimedia objects extracted from the multimedia content;

generating, from the multimedia object descriptions, non-hierarchical entity relation graph descriptions for the at least one identified multimedia type, wherein the non-hierarchical entity relation graph descriptions are associated with communication between multimedia objects; and

integrating the multimedia object descriptions and the entity relation graph descriptions to generate at least one description record to represent content embedded within the multimedia content;

extracting a description record from the multimedia content;

displaying the multimedia content; and

displaying the extracted description record in connection with the displayed multimedia content.

2. The method of claim 1 , wherein generating non-hierarchical entity relation graph descriptions for at least one multimedia type based on media content further comprises:

identifying multimedia types in multimedia content; and

extracting multimedia objects to generate multimedia object descriptions from the multimedia content for at least one of the multimedia types, wherein generating non-hierarchical entity relation graph is based on at least in part on the multimedia object descriptions.

3. The method of claim 1 , the steps for generating description records further comprising generating, from the multimedia content, multimedia object hierarchy descriptions by object hierarchy construction and extraction processing for at least one of the multimedia types.

4. The method of claim 1 , wherein multimedia types include image, audio, video, synthetic and text.

5. The method of claim 2 , wherein the multimedia objects are extracted from the multimedia content by steps comprising:

segmenting multimedia content into segments including content from at least one of the multimedia types for the multimedia content; and

generating at least one feature description for at least one of the segments by feature extraction and annotation, wherein the generated multimedia object descriptions comprise the at least one feature description for the at least one segment.

6. The method of claim 5 , wherein the segments are selected from the group consisting of local segments and global segments.

7. The method of claim 5 , the multimedia objects are extracted from the multimedia content by steps further comprising selecting the at least one feature description from the group consisting of media, semantic, and temporal features.

8. The method of claim 7 , wherein the media features are further defined by at least one feature description selected from the group comprising data location, scalable representation, and modality transcoding.

9. The method of claim 7 , wherein the semantic features are further defined by at least one feature description selected from the group comprising keywords, who, what object, what action, why, when, where, and text annotation.

10. The method of claim 7 , wherein the temporal features are further defined by at least one feature description consisting of duration.

11. The method of claim 5 , wherein the multimedia objects are extracted from the multimedia content by steps further comprising:

generating media object descriptions from the multimedia segment for one of the multimedia types by media object extraction processing;

generating media object hierarchy descriptions from the generated media object descriptions by object hierarchy construction and extraction processing; and

generating media entity relation graph descriptions from the generated media object descriptions by entity relation graph generation processing.

12. The method of claim 11 , wherein the multimedia objects are extracted from the multimedia content by steps further comprising:

segmenting the content of each multimedia type in the multimedia object into segments within the multimedia object by media segmentation processing; and

generating at least one feature description for at least one of the segments by feature extraction and annotation, wherein the generated media object descriptions comprise the at least one feature description for the at least one of the segments.

13. The method of claim 12 , wherein the multimedia objects are extracted from the multimedia content by steps further comprising selecting the at least one feature description from the group consisting of media, semantic and temporal.

14. The method of claim 12 , wherein generating media object hierarchy descriptions generates media object hierarchy descriptions of the media object descriptions based on media feature relationships of media objects represented by the media object descriptions.

15. The method of claim 12 , wherein generating media object hierarchy descriptions generates semantic object hierarchy descriptions of the media object descriptions based on semantic feature relationships of media objects represented by the media object descriptions.

16. The method of claim 12 , wherein generating media object hierarchy descriptions generates temporal object hierarchy descriptions of the media object descriptions based on temporal features relationships of media objects represented by the media object descriptions.

17. The method of claim 12 , wherein generating media object hierarchy descriptions generates media object hierarchy descriptions of the media object descriptions based on relationships of media objects represented by the media object descriptions, and wherein the relationships are selected from the group comprising media feature relationships, semantic feature relationships, temporal feature relationships, and spatial feature relationships.

18. The method of claim 12 , wherein generating media entity relation graph descriptions generates entity relation graph descriptions of the media object descriptions based on relationship of the media objects represented by the media object descriptions, wherein the relationships are selected from the group comprising media feature relationships, semantic feature relationships, temporal feature relationships, and spatial feature relationships.

19. The method of claim 1 , wherein generating multimedia object hierarchy descriptions further generates multimedia object hierarchy descriptions of the multimedia object descriptions based on media feature relationships of multimedia objects represented by the multimedia object descriptions.

20. The method of claim 1 , wherein generating multimedia object hierarchy descriptions further generates semantic object hierarchy descriptions of the multimedia object descriptions based on semantic feature relationships of multimedia objects represented by the multimedia object descriptions.

21. The method of claim 1 , wherein generating multimedia object hierarchy descriptions further generates temporal object hierarchy descriptions of the multimedia object descriptions based on temporal feature relationships of multimedia objects represented by the multimedia object descriptions.

22. The method of claim 1 , wherein generating multimedia object hierarchy descriptions further generates multimedia object hierarchy descriptions of the multimedia object descriptions based on relationships of multimedia objects represented by the multimedia object descriptions, wherein the relationships are selected from the group comprising media feature relationships, semantic feature relationships, temporal feature relationships, and spatial feature relationships.

23. The method of claim 1 , wherein generating entity relation graph descriptions further generates the entity relation graph descriptions of the multimedia object descriptions based on relationships of multimedia objects represented by the multimedia object descriptions, wherein the relationships are selected from the group comprising media feature relationships, semantic feature relationships, temporal feature relationships, and spatial feature relationships.

24. The method of claim 1 , wherein description records are generated by steps further comprising:

receiving and encoding the multimedia object descriptions into encoded description information; and

storing the encoded description information as the at least one description record.

25. The method of claim 2 , wherein description records are generated by steps further comprising:

combining the multimedia object description, the multimedia object hierarchy descriptions, and the entity relation graph description to form a multimedia description;

receiving and encoding the multimedia description into encoded description information; and

storing the encoded description information as the at least one description record.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 22, 2014
From: HUANG, QIAN
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 034006/0242 →
Continuity (4)
Continuation 1161631500 · Dec 27, 2006
Continuation 0949517500 · Feb 1, 2000
Provisional Application 6011802200 · Feb 1, 1999
Related Publication 20100005121A1 · Jan 7, 2010