IP Library Granted Patent US 9,477,664
Granted Patent B2
US 9,477,664 · App. 13/285,429 · Granted Oct 25, 2016

Method and apparatus for querying media based on media characteristics

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,477,664
App. No.
13/285,429
Granted
Oct 25, 2016
Kind
B2
Abstract

An approach is provided for querying media based on media characteristics. A media platform processes and/or facilitates a processing of one or more images, one or more videos, or a combination thereof to determine one or more latent vectors associated with the one or more images, the one or more videos, or the combination thereof. The media platform further causes, at least in part, a comparison of the one or more latent vectors to one or more models. The media platform also causes, at least in part, an indexing of the one or more images, the one or more videos, or the combination thereof based, at least in part, on the one or more latent vectors, the one or more models, or a combination thereof.

Claims (70)

1. A method comprising facilitating a processing of or processing (1) data or (2) information or (3) at least one signal, the (1) data or (2) information or (3) at least one signal based, at least in part, on the following:

a processing of one or more images, one or more videos, one or more segments of the one or more images, one or more segments of the one or more videos, or a combination thereof to determine one or more media latent vectors associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof;

a processing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof to determine metadata associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof;

an indexing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof based, at least in part, on the metadata;

a comparison of the one or more media latent vectors to one or more model latent vectors associated with one or more models,

wherein the one or more models are predetermined media,

wherein the one or more model latent vectors are based on a factorization of the predetermined media using one or more predetermined latent parameters; and

an indexing to determine an index of indexed latent vectors associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof based, at least in part, on the comparison.

2. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

a rendering of a user interface for determining a selection of at least one of the one or more models, one or more objects represented by the one or more models, or a combination thereof;

a querying of the index for the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or a combination thereof based, at least in part, on the selection; and

a rendering of one or more results of the query in the user interface,

wherein the one or more media latent vectors represent one or more objects and/or topics associated with the one or more images, one or more videos, one or more segments of the one or more images, one or more segments of the one or more videos, or a combination thereof.

3. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

a processing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof to determine one or more sets of latent parameters; and

at least one determination of the one or more media latent vectors based, at least in part, on the one or more sets of the latent parameters.

4. A method of claim 1 , wherein the metadata includes location information, and the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof are indexed based, at least in part, on the location information.

5. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

a processing of the one or more images, the one or more videos, or the combination thereof to determine the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof and respective segment latent vectors.

6. A method of claim 5 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

a synchronization of at least part of the metadata associated with the one or more segments of the one or more images, the one or more segments of the one or more videos, the one or more segment latent vectors, or a combination thereof; and

an indexing of the one or more segments based, at least in part, on the synchronized metadata.

7. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

at least one image, at least one video, at least one segment of at least one image, at least one segment of at least one video, or a combination thereof match at least one model associated with at least one landmark; and

an indexing of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof based, at least in part, on the at least one landmark.

8. A method of claim 7 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

an association of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof with metadata associated with the at least one landmark; and

an indexing of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof based, at least in part, on the at least one landmark metadata,

wherein the landmark metadata includes location information, orientation information, or a combination thereof.

9. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

one or more image queries, one or more video queries, or a combination thereof; and

at least one image, at least one video, at least one segment of at least one image, at least one segment of at least one video, or a combination thereof that satisfies the one or more image queries, the one or more video queries, or the combination thereof based, at least in part, on the one or more media latent vectors.

10. A method of claim 1 , wherein the (1) data or (2) information or (3) at least one signal are further based, at least in part, on the following:

determining at least one distance between the one or more media latent vectors and the one or more model latent vectors based on the comparison; and

matching at least one of the media latent vectors with at least one of the one or more model latent vectors based on the at least one distance, wherein the indexing is further based on the matching.

11. An apparatus comprising:

at least one processor; and

at least one memory including computer program code for one or more programs,

the at least one memory and the computer program code configured to, with the at least one processor, cause the apparatus to perform at least the following,

process or facilitate a processing of one or more images, one or more videos, one or more segments of the one or more images, one or more segments of the one or more videos, or a combination thereof to determine one or more media latent vectors associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof,

process or facilitate a processing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof to determine metadata associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof,

cause, at least in part, an indexing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof based, at least in part, on the metadata,

cause, at least in part, a comparison of the one or more media latent vectors to one or more model latent vectors associated with one or more models,

wherein the one or more models are predetermined media,

wherein the one or more model latent vectors are based on a factorization of the predetermined media using one or more latent parameters, and

cause, at least in part, an indexing to determine an index of indexed latent vectors associated with the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof based, at least in part, on the comparison to determine an index.

12. An apparatus of claim 11 , wherein the apparatus is further caused to:

cause, at least in part, a rendering of a user interface for determining a selection of at least one of the one or more models, one or more objects represented by the one or more models, or a combination thereof,

cause, at least in part, a querying of the index for the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or a combination thereof based, at least in part, on the selection, and

cause, at least in part, a rendering of one or more results of the query in the user interface,

wherein the one or more media latent vectors represent one or more objects or topics associated with the one or more images, one or more videos, one or more segments of the one or more images, one or more segments of the one or more videos, or a combination thereof.

13. An apparatus of claim 11 , wherein the apparatus is further caused to:

process or facilitate a processing of the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof to determine one or more sets of latent parameters, and

determine the one or more latent vectors based, at least in part, on the one or more sets of the latent parameters.

14. An apparatus of claim 11 , wherein the metadata includes location information, and the one or more images, the one or more videos, the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof are indexed based, at least in part, on the location information.

15. An apparatus of claim 11 , wherein the apparatus is further caused to:

process or facilitate a processing of the one or more images, the one or more videos, or the combination thereof to determine the one or more segments of the one or more images, the one or more segments of the one or more videos, or the combination thereof and respective segment latent vectors.

16. An apparatus of claim 15 , wherein the apparatus is further caused to:

cause, at least in part, a synchronization of at least part of the metadata associated with the one or more segments of the one or more images, the one or more segments of the one or more videos, the one or more segment latent vectors, or a combination thereof, and

cause, at least in part, an indexing of the one or more segments based, at least in part, on the synchronized metadata.

17. An apparatus of claim 11 , wherein the apparatus is further caused to:

determine at least one image, at least one video, at least one segment of at least one image, at least one segment of at least one video, or a combination thereof match at least one model associated with at least one landmark, and

cause, at least in part, an indexing of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof based, at least in part, on the at least one landmark.

18. An apparatus of claim 17 , wherein the apparatus is further caused to:

cause, at least in part, an association of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof with metadata associated with the at least one landmark, and

cause, at least in part, an indexing of the at least one image, the at least one video, the at least one segment of the at least one image, the at least one segment of the at least one video, or the combination thereof based, at least in part, on the at least one landmark metadata,

wherein the landmark metadata includes location information, orientation information, or a combination thereof.

19. An apparatus of claim 11 , wherein the apparatus is further caused to:

receive one or more image queries, one or more video queries, or a combination thereof, and

determine at least one image, at least one video, at least one segment of at least one image, at least one segment of at least one video, or a combination thereof that satisfies the one or more image queries, the one or more video queries, or the combination thereof based, at least in part, on the one or more media latent vectors.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 1, 2015
From: NOKIA CORPORATION
To: NOKIA TECHNOLOGIES OY
Reel/Frame 035313/0218 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 22, 2012
From: SATHISH, SAILESH KUMAR; CURCIO, IGOR DANILO DIEGO
To: NOKIA CORPORATION
Reel/Frame 028826/0819 →