IP Library Granted Patent US 7,961,949
Granted Patent B2
US 7,961,949 · App. 12/577,487 · Granted Jun 14, 2011

Extracting multiple identifiers from audio and video content

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,961,949
App. No.
12/577,487
Filed
Oct 12, 2009
Granted
Jun 14, 2011
Kind
B2
Examiner
DO, ANH HONG
Art Unit
2624
USPC
382/190
Abstract

The disclosure concerns content identification, such as extracting identifying information from content itself. One combination described in the disclosure is a method including: extracting first identifying information from data representing audio elements of an audio signal, the act of extracting first identifying information from data representing audio elements of the audio signal utilizes a programmed electronic processor; extracting second identifying information from data representing picture elements of a video signal that is associated with the audio signal, the act of extracting second identifying information from data representing picture elements of the video signal utilizes a programmed electronic processor; and utilizing the first identifying information or the second identifying information in a synchronization process, the synchronization process controls content synchronization during rendering of the audio signal or the video signal. Of course, other combinations are provided as well.

Claims (76)

1. A method comprising:

extracting first identifying information from data representing audio elements of an audio signal, wherein the act of extracting first identifying information from data representing audio elements of the audio signal utilizes a processor;

extracting second identifying information from data representing picture elements of a video signal that is associated with the audio signal, wherein the act of extracting second identifying information from data representing picture elements of the video signal utilizes a processor; and

utilizing the first identifying information or the second identifying information in a synchronization process, wherein the synchronization process controls content synchronization during rendering of the audio signal via a speaker or during rendering of the video signal via a display.

2. The method of claim 1 , wherein the video signal comprises a time-compressed format.

3. The method of claim 1 , wherein the act of extracting first identifying information comprises decoding steganographically hidden information from the data representing audio elements of the audio signal.

4. The method of claim 1 , wherein the act of extracting second identifying information comprises decoding steganographically hidden information from the data representing picture elements of the video signal that is associated with the audio signal.

5. The method of claim 1 , wherein the content comprises metadata associated with the audio signal or the video signal.

6. The method of claim 5 , wherein the metadata includes a URL link.

7. The method of claim 1 , wherein the content comprises audio or video.

8. The method of claim 1 , wherein the content comprises ownership information.

9. The method of claim 1 , wherein the content comprises purchase information.

10. The method of claim 1 , further comprising utilizing the first identifying information or the second identifying information to track rendering of the audio signal or the video signal.

11. The method of claim 1 , wherein the act of extracting second identifying information comprises low pass filtering the data representing picture elements.

12. The method of claim 1 , wherein the act of extracting first identifying information comprises low pass filtering the data representing audio elements.

13. The method of claim 1 , wherein the act of extracting second identifying information comprises hashing the data representing picture elements.

14. The method of claim 1 , wherein the act of extracting first identifying information comprises hashing the data representing audio elements.

15. The method of claim 1 , wherein the act of extracting first identifying information comprises analyzing aural attributes of the data representing audio elements.

16. The method of claim 15 , wherein a temporal location of the aural attributes is analyzed.

17. The method of claim 15 , wherein spectral energy of the aural attributes is analyzed.

18. The method of claim 1 , wherein the act of extracting first identifying information transforms the data representing audio elements into a transform domain.

19. The method of claim 1 , wherein the act of extracting second identifying information transforms the data representing pictures elements into a transform domain.

20. A non-transitory computer-readable medium comprising instructions stored thereon, the instructions comprising:

instructions to extract first identifying information from data representing audio elements of an audio signal;

instructions to extract second identifying information from data representing picture elements of a video signal that is associated with the audio signal; and

instructions to utilize the first identifying information or the second identifying information in a synchronization process, wherein the synchronization process controls content synchronization during rendering of the audio signal via a speaker or during rendering of the video signal via a display.

21. The non-transitory computer readable medium of claim 20 , wherein the act of extracting first identifying information comprises decoding steganographically hidden information from the data representing audio elements of the audio signal.

22. An apparatus comprising:

a memory configured to:

buffer data representing audio elements of an audio signal; and

buffer data representing picture elements of a video signal; and one or more processors configured to:

extract first identifying information from data representing audio elements of an audio signal;

extract second identifying information from data representing picture elements of a video signal that is associated with the audio signal; and

utilize the first identifying information or the second identifying information in a synchronization process, wherein the synchronization process controls content synchronization during rendering of the audio signal or the video signal.

23. The apparatus of claim 22 , wherein the extracting first identifying information comprises decoding steganographically hidden information from the data representing audio elements of the audio signal.

24. The apparatus of claim 22 , wherein the extracting second identifying information comprises decoding steganographically hidden information from the data representing picture elements of the video signal that is associated with the audio signal.

25. The apparatus of claim 22 , wherein the content comprises metadata associated with the audio signal or the video signal.

26. The apparatus of claim 25 , wherein the metadata includes a URL link.

27. The apparatus of claim 22 , wherein the content comprises audio or video.

28. The apparatus of claim 22 , wherein the content comprises ownership information.

29. The apparatus of claim 22 , wherein the content comprises purchase information.

30. The apparatus of claim 22 , wherein the one or more processors are further configured to utilize the first identifying information or the second identifying information to track rendering of the audio signal or the video signal.

31. The apparatus of claim 22 , wherein the apparatus comprises an electronic handheld media player.

32. The apparatus of claim 22 , wherein the content comprises purchase information.

33. The apparatus of claim 31 , wherein the one or more processors are further configured to utilize the first identifying information or the second identifying information to track rendering of the audio signal or the video signal.

34. The apparatus of claim 31 , wherein the one or more processors are configured to extract the second identifying information by low pass filtering the data representing picture elements.

35. The apparatus of claim 31 , wherein the one or more processors are further configured to analyze a temporal location of aural attributes of the data representing audio elements.

36. The apparatus of claim 31 , wherein the one or more processors are further configured to analyze a spectral energy of aural attributes of the data representing audio elements.

37. The apparatus of claim 22 , wherein the one or more processors are configured to extract the second identifying information by low pass filtering the data representing picture elements.

38. The apparatus of claim 22 , wherein the one or more processors are configured to extract the first identifying information by low pass filtering the data representing audio elements.

39. The apparatus of claim 22 , wherein the one or more processors are configured to extract the second identifying information by hashing the data representing picture elements.

40. The apparatus of claim 22 , wherein the one or more processors are configured to extract the first identifying information by hashing the data representing audio elements.

41. The apparatus of claim 22 , wherein the one or more processors are configured to extract the first identifying information by analyzing aural attributes of the data representing audio elements.

42. The apparatus of claim 41 , wherein the one or more processors are configured to analyze a temporal location of the aural attributes.

43. The apparatus of claim 41 , wherein the one or more processors are configured to a spectral energy of the aural attributes.

44. The apparatus of claim 22 , wherein the one or more processors are configured to extract the first identifying information by transforming the data representing audio elements into a transform domain.

45. The apparatus of claim 22 , wherein the one or more processors are configured to extract the second identifying information by transforming the data representing pictures elements into a transform domain.

46. The apparatus of claim 22 , further comprising a speaker configured to render the audio signal.

47. The apparatus of claim 22 , further comprising a display configured to render the video signal.

48. The apparatus of claim 22 , wherein the apparatus comprises a server.

49. A system comprising:

means for buffering data representing audio elements of an audio signal;

means for buffering data representing picture elements of a video signal;

means for extracting first identifying information from data representing audio elements of an audio signal;

means for extracting second identifying information from data representing picture elements of a video signal that is associated with the audio signal; and

means for utilizing the first identifying information or the second identifying information in a synchronization process, wherein the synchronization process controls content synchronization during rendering of the audio signal or the video signal.

50. The system of claim 49 , wherein the means for extracting first identifying information comprises means for decoding steganographically hidden information from the data representing audio elements of the audio signal.

51. The system of claim 49 , wherein the means for extracting second identifying information comprises means for decoding steganographically hidden information from the data representing picture elements of the video signal that is associated with the audio signal.

52. The system of claim 49 , wherein the content comprises metadata associated with the audio signal or the video signal.

53. The system of claim 52 , wherein the metadata includes a URL link.

54. The system of claim 49 , wherein the content comprises audio or video.

55. The system of claim 49 , wherein the content comprises ownership information.

56. The system of claim 49 , wherein the content comprises purchase information.

57. The system of claim 49 , further comprising means for utilizing the first identifying information or the second identifying information to track rendering of the audio signal or the video signal.

58. The system of claim 49 , further comprising means for rendering the audio signal.

59. The system of claim 49 , further comprising a means for rendering the video signal.

Assignments (5)
MERGER Recorded Nov 2, 2010
From: DMRC LLC
To: DMRC CORPORATION
Reel/Frame 025227/0808 →
MERGER Recorded Nov 2, 2010
From: DMRC CORPORATION
To: DIGIMARC CORPORATION
Reel/Frame 025227/0832 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 29, 2010
From: DIGIMARC CORPORATION (A DELAWARE CORPORATION)
To: DMRC LLC
Reel/Frame 025217/0508 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 25, 2010
From: LEVY, KENNETH L.; HANNIGAN, BRETT T.; BRADLEY, BRETT A.; RHOADS, GEOFFREY B.
To: DIGIMARC CORPORATION
Reel/Frame 025187/0404 →
MERGER Recorded May 12, 2010
From: DIGIMARC CORPORATION (A DELAWARE CORPORATION)
To: DIGIMARC CORPORATION (AN OREGON CORPORATION)
Reel/Frame 024369/0582 →