IP Library Granted Patent US 7,277,766
Granted Patent B1
US 7,277,766 · App. 09/695,457 · Granted Oct 2, 2007

Method and system for analyzing digital audio files

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,277,766
App. No.
09/695,457
Granted
Oct 2, 2007
Kind
B1
Abstract

A method and system for analyzing audio files is provided. Plural audio file feature vector values based on an audio file's content are determined and the audio file feature vectors are stored in a database that also stores other pre-computed audio file features. The process determines if the audio files feature vectors match the stored audio file vectors. The process also associates a plurality of known attributes to the audio file.

Claims (22)

1. An apparatus for fingerprinting an audio waveform, comprising:

a memory that stores a codebook comprising one or more multivariate vectors, each multivariate vector comprising one or more spectral features and represents a corresponding one of a plurality of codes; and

a processor configured to divide the audio waveform into a plurality of bins, compute one or more spectral features for each bin, and generate a fingerprint representing the audio waveform with a string of codes from the codebook based on the computed one or more spectral features for each bin.

2. The apparatus of claim 1 , wherein the string of codes are aligned in a time series.

3. The apparatus of claim 1 wherein the processor compresses the string of codes to form a compressed string of codes.

4. The apparatus of claim 1 wherein each code in the string of codes is temporally aligned with the audio waveform such that the position of a code within the string corresponds to a time period of the audio waveform.

5. The apparatus of claim 1 , wherein the processor compresses the string of codes such that temporal alignment between the string of codes and the audio waveform is maintained.

6. The apparatus of claim 1 wherein each code is a hash code.

7. The apparatus of claim 1 , wherein the fingerprint is a unique identifier.

8. The apparatus of claim 1 , wherein the fingerprint is a unique audio signature.

9. The apparatus of claim 1 , wherein the processor computes the one or more spectral features for a first group of data points within each bin, shifts one or more data points within each bin, and computes the one or more spectral features for a second group of data points within each bin.

10. The apparatus of claim 1 , wherein the codebook is predefined prior to representing the audio waveform with the string of codes from the codebook.

11. The apparatus of claim 1 , wherein the processor generates the fingerprint at a given time within the audio waveform.

12. The apparatus of claim 1 , wherein the processor queries a database of audio files using the fingerprint.

13. The apparatus of claim 1 , wherein the processor queries a database of pre-computed audio signatures.

14. The apparatus of claim 13 , wherein the pre-computed audio signatures correspond respectively to a plurality of audio files.

15. The apparatus of claim 9 , wherein the processor receives at least one attribute associated with the fingerprint.

16. The apparatus of claim 9 , wherein the processor receives at least one attribute, wherein the at least one attribute is tagged to an audio file.

17. The apparatus of claim 9 , wherein the processor receives a customized playlist based on the fingerprint.

18. The apparatus of claim 17 , wherein the customized playlist is further based on a predetermined set of user preferences.

19. The apparatus of claim 9 , wherein the processor receives at least one feature value based on the fingerprint.

20. The apparatus of claim 19 , wherein the at least one feature value is at least one of an emotional quality vector value, a vocal vector value, a sound quality vector values, a situational quality vector value, an ensemble vector value, a genre vector value, and an instrument vector value.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 24, 2009
From: MOODLOGIC, INC.
To: ROVI TECHNOLOGIES CORPORATION
Reel/Frame 023273/0821 →
CHANGE OF NAME Recorded Mar 12, 2001
From: EMOTIONEERING, INC.
To: MOODLOGIC, INC.
Reel/Frame 011598/0280 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 24, 2000
From: KHAN, REHAN M.; TZANETAKIS, GEORGE; MATHYS, MARC M.; PIRKNER, CHRISTIAN D.; SULZER, THOMAS R.
To: MOODLOGIC, INC. A DELAWARE CORPORATION
Reel/Frame 011285/0016 →