IP Library Granted Patent US 11,037,258
Granted Patent B2
US 11,037,258 · App. 16/288,857 · Granted Jun 15, 2021

Media content processing techniques using fingerprinting and heuristics

Inventors: Vadim Brenner (San Francisco, CA); Stephen White (San Francisco, CA); Ryan Wigley (Brooklyn, NY); Mijat Nenezic (Miami Beach, FL); Jay Juilin Hung (San Francisco, CA)
G06Q50/184G06F40/30G10L19/018G10L25/51G10L15/10G10L2015/025
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,037,258
App. No.
16/288,857
Granted
Jun 15, 2021
Kind
B2
Abstract

Systems and methods in accordance with various embodiments of the present disclosure provide improved techniques to process and identify segments of media content and associated intellectual property rights associated with the media content. Intellectual property rights associated with media content may include copyright, trademarks, licenses to composition, synchronization, performance, recordings, etc. In particular, various embodiments provide improved techniques to identify segments of media content using fingerprinting and/or other heuristic rules to identify the original source of the segments of media content, the rights holders, and/or rights associated with the segments of media content.

Claims (52)

1. A computer-implemented method for processing media content, comprising:

receiving a request with media content from a plurality of sources, the media content including at least audio information;

scanning the media request;

performing an audio fingerprinting of the media content to generate a set of audio fingerprints for the media content;

identifying a plurality of segments from the media content based at least in part on the set of audio fingerprints corresponding to each segment matching data in a media content database;

grading the plurality of segments;

merging the plurality of segments;

purging redundant segments from the plurality of segments to generate a remainder of segments;

merging the remainder of the segments;

applying a phonetic fingerprinting on the remainder of the segments to generate a set of phonetic fingerprints for each segment in the remainder of the segments;

applying a Levenshtein distance algorithm on the set of phonetic fingerprints compared to data in the media content database to update the remainder of the segments as having a Levenshtein distance that satisfies a threshold distance;

applying a N-gram fingerprinting on the remainder of the segments as having the Levenshtein distance satisfying the threshold distance to generate a set of N-gram fingerprints for each segment in the remainder of the segments;

merging the set of N-gram fingerprints for each segment in the remainder of the segments;

performing a textual lookup of each segment in the remainder of the segments to data in the media content database; and

generating a list of asset information associated with each segment in the remainder of the segments based at least in part on the textual lookup.

2. The computer-implemented method of claim 1 , wherein the list of asset information includes one or more rights holders associated with the segment, tracklist data associated with the segment, a publisher or label associated with the segment, licensing rights associated with the segment, or royalty payments associated with the segment.

3. The computer-implemented method of claim 1 , further comprising:

dividing the media content into segments based at least in part on a length of each segment.

4. The computer-implemented method of claim 1 , further comprising:

determining the list of asset information associated with each segment complies with one or more business rules.

5. The computer-implemented method of claim 4 , wherein the one or more business rules includes usage rules, distribution rule, territory rules, or program rules.

6. The computer-implemented method of claim 1 , wherein the purging redundant segments from the plurality of segments is based at least in part on purging segments from the plurality of segments that are below a predetermined duration.

7. The computer-implemented method of claim 1 , wherein the applying the phonetic fingerprinting on the remainder of the segments comprises performing a plurality of text string normalizing steps in a predetermined order.

8. The computer-implemented method of claim 1 , wherein the media content database includes normalized publisher supplied music related Intellectual Property right information, third-party information related to music, and editorial search information related to music.

9. A computer-implemented method for processing media content, comprising:

receiving media content from a plurality of sources, the media content including audio information;

performing an audio fingerprinting of the media content to generate a set of audio fingerprints for the media content;

identifying, based at least in part on the set of audio fingerprints for the media content corresponding to matching data in a media content database, a plurality of segments from the media content;

applying a phonetic fingerprinting on a set of the plurality of segments to generate a set of phonetic fingerprints for each segment of the set of the plurality of segments;

processing, using the set of phonetic fingerprints, the set of the plurality of segments through at least one textual fingerprinting technique to determine a subset of segments matching a set of predetermined criteria; and

generating a list of asset information associated with each segment in the subset of the segments based at least in part on a textual lookup of each segment in the subset of segments in the media content database.

10. The computer-implemented method of claim 9 , wherein the list of asset information includes one or more rights holders associated with the segment, tracklist data associated with the segment, a publisher or label associated with the segment, licensing rights associated with the segment, or royalty payments associated with the segment.

11. The computer-implemented method of claim 9 , further comprising:

dividing the media content into segments based at least in part on a length of each segment.

12. The computer-implemented method of claim 9 , further comprising:

determining the list of asset information associated with each segment complies with one or more business rules.

13. The computer-implemented method of claim 12 , wherein the one or more business rules includes usage rules, distribution rule, territory rules, or program rules.

14. The computer-implemented method of claim 9 , wherein the set of predetermined criteria is based at least in part on a Levenshtein distance that satisfies a threshold distance.

15. The computer-implemented method of claim 9 , wherein the processing the plurality of segments through at least one textual fingerprinting technique includes performing a plurality of text string normalizing steps in a predetermined order.

16. The computer-implemented method of claim 9 , wherein the media content database includes normalized publisher supplied music related Intellectual Property right information, third-party information related to music, and editorial search information related to music.

17. A system, comprising:

at least one processor; and

memory storing instructions that, when executed by the at least one processor, cause the system to:

receive media content from a plurality of sources, the media content including audio information;

perform an audio fingerprinting of the media content to generate a set of audio fingerprints for the media content;

identify, based at least in part on the set of audio fingerprints for the media content corresponding to matching data in a media content database, a plurality of segments from the media content;

apply a phonetic fingerprinting on a set of the plurality of segments to generate a set of phonetic fingerprints for each segment of the set of the plurality of segments;

process, using the set of phonetic fingerprints, the set of the plurality of segments through at least one textual fingerprinting technique to determine a subset of segments matching a set of predetermined criteria; and

generate a list of asset information associated with each segment in the subset of the segments based at least in part on a textual lookup of each segment in the subset of segments in the media content database.

18. The system of claim 17 , wherein the set of predetermined criteria is based at least in part on a Levenshtein distance that satisfies a threshold distance.

19. The system of claim 17 , wherein the instructions that, when executed by the at least one processor, further cause the system to perform a plurality of text string normalizing steps in a predetermined order.

20. The system of claim 17 , wherein the media content database includes normalized publisher supplied music related Intellectual Property right information, third-party information related to music, and editorial search information related to music.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 28, 2025
From: PEXESO, INC.
To: VOBILE ACQUISITION, INC.
Reel/Frame 070957/0488 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 30, 2025
From: DUBSET MEDIA HOLDINGS, INC.
To: PEXESO, INC
Reel/Frame 070063/0728 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 28, 2019
From: BRENNER, VADIM; WHITE, STEPHEN; WIGLEY, RYAN; NENEZIC, MIJAT; HUNG, JAY JUILIN
To: DUBSET MEDIA HOLDIGS, INC.
Reel/Frame 048469/0157 →
Continuity (2)
Provisional Application 62637570 · Mar 2, 2018
Related Publication 20190272834A1 · Sep 5, 2019
Cited By (2)
US 12,314,263 US 12,675,484