IP Library Granted Patent US 10,321,167
Granted Patent B1
US 10,321,167 · App. 15/413,365 · Granted Jun 11, 2019

Method and system for determining media file identifiers and likelihood of media file relationships

Inventors: Aaron Edell (Walnut Creek, CA); Sek Chai (Valencia, CA); Mat Ryer (London, GB)
Assignee: GRAYMETA, INC.
H04N21/23418G06F16/783G06F16/901G06F21/10H04L63/30H04N21/2541H04N21/2743
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,321,167
App. No.
15/413,365
Granted
Jun 11, 2019
Kind
B1
Abstract

A method and system for determining the likelihood or similarity ratio that a selected media file of interest is related to one or more predetermined media files is provided that utilizes, combines, analyzes, and evaluates different categories of data and metadata extracted from each media file to generate a media file identifier for each media file that can then be used as a basis to compare any two media files to each other.

Claims (60)

1. A method for determining a similarity ratio that a selected media file of interest is a derivative work of or is derived from one or more predetermined media files comprising:

providing a selected media file of interest and one or more predetermined media files to be compared with the selected media file of interest,

obtaining media type classifications for the selected media file of interest and for each of the one or more predetermined media files,

extracting data or metadata from the selected media file of interest and the one or more predetermined files, wherein at least two different categories of data or metadata are extracted from the selected media file of interest and the one or more predetermined media files by engaging at least two extraction engines,

harvesting the data or metadata extracted from the selected media file of interest and the one or more predetermined media files,

storing the data or metadata extracted from the selected media file of interest and the one or more predetermined media files,

selecting, based on the media type classifications, two or more ranked categories of data or metadata to be used for generating media file identifiers for the selected media file of interest and for each of the one or more predetermined media files,

generating, based on the selected two or more ranked categories of data or metadata, the media file identifiers for the selected media file of interest and for each of the one or more predetermined media files,

storing the media file identifier generated for the selected media file of interest and the media file identifiers generated for each of the one or more predetermined media files,

comparing the media file identifier generated for the selected media file of interest to the media file identifier generated for each of the one or more predetermined media files, and

determining a similarity ratio that the selected media file of interest is a derivative work of or is derived from each of the one or more predetermined media files based on comparing the media file identifier generated for the selected media file of interest to the media file identifiers generated for each of the one or more predetermined media files, said similarity ratio indicative of whether the selected media file of interest is derived from one or more predetermined media files without regard to the media type classifications of selected media file of interest and the one or more predetermined media files.

2. The method of claim 1 , comprising storing media file identifiers that have been generated for each of the one or more predetermined media files to form a library of media file identifiers and retrieving the stored media file identifiers that have been generated for each of the one or more predetermined media files from the library of media file identifiers to compare the media file identifier generated for the selected media file of interest to the media file identifier generated for each of the one or more predetermined media files.

3. The method of claim 1 , comprising assigning a weight to the one or more harvested subsets of data or metadata, wherein the weight assigned to the one or more harvested subsets of data or metadata is based on the ranking of the categories for the data or metadata, and wherein the assigned weight is used to generate the media file identifiers for the selected media file of interest and for each of the one or more predetermined media files.

4. The method of claim 1 , wherein the at least two data extraction engines are configured to use the same data extraction process to extract different categories of data or metadata from the selected media file of interest and from each of the one or more predetermined media files.

5. The method of claim 1 , wherein the similarity ratios are determined as a percentage.

6. The method of claim 1 , wherein at least one of the one or more predetermined media files is not the same type of media file as the selected media file of interest.

7. A system for determining a similarity ratio that a selected media file of interest is a derivative work of or is derived from one or more predetermined media files comprising:

a data receiving and input device for receiving a selected media file of interest and one or more predetermined media files to be compared with the selected media file of interest;

a data receiving and output device for providing similarity ratios quantifying how similar a selected media file of interest is to each of the one or more predetermined media files; and

a data and metadata harvesting, extraction, analysis, evaluation, and storage system, the data and metadata harvesting, extraction, analysis, evaluation, and storage system further comprising:

at least two data extraction engines configured to extract different categories of data or metadata from the selected media file of interest and from each of the one or more predetermined media files;

a data and metadata harvesting engine configured to manage the at least two data extraction engines and to collect and harvest data or metadata extracted from the at least two data extraction engines,

wherein the harvested data or metadata is stored in a data store as data or metadata subsets within each category of data or metadata extracted from the selected media file of interest and each of the one or more predetermined media files;

an analysis engine configured to:

obtain media type classifications for the selected media file of interest and for each of the one or more predetermined media files;

select, based on the media type classifications, at least one category of data or metadata to be used for generating media file identifiers for the selected media file of interest and for each of the one or more predetermined media files;

generate, based on the selected at least one category of data or metadata, the media file identifiers for the selected media file of interest and for each of the one or more predetermined media files;

compare the generated media file identifier for the selected media file of interest with each of the generated media file identifiers for each of the one or more predetermined media files; and

determine similarity ratios that a selected media file of interest is a derivative work of or is derived from one or more predetermined media files based on the comparison of the media file identifier for the selected media file of interest with each of the generated media file identifiers for each of the one or more predetermined media files, said similarity ratio indicative of whether the selected media file of interest is derived from one or more predetermined media files without regard to the media type classifications of selected media file of interest and the one or more predetermined media files; and

a user interface configured to provide a user access to a set of features and functionality of the system and to enable the user to select and rank one or more harvested subsets of data or metadata.

8. The system of claim 7 , comprising a data store configured to store:

the data and metadata extracted from the selected media file of interest and each of the one or more predetermined media files;

the media file identifiers generated for the selected media file of interest and each of the one or more predetermined media files; and

the similarity ratios determined for each comparison made between the media file identifier of the selected media file of interest and the media file identifiers of each of the one or more predetermined media files.

9. The system of claim 8 , wherein the data store is configured to store media file identifiers that have been generated for each of the one or more predetermined media files to form a library of media file identifiers for a set of predetermined media files prior to receiving a selected media file of interest.

10. The system of claim 7 , wherein the one or more harvested subsets of data or metadata are ranked according to their importance by the user, wherein each ranking is used to assign a weight to the one or more harvested subsets of data or metadata, and wherein the weight assigned to the one or more harvested subsets of data or metadata is used in generating media file identifiers for the selected media file of interest and for each of the one or more predetermined media files.

11. The system of claim 7 , wherein the at least two data extraction engines are configured to use the same data extraction process to extract different categories of data or metadata from the selected media file of interest and from each of the one or more predetermined media files.

12. The system of claim 7 , wherein the selected media file of interest has been selected by an automated selection system.

13. The system of claim 7 , wherein the data and metadata harvesting engine comprises an extraction engine manager configured to engage a first extraction engine and one or more other extraction engines in parallel.

14. The system of claim 7 , wherein the data and metadata harvesting engine comprises an extraction engine manager configured to engage a first extraction engine and one or more other extraction engines in sequence.

15. The system of claim 7 , wherein the similarity ratios are determined as a percentage.

16. The system of claim 7 , wherein at least one of the one or more predetermined media files is not the same type of media file as the selected media file of interest.

17. Non-transitory computer-readable storage media encoded with a computer program including instructions executable by a processor for determining a similarity ratio that a selected media file of interest is a derivative work or is derived from one or more predetermined media files, the media comprising:

a database, recorded on the media, comprising different types of data and metadata extracted from each of the selected media file of interest and the one or more predetermined media files;

an evaluation software module comprising instructions for:

obtaining media type classifications for the selected media file of interest and for each of the one or more predetermined media files,

extracting data or metadata from the selected media file of interest and the one or more predetermined files, wherein at least two different categories of data or metadata are extracted from the selected media file of interest and the one or more predetermined media files by engaging at least two extraction engines,

harvesting the data or metadata extracted from the selected media file of interest and the one or more predetermined media files,

storing the data or metadata extracted from the selected media file of interest and the one or more predetermined media files,

selecting, based on the media type classifications, two or more ranked categories of data or metadata to be used for generating media file identifiers for the selected media file of interest and for each of the one or more predetermined media files,

generating, based on the selected two or more ranked categories of data or metadata, the media file identifiers for the selected media file of interest and for each of the one or more predetermined media files,

storing the media file identifiers generated for the selected media file of interest and for each of the one or more predetermined media files,

comparing the media file identifier generated for the selected media file of interest to each of the media file identifier generated for each of the one or more predetermined media files, and

determining a similarity ratio that the selected media file of interest is a derivative work of or is derived from each of the one or more predetermined media files based on comparing the media file identifier generated for the selected media file of interest to each of the media file identifiers generated for each of the one or more predetermined media files, said similarity ratio indicative of whether the selected media file of interest is derived from one or more predetermined media files without regard to the media type classifications of selected media file of interest and the one or more predetermined media files.

18. The media of claim 17 , wherein the evaluation software module comprises instructions for:

storing the data or metadata extracted from the selected media file of interest and the one or more predetermined media files in the database;

storing the media file identifiers generated for the selected media file of interest and for each of the one or more predetermined media files in the database; and

storing the similarity ratio in the database.

19. The media of claim 17 , wherein the one or more harvested subsets of data or metadata are ranked according to their importance, wherein each ranking is used to assign a weight to the one or more harvested subsets of data or metadata, and wherein the weight assigned to the one or more harvested subsets of data or metadata is used in generating media file identifiers for the selected media file of interest and for each of the one or more predetermined media files.

20. The media of claim 17 , wherein at least one of the one or more predetermined media files is not the same type of media file as the selected media file of interest.

Assignments (4)
SECURITY INTEREST Recorded Apr 8, 2026
From: WASABI TECHNOLOGIES LLC
To: BAIN CAPITAL CREDIT, LP AS AGENT FOR THE LENDERS
Reel/Frame 074307/0425 →
CHANGE OF NAME Recorded Aug 22, 2025
From: WASABI TECHNOLOGIES, INC.
To: WASABI TECHNOLOGIES LLC
Reel/Frame 072532/0231 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 5, 2024
From: GRAYMETA, INC.
To: WASABI TECHNOLOGIES LLC
Reel/Frame 066652/0457 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 4, 2017
From: EDELL, AARON; CHAI, SEK; RYER, MAT
To: GRAYMETA, INC.
Reel/Frame 042399/0719 →
Continuity (1)
Provisional Application 62281711 · Jan 21, 2016
Cited By (1)
US 12,468,505