IP Library Granted Patent US 9,626,456
Granted Patent B2
US 9,626,456 · App. 12/901,321 · Granted Apr 18, 2017

Crowd sourcing for file recognition

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,626,456
App. No.
12/901,321
Granted
Apr 18, 2017
Kind
B2
Abstract

Methods for identifying content in encrypted or otherwise protected files utilize crowd sourcing for content identification. One such method includes, using a computer, selecting defined content titles to be presented with identifiers for data files for use in obtaining user selection data. The method may also include receiving the user selection data from multiple independent sources, the user selection data indicating users' selections of single ones of the content titles for respective single ones of the data files. The method may also include determining for ones of the identifiers, using the one or more computers processing the user selection data, respective ones of the content titles satisfying a minimum confidence threshold for association with the ones of the identifiers. An apparatus for performing the method comprises a processor coupled to a memory, the memory holding instructions for performing steps of the method as summarized above.

Claims (27)

1. A method for identifying encrypted media content contained within each file of a library of encrypted data files without decrypting the data files, comprising:

selecting, by one or more computers, content titles from a database of content titles to be presented with identifiers for the data files, including selecting multiple different ones of the content titles for singular ones of the identifiers, wherein each of the identifiers comprises a hash of a set of unique file metadata, and each of the content titles is a character string used to uniquely identify the encrypted media content contained within the each file of the library;

sending the multiple different ones of the content titles with the singular ones of the identifiers to different clients each operated by an independent user for presentation by the different clients with a request that a user identify a correct one of the multiple different ones of the content titles for corresponding ones of the data files;

receiving user selection data from the multiple independent sources in response to the sending, the user selection data indicating users' selections of a user-selected correct one of the content titles for each respective one of the data files responsive to presentations of the multiple different ones of the content titles with the singular ones of the identifiers by the different clients;

determining for each one of the identifiers, using the one or more computers processing the user selection data, a respective one of the content tides satisfying a minimum confidence threshold for association with the each one of the identifiers, based on at least one of a quality or quantity of the multiple independent sources supplying the user selection data for each of the content titles; and

providing the respective one of the content titles satisfying the minimum confidence threshold for recording as associated with the each one of the identifiers in a data structure, the data structure is used for querying using one of the identifiers to provide an associated one of the content tides for use in identifying a data file and for use in providing one of the content titles for the data file in response to determining that the one of the content titles satisfies the minimum confidence threshold and is associated with the one of the identifiers for the data file.

2. The method of claim 1 , further comprising querying the data structure using a content title to identify a data file containing content titled by the content title.

3. The method of claim 1 , further comprising querying the data structure using an identifier to provide an associated one of the content titles for use in identifying a data file.

4. The method of claim 3 , further comprising providing one of the content titles for the data file in response to determining that the one of the content titles satisfies the minimum confidence threshold and is associated with an identifier for the data file.

5. The method of claim 3 , further comprising providing a message indicating that user input is needed to identify the data file, in response to determining that the identifier is not associated with any content title satisfying the minimum confidence threshold.

6. The method of claim 1 , further comprising automatically organizing a directory of the data files in response to the respective one of the content titles associated with the each one of the identifiers for the data files.

7. The method of claim 1 , further comprising processing data files stored in a computer-readable storage medium to automatically generate the identifiers using a hashing algorithm.

8. The method of claim 1 , wherein the one or more computers comprise multiple computer servers operatively coupled to each other.

9. The method of claim 1 , further comprising generating the identifiers by hashing the set of unique metadata for each of the files.

10. An apparatus for identifying, encrypted media content contained within each file of a library of encrypted data files without decrypting the data files comprising a processor coupled to a memory, the memory holding instructions for:

content titles from a database of content titles to be presented with identifiers for the data files, including selecting multiple different ones of the content titles for singular ones of the identifiers, wherein each of the identifiers comprises a hash of a set of unique file metadata, and each of the content titles is a character string used to uniquely identify the encrypted media content contained within the each file of the library;

sending the multiple different ones of the content titles with the singular ones of the identifiers to different clients each operated by an independent user for presentation by the different clients with a request that a user identify a correct one of the multiple different ones of the content titles for corresponding ones of the data files;

receiving user selection data from the multiple independent sources in response to the sending, the user selection data indicating users' selections of a user-selected correct one of the content titles for each respective one of the data files responsive to presentations of the multiple different ones of the content titles with the singular ones of the identifiers by the different clients;

determining for each one of the identifiers, using the one or more computers processing the user selection data, a respective one of the content titles satisfying a minimum confidence threshold for association with the each one of the identifiers, based on at least one of a quality or quantity of the multiple independent sources supplying the user selection data for each of the content titles; and

providing the respective one of the content tides satisfying the minimum confidence threshold for recording as associated with the each one of the identifiers in a data structure, the data structure is used for querying using one of the identifiers to provide an associated one of the content titles for use in identifying a data file and for use in providing one of the content tides for the data file in response to determining that the one of the content titles satisfies the minimum confidence threshold and is associated with the one of the identifiers for the data file.

11. The apparatus of claim 10 , wherein the memory further holds instructions for querying the data structure using a content title to identify a data file containing content titled by the content title.

12. The apparatus of claim 10 , wherein the memory further holds instructions for querying the data structure using an identifier to provide an associated one of the content titles for use in identifying a data file.

13. The apparatus of claim 12 , wherein the memory further holds instructions for providing one of the content titles for the data file in response to determining that the one of the content titles satisfies the minimum confidence threshold and is associated with an identifier for the data file.

14. The apparatus of claim 12 , wherein the memory further holds instructions for providing a message indicating that user input is needed to identify the data file, in response to determining that the identifier is not associated with any content title satisfying the minimum confidence threshold.

15. The apparatus of claim 10 , wherein the memory further holds instructions for automatically organizing a directory of the data files in response to the respective one of the content titles associated with the each one of the identifiers for the data files.

16. The apparatus of claim 10 , wherein the memory further holds instructions for processing data files stored in a computer-readable storage medium to automatically generate the identifiers using a hashing algorithm.

17. The apparatus of claim 10 , wherein the memory further holds instructions for generating the identifiers by hashing the set of unique metadata for each of the files.

Assignments (1)
SECURITY INTEREST Recorded Oct 1, 2025
From: WARNER BROS. DISCOVERY, INC.; WARNER MEDIA, LLC; TURNER BROADCASTING SYSTEM, INC.; HOME BOX OFFICE, INC.; DISCOVERY COMMUNICATIONS, LLC; WARNERMEDIA DIRECT LLC; DISCOVERY.COM LLC; WARNER BROS. ENTERTAINMENT INC.; CNN INTERACTIVE GROUP, INC.
To: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
Reel/Frame 072995/0858 →