IP Library Granted Patent US 11,200,299
Granted Patent B2
US 11,200,299 · App. 15/489,559 · Granted Dec 14, 2021

Crowd sourcing for file recognition

Inventor: Kevin Michael Kozan (Seattle, WA)
Assignee: WARNER BROS. ENTERTAINMENT INC.
G06F21/10G06F16/24578G06F16/907G06F16/951G06F21/6209G06F21/6218G11B20/0021G11B20/00086H04L63/0428G06F16/24558G06F2221/2107
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,200,299
App. No.
15/489,559
Granted
Dec 14, 2021
Kind
B2
Abstract

Methods for identifying encrypted content in ones of a plurality of encrypted data files in a library of encrypted files without decrypting the data files utilize crowd sourcing for content identification. A method includes selecting, by a computer, content titles for presenting with ones of identifiers for the data files in a data structure. Each of the identifiers includes a hash of metadata for one of the data files and the content titles include a character string that identifies each file's content. The user selection data identifies the content titles that correspond to the data files. The computer determines which content titles satisfy a minimum confidence threshold for associating with one of the identifiers, based on a quality or quantity of the multiple independent clients supplying the user selection data. An apparatus for performing the method includes a memory holding instructions for performing steps of the method as summarized above.

Claims (29)

1. A method for identifying encrypted content in ones of a plurality of encrypted data files in a library of encrypted data files without decrypting the encrypted data files, the method comprising:

selecting, by one or more computers, multiple different ones of content titles from a plurality of content titles for presenting with singular ones of identifiers for the encrypted data files in a data structure, wherein each of the identifiers includes or is associated with a hash of metadata for a corresponding one of the encrypted data files, each of the content titles comprises a human-recognizable character string that identifies media content encrypted in the each file of the library, and ones of the encrypted data files are associated with plural ones of the multiple different content titles;

receiving user selection data from multiple independent clients indicating users' selections of single ones of the content titles for corresponding ones of the encrypted data files;

determining for ones of the identifiers, using the one or more computers processing the user selection data, respective ones of the content titles satisfying a minimum confidence threshold for associating as a most correct one of the multiple different content titles with the ones of the identifiers, based on at least one of a quality or quantity of the multiple independent clients supplying the user selection data;

recording the content titles satisfying a minimum confidence threshold and associated identifiers for the plurality of encrypted data files in the data structure; and

providing content from at least one of the encrypted data files to a client device, based at least in part on an associated one of the content titles satisfying the minimum confidence threshold for the at least one of the encrypted data files.

2. The method of claim 1 , further comprising providing the respective ones of the content titles satisfying the minimum confidence threshold for recording in the data structure associated with the respective ones of the identifiers in a data structure.

3. The method of claim 2 , further comprising querying the data structure using a content title to identify the at least one of the encrypted data files containing content titled by the content title.

4. The method of claim 2 , further comprising querying the data structure using an identifier to provide an associated one of the content titles for use in identifying the at least one of the encrypted data files.

5. The method of claim 4 , further comprising providing a content title for the at least one of the encrypted data files based on the content title satisfying the minimum confidence threshold associated with an identifier for the data file.

6. The method of claim 4 , further comprising providing a message indicating that user input is needed to identify the at least one of the encrypted data files, based on determining that the identifier is not associated with any content title satisfying the minimum confidence threshold.

7. The method of claim 2 , further comprising automatically organizing a directory of the encrypted data files based on the respective ones of the content titles being associated with the respective ones of the identifiers for the encrypted data files.

8. The method of claim 1 , further comprising processing the encrypted data files stored in a computer-readable storage medium to automatically generate the identifiers using a hashing algorithm.

9. The method of claim 1 , wherein the one or more computers comprise multiple computer servers operatively coupled to each other.

10. The method of claim 1 , further comprising generating the identifiers using a one-way hashing algorithm operating on respective ones of the encrypted data files.

11. An apparatus comprising a processor coupled to a memory, the memory holding instructions for identifying encrypted content in ones of a plurality of encrypted data files in a library of encrypted data files without decrypting the encrypted data files, at least in part by:

selecting multiple different ones of content titles from a plurality of content titles for presenting with singular ones of identifiers for the encrypted data files in a data structure, wherein each of the identifiers includes or is associated with a hash of metadata for a corresponding one of the encrypted data files, each of the content titles comprises a human-recognizable character string that identifies media content encrypted in the each file of the library, and ones of the encrypted data files are associated with plural ones of the multiple different content titles;

receiving user selection data from multiple independent clients indicating users' selections of single ones of the content titles for corresponding ones of the encrypted data files;

determining for ones of the identifiers, using the one or more computers processing the user selection data, respective ones of the content titles satisfying a minimum confidence threshold for associating as a most correct one of the multiple different content titles with the ones of the identifiers, based on at least one of a quality or quantity of the multiple independent clients supplying the user selection data;

recording the content titles satisfying a minimum confidence threshold and associated identifiers for the plurality of encrypted data files in the data structure; and

providing content from at least one of the encrypted data file to a client device, based at least in part on an associated one of the content titles satisfying the minimum confidence threshold for the at least one of the encrypted data files.

12. The apparatus of claim 11 , wherein the memory further holds instructions for providing the respective ones of the content titles satisfying the minimum confidence threshold for recording in the data structure associated with the respective ones of the identifiers in a data structure.

13. The apparatus of claim 12 , wherein the memory further holds instructions for querying the data structure using a content title to identify the at least one of the encrypted data files containing content titled by the content title.

14. The apparatus of claim 12 , wherein the memory further holds instructions for querying the data structure using an identifier to provide an associated one of the content titles for use in identifying at least one of the encrypted data files.

15. The apparatus of claim 14 , wherein the memory further holds instructions for providing a content title for the at least one of the encrypted data files based on the content title satisfying the minimum confidence threshold associated with an identifier for the at least one of the encrypted data files.

16. The apparatus of claim 14 , wherein the memory further holds instructions for providing a message indicating that user input is needed to identify the at least one of the encrypted data files, based on determining that the identifier is not associated with any content title satisfying the minimum confidence threshold.

17. The apparatus of claim 12 , wherein the memory further holds instructions for automatically organizing a directory of the encrypted data files based on the respective ones of the content titles being associated with the respective ones of the identifiers for the encrypted data files.

18. The apparatus of claim 11 , wherein the memory further holds instructions for processing encrypted data files stored in a computer-readable storage medium to automatically generate the identifiers using a hashing algorithm.

19. The apparatus of claim 11 , wherein the memory further holds instructions for generating the identifiers using a one-way hashing algorithm operating on respective ones of the encrypted data files.

Assignments (2)
SECURITY INTEREST Recorded Oct 1, 2025
From: WARNER BROS. DISCOVERY, INC.; WARNER MEDIA, LLC; TURNER BROADCASTING SYSTEM, INC.; HOME BOX OFFICE, INC.; DISCOVERY COMMUNICATIONS, LLC; WARNERMEDIA DIRECT LLC; DISCOVERY.COM LLC; WARNER BROS. ENTERTAINMENT INC.; CNN INTERACTIVE GROUP, INC.
To: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
Reel/Frame 072995/0858 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 1, 2021
From: KOZEN, KEVIN MICHAEL
To: WARNER BROS. ENTERTAINMENT INC.
Reel/Frame 058260/0095 →
Continuity (2)
Continuation 12901321 · Oct 8, 2010
Related Publication 20170220776A1 · Aug 3, 2017