IP Library Granted Patent US 12,354,404
Granted Patent B2
US 12,354,404 · App. 18/755,109 · Granted Jul 8, 2025

Obtaining artist imagery from video content using facial recognition

Inventors: Jeffrey Scott (Oakland, CA); Aneesh Vartakavi (Emeryville, CA)
Assignee: GRACENOTE, INC.
G06V40/173G06F16/784G06T7/75G06V40/171
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,354,404
App. No.
18/755,109
Granted
Jul 8, 2025
Kind
B2
Abstract

An example method may include receiving, at a computing device, a digital image associated with a particular media content program, the digital image containing one or more faces of particular people associated with the particular media content program. A computer-implemented automated face recognition program may be applied to the digital image to recognize, based on at least one feature vector from a prior-determined set of feature vectors, one or more of the particular people in the digital image, together with respective geometric coordinates for each of the one or more detected faces. At least a subset of the prior-determined set of feature vectors may be associated with a respective one of the particular people. The digital image together may be stored in non-transitory computer-readable memory, together with information assigning respective identities of the recognized particular people, and associating with each respective assigned identity geometric coordinates in the digital image.

Claims (55)

1. A tangible, non-transitory computer readable medium comprising instructions that, when executed, cause at least one processor to perform a set of operations comprising:

receiving a digital image associated with a particular media content program, the digital image containing one or more faces of particular people associated with the particular media content program;

applying a face recognition program together with a set of computational models to the digital image to recognize one or more of the particular people in the digital image from among one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, wherein each of at least a subset of the set of the computational models is further associated with a respective one of the particular people; and

storing, in non-transitory computer-readable memory, the digital image together with information (i) assigning respective identities of the recognized one or more of the particular people in the digital image, and (ii) associating with each respective assigned identity geometric coordinates in the digital image of a face to which the identity is assigned.

2. The tangible, non-transitory computer readable medium of claim 1 , wherein applying the face recognition program together with the set of computational models to the digital image to recognize the one or more of the particular people in the digital image from among the one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, comprises:

determining a feature vector corresponding to at least one of the one or more faces detected, together with respective geometric coordinates, in the digital image;

applying the face recognition program together with the set of computational models to the feature vector; and

determining that applying the face recognition program together with a particular model of the set of computational models to the feature vector yields a probability that both exceeds a threshold and is greater than probabilities yielded from applying the face recognition program together with any of the other computational models of the set.

3. The tangible, non-transitory computer readable medium of claim 2 , wherein determining the feature vector corresponding to at least one of the one or more faces detected, together with respective geometric coordinates, in the digital image comprises:

applying a face detection program to the digital image to detect a spatial region of the digital image that includes the at least one of the one or more faces detected, together with respective geometric coordinates of the spatial region; and

applying a computer-implemented feature extraction program to the spatial region of the digital image to generate the feature vector.

4. The tangible, non-transitory computer readable medium of claim 1 , wherein at least one of the one or more of the particular people in the digital image is a cast member of the particular media content program.

5. The tangible, non-transitory computer readable medium of claim 1 , wherein the particular media content program is one of: a television program, a movie, a sporting event, or a web-based user-hosted and/or user-generated content program.

6. The tangible, non-transitory computer readable medium of claim 1 , wherein the set of operations further comprises:

receiving a further digital image associated with a further particular media content program, the further digital image containing one or more faces of further particular people associated with the further particular media content program;

applying the computer-implemented face recognition program together with a further set of computational models to the further digital image to recognize one or more of the further particular people in the further digital image from among one or more further faces detected, together with respective geometric coordinates for each of the one or more further detected faces in the further digital image, wherein each of at least a further subset of the further set of the computational models is additionally associated with a respective one of the further particular people; and

storing, in non-transitory computer-readable memory, the further digital image together with information (i) assigning respective identities of the recognized one or more of the further particular people in the further digital image, and (ii) associating with each respective assigned identity geometric coordinates in the further digital image of a further face to which the identity is assigned.

7. The tangible, non-transitory computer readable medium of claim 6 , wherein the further digital image is different from the digital image,

wherein the further particular media content program is one of: different from the particular media content program, or the same as the particular media content program, and wherein at least one of the recognized further particular people is one of:

different from any of the recognized particular people, or the same as one of the recognized particular people.

8. The tangible, non-transitory computer readable medium of claim 1 , wherein each of the computational models of the set has been trained for recognizing a different one a plurality of people associated with the particular media content program, wherein the plurality includes at least the particular people.

9. A computer-implemented method comprising:

receiving, at a computing device, a digital image associated with a particular media content program, the digital image containing one or more faces of particular people associated with the particular media content program;

applying a computer-implemented face recognition program together with a set of computational models to the digital image to recognize one or more of the particular people in the digital image from among one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, wherein each of at least a subset of the set of the computational models is further associated with a respective one of the particular people; and

storing, in non-transitory computer-readable memory, the digital image together with information (i) assigning respective identities of the recognized one or more of the particular people in the digital image, and (ii) associating with each respective assigned identity geometric coordinates in the digital image of a face to which the identity is assigned.

10. The computer-implemented method of claim 9 , wherein applying the face recognition program together with the set of computational models to the digital image to recognize the one or more of the particular people in the digital image from among the one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, comprises:

determining a feature vector corresponding to at least one of the one or more faces detected, together with respective geometric coordinates, in the digital image;

applying the face recognition program together with the set of computational models to the feature vector; and

determining that applying the face recognition program together with a particular model of the set of computational models to the feature vector yields a probability that both exceeds a threshold and is greater than probabilities yielded from applying the face recognition program together with any of the other computational models of the set.

11. The computer-implemented method of claim 10 , wherein determining the feature vector corresponding to at least one of the one or more faces detected, together with respective geometric coordinates, in the digital image comprises:

applying a face detection program to the digital image to detect a spatial region of the digital image that includes the at least one of the one or more faces detected, together with respective geometric coordinates of the spatial region; and

applying a computer-implemented feature extraction program to the spatial region of the digital image to generate the feature vector.

12. The computer-implemented method of claim 9 , wherein at least one of the one or more of the particular people in the digital image is a cast member of the particular media content program.

13. The computer-implemented method of claim 9 , wherein the particular media content program is one of: a television program, a movie, a sporting event, or a web-based user-hosted and/or user-generated content program.

14. The computer-implemented method of claim 9 , further comprising:

receiving a further digital image associated with a further particular media content program, the further digital image containing one or more faces of further particular people associated with the further particular media content program;

applying the computer-implemented face recognition program together with a further set of computational models to the further digital image to recognize one or more of the further particular people in the further digital image from among one or more further faces detected, together with respective geometric coordinates for each of the one or more further detected faces in the further digital image, wherein each of at least a further subset of the further set of the computational models is additionally associated with a respective one of the further particular people; and

storing, in non-transitory computer-readable memory, the further digital image together with information (i) assigning respective identities of the recognized one or more of the further particular people in the further digital image, and (ii) associating with each respective assigned identity geometric coordinates in the further digital image of a further face to which the identity is assigned.

15. The computer-implemented method of claim 14 , wherein the further digital image is different from the digital image,

wherein the further particular media content program is one of: different from the particular media content program, or the same as the particular media content program,

and wherein at least one of the recognized further particular people is one of:

different from any of the recognized particular people, or the same as one of the recognized particular people.

16. The computer-implemented method of claim 9 , wherein each of the computational models of the set has been trained for recognizing a different one a plurality of people associated with the particular media content program, wherein the plurality includes at least the particular people.

17. A computing device comprising:

at least one processors; and

tangible, non-transitory computer readable medium comprising instructions that, when executed, cause the at least one processor to perform a set of operations comprising:

receiving a digital image associated with a particular media content program, the digital image containing one or more faces of particular people associated with the particular media content program;

applying a face recognition program together with a set of computational models to the digital image to recognize one or more of the particular people in the digital image from among one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, wherein each of at least a subset of the set of the computational models is further associated with a respective one of the particular people; and

storing, in non-transitory computer-readable memory, the digital image together with information (i) assigning respective identities of the recognized one or more of the particular people in the digital image, and (ii) associating with each respective assigned identity geometric coordinates in the digital image of a face to which the identity is assigned.

18. The computing device of claim 17 , wherein applying the face recognition program together with the set of computational models to the digital image to recognize the one or more of the particular people in the digital image from among the one or more faces detected, together with respective geometric coordinates for each of the one or more detected faces in the digital image, comprises:

determining a feature vector corresponding to at least one of the one or more faces detected, together with respective geometric coordinates, in the digital image;

applying the face recognition program together with the set of computational models to the feature vector; and

determining that applying the face recognition program together with a particular model of the set of computational models to the feature vector yields a probability that both exceeds a threshold and is greater than probabilities yielded from applying the face recognition program together with any of the other computational models of the set.

19. The computing device of claim 17 , wherein at least one of the one or more of the particular people in the digital image is a cast member of the particular media content program.

20. The computing device of claim 17 , wherein the particular media content program is one of: a television program, a movie, a sporting event, or a web-based user- hosted and/or user-generated content program.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 26, 2024
From: SCOTT, JEFFREY; VARTAKAVI, ANEESH
To: GRACENOTE, INC.
Reel/Frame 067873/0377 →
Continuity (5)
Continuation 18244086 · Sep 8, 2023
Continuation 17340682 · Jun 7, 2021
Continuation 16720200 · Dec 19, 2019
Provisional Application 62906238 · Sep 26, 2019
Related Publication 20240346847A1 · Oct 17, 2024
References Cited (35)
US 5574799A · Bankman · 1996 [cited by applicant]
US 8442384B2 · Bronstein et al. · 2013 [cited by applicant]
US 9271035B2 · Mei et al. · 2016 [cited by applicant]
US 9502073B2 · Boiman et al. · 2016 [cited by applicant]
US 10107594B1 · Li et al. · 2018 [cited by applicant]
US 10657730B2 · Fyke · 2020 [cited by applicant]
US 10762608B2 · Shen et al. · 2020 [cited by applicant]
US 20030139840A1 · Magee et al. · 2003 [cited by applicant]
US 20030198368A1 · Kee · 2003 [cited by applicant]
US 20060064716A1 · Sull et al. · 2006 [cited by applicant]
US 20080273766A1 · Kim et al. · 2008 [cited by applicant]
US 20090209857A1 · Secretain et al. · 2009 [cited by applicant]
US 20120106806A1 · Folta et al. · 2012 [cited by applicant]
US 20120254791A1 · Jackson et al. · 2012 [cited by applicant]
US 20130027532A1 · Hirota et al. · 2013 [cited by applicant]
US 20150128160A1 · Benea · 2015 [cited by examiner]
US 20160379090A1 · Shah · 2016 [cited by applicant]
US 20170154237A1 · Shen et al. · 2017 [cited by applicant]
US 20180005026A1 · Shaburov et al. · 2018 [cited by applicant]
US 20180121762A1 · Han et al. · 2018 [cited by applicant]
US 20190073520A1 · Ayyar et al. · 2019 [cited by applicant]
US 20210374391A1 · Jorasch et al. · 2021 [cited by applicant]
US 20210399911A1 · Jorasch et al. · 2021 [cited by applicant]
US 20210400142A1 · Jorasch et al. · 2021 [cited by applicant]
CN 107750460A · 2018 [cited by applicant]
CN 113010711A · 2021 [cited by applicant]
FR 2680931A1 · 1993 [cited by applicant]
KR 1020170082025 · 2017 [cited by applicant]
KR 1020180079894 · 2018 [cited by applicant]
WO 2016205432A1 · 2016 [cited by applicant]
International Search Report for PCT/US2020/051414 mailed Dec. 30, 2020. [cited by applicant]
International Searching Authority Written Opinion for PCT/US2020/051414 mailed Dec. 30, 2020. [cited by applicant]
Search Report (English Translation), China National Intellectual Property Administration, Application No. CN 202080067828.2, dated Sep. 30, 2022, 2 pages. [cited by applicant]
The First Office Action (English Translation), China National Intellectual Property Administration, Application No. CN 202080067828.2, dated Nov. 15, 2022, 6 pages. [cited by applicant]
Bradley, Nick, “Automated Generation of Intelligent Video Previews on Cloudinary's Dynamic Video Platform”, Apr. 13, 2020, 15 pages. [cited by applicant]