IP Library Granted Patent US 12,474,588
Granted Patent B2
US 12,474,588 · App. 18/981,592 · Granted Nov 18, 2025

Face model capture by a wearable device

Inventors: Gholamreza Amayeh (Santa Clara, CA); Adrian Kaehler (Los Angeles, CA); Douglas Lee (Redwood City, CA)
Assignee: Magic Leap, Inc.
G02B27/0172G02B27/0093G02B27/0103G06V20/64G06V40/169G02B2027/0105G02B2027/0138G02B2027/0174G02B2027/0178G02B2027/0187G06V40/19G06V2201/12
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,474,588
App. No.
18/981,592
Granted
Nov 18, 2025
Kind
B2
Abstract

Systems and methods for generating a face model for a user of a head-mounted device are disclosed. The head-mounted device can include one or more eye cameras configured to image the face of the user while the user is putting the device on or taking the device off. The images obtained by the eye cameras may be analyzed using a stereoscopic vision technique, a monocular vision technique, or a combination, to generate a face model for the user. The face model can be used to generate a virtual image of at least a portion of the user's face, for example to be presented as an avatar.

Claims (39)

1 . A display system comprising:

a wearable display configured to present a three-dimensional (3D) environment to a user of the wearable display;

an imaging system coupled to the wearable display, the imaging system including at least one inward-facing camera configured to capture images of at least a portion of a face of the user; and

at least one processor communicatively coupled to the wearable display and the imaging system, the at least one processor configured to execute software to perform operations comprising:

analyzing the images captured by the at least one inward-facing camera to determine that at least a threshold number of consecutively captured images have substantially the same content;

determining a stopping trigger based on determining that at least the threshold number of consecutively captured images have substantially the same content;

generating a face model of at least the portion of the face of the user, based at least partly on further analyzing the images captured during a time period bounded by the stopping trigger; and

using the face model to generate, for presentation, a virtual image of at least the portion of the face.

2 . The display system of claim 1 , wherein the images used to generate the face model are captured while the at least one inward-facing camera is at a distance from the face that is greater than the distance between the at least one inward-facing camera and the face when the wearable display is being worn.

3 . The display system of claim 1 , wherein the display system further comprises at least one sensor, and wherein the time period is further bounded by a starting trigger that is determined based at least partly on data received from the at least one sensor.

4 . The display system of claim 3 , wherein the at least one sensor includes an inertial measurement unit, wherein the data describes an acceleration of the wearable display, and wherein the starting trigger is determined based at least partly on the acceleration exceeding a threshold acceleration.

5 . The display system of claim 1 , wherein analyzing the images to generate the face model comprises converting the one or more images to point clouds using a stereo vision algorithm.

6 . The display system of claim 5 , wherein analyzing the images to generate the face model further comprises combining at least two of the point clouds using an iterative closest point algorithm.

7 . The display system of claim 5 , wherein the stereo vision algorithm comprises at least one of a block-matching algorithm, a semi-global matching algorithm, a semi-global block-matching algorithm, or a neural network algorithm.

8 . The display system of claim 1 , wherein generating the face model comprises:

accessing a pre-existing face model; and

updating the pre-existing face model based at least partly on the analyzing of the images.

9 . The display system of claim 8 , wherein the pre-existing face model comprises at least one of a generic face model or a previously generated face model of the user.

10 . The display system of claim 1 , wherein generating the face model further comprises:

accessing one or more images of the portion of the face previously acquired by at least one of the imaging system or another imaging device;

wherein generating the face model is further based on analyzing the one or more accessed images.

11 . The display system of claim 1 , wherein determining the stopping trigger is further based on determining that at least the threshold number of consecutively captured images are captured within a threshold duration of time and have substantially the same content.

12 . The display system of claim 1 , wherein determining the stopping trigger is further based on determining that at least the threshold number of consecutively captured images each include at least one same feature of the face.

13 . The display system of claim 12 , wherein the at least one same feature includes at least one eye of the user.

14 . A computer-implemented method for generating a virtual image of at least a portion of a face of a user of a wearable display, the method performed by at least one processor executing instructions stored on non-transitory computer storage media, the method comprising:

analyzing images captured by at least one inward-facing camera coupled to the wearable display to determine that at least a threshold number of consecutively captured images have substantially the same content;

determining a stopping trigger based on determining that at least the threshold number of consecutively captured images have substantially the same content;

generating a face model of at least the portion of the face of the user, based at least partly on further analyzing the images captured during a time period bounded by the stopping trigger; and

using the face model to generate, for presentation, the virtual image of at least the portion of the face.

15 . The computer-implemented method of claim 14 , wherein the images used to generate the face model are captured while the at least one inward-facing camera is at a distance from the face that is greater than the distance between the at least one inward-facing camera and the face when the wearable display is being worn.

16 . The computer-implemented method of claim 14 , wherein the time period is further bounded by a starting trigger that is determined based at least partly on data received from the at least one sensor.

17 . The computer-implemented method of claim 16 , wherein the at least one sensor includes an inertial measurement unit, wherein the data describes an acceleration of the wearable display, and wherein the starting trigger is determined based at least partly on the acceleration exceeding a threshold acceleration.

18 . The computer-implemented method of claim 14 , wherein determining the stopping trigger is further based on determining that at least the threshold number of consecutively captured images each include at least one same feature of the face.

19 . The computer-implemented method of claim 18 , wherein the at least one same feature includes at least one eye of the user.

20 . One or more non-transitory computer-readable storage media storing computer instructions which, when executed by at least one processor, instruct the processor to perform operations for generating a virtual image of at least a portion of a face of a user of a wearable display, the operations comprising:

analyzing images captured by at least one inward-facing camera coupled to the wearable display to determine that at least a threshold number of consecutively captured images have substantially the same content;

determining a stopping trigger based on determining that at least the threshold number of consecutively captured images have substantially the same content;

generating a face model of at least the portion of the face of the user, based at least partly on further analyzing the images captured during a time period bounded by the stopping trigger; and

using the face model to generate, for presentation, the virtual image of at least the portion of the face.

Assignments (4)
SECURITY INTEREST Recorded Oct 29, 2025
From: MAGIC LEAP, INC.; MENTOR ACQUISITION ONE, LLC; MOLECULAR IMPRINTS, INC.
To: CITIBANK, N.A., AS COLLATERAL AGENT
Reel/Frame 073438/0463 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 16, 2024
From: LEE, DOUGLAS
To: MAGIC LEAP, INC.
Reel/Frame 069595/0600 →
PROPRIETARY INFORMATION AND INVENTIONS AGREEMENT Recorded Dec 16, 2024
From: AMAYEH, GHOLAMREZA
To: MAGIC LEAP, INC.
Reel/Frame 069712/0687 →
PROPRIETARY INFORMATION AND INVENTIONS AGREEMENT Recorded Dec 16, 2024
From: KAEHLER, ADRIAN
To: MAGIC LEAP, INC.
Reel/Frame 069713/0681 →
Continuity (6)
Continuation 18345396 · Jun 30, 2023
Continuation 17872443 · Jul 25, 2022
Continuation 17196394 · Mar 9, 2021
Continuation 15717223 · Sep 27, 2017
Provisional Application 62400907 · Sep 28, 2016
Related Publication 20250116868A1 · Apr 10, 2025
References Cited (107)
US 6556196B1 · Blanz et al. · 2003 [cited by applicant]
US 6850221B1 · Tickle · 2005 [cited by applicant]
US D514570S · Ohta · 2006 [cited by applicant]
US 7481531B2 · Howell et al. · 2009 [cited by applicant]
US 8223024B1 · Petrou · 2012 [cited by applicant]
US 8248458B2 · Schowengerdt et al. · 2012 [cited by applicant]
US 9191658B2 · Kato et al. · 2015 [cited by applicant]
US 9264803B1 · Johnson et al. · 2016 [cited by applicant]
US D752529S · Loretan et al. · 2016 [cited by applicant]
US 9317126B2 · Fujimaki · 2016 [cited by examiner]
US D759657S · Kujawski et al. · 2016 [cited by applicant]
US 9971937B1 · Ovsiannikov et al. · 2018 [cited by applicant]
US 9996150B2 · Swaminathan et al. · 2018 [cited by applicant]
US 10976549B2 · Amayeh et al. · 2021 [cited by applicant]
US 11428941B2 · Amayeh et al. · 2022 [cited by applicant]
US 11740474B2 · Amayeh et al. · 2023 [cited by applicant]
US 20060028436A1 · Armstrong · 2006 [cited by applicant]
US 20070081123A1 · Lewis · 2007 [cited by applicant]
US 20070120986A1 · Nunomaki · 2007 [cited by applicant]
US 20110150340A1 · Gotoh · 2011 [cited by examiner]
US 20120127062A1 · Bar-Zeev et al. · 2012 [cited by applicant]
US 20120127284A1 · Bar-Zeev et al. · 2012 [cited by applicant]
US 20120146894A1 · Yang et al. · 2012 [cited by applicant]
US 20120162549A1 · Gao et al. · 2012 [cited by applicant]
US 20130082922A1 · Miller · 2013 [cited by applicant]
US 20130083003A1 · Perez et al. · 2013 [cited by applicant]
US 20130117377A1 · Miller · 2013 [cited by applicant]
US 20130125027A1 · Abovitz · 2013 [cited by applicant]
US 20130154906A1 · Braun et al. · 2013 [cited by applicant]
US 20130169683A1 · Perez et al. · 2013 [cited by applicant]
US 20130208234A1 · Lewis · 2013 [cited by applicant]
US 20130235169A1 · Kato et al. · 2013 [cited by applicant]
US 20130242262A1 · Lewis · 2013 [cited by applicant]
US 20130278631A1 · Border · 2013 [cited by examiner]
US 20130339433A1 · Li et al. · 2013 [cited by applicant]
US 20140016056A1 · Miyake et al. · 2014 [cited by applicant]
US 20140071163A1 · Kinnebrew et al. · 2014 [cited by applicant]
US 20140071539A1 · Gao · 2014 [cited by applicant]
US 20140104143A1 · Benson · 2014 [cited by examiner]
US 20140177023A1 · Gao et al. · 2014 [cited by applicant]
US 20140218468A1 · Gao et al. · 2014 [cited by applicant]
US 20140253589A1 · Tout et al. · 2014 [cited by applicant]
US 20140254939A1 · Kimura et al. · 2014 [cited by applicant]
US 20140267420A1 · Schowengerdt et al. · 2014 [cited by applicant]
US 20140306866A1 · Miller et al. · 2014 [cited by applicant]
US 20140361976A1 · Osman et al. · 2014 [cited by applicant]
US 20140375542A1 · Robbins et al. · 2014 [cited by applicant]
US 20140375680A1 · Ackerman et al. · 2014 [cited by applicant]
US 20150016777A1 · Abovitz et al. · 2015 [cited by applicant]
US 20150103306A1 · Kaji et al. · 2015 [cited by applicant]
US 20150178939A1 · Bradski et al. · 2015 [cited by applicant]
US 20150205126A1 · Schowengerdt · 2015 [cited by applicant]
US 20150222883A1 · Welch · 2015 [cited by applicant]
US 20150222884A1 · Cheng · 2015 [cited by applicant]
US 20150268415A1 · Schowengerdt et al. · 2015 [cited by applicant]
US 20150302652A1 · Miller et al. · 2015 [cited by applicant]
US 20150309263A2 · Abovitz et al. · 2015 [cited by applicant]
US 20150310263A1 · Zhang et al. · 2015 [cited by applicant]
US 20150326570A1 · Publicover et al. · 2015 [cited by applicant]
US 20150346490A1 · TeKolste et al. · 2015 [cited by applicant]
US 20150346495A1 · Welch et al. · 2015 [cited by applicant]
US 20160007934A1 · Arnold et al. · 2016 [cited by applicant]
US 20160011419A1 · Gao · 2016 [cited by applicant]
US 20160025971A1 · Crow · 2016 [cited by examiner]
US 20160026253A1 · Bradski et al. · 2016 [cited by applicant]
US 20160041048A1 · Blum et al. · 2016 [cited by applicant]
US 20160078278A1 · Moore et al. · 2016 [cited by applicant]
US 20160189426A1 · Thomas et al. · 2016 [cited by applicant]
US 20160217621A1 · Raghoebardajal et al. · 2016 [cited by applicant]
US 20170092235A1 · Osman et al. · 2017 [cited by applicant]
US 20170124713A1 · Jurgenson et al. · 2017 [cited by applicant]
US 20170206691A1 · Harrises et al. · 2017 [cited by applicant]
US 20170277254A1 · Osman · 2017 [cited by examiner]
US 20180088340A1 · Amayeh et al. · 2018 [cited by applicant]
US 20180253897A1 · Satake · 2018 [cited by applicant]
US 20210223552A1 · Amayeh et al. · 2021 [cited by applicant]
US 20220357582A1 · Amayeh et al. · 2022 [cited by applicant]
US 20230359044A1 · Amayeh et al. · 2023 [cited by applicant]
EP 2887639A1 · 2015 [cited by applicant]
JP H09193803A · 1997 [cited by applicant]
JP 2014064248A · 2014 [cited by applicant]
WO WO2015192117A1 · 2015 [cited by examiner]
WO 2016129156A1 · 2016 [cited by applicant]
WO 2018064169A1 · 2018 [cited by applicant]
“Introduction to SURF (Speeded-Up Robust Features)”, OpenCV, accessed Apr. 29, 2016, in 5 pages. URL: http://docs.opencv.org/3.0-beta/doc/py_tutorials/py_feature2d/py_surf_intro/py_surf_intro.html. [cited by applicant]
“Scale-invariant feature transform”, Wikipedia, printed Feb. 29, 2016, in 18 pages. URL: http://en.wikipedia.org/wiki/Scale-invariant_feature_transform. [cited by applicant]
ARToolKit: https://web.archive.org/web/20051013062315/http://www.hitl.washington.edu:80/artoolkit/documentation/hardware.htm, archived Oct. 13, 2005. [cited by applicant]
Azuma, “A Survey of Augmented Reality,” Teleoperators and Virtual Environments 6, 4 (Aug. 1997), pp. 355-385. https://web.archive.org/web/20010604100006/http://www.cs.unc.edu/ azuma/ARpresence.pdf. [cited by applicant]
Azuma, “Predictive Tracking for Augmented Realty,” TR95-007, Department of Computer Science, UNC-Chapel Hill, NC, Feb. 1995. [cited by applicant]
Bimber, et al., “Spatial Augmented Reality—Merging Real and Virtual Worlds,” 2005 https://web.media.mit.edu/raskar/book/BimberRaskarAugmentedRealityBook.pdf. [cited by applicant]
EP23168280.8 Extended European Search Report dated Jul. 6, 2023. [cited by applicant]
International Preliminary Report on Patentability for PCT Application No. PCT/US2017/053729, dated Apr. 2, 2019. [cited by applicant]
International Search Report and Written Opinion for PCT Application No. PCT/US2017/053729, dated Nov. 30, 2017. [cited by applicant]
Jacob, “Eye Tracking in Advanced Interface Design,” Human-Computer Interaction Lab Naval Research Laboratory, Washington, D.C. / paper/ in Virtual Environments and Advanced Interface Design, ed. by W. Barfield and T.A. … [cited by applicant]
JP2022-34301 Office Action dated May 9, 2023. [cited by applicant]
KR2022-7017129 Final Office Action dated Sep. 21, 2023. [cited by applicant]
Liu Z. et al., “Face Geometry and Appearance Modeling: Concepts and Applications,” Cambridge University Press, Apr. 2011, in 172 pages (uploaded in two parts). [cited by applicant]
Mattoccia, S., “Stereo Vision: Algorithms and Applications,” University of Bologna, Jan. 2013, in 208 pages. [cited by applicant]
Rublee, E et al., “ORB: an efficient alternative to SIFT or SURF”, Menlo Park, California, Nov. 2011, in 8 pages. [cited by applicant]
Strub, et al.: “Automated Facial Conformation for Model-Based Videophone Coding,” Proceedings of the International Conference on Image Processing (Icip). IEEE Comp. Soc. Press, US, vol. 2, Oct. 23, 1995, pp. 587-590. [cited by applicant]
Tanriverdi and Jacob, “Interacting With Eye Movements in Virtual Environments,” Department of Electrical Engineering and Computer Science, Tufts University, Medford, MA-paper/Proc. AMC CHI 2000 Human Factors in Computin… [cited by applicant]
U.S. Appl. No. 18/401,020 Office Action dated Dec. 6, 2024. [cited by applicant]
Wikipedia: “Iterative closest point”, Wikipedia, printed Feb. 29, 2016, in 3 pages. URL: https://en.wikipedia.org/wiki/Iterative_closest_point. [cited by applicant]
Wikipedia: “Lucky Imaging,” Wikipedia, printed Jul. 6, 2016, in 6 pages. URL: https://en.wikipedia.org/wiki/Lucky_imaging. [cited by applicant]
Wikipedia: “Super-resolution imaging”, Wikipedia, printed Jul. 6, 2016, in 6 pages. URL: https://en.wikipedia.org/wiki/Super-resolution_imaging. [cited by applicant]
Zhang, et al., “Occlusion-free Face Alignment: Deep Regression Networks Coupled with De-corrupt AutoEncoders,” 2016 IEEE Conference on Computer Vision and Pattern Recognition, The Computer Society, Jun. 2016. [cited by applicant]
EP24205698.4 Extended European Search Report dated Dec. 17, 2024. [cited by applicant]