IP Library Granted Patent US 12,488,488
Granted Patent B2
US 12,488,488 · App. 17/221,250 · Granted Dec 2, 2025

Personalized neural network for eye tracking

Inventors: Adrian Kaehler (Los Angeles, CA); Douglas Bertram Lee (Redwood City, CA); Vijay Badrinarayanan (Mountain View, CA)
Assignee: MAGIC LEAP, INC.
G06T7/70G02B27/0093G02B27/0172G06F1/163G06F3/011G06F3/013G06F3/0346G06F3/04815G06N3/02G06T7/20G06T7/246G02B2027/0185G06T19/006G06T2207/10016G06T2207/10048G06T2207/20081G06T2207/20084G06T2207/30041G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,488,488
App. No.
17/221,250
Granted
Dec 2, 2025
Kind
B2
Abstract

Disclosed herein is a wearable display system for capturing retraining eye images of an eye of a user for retraining a neural network for eye tracking. The system captures retraining eye images using an image capture device when user interface (UI) events occur with respect to UI devices displayed at display locations of a display. The system can generate a retraining set comprising the retraining eye images and eye poses of the eye of the user in the retraining eye images (e.g., related to the display locations of the UI devices) and obtain a retrained neural network that is retrained using the retraining set.

Claims (39)

1 . A computing system comprising:

a display device;

a non-transitory computer-readable storage medium configured to store software instructions;

a hardware processor configured to execute the software instructions to cause the computing system to:

capture one or more first images of an eye of a user during or immediately after a first user interface event in which the user activates or deactivates a virtual button of a virtual remote control in a first position, the first images reflecting eye poses of the user which are associated with a particular first portion of a user interface rendered as virtual content;

capture one or more second images of an eye of a user during or immediately after a second user interface event in which the user activates or deactivates the virtual button of the virtual remote control in a second position, the second images reflecting eye poses of the user which are associated with a particular second portion of a user interface, different than the particular first portion, rendered as virtual content;

cause update, based on the obtained first and second images as a set of retraining eye images, of a machine learning model configured to output an eye pose based on an input image related to the particular portion of the user interface, wherein the eye pose indicates a plurality of angular parameters relative to a natural resting direction of the eye and wherein the angular parameters indicate an azimuthal deflection and a zenithal deflection; and

identify, during operation of the computing system, a particular eye pose of the user via applying the updated machine learning model to an input image.

2 . The computing system of claim 1 , wherein the particular portion of the user interface corresponds to a location of the user interface event.

3 . The computing system of claim 1 , wherein the computing system is further configured to: transmit the one or more images and the associated particular portion of the user interface to a remote server configured to update the neural network.

4 . The computing system of claim 1 , wherein the machine learning model is updated to personalize the machine learning model to the user at or proximate to when the one or more images of the eye of the user are obtained during or immediately after the user interface event.

5 . The computing system of claim 1 , wherein the eye poses are determined based on the location of the particular portion and the location of the eye in the one or more images.

6 . The computing system of claim 1 , wherein the computing system comprises a wearable augmented reality headset and the user interface is rendered in a three-dimensional environment, and wherein the input image is obtained via an inward-facing camera.

7 . A computerized method, performed by a computing system having one or

more hardware computer processors and one or more non-transitory computer readable storage device storing software instructions executable by the computing system to perform the computerized method comprising:

capturing one or more first images of an eye of a user during or immediately after a first user interface event in which the user activates or deactivates a virtual button of a virtual remote control, the first images reflecting eye poses of the user which are associated with a particular first portion of a user interface rendered as virtual content;

capturing one or more additional images of the eye of the user during or immediately after respective further user interface events in which the user activates or deactivates the virtual button of the virtual remote control, the additional images reflecting eye poses of the user which are associated with respective particular further portions of the user interface rendered as virtual content,

wherein the first images and the particular first portion and the additional images and the respective particular further portions form a retraining set with retraining input data and corresponding retraining target output data with an eye pose of the eye of the user in each eye image of the eye images related to a display location of the virtual button with respect to the eye image;

determining a distribution probability of the virtual button in a first eye pose region of a plurality of eye pose regions according to a probability distribution function; and generating the retraining input data comprising the retraining eye image at an inclusion probability related to the distribution probability of display locations of the virtual button;

causing update, based on the obtained images, of a machine learning model configured to output an eye pose based on an input image, wherein the eye pose indicates a plurality of angular parameters relative to a natural resting direction of the eye and wherein the angular parameters indicate an azimuthal deflection and a zenithal deflection; and

identifying, during operation of the computing system, a particular eye pose of the user via applying the updated machine learning model to an input image.

8 . The method of claim 7 , wherein the particular portion of the user interface corresponds to the location of a user interface event.

9 . The method of claim 7 , further comprising:

transmitting the one or more images and the associated particular portion of the user interface to a remote server configured to update the neural network.

10 . The method of claim 7 , wherein the machine learning model is updated to personalize the machine learning model to the user, and wherein weights of the updated machine learning model are set to initial weights corresponding to the weights prior to updating the machine learning model.

11 . The method of claim 7 , wherein the eye poses are determined based on the location of the particular portion and the location of the eye in the one or more images.

12 . A non-transitory computer readable medium having software instructions stored thereon, the software instructions executable by a hardware computer processor to cause a computing system to perform operations comprising:

capturing one or more first images of an eye of a user during or immediately after a first user interface event in which the user activates or deactivates a virtual button of a virtual remote control, the one or more first images reflecting eye poses of the user which are associated with a particular first portion of a user interface rendered as virtual content, wherein the virtual remote control during the first user interface event has a primary function other than capturing the one or more first images of the user;

capture one or more second images of an eye of a user during or immediately after a second user interface event in which the user activates or deactivates the virtual button of the virtual remote control, the one or more second images reflecting eye poses of the user which are associated with a particular second portion of a user interface, different than the particular first portion, rendered as virtual content;

causing update, based on the obtained one or more first and second images as a set of retraining eye images, of a machine learning model configured to output an eye pose based on an input image, wherein the eye pose indicates a plurality of angular parameters relative to a natural resting direction of the eye and wherein the angular parameters indicate an azimuthal deflection and a zenithal deflection; and

identifying, during operation of the computing system, a particular eye pose of the user via applying the updated machine learning model to an input image.

13 . The non-transitory storage media of claim 12 , wherein the particular first portion of the user interface corresponds to the location of the first user interface event.

14 . The non-transitory storage media of claim 12 , wherein the operations further comprise:

transmitting the one or more first and second images and the associated first and second particular portions of the user interface to a remote server configured to update the neural network.

15 . The non-transitory storage media of claim 12 , wherein the machine learning model is updated to personalize the machine learning model to the user.

16 . The non-transitory storage media of claim 12 , wherein the eye poses associated with the one or more first images are determined based on the location of the particular first portion and the location of the eye in the one or more first images.

17 . The computing system of claim 1 , wherein the capture of the one or more images and the update of the machine learning model occurs independently of an initial training of the machine learning model, and wherein the computing system comprises a wearable augmented reality headset and the user interface is rendered in a three-dimensional environment, wherein the input image is obtained via an inward-facing camera, and wherein the hardware processor is configured to execute the software instructions to cause the computing system to perform the update of the machine learning model partially or entirely locally.

18 . The computing system of claim 1 , wherein the eye pose further indicates an angular roll of the eye.

19 . The method of claim 7 , wherein the virtual remote control comprises a first component and a second component, wherein the probability distribution function comprises a combined probability distribution of a distribution probability distribution function with respect to the first component and a second probability distribution function with respect to the second component.

Assignments (3)
SECURITY INTEREST Recorded May 24, 2022
From: MOLECULAR IMPRINTS, INC.; MENTOR ACQUISITION ONE, LLC; MAGIC LEAP, INC.
To: CITIBANK, N.A., AS COLLATERAL AGENT
Reel/Frame 060338/0665 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 27, 2022
From: LEE, DOUGLAS BERTRAM; BADRINARAYANAN, VIJAY
To: MAGIC LEAP, INC.
Reel/Frame 058799/0577 →
PROPRIETARY INFORMATION AND INVENTIONS AGREEMENT Recorded Jan 27, 2022
From: KAEHLER, ADRIAN
To: MAGIC LEAP, INC.
Reel/Frame 058889/0299 →
Continuity (4)
Continuation 16880752 · May 21, 2020
Continuation 16134600 · Sep 18, 2018
Provisional Application 62560898 · Sep 20, 2017
Related Publication 20210327085A1 · Oct 21, 2021
References Cited (327)
US 5285297A · Rose · 1994 [cited by examiner]
US 5291560A · Daugman · 1994 [cited by applicant]
US 5583795A · Smyth · 1996 [cited by applicant]
US 6246779B1 · Fukui et al. · 2001 [cited by applicant]
US 6850221B1 · Tickle · 2005 [cited by applicant]
US D514570S · Ohta · 2006 [cited by applicant]
US 7771049B2 · Knaan et al. · 2010 [cited by applicant]
US 7970179B2 · Tosa · 2011 [cited by applicant]
US 8098891B2 · Lv et al. · 2012 [cited by applicant]
US 8341100B2 · Miller et al. · 2012 [cited by applicant]
US 8345984B2 · Ji et al. · 2013 [cited by applicant]
US 8363783B2 · Gertner et al. · 2013 [cited by applicant]
US 8845625B2 · Angeley et al. · 2014 [cited by applicant]
US 8950867B2 · Macnamara · 2015 [cited by applicant]
US 9081426B2 · Armstrong · 2015 [cited by applicant]
US 9141916B1 · Corrado et al. · 2015 [cited by applicant]
US 9207760B1 · Wu et al. · 2015 [cited by applicant]
US 9215293B2 · Miller · 2015 [cited by applicant]
US 9262680B2 · Nakazawa et al. · 2016 [cited by applicant]
US D752529S · Loretan et al. · 2016 [cited by applicant]
US 9310559B2 · Macnamara · 2016 [cited by applicant]
US 9348143B2 · Gao et al. · 2016 [cited by applicant]
US D758367S · Natsume · 2016 [cited by applicant]
US D759657S · Kujawski et al. · 2016 [cited by applicant]
US 9417452B2 · Schowengerdt et al. · 2016 [cited by applicant]
US 9430829B2 · Madabhushi et al. · 2016 [cited by applicant]
US 9470906B2 · Kaji et al. · 2016 [cited by applicant]
US 9547174B2 · Gao et al. · 2017 [cited by applicant]
US 9671566B2 · Abovitz et al. · 2017 [cited by applicant]
US D794288S · Beers et al. · 2017 [cited by applicant]
US 9740006B2 · Gao · 2017 [cited by applicant]
US 9791700B2 · Schowengerdt et al. · 2017 [cited by applicant]
US D805734S · Fisher et al. · 2017 [cited by applicant]
US 9851563B2 · Gao et al. · 2017 [cited by applicant]
US 9857591B2 · Welch et al. · 2018 [cited by applicant]
US 9874749B2 · Bradski · 2018 [cited by applicant]
US 10241572B2 · Vidal et al. · 2019 [cited by applicant]
US 10650432B1 · Joseph et al. · 2020 [cited by applicant]
US 10686984B1 · Schmidt · 2020 [cited by examiner]
US 10719951B2 · Kaehler · 2020 [cited by examiner]
US 10977820B2 · Kaehler · 2021 [cited by examiner]
US 20030020755A1 · Lemelson et al. · 2003 [cited by applicant]
US 20040130680A1 · Zhou et al. · 2004 [cited by applicant]
US 20060028436A1 · Armstrong · 2006 [cited by applicant]
US 20060088193A1 · Muller et al. · 2006 [cited by applicant]
US 20060147094A1 · Yoo · 2006 [cited by applicant]
US 20070052672A1 · Ritter et al. · 2007 [cited by applicant]
US 20070081123A1 · Lewis · 2007 [cited by applicant]
US 20070140531A1 · Hamza · 2007 [cited by applicant]
US 20070164990A1 · Bjorklund et al. · 2007 [cited by applicant]
US 20070189742A1 · Knaan et al. · 2007 [cited by applicant]
US 20080278682A1 · Huxlin et al. · 2008 [cited by applicant]
US 20080292146A1 · Breed et al. · 2008 [cited by applicant]
US 20090129591A1 · Hayes et al. · 2009 [cited by applicant]
US 20090141947A1 · Kyyko et al. · 2009 [cited by applicant]
US 20090163898A1 · Gertner et al. · 2009 [cited by applicant]
US 20100014718A1 · Savvides et al. · 2010 [cited by applicant]
US 20100131096A1 · Koyano · 2010 [cited by applicant]
US 20100208951A1 · Williams et al. · 2010 [cited by applicant]
US 20100232654A1 · Rahmes et al. · 2010 [cited by applicant]
US 20100284576A1 · Tosa et al. · 2010 [cited by applicant]
US 20100316263A1 · Hamza · 2010 [cited by examiner]
US 20110182469A1 · Ji et al. · 2011 [cited by applicant]
US 20110202046A1 · Angeley et al. · 2011 [cited by applicant]
US 20120127062A1 · Bar-Zeev et al. · 2012 [cited by applicant]
US 20120162549A1 · Gao et al. · 2012 [cited by applicant]
US 20120163678A1 · Du et al. · 2012 [cited by applicant]
US 20120164618A1 · Kullok et al. · 2012 [cited by applicant]
US 20130082922A1 · Miller · 2013 [cited by applicant]
US 20130117377A1 · Miller · 2013 [cited by applicant]
US 20130125027A1 · Abovitz · 2013 [cited by applicant]
US 20130159939A1 · Krishnamurthi · 2013 [cited by applicant]
US 20130208234A1 · Lewis · 2013 [cited by applicant]
US 20130242262A1 · Lewis · 2013 [cited by applicant]
US 20140071539A1 · Gao · 2014 [cited by applicant]
US 20140126782A1 · Takai et al. · 2014 [cited by applicant]
US 20140161325A1 · Bergen · 2014 [cited by applicant]
US 20140177023A1 · Gao et al. · 2014 [cited by applicant]
US 20140218468A1 · Gao et al. · 2014 [cited by applicant]
US 20140267420A1 · Schowengerdt · 2014 [cited by applicant]
US 20140270405A1 · Derakhshani et al. · 2014 [cited by applicant]
US 20140279774A1 · Wang et al. · 2014 [cited by applicant]
US 20140306866A1 · Miller et al. · 2014 [cited by applicant]
US 20140380249A1 · Fleizach · 2014 [cited by applicant]
US 20150016777A1 · Abovitz et al. · 2015 [cited by applicant]
US 20150103306A1 · Kaji et al. · 2015 [cited by applicant]
US 20150117760A1 · Wang et al. · 2015 [cited by applicant]
US 20150125049A1 · Taigman et al. · 2015 [cited by applicant]
US 20150134583A1 · Tamatsu et al. · 2015 [cited by applicant]
US 20150154758A1 · Nakazawa et al. · 2015 [cited by applicant]
US 20150170002A1 · Szegedy et al. · 2015 [cited by applicant]
US 20150178939A1 · Bradski et al. · 2015 [cited by applicant]
US 20150205126A1 · Schowengerdt · 2015 [cited by applicant]
US 20150222883A1 · Welch · 2015 [cited by applicant]
US 20150222884A1 · Cheng · 2015 [cited by applicant]
US 20150268415A1 · Schowengerdt et al. · 2015 [cited by applicant]
US 20150278642A1 · Chertok et al. · 2015 [cited by applicant]
US 20150302652A1 · Miller et al. · 2015 [cited by applicant]
US 20150309263A2 · Abovitz et al. · 2015 [cited by applicant]
US 20150326570A1 · Publicover et al. · 2015 [cited by applicant]
US 20150338915A1 · Publicover et al. · 2015 [cited by applicant]
US 20150346490A1 · TeKolste et al. · 2015 [cited by applicant]
US 20150346495A1 · Welch et al. · 2015 [cited by applicant]
US 20160011419A1 · Gao · 2016 [cited by applicant]
US 20160012292A1 · Perna et al. · 2016 [cited by applicant]
US 20160012304A1 · Mayle et al. · 2016 [cited by applicant]
US 20160026253A1 · Bradski et al. · 2016 [cited by applicant]
US 20160034679A1 · Yun et al. · 2016 [cited by applicant]
US 20160034811A1 · Paulik et al. · 2016 [cited by applicant]
US 20160035078A1 · Lin et al. · 2016 [cited by applicant]
US 20160098844A1 · Shaji et al. · 2016 [cited by applicant]
US 20160104053A1 · Yin et al. · 2016 [cited by applicant]
US 20160104056A1 · He et al. · 2016 [cited by applicant]
US 20160135675A1 · Du et al. · 2016 [cited by applicant]
US 20160162782A1 · Park · 2016 [cited by applicant]
US 20160180722A1 · Yehezkel · 2016 [cited by examiner]
US 20160189027A1 · Graves et al. · 2016 [cited by applicant]
US 20160216761A1 · Klingström · 2016 [cited by examiner]
US 20160291327A1 · Kim et al. · 2016 [cited by applicant]
US 20160299685A1 · Zhai et al. · 2016 [cited by applicant]
US 20160335795A1 · Flynn et al. · 2016 [cited by applicant]
US 20160377864A1 · Moran · 2016 [cited by examiner]
US 20170053165A1 · Kaehler · 2017 [cited by applicant]
US 20170061330A1 · Kurata · 2017 [cited by applicant]
US 20170061625A1 · Estrada et al. · 2017 [cited by applicant]
US 20170061688A1 · Miller · 2017 [cited by applicant]
US 20170068322A1 · Steinberg · 2017 [cited by examiner]
US 20170161506A1 · Gates et al. · 2017 [cited by applicant]
US 20170168566A1 · Osterhout et al. · 2017 [cited by applicant]
US 20170180721A1 · Parker · 2017 [cited by examiner]
US 20170186236A1 · Kawamoto · 2017 [cited by applicant]
US 20170212583A1 · Krasadakis · 2017 [cited by examiner]
US 20170262737A1 · Rabinovich et al. · 2017 [cited by applicant]
US 20170308734A1 · Chalom et al. · 2017 [cited by applicant]
US 20180008141A1 · Krueger · 2018 [cited by applicant]
US 20180018451A1 · Spizhevoy et al. · 2018 [cited by applicant]
US 20180018515A1 · Spizhevoy et al. · 2018 [cited by applicant]
US 20180053056A1 · Rabinovich et al. · 2018 [cited by applicant]
US 20180082172A1 · Patel et al. · 2018 [cited by applicant]
US 20180089834A1 · Spizhevoy et al. · 2018 [cited by applicant]
US 20180096226A1 · Aliabadi et al. · 2018 [cited by applicant]
US 20180137642A1 · Malisiewicz et al. · 2018 [cited by applicant]
US 20180184002A1 · Thukral · 2018 [cited by examiner]
US 20180239412A1 · Elvesjöet al. · 2018 [cited by applicant]
US 20180268220A1 · Lee et al. · 2018 [cited by applicant]
US 20180367752A1 · Donsbach · 2018 [cited by examiner]
US 20200286251A1 · Kaehler et al. · 2020 [cited by applicant]
US 20210241424A1 · Peuhkurinen · 2021 [cited by examiner]
CN 105607255A · 2016 [cited by examiner]
CN 104299245A · 2017 [cited by applicant]
CN 105247539A · 2018 [cited by applicant]
JP 11175246A · 1999 [cited by applicant]
JP 2008502990A · 2008 [cited by applicant]
JP 2012530305A · 2012 [cited by applicant]
KR 20100105591A · 2010 [cited by applicant]
KR 20140102486A · 2014 [cited by applicant]
KR 20170029166A · 2017 [cited by applicant]
WO WO2014182769 · 2014 [cited by applicant]
WO WO2015161307 · 2015 [cited by applicant]
WO WO2015164807 · 2015 [cited by applicant]
WO WO2018013199 · 2018 [cited by applicant]
WO WO2018013200 · 2018 [cited by applicant]
WO WO2018039269 · 2018 [cited by applicant]
WO WO2018063451 · 2018 [cited by applicant]
WO WO2018067603 · 2018 [cited by applicant]
WO WO2018093796 · 2018 [cited by applicant]
WO WO2018170421 · 2018 [cited by applicant]
WO WO2019060283 · 2019 [cited by applicant]
International Search Report and Written Opinion for PCT Application No. PCT/US18/51461, mailed Nov. 8, 2018. [cited by applicant]
International Preliminary Report on Patentability for PCT Application No. PCT/US18/51461, issued Mar. 24, 2020. [cited by applicant]
“Camera calibration with OpenCV”, OpenCV, retrieved May 5, 2016, in 7 pages. URL: http://docs.opencv.org/3.1.0/d4/d94/tutorial_camera_calibration.html#gsc.tab=0. [cited by applicant]
“Feature Extraction Using Convolution”, Ufldl, printed Sep. 1, 2016, in 3 pages. URL:http://deeplearning.stanford.edu/wiki/index.php/Feature_extraction_using_convolution. [cited by applicant]
“Machine Learning”, Wikipedia, printed Oct. 3, 2017, in 14 pages. URL: https://en.wikipedia.org/wiki/Machine_learning. [cited by applicant]
“Transfer Function Layers”, GitHub, Dec. 1, 2015, in 13 pages; accessed URL: http://github.com/torch/nn/blob/master/doc/transfer.md. [cited by applicant]
Adegoke et al., “Iris Segmentation: A Survey”, Int J Mod Engineer Res. (IJMER) (Jul./Aug. 2013) 3(4): 1885-1889. [cited by applicant]
Anthony, S., “MIT releases open-source software that reveals invisible motion and detail in video”, Extreme Tech, Feb. 28, 2013, as accessed Aug. 4, 2017, in 5 pages. [cited by applicant]
Arevalo J. et al., “Convolutional neural networks for mammography mass lesion classification”, in Engineering in Medicine and Biology Society (EMBC); 37th Annual International Conference IEEE, Aug. 25-29, 2015, pp. 797-… [cited by applicant]
ARToolKit: https://web.archive.org/web/20051013062315/http://www.hitl.washington.edu:80/artoolkit/documentation/hardware.htm, archived Oct. 13, 2005. [cited by applicant]
Aubry M. et al., “Seeing 3D chairs: exemplar part-based 2D-3D alignment using a large dataset of CAD models”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (Jun. 23-28, 2014); Computer Vi… [cited by applicant]
Azizpour et al., “From Generic to Specific Deep Representations for Visual Recognition,” ResearchGate, Jun. 2014. Https://www.researchgate.net/publications/263352539. arXiv:1406.5774v1 [cs.CV] Jun. 22, 2014. [cited by applicant]
Azuma, “A Survey of Augmented Reality,” Teleoperators and Virtual Environments 6, 4 (Aug. 1997), pp. 355-385. https://web.archive.org/web/20010604100006/http://www.cs.unc.edu/˜azuma/ARpresence.pdf. [cited by applicant]
Azuma, “Predictive Tracking for Augmented Realty,” TR95-007, Department of Computer Science, UNC-Chapel Hill, NC, Feb. 1995. [cited by applicant]
Badrinarayanan et al., “SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation”, IEEE (Dec. 8, 2015) arXiv:1511.00561v2 in 14 pages. [cited by applicant]
Badrinarayanan et al., “SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation”, TPAMI, vol. 39, No. 12, Dec. 2017. [cited by applicant]
Bansal A. et al., “Marr Revisited: 2D-3D Alignment via Surface Normal Prediction”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (Jun. 27-30, 2016) pp. 5965-5974. [cited by applicant]
Belagiannis V. et al., “Recurrent Human Pose Estimation”, In Automatic Face & Gesture Recognition; 12th IEEE International Conference—May 2017, ar Xiv: 1605.02914v3; (Aug. 5, 2017) Open Access Version in 8 pages. [cited by applicant]
Bell S. et al., “Inside-Outside Net: Detecting Objects in Context with Skip Pooling and Recurrent Neural Networks”, In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Jun. 27-30, 2016; pp.… [cited by applicant]
Biederman I., “Recognition-by-Components: A Theory of Human Image Understanding”, Psychol Rev. (Apr. 1987) 94(2): 115-147. [cited by applicant]
Bimber, et al., “Spatial Augmented Reality—Merging Real and Virtual Worlds,” 2005 https://web.media.mit.edu/˜raskar/book/BimberRaskarAugmentedRealityBook.pdf. [cited by applicant]
Bouget, J., “Camera Calibration Toolbo for Matlab” Cal-Tech, Dec. 2, 2013, in 5 pages. URL: https://www.vision.caltech.edu/bouguetj/calib_doc/inde .html#parameters. [cited by applicant]
Bulat A. et al., “Human pose estimation via Convolutional Part Heatmap Regression”, arXiv e-print arXiv:1609.01743v1, Sep. 6, 2016 in 16 pages. [cited by applicant]
Carreira J. et al., “Human Pose Estimation with Iterative Error Feedback”, In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Jun. 27-30, 2016, pp. 4733-4742. [cited by applicant]
Chatfield et al., “Return of the Devil in the Details: Delving Deep into Convolutional Nets”, arXiv e-print arXiv:1405.3531v4, Nov. 5, 2014 in 11 pages. [cited by applicant]
Chen et al., “Semantic Image Segmentation With Deep Convolutional Nets and Fully Connected CRFs,” In ICLR, arXiv:1412.7062v3 [cs.CV] Apr. 9, 2015. [cited by applicant]
Chen X. et al., “3D Object Proposals for Accurate Object Class Detection”, in Advances in Neural Information Processing Systems, (2015) Retrieved from <http://papers.nips.cc/paper/5644-3d-objectproposals-for-accurate-ob… [cited by applicant]
Choy et al., “3D-R2N2: A Unified Approach for Single and Multi-view 3D Object Reconstruction”, arXiv; e-print arXiv:1604.00449v1, Apr. 2, 2016 in 17 pages. [cited by applicant]
Collet et al., “The MOPED framework: Object Recognition and Pose Estimation for Manipulation”, The International Journal of Robotics Research. (Sep. 2011) 30(10):1284-306; preprint Apr. 11, 2011 in 22 pages. [cited by applicant]
Coughlan et al., “The Manhattan World Assumption: Regularities in scene statistics which enable bayesian inference,” In NIPS, 2000. [cited by applicant]
Crivellaro A. et al., “A Novel Representation of Parts for Accurate 3D Object Detection and Tracking in Monocular Images”, In Proceedings of the IEEE International Conference on Computer Vision; Dec. 7-13, 2015 (pp. 439… [cited by applicant]
Dai J. et al., “Instance-aware Semantic Segmentation via Multi-task Network Cascades”, In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition; Jun. 27-30, 2016 (pp. 3150-3158). [cited by applicant]
Dai J. et al., “R-FCN: Object Detection via Region-based Fully Convolutional Networks”, in Advances in neural information processing systems; (Jun. 21, 2016) Retrieved from <https://arxiv.org/pdf/1605.06409.pdf in 11 pa… [cited by applicant]
Dasgupta et al., “Delay: Robust Spatial Layout Estimation for Cluttered Indoor Scenes,” In CVPR, 2016. [cited by applicant]
Daugman, J. et al., “Epigenetic randomness, compleity and singularity of human iris patterns”, Proceedings of Royal Society: Biological Sciences, vol. 268, Aug. 22, 2001, in 4 pages. [cited by applicant]
Daugman, J., “How Iris Recognition Works”, IEEE Transactions on Circuits and Systems for Video Technology, vol. 14, No. 1, Jan. 2004, in 10 pages. [cited by applicant]
Daugman, J., “New Methods in Iris Recognition,” IEEE Transactions on Systems, Man, and Cybernetics—Part B: Cybernetics, vol. 37, No. 5, Oct. 2007, in 9 pages. [cited by applicant]
Daugman, J., “Probing the Uniqueness and Randomness of IrisCodes: Results From 200 Billion Iris Pair Comparisons,” Proceedings of the IEEE, vol. 94, No. 11, Nov. 2006, in 9 pages. [cited by applicant]
Del Pero et al., “Bayesian geometric modeling of indoor scenes,” In CVPR, 2012. [cited by applicant]
Del Pero et al., “Understanding bayesian rooms using composite 3d object models,” In CVPR, 2013. [cited by applicant]
Detone D. et al., “Deep Image Homography Estimation”, arXiv e-print arXiv:1606.03798v1, Jun. 13, 2016 in 6 pages. [cited by applicant]
Dwibedi et al., “Deep Cuboid Detection: Beyond 2D Bounding Bo es”, arXiv e-print arXiv:1611.10010v1; Nov. 30, 2016 in 11 pages. [cited by applicant]
Everingham M. et al., “The PASCAL Visual Object Classes (VOC) Challenge”, Int J Comput Vis (Jun. 2010) 88(2):303-38. [cited by applicant]
Farabet, C. et al., “Hardware Accelerated Convolutional Neural Networks for Synthetic Vision Systems”, Proceedings of the 2010 IEEE International Symposium (May 30-Jun. 2, 2010) Circuits and Systems (ISCAS), pp. 257-260. [cited by applicant]
Fidler S. et al., “3D Object Detection and Viewpoint Estimation with a Deformable 3D Cuboid Model”, in [cited by applicant]
Fouhey D. et al., “Data-Driven 3D Primitives for Single Image Understanding”, Proceedings of the IEEE International Conference on Computer Vision, Dec. 1-8, 2013; pp. 3392-3399. [cited by applicant]
Geiger A. et al., “Joint 3D Estimation of Objects and Scene Layout”, In Advances in Neural Information Processing Systems 24; (Dec. 12-17, 2011) in 9 pages. [cited by applicant]
Gidaris S. et al., “Object detection via a multi-region & semantic segmentation-aware CNN model”, in Proceedings of the IEEE International Conference on Computer Vision; Dec. 7-13, 2015 (pp. 1134-1142). [cited by applicant]
Girshick R. et al., “Fast R-CNN”, Proceedings of the IEEE International Conference on Computer Vision; Dec. 7-13, 2015 (pp. 1440-1448). [cited by applicant]
Girshick R. et al., “Rich feature hierarchies for accurate object detection and semantic segmentation”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Jun. 23-28, 2014 (pp. 580-587). [cited by applicant]
Gupta A. et al., “Blocks World Revisited: Image Understanding Using Qualitative Geometry and Mechanics”, in European Conference on Computer Vision; Sep. 5, 2010 in 14 pages. [cited by applicant]
Gupta A. et al., “From 3D Scene Geometry to Human Workspace”, in Computer Vision and Pattern Recognition (CVPR); IEEE Conference on Jun. 20-25, 2011 (pp. 1961-1968). [cited by applicant]
Gupta et al., “Perceptual Organization and Recognition of Indoor Scenes from RGB-D Images,” In CVPR, 2013. [cited by applicant]
Gupta S. et al., “Aligning 3D Models to RGB-D Images of Cluttered Scenes”, in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Jun. 7-12, 2015 (pp. 4731-4740). [cited by applicant]
Gupta S. et al., “Inferring 3D Object Pose in RGB-D Images”, arXiv e-print arXiv:1502.04652v1, Feb. 16, 2015 in 13 pages. [cited by applicant]
Gupta S. et al., “Learning Rich Features from RGB-D Images for Object Detection and Segmentation”, in European Conference on Computer Vision; (Jul. 22, 2014); Retrieved from <https://arxiv.org/pdf/1407.5736.pdf> in 16 p… [cited by applicant]
Han et al., “Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding”, arXiv e-print arX iv:1510.00149v5, Feb. 15, 2016 in 14 pages. [cited by applicant]
Hansen, D. et al., “In the Eye of the Beholder: A Survey of Models for Eyes and Gaze”, IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 32, No. 3, Mar. 2010, in 23 pages. [cited by applicant]
Hartley R. et al., [cited by applicant]
He et al., “Deep Residual Learning for Image Recognition,” In CVPR, 2016. [cited by applicant]
He et al., “Delving Deep into Rectifiers: Surpassing Human-level Performance on ImageNet Classification”, arXiv: e-print arXiv:1502.01852v1, Feb. 6, 2015. [cited by applicant]
He et al., “Spatial Pyramid Pooling in Deep Convolutional Networks for Visual Recognition”, arXiv e-print arXiv:1406.4729v2; Aug. 29, 2014 in 14 pages. [cited by applicant]
Hedau et al., “Recovering the Spatial Layout of Cluttered Rooms,” In ICCV, 2009. [cited by applicant]
Hedau V. et al., “Recovering Free Space of Indoor Scenes from a Single Image”, in [cited by applicant]
Hejrati et al., “Categorizing Cubes: Revisiting Pose Normalization”, Applications of Computer Vision (WACV), 2016 IEEE Winter Conference, Mar. 7-10, 2016 in 9 pages. [cited by applicant]
Hijazi, S. et al., “Using Convolutional Neural Networks for Image Recognition”, Tech Rep. (Sep. 2015) available online URL: http://ip. cadence. com/uploads/901/cnn-wp-pdf, in 12 pages. [cited by applicant]
Hochreiter et al., “Long Short-Term Memory,” Neural computation, 9, 1735-1780, 1997. [cited by applicant]
Hoffer et al., “Deep Metric Learning Using Triplet Network”, International Workshop on Similarity-Based Pattern Recognition [ICLR]; Nov. 25, 2015; [online] retrieved from the internet <https://arxv.org/abs/1412.6622>; p… [cited by applicant]
Hoiem D. et al., “Representations and Techniques for 3D Object Recognition and Scene Interpretation”, Synthesis Lectures on Artificial Intelligence and Machine Learning, Aug. 2011, vol. 5, No. 5, pp. 1-169; Abstract in … [cited by applicant]
Hsiao E. et al., “Making specific features less discriminative to improve point-based 3D object recognition”, in [cited by applicant]
Huang et al., “Sign Language Recognition Using 3D Convolutional Neural Networks”, University of Science and Technology of China, 2015 IEEE International Conference on Multimedia and Expo. Jun. 29-Jul. 3, 2015, in 6 page… [cited by applicant]
Iandola F. et al., “SqueezeNet: Ale Net-level accuracy with 50 fewer parameters and <1MB model size”, arXiv e-print arXiv:1602.07360v1, Feb. 24, 2016 in 5 pages. [cited by applicant]
Ioffe S. et al., “Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift”, arXiv:1502.03167v3 [cs.LG] Mar. 2, 2015. [cited by applicant]
Izadinia et al., “IM2CAD,” arXiv preprint arXiv:1608.05137, 2016. [cited by applicant]
Jacob, “Eye Tracking in Advanced Interface Design,” Human-Computer Interaction Lab Naval Research Laboratory, Washington, D.C. / paper/ in Virtual Environments and Advanced Interface Design, ed. by W. Barfield and T.A. … [cited by applicant]
Ji, H. et al., “3D Convolutional Neural Networks for Human Action Recognition”, IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 35:1, Jan. 2013, in 11 pages. [cited by applicant]
Jia et al., “3D-Based Reasoning with Blocks, Support, and Stability”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition; Jun. 23-28, 2013 in 8 pages. [cited by applicant]
Jia et al., “Caffe: Convolutional Architecture for Fast Feature Embedding”, arXiv e-print arXiv:1408.5093v1, Jun. 20, 2014 in 4 pages. [cited by applicant]
Jiang H. et al., “A Linear Approach to Matching Cuboids in RGBD Images”, in [cited by applicant]
Jillela et al., “An Evaluation of Iris Segmentation Algorithms in Challenging Periocular Images”, Handbook of Iris Recognition, Springer Verlag, Heidelberg (Jan. 12, 2013) in 28 pages. [cited by applicant]
Kar A. et al., “Category-specific object reconstruction from a single image”, in [cited by applicant]
Krizhevsky et al., “ImageNet Classification with Deep Convolutional Neural Networks”, Advances in Neural Information Processing Systems. Apr. 25, 2013, pp. 1097-1105. [cited by applicant]
Lavin, A. et al.: “Fast Algorithms for Convolutional Neural Networks”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (Nov. 2016) arXiv:1509.09308v2, Nov. 10, 2015 in 9 pages. [cited by applicant]
Le, et al., Smooth Skinning Decomposition with Rigid Bones, ACM Transactions on Graphics (TOG). vol. 31(6), pp. 199-209, Nov. 2012. [cited by applicant]
Lee D. et al., “Geometric Reasoning for Single Image Structure Recovery”, in IEEE Conference Proceedings in Computer Vision and Pattern Recognition (CVPR) Jun. 20-25, 2009, pp. 2136-2143. [cited by applicant]
Lee et al., “Deeply-Supervised Nets,” In AISTATS, San Diego, CA 2015, JMLR: W&CP vol. 38. [cited by applicant]
Lee et al., “Estimating Spatial Layout of Rooms using Volumetric Reasoning about Objects and Surfaces,” In NIPS, 2010. [cited by applicant]
Lee et al., “Generalizing Pooling Functions in Convolutional Neural Networks: Mixed, Gated, and Tree,” In AISTATS, Càdiz, Spain, JMLR: W&CP vol. 51, 2016. [cited by applicant]
Lee et al., “Recursive Recurrent Nets with Attention Modeling for OCR in the Wild,” In CVPR, 2016. [cited by applicant]
Liang et al., “Recurrent Convolutional Neural Network for Object Recognition,” In CVPR, 2015. [cited by applicant]
Lim J. et al., “FPM: Fine pose Parts-based Model with 3D CAD models”, European Conference on Computer Vision; Springer Publishing, Sep. 6, 2014, pp. 478-493. [cited by applicant]
Liu et al., “ParseNet: Looking Wider to See Better”, arXiv e-print arXiv:1506.04579v1; Jun. 15, 2015 in 9 pages. [cited by applicant]
Liu et al., “Rent3d: Floor-Plan Priors for Monocular Layout Estimation,” In CVPR, 2015. [cited by applicant]
Liu W. et al., “SSD: Single Shot MultiBo Detector”, arXiv e-print arXiv:1512.02325v5, Dec. 29, 2016 in 17 pages. [cited by applicant]
Long et al., “Fully Convolutional Networks for Semantic Segmentation”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (Jun. 7-12, 2015) in 10 pages. [cited by applicant]
Mallya et al., “Learning Informative Edge Maps for Indoor Scene Layout Prediction,” In ICCV, 2015. [cited by applicant]
Mirowski et al., “Learning to Navigate in Complex Environments,” In ICLR, 2017. [cited by applicant]
Nair et al., “Rectified Linear Units Improve Restricted Boltzmann Machines,” In ICML, Haifa, Israel Jun. 2010. [cited by applicant]
Newell et al., “Stacked Hourglass Networks for Human Pose Estimation,” In ECCV, ArXiv:1603.06937v2 [cs.CV] 2016. [cited by applicant]
Noh et al., “Learning Deconvolution Network for Semantic Segmentation,” In ICCV, 2015. [cited by applicant]
Oberweger et al., “Training a Feedback Loop for Hand Pose Estimation,” In ICCV, 2015. [cited by applicant]
Open CV: “Camera calibration with OpenCV”, OpenCV, retrieved May 5, 2016, in 12 pages. URL: http://docs.opencv.org/2.4/doc/tutorials/calib3d/camera_calibration/camera_calibration.html. [cited by applicant]
OpenCV: “Camera Calibration and 3D Reconstruction”, OpenCV, retrieved May 5, 2016, in 51 pages. URL: http://docs.opencv.org/2.4/modules/calib3d/doc/camera_calibration_and_3d_reconstruction.html. [cited by applicant]
Pavlakos G. et al., “6-dof object pose from semantic keypoints”, in arXiv preprint Mar. 14, 2017; Retrieved from <http://www.cis.upenn.edu/˜kostas/mypub.dir/pavlakos17icra.pdf> in 9 pages. [cited by applicant]
Peng et al., “A Recurrent Encoder-Decoder Network for Sequential Face Alignment,” In ECCV, arXiv:1608.05477v2 [cs.CV] 2016. [cited by applicant]
Pfister et al., “Flowing Convnets for Human Pose Estimation in Videos,” In ICCV, 2015. [cited by applicant]
Ramalingam et al., “Manhattan Junction Catalogue for Spatial Reasoning of Indoor Scenes,” In CVPR, 2013. [cited by applicant]
Rastegari et al., “XNOR-Net: ImageNet Classification Using Binary Convolutional Neural Networks”, arXiv e-print arXiv:1603.05279v4; Aug. 2, 2016 in 17 pages. [cited by applicant]
Redmon et al., “You Only Look Once: Unified, Real-Time Object Detection”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (Jun. 27-30, 2016) pp. 779-788. [cited by applicant]
Ren et al., “A Coarse-to-Fine Indoor Layout Estimation (CFILE) Method,” In ACCV, arXiv:1607.00598v1 [cs.CV] 2016. [cited by applicant]
Ren et al., “Faster R-CNN: Towards real-time object detection with region proposal networks”, arxiv e-print arXiv:1506.01497v3; Jan. 6, 2016 in 14 pages. [cited by applicant]
Ren et al.: “On Vectorization of Deep Convolutional Neural Networks for Vision Tasks,” AAAI, arXiv: e-print arXiv:1501.07338v1, Jan. 29, 2015 in 8 pages. [cited by applicant]
Roberts L. et al., “Machine Perception of Three-Dimensional Solids”, Doctoral Thesis MIT; Jun. 1963 in 82 pages. [cited by applicant]
Rubinstein, M., “Eulerian Video Magnification”, YouTube, published May 23, 2012, as archived Sep. 6, 2017, in 13 pages (with video transcription). URL: https://web.archive.org/web/20170906180503/https://www.youtube.com/… [cited by applicant]
Russell et al., “Labelme: a database and web-based tool for image annotation,” IJCV, vol. 77, Issue 1-3, pp. 157-173, May 2008. [cited by applicant]
Savarese et al., “3D generic object categorization, localization and pose estimation”, in [cited by applicant]
Saxena A., “Convolutional Neural Networks (CNNS): An Illustrated E planation”, Jun. 29, 2016 in 16 pages; Retrieved from <http://xrds.acm.org/blog/2016/06/convolutional-neural-networks-cnns-illustrated-explanation/>. [cited by applicant]
Schroff et al., “FaceNet: A unified embedding for Face Recognition and Clustering”, arXiv eprint arXiv:1503.03832v3, Jun. 17, 2015 in 10 pages. [cited by applicant]
Shafiee et al., “ISAAC: A Convolutional Neural Network Accelerator with In-Situ Analog Arithmetic in Crossbars”, ACM Sigarch Comp. Architect News (Jun. 2016) 44(3):14-26. [cited by applicant]
Shao et al., “Imagining the Unseen: Stability-based Cuboid Arrangements for Scene Understanding”, ACM Transactions on Graphics. (Nov. 2014) 33(6) in 11 pages. [cited by applicant]
Shi et al., “Convolutional LSTM Network: A Machine Learning Approach for Precipitation Nowcasting,” in NIPS, 2015. [cited by applicant]
Simonyan et al., “Very deep convolutional networks for large-scale image recognition”, arXiv e-print arXiv:1409.1556v6, Apr. 10, 2015 in 14 pages. [cited by applicant]
Song et al., “Deep Sliding Shapes for Amodal 3D Object Detection in RGB-D Images”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. Jun. 27-30, 2016 (pp. 808-816). [cited by applicant]
Song et al., “Sliding Shapes for 3D Object Detection in Depth Images”, in European Conference on Computer Vision, (Sep. 6, 2014) Springer Publishing (pp. 634-651). [cited by applicant]
Song et al., “Sun RGB-D: A RGB-D Scene Understanding Benchmark Suite,” In CVPR, 2015. [cited by applicant]
Su et al., “Render for CNN: Viewpoint Estimation in Images Using CNNs Trained with Rendered 3D Model Views”, in Proceedings of the IEEE International Conference on Computer Vision, Dec. 7-13, 2015 (pp. 2686-2694). [cited by applicant]
Szegedy et al., “Going deeper with convolutions”, arXiv:1409.4842v1, Sep. 17, 2014 in 12 pages. [cited by applicant]
Szegedy et al., “Going Deeper with Convolutions,” In CVPR, 2015 in 9 pages. [cited by applicant]
Szegedy et al., “Rethinking the Inception Architecture for Computer Vision”, arXiv e-print arXIV:1512.00567v3, Dec. 12, 2015 in 10 pages. [cited by applicant]
Tanriverdi and Jacob, “Interacting With Eye Movements in Virtual Environments,” Department of Electrical Engineering and Computer Science, Tufts University, Medford, MA—paper/Proc. ACM CHI 2000 Human Factors in Computin… [cited by applicant]
Tompson et al., “Joint Training of a Convolutional Network and a Graphical Model for Human Pose Estimation,” In NIPS, 2014. [cited by applicant]
Tu et al., “Auto-context and Its Application to High-level Vision Tasks,” In CVPR, 2008. 978-1-4244-2243-2/08, IEEE. [cited by applicant]
Tulsiani S. et al., “Viewpoints and Keypoints”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition; Jun. 7-12, 2015 (pp. 1510-1519). [cited by applicant]
Villanueva, A. et al., “A Novel Gaze Estimation System with One Calibration Point”, IEEE Transactions on Systems, Man, and Cybernetics—Part B:Cybernetics, vol. 38:4, Aug. 2008, in 16 pages. [cited by applicant]
Wikipedia: “Convolution”, Wikipedia, accessed Oct. 1, 2017, in 17 pages. URL: https://en.wikipedia.org/wiki/Convolution. [cited by applicant]
Wikipedia: “Deep Learning”, Wikipedia, printed Apr. 27, 2016, in 40 pages. URL: https://en.wikipedia.org/wiki/Deep_learning#Deep_neural_networks. [cited by applicant]
Wikipedia: “Deep Learning”, Wikipedia, printed Oct. 3, 2017, in 23 pages. URL: https://en.wikipedia.org/wiki/Deep_learning. [cited by applicant]
Wikipedia: Inductive transfer, Wikipedia, retrieved , Apr. 27, 2016. https://en.wikipedia.org/wiki/Inductive_transfer, in 3 pages. [cited by applicant]
Wilczkowiak et al., “Using Geometric Constraints Through Parallelepipeds for Calibration and 3D Modelling”, IEEE Transactions on Pattern Analysis and Machine Intelligence—No. 5055 (Nov. 2003) 27(2) in 53 pages. [cited by applicant]
Wu et al., “Single Image 3D Interpreter Network”, arXiv e-print arXiv:1604.08685v2, Oct. 4, 2016 in 18 pages. [cited by applicant]
Xiang et al., “Data-Driven 3D Voxel Patterns for Object Category Recognition”, in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Jun. 7-12, 2015 (pp. 1903-1911). [cited by applicant]
Xiao et al., “Localizing 3D cuboids in single-view images”, in Advances in Neural Information Processing Systems 25. F. Pereira et al. [Eds.] Apr. 2013 in 9 pages. [cited by applicant]
Xiao et al., “Reconstructing the Worlds Museums,” IJCV, 2014. [cited by applicant]
Xiao et al., “Sun database: Large-scale scene recognition from abbey to zoo,” In CVPR, 2010 IEEE Conference on 2010, 3485-3492. [cited by applicant]
Yang et al., “Articulated human detection with flexible mixtures of parts”, IEEE Transactions on Pattern Analysis and Machine Intelligence. Dec. 2013; 35(12):2878-90. [cited by applicant]
Yosinski, et al., “How transferable are features in deep neural networks?,” In Advances in Neural Information Processing Systems 27 (NIPS '14), NIPS Foundation, 2014. [cited by applicant]
Zhang et al., “Estimating the 3D Layout of Indoor Scenes and its Clutter from Depth Sensors,” In ICCV, 2013. [cited by applicant]
Zhang et al., Large-scale Scene Understanding Challenge: Room Layout Estimation, 2016. [cited by applicant]
Zhao et al., “Scene Parsing by Integrating Function, Geometry and Appearance Models,” In CVPR, 2013. [cited by applicant]
Zheng et al., “Conditional Random Fields as Recurrent Neural Networks,” In CVPR, 2015. [cited by applicant]
Zheng et al., “Interactive Images: Cuboid Pro ies for Smart Image Manipulation”, ACM Trans Graph. (Jul. 2012) 31(4):99-109. [cited by applicant]
Baluja, et al., “Non-Intrusive Gaze Tracking Using Artifical Neural Networks,” Jan. 5, 1994, URL:https://apps.dtic.mil/sti/pdfs/ADA275186.pdf. [cited by applicant]
Doulamis, et al., “An Efficient Fully Unsupervised Video Object Segmentation Scheme Using an Adaptive Neural-Network Classifier Architecture,” IEEE Transactions on Neural Networks, vol. 14, No. 3, May 2003. [cited by applicant]
Sesin, et al., “Adaptive eye-gaze tracking using neural-network-based user profiles to assist people with motor disability,” Florida International University, FIU Digital Commons, 2008. JRRD, vol. 45, pp. 801-818, No. 6… [cited by applicant]
Kannan, “Eye Tracking for the iPhone using Deep Learning,” https://dspace.mit.edu/bitstream/handle/1721.1/113142/1017990444- MIT.pdf?sequence=1. Massachusetts Institute of Technology Feb. 2017. [cited by applicant]
Lam, et al., “Convolutional Neural Networks for Eye Detection in Remote Gaze Estimation Systems,” [https://www.researchgate.net/publication/44261632_Convolutional_Neural_Networks_for_Eye_Detection_in_Remote_Gaze_Estimat… [cited by applicant]
Office Action issued in counterpart Chinese Patent Application No. 201880052812.7 dated Dec. 11, 2023. (16 pages). [cited by applicant]
Office Action issued in counterpart Japanese Patent Application No. 2022-166262 dated Aug. 14, 2023. (4 pages). [cited by applicant]
Office Action issued in counterpart Korean Patent Application No. 10-2020-7004193 dated Apr. 16, 2024. (16 pages). [cited by applicant]