IP Library Granted Patent US 12,694,549
Granted Patent B2
US 12,694,549 · App. 18/527,756 · Granted Jul 28, 2026

Depth estimation using variant features

Inventors: Akash Mittal (Champaign, IL); Brian Harms (San Jose, CA); Nigel Clarke (Sunnyvale, CA)
Assignee: Samsung Electronics Co., Ltd.
G06T7/55G06T7/74G06T2207/10016G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,694,549
App. No.
18/527,756
Granted
Jul 28, 2026
Kind
B2
Abstract

In one embodiment, an apparatus includes accessing an image, captured by a first camera, of an object that includes a size-invariant feature and a size-variant feature and determining a size of the size-invariant feature in the image. The method further includes determining, based on the determined size of the size-invariant feature in the image, a distance between the camera and the object and determining a size of the size-variant feature in the image. The method further includes determining, based on the determined distance between the camera and the object and based on the determined size of the size-variant feature in the image, an actual size of the size-variant feature.

Claims (55)

1 . A method comprising:

accessing an initial image, captured by a first camera, of an object comprising a size-invariant feature and a size-variant feature;

determining a size of the size-invariant feature in the initial image;

determining, based on the determined size of the size-invariant feature in the initial image, a distance between the camera and the object;

determining a size of the size-variant feature in the initial image;

determining, based on the determined distance between the camera and the object and based on the determined size of the size-variant feature in the initial image, an actual size of the size-variant feature;

accessing a subsequent image, captured by a second camera, of at least the size-variant feature;

determining a size of the size-variant feature in the subsequent image; and

determining, based on the determined actual size of the size-variant feature and the determined size of the size-variant feature in the subsequent image, a distance between the second camera and an object comprising the size-variant feature.

2 . The method of claim 1 , further comprising:

determining that an appearance of the size-invariant feature in the initial image satisfies one or more visibility criteria; and

determining the size of the size-invariant feature in the initial image in response to the determination that the size-invariant feature in the initial image satisfies the one or more visibility criteria.

3 . The method of claim 1 , wherein the accessed initial image is a first one of a plurality of accessed initial images of the object, the method further comprising:

determining, for each of the plurality of accessed initial images, a visibility of the size-invariant feature in that image;

selecting, based on the determined visibility of the size-invariant feature in each of the plurality of accessed initial images, a particular initial image as the first one of the plurality initial images.

4 . The method of claim 1 , wherein the object comprising the size-variant feature in the subsequent image is the same object comprising the size-invariant feature and the size-variant feature in the initial image.

5 . The method of claim 1 , wherein the first camera is the same as the second camera.

6 . The method of claim 4 , wherein the size-variant feature is a first size-variant feature of a plurality of size-variant features of the object, the method further comprising:

determining, for each of the plurality of size-variant features, a visibility of the size-variant feature in the subsequent image; and

selecting, based on the determined visibility of the size-variant features, a particular size-variant feature as the first size-variant feature.

7 . The method of claim 6 , wherein the determined visibility is based on one or more of:

an occlusion of the respective size-variant feature;

a coverage area of the respective size-variant feature relative to a coverage threshold;

an angle between a plane of the subsequent image and the respective size-variant feature; or

a rotation of the object.

8 . The method of claim 4 , further comprising determining, based on the determined distance between the second camera and the object, a distance between the object and a device.

9 . The method of claim 8 , further comprising automatically adjusting a position of the device based on the determined distance between the object and the device.

10 . The method of claim 9 , wherein the object comprises a person's face, and the device comprises a display screen.

11 . The method of claim 10 , further comprising automatically adjusting the position of the display screen based on one or more preferences of the person.

12 . The method of claim 10 , wherein the display screen is part of a computer monitor.

13 . The method of claim 10 , further comprising:

identifying the person; and

associating the actual size of the size-variant feature with the identity of the person.

14 . An apparatus comprising one or more non-transitory computer readable storage media storing instructions; and one or more processors coupled to the non-transitory computer readable storage media, the one or more processors operable to execute the instructions to:

access an initial image, captured by a first camera, of an object comprising a size-invariant feature and a size-variant feature;

determine a size of the size-invariant feature in the initial image;

determine, based on the determined size of the size-invariant feature, a distance between the camera and the object;

determine a size of the size-variant feature in the initial image;

determine, based on the determined distance between the camera and the object and based on the determined size of the size-variant feature, an actual size of the size-variant feature

access a subsequent image, captured by a second camera, of at least the size-variant feature;

determine a size of the size-variant feature in the subsequent image; and

determine, based on the determined actual size of the size-variant feature and the determined size of the size-variant feature in the subsequent image, a distance between the second camera and an object comprising the size-variant feature.

15 . The apparatus of claim 14 , wherein the object comprising the size-variant feature in the subsequent image is the same object comprising the size-invariant feature and the size-variant feature in the initial image.

16 . A method comprising:

accessing a subsequent image, captured by a second camera, of an object comprising a size-variant feature;

determining a size of the size-variant feature in the subsequent image;

accessing an actual size of the size-variant feature, wherein the actual size of the size-variant feature is based on:

a determined size of the size-invariant feature in an initial image, captured by a first camera, of an object comprising a size-invariant feature and the size-variant feature;

a distance, determined based on the determined size of the size of the size-invariant feature in the initial image captured by the first camera, between the first camera and the object in the initial image; and

a size of the size-variant feature in the initial image captured by the first camera; and

determining, based on the actual size of the size-variant feature and the determined size of the size-variant feature in the subsequent image, a distance between the second camera and the object comprising the size-variant feature.

17 . The method of claim 16 , wherein the object comprising the size-variant feature in the subsequent image is the same object comprising the size-invariant feature and the size-variant feature in the initial image.

18 . The method of claim 16 , further comprising determining, based on the determined distance between the second camera and the object, a distance between the object and a device.

19 . The method of claim 18 , further comprising automatically adjusting a position of the device based on the determined distance between the object and the device.

20 . The method of claim 19 , wherein the object comprises a person's face, and the device comprises a display screen.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 4, 2023
From: MITTAL, AKASH; HARMS, BRIAN; CLARKE, NIGEL
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 065749/0701 →
Continuity (1)
Related Publication 20250182309A1 · Jun 5, 2025
References Cited (32)
US 8091842B2 · Thomas · 2012 [cited by applicant]
US 8939500B2 · Voigt · 2015 [cited by applicant]
US 9044172B2 · Baxi · 2015 [cited by applicant]
US 9703444B2 · Nicholson · 2017 [cited by applicant]
US 11004422B1 · Bull · 2021 [cited by applicant]
US 11734854B2 · Huelsdunk · 2023 [cited by applicant]
US 20150293588A1 · Strupczewski · 2015 [cited by examiner]
US 20170169570A1 · Vashishtha · 2017 [cited by examiner]
US 20180253143A1 · Saleem · 2018 [cited by applicant]
US 20200380901A1 · Ryu · 2020 [cited by applicant]
US 20220375126A1 · Huelsdunk · 2022 [cited by examiner]
US 20230103129A1 · Morrell · 2023 [cited by examiner]
US 20230116779A1 · Hoyle · 2023 [cited by examiner]
US 20230127218A1 · Hsieh · 2023 [cited by applicant]
US 20240031619A1 · Jayaram · 2024 [cited by applicant]
CN 111815757B · 2023 [cited by applicant]
CN 116091577A · 2023 [cited by applicant]
CN 116402870A · 2023 [cited by applicant]
KR 1020200049958A · 2020 [cited by applicant]
KR 1020220048109A · 2022 [cited by applicant]
WO WO2022225115A1 · 2022 [cited by examiner]
“MediaPipe Iris: Real-time Iris Tracking & Depth Estimation,” Posted by Andrey Vakunov and Dmitry Lagun, Research Engineers, Google Research (https://ai.googleblog.com/2020/08/mediapipe-iris-real-time-iris-tracking.html… [cited by applicant]
Shin et al., “Slow Robots for Unobtrusive Posture Correction” (CHI '19: ACM CHI Conference on Human Factors in Computing Systems), Apr. 2019. [cited by applicant]
Screen captures from YouTube video clip entitled “Roco—Robotic Computer Monitor” 4 pages, uploaded on Jan. 24, 2014 by user “gerbilproductions.” Retrieved from internet: <https://www.youtube.com/watch?v=ljim-BW8Y8E&t=1s… [cited by applicant]
Screen captures from YouTube video clip entitled “DOT Stand V1 (English Ver.),” 32 pages, uploaded on Aug. 7, 2022 by user “DOT Heal”. Retrieved from internet: <https://www.youtube.com/watch?v=Xcs3YkJeCLQ>. [cited by applicant]
Kan, “This LG Monitor Can Continuously Move Itself to Meet Your Eye Level” PCMag (available at: https://www.pcmag.com/news/this-lg-monitor-can-continuously-move-itself-to-meet-your-eye-level), Aug. 2022. [cited by applicant]
Non-final office action in U.S. Appl. No. 18/527,899, Aug. 26, 2024. [cited by applicant]
Notice of Allowance in U.S. Appl. No. 18/527,899, Dec. 11, 2024. [cited by applicant]
International Search Report in Application No. PCT/KR2024/019677, Mar. 12, 2025. [cited by applicant]
Written Opinion of the International Searching Authority in Application No. PCT/KR2024/019677, Mar. 12, 2025. [cited by applicant]
P.J.A. Alphonse et al., Depth perception in single rgb camera system using lens aperture and object size: a geometrical approach for depth estimation, SN Appl. Sci. vol. 3, 595 (2021), May 1, 2021. [cited by applicant]
Yuan Cui et al., Camera distance helps 3D hand pose estimated from a single RGB image, ScienceDirect, Graphical Models, vol. 127, 101179, May 16, 2023. [cited by applicant]