IP Library Granted Patent US 12,367,626
Granted Patent B2
US 12,367,626 · App. 18/055,584 · Granted Jul 22, 2025

Modifying two-dimensional images utilizing three-dimensional meshes of the two-dimensional images

Inventors: Radomir Mech (Mountain View, CA); Nathan Carr (San Jose, CA); Matheus Gadelha (San Jose, CA)
Assignee: Adobe Inc.
G06T11/60G06T7/70G06T17/20G06T19/20G06T2219/2004
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,367,626
App. No.
18/055,584
Granted
Jul 22, 2025
Kind
B2
Abstract

Methods, systems, and non-transitory computer readable storage media are disclosed for generating three-dimensional meshes representing two-dimensional images for editing the two-dimensional images. The disclosed system utilizes a first neural network to determine density values of pixels of a two-dimensional image based on estimated disparity. The disclosed system samples points in the two-dimensional image according to the density values and generates a tessellation based on the sampled points. The disclosed system utilizes a second neural network to estimate camera parameters and modify the three-dimensional mesh based on the estimated camera parameters of the pixels of the two-dimensional image. In one or more additional embodiments, the disclosed system generates a three-dimensional mesh to modify a two-dimensional image according to a displacement input. Specifically, the disclosed system maps the three-dimensional mesh to the two-dimensional image, modifies the three-dimensional mesh in response to a displacement input, and updates the two-dimensional image.

Claims (71)

1. A method comprising:

determining, by at least one processor utilizing one or more neural networks, pixel depth values representing estimated depths for pixels of a two-dimensional image;

generating, by the at least one processor utilizing the one or more neural networks, a three-dimensional mesh by incorporating depth information from the pixel depth values representing the estimated depths at the pixels of the two-dimensional image into a tessellation of the two-dimensional image according to camera parameters extracted from the two-dimensional image;

detecting, by the at least one processor, a displacement input at a two-dimensional position on the two-dimensional image;

modifying, by the at least one processor in response to the displacement input at the two-dimensional position on the two-dimensional image, the three-dimensional mesh by determining a displaced portion of the three-dimensional mesh according to a mapping of the two-dimensional position of the displacement input to a three-dimensional position of the three-dimensional mesh; and

generating, by the at least one processor, a modified two-dimensional image comprising a displacement of a portion of the two-dimensional image by re-rendering the two-dimensional image according to the displaced portion of the three-dimensional mesh and the camera parameters extracted from the two-dimensional image.

2. The method of claim 1 , wherein:

detecting the displacement input comprises determining the two-dimensional position of the displacement input within the two-dimensional image comprising two-dimensional coordinates of the displacement input relative to the two-dimensional image; and

modifying the three-dimensional mesh comprises determining three-dimensional coordinates of the displacement input relative to the three-dimensional mesh based on the mapping of the two-dimensional coordinates of the displacement input to the three-dimensional position according to the camera parameters extracted from the two-dimensional image.

3. The method of claim 1 , wherein modifying the three-dimensional mesh comprises:

determining, based on an attribute of the displacement input, that the displacement input indicates a displacement direction for a portion of the three-dimensional mesh; and

displacing the displaced portion of the three-dimensional mesh in the displacement direction.

4. The method of claim 3 , wherein modifying the three-dimensional mesh comprises:

determining, based on an additional attribute of the displacement input, that the displacement input indicates an additional displacement direction of the portion of the three-dimensional mesh; and

displacing the displaced portion of the three-dimensional mesh according to the additional displacement direction.

5. The method of claim 4 , wherein modifying the three-dimensional mesh comprises:

selecting, in response to an additional input in connection with the displacement input, a new portion of the three-dimensional mesh to displace according to a mapping of an additional two-dimensional position of the additional input to an additional three-dimensional position of the three-dimensional mesh; and

displacing the new portion of the three-dimensional mesh according to the displacement input in the displacement direction.

6. The method of claim 5 , wherein modifying the three-dimensional mesh comprises displacing, based on movement of the displacement input, one or more additional portions of the three-dimensional mesh from the portion of the three-dimensional mesh to the new portion of the three-dimensional mesh in the displacement direction.

7. The method of claim 1 , wherein:

generating the three-dimensional mesh comprises generating a mapping of the pixels of the two-dimensional image to three-dimensional coordinates of the three-dimensional mesh based on the pixel depth values and the camera parameters extracted from the two-dimensional image; and

modifying the three-dimensional mesh comprises determining the mapping of the two-dimensional position of the displacement input to the three-dimensional position based on the mapping of the pixels of the two-dimensional image to the three-dimensional coordinates of the three-dimensional mesh.

8. The method of claim 1 , wherein generating the modified two-dimensional image comprises:

providing a preview two-dimensional image comprising the displacement of the portion of the two-dimensional image in response to the displacement input;

providing, for display within a graphical user interface, an option to commit the displacement of the portion of the two-dimensional image; and

generating the modified two-dimensional image comprising the displacement of the portion of the two-dimensional image in response to detecting a selection of the option to commit the displacement of the portion of the two-dimensional image.

9. The method of claim 1 , wherein modifying the three-dimensional mesh comprises:

determining a displacement filter indicating a shape associated with the displacement input; and

displacing the displaced portion of the three-dimensional mesh according to the shape of the displacement input in a direction of the displacement input.

10. A system comprising:

a memory component; and

a processing device coupled to the memory component, the processing device to perform operations comprising:

determining, utilizing one or more neural networks, pixel depth values representing estimated depths for pixels of a two-dimensional image based on objects of the two-dimensional image;

generating, utilizing the one or more neural networks, a three-dimensional mesh by incorporating depth information from the pixel depth values corresponding to the objects of the two-dimensional image into a tessellation of the two-dimensional image according to camera parameters extracted from the two-dimensional image;

determining a three-dimensional position of the three-dimensional mesh based on a corresponding position of a displacement input at a two-dimensional position on the two-dimensional image according to a mapping of the two-dimensional image to the three-dimensional mesh;

modifying, in response to the displacement input at the two-dimensional position on the two-dimensional image, the three-dimensional mesh by determining a displaced portion of the three-dimensional mesh at the three-dimensional position of the three-dimensional mesh; and

generating a modified two-dimensional image comprising a displacement of a portion of the two-dimensional image by re-rendering the two-dimensional image according to the displaced portion of the three-dimensional mesh at the three-dimensional position of the three-dimensional mesh and the camera parameters extracted from the two-dimensional image.

11. The system of claim 10 , wherein determining the three-dimensional position of the three-dimensional mesh comprises:

determining the two-dimensional position of the displacement input comprises determining two-dimensional coordinates of the displacement input relative to the two-dimensional image; and

determining the three-dimensional position of the three-dimensional mesh based on the two-dimensional coordinates of the displacement input relative to the two-dimensional image and a projection from the two-dimensional image onto the three-dimensional mesh corresponding to the mapping of the two-dimensional image to the three-dimensional mesh.

12. The system of claim 11 , wherein modifying the three-dimensional mesh comprises:

determining that the displacement input indicates a displacement of a portion of the three-dimensional mesh according to the three-dimensional position of the three-dimensional mesh; and

determining the displaced portion of the three-dimensional mesh according to a movement of the displacement input and the projection from the two-dimensional image onto the three-dimensional mesh.

13. The system of claim 12 , wherein generating the modified two-dimensional image comprises:

determining, based on the projection from the two-dimensional image onto the three-dimensional mesh, an additional two-dimensional position of the two-dimensional image corresponding to the displaced portion of the three-dimensional mesh; and

generating, based on the additional two-dimensional position of the two-dimensional image, the modified two-dimensional image comprising the displacement of the portion of the two-dimensional image according to the displaced portion of the three-dimensional mesh.

14. The system of claim 10 , wherein modifying the three-dimensional mesh comprises:

determining a direction of movement of the displacement input within a graphical user interface displaying the two-dimensional image;

determining a displacement height and a displacement radius based on the direction of movement of the displacement input; and

determining the displaced portion of the three-dimensional mesh based on the displacement height and the displacement radius.

15. The system of claim 10 , wherein modifying the three-dimensional mesh comprises:

determining one or more normal values corresponding to one or more vertices or one or more faces at the three-dimensional position of the three-dimensional mesh; and

determining, in response to the displacement input, the displaced portion of the three-dimensional mesh in one or more directions corresponding to the one or more normal values corresponding to the one or more vertices or the one or more faces.

16. The system of claim 10 , wherein modifying the three-dimensional mesh comprises:

determining a displacement direction based on a surface of a selected portion of the three-dimensional mesh; and

determining, in response to the displacement input, the displaced portion of the three-dimensional mesh in the displacement direction.

17. A non-transitory computer readable medium comprising instructions, which when executed by a processing device, cause the processing device to perform operations comprising:

determining, utilizing one or more neural networks, pixel depth values representing estimated depths for pixels of a two-dimensional image;

generating, utilizing the one or more neural networks, a three-dimensional mesh by incorporating depth information from the pixel depth values representing the estimated depths at the pixels of the two-dimensional image into a tessellation of the two-dimensional image according to camera parameters extracted from the two-dimensional image;

detecting a displacement input at a two-dimensional position on the two-dimensional image;

modifying, in response to the displacement input at the two-dimensional position on the two-dimensional image, the three-dimensional mesh by determining a displaced portion of the three-dimensional mesh according to a mapping of the two-dimensional position of the displacement input to a three-dimensional position of the three-dimensional mesh; and

generating a modified two-dimensional image comprising a displacement of a portion of the two-dimensional image by re-rendering the two-dimensional image according to the displaced portion of the three-dimensional mesh and the camera parameters extracted from the two-dimensional image.

18. The non-transitory computer readable medium of claim 17 , wherein generating the three-dimensional mesh comprises:

generating the three-dimensional mesh by determining displacement of vertices of the tessellation of the two-dimensional image based on the pixel depth values and the camera parameters extracted from the two-dimensional image; or

generating the three-dimensional mesh by generating the tessellation based on a plurality of points sampled in the two-dimensional image according to density values determined from the pixel depth values of the two-dimensional image.

19. The non-transitory computer readable medium of claim 17 , wherein modifying the three-dimensional mesh comprises:

determining a projection from the two-dimensional image onto the three-dimensional mesh; and

determining the mapping of the two-dimensional position of the displacement input to the three-dimensional position of the three-dimensional mesh, according to the projection from the two-dimensional image onto the three-dimensional mesh.

20. The non-transitory computer readable medium of claim 19 , wherein modifying the three-dimensional mesh comprises:

determining, based on the projection from the two-dimensional image onto the three-dimensional mesh, movement of the displacement input relative to the two-dimensional image and a corresponding movement of the displacement input relative to the three-dimensional mesh; and

determining the displaced portion of the three-dimensional mesh based on the corresponding movement of the displacement input relative to the three-dimensional mesh.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 15, 2022
From: MECH, RADOMIR; CARR, NATHAN; GADELHA, MATHEUS
To: ADOBE INC.
Reel/Frame 061775/0790 →
Continuity (1)
Related Publication 20240161366A1 · May 16, 2024
References Cited (130)
US 8289318B1 · Hadap · 2012 [cited by examiner]
US 8830237B2 · Zimmerman · 2014 [cited by applicant]
US 9153209B2 · Dmitriev · 2015 [cited by applicant]
US 10043279B1 · Eshet · 2018 [cited by applicant]
US 10460214B2 · Lu et al. · 2019 [cited by applicant]
US 10679046B1 · Black et al. · 2020 [cited by applicant]
US 10691286B2 · King et al. · 2020 [cited by applicant]
US 10930075B2 · Costa et al. · 2021 [cited by applicant]
US 11094083B2 · Eisenmann · 2021 [cited by examiner]
US 11217035B2 · Hajjar · 2022 [cited by applicant]
US 11263823B2 · Gauseback et al. · 2022 [cited by applicant]
US 11270507B1 · Lambert · 2022 [cited by examiner]
US 11328535B1 · Guo · 2022 [cited by examiner]
US 11494995B2 · Berkebile · 2022 [cited by applicant]
US 11514638B2 · Lafter et al. · 2022 [cited by applicant]
US 11741668B2 · Jones · 2023 [cited by examiner]
US 11869152B2 · Koh · 2024 [cited by examiner]
US 11881049B1 · Soltz · 2024 [cited by applicant]
US 20110018873A1 · Chang · 2011 [cited by examiner]
US 20110298799A1 · Mariani · 2011 [cited by examiner]
US 20120081357A1 · Habbecke · 2012 [cited by examiner]
US 20130135305A1 · Bystrov · 2013 [cited by examiner]
US 20140359536A1 · Cheng et al. · 2014 [cited by applicant]
US 20160035068A1 · Wilensky · 2016 [cited by examiner]
US 20160063669A1 · Wilensky · 2016 [cited by examiner]
US 20160119670A1 · Izutsu et al. · 2016 [cited by applicant]
US 20180158230A1 · Yan · 2018 [cited by examiner]
US 20180218535A1 · Ceylan · 2018 [cited by examiner]
US 20180329485A1 · Carothers et al. · 2018 [cited by applicant]
US 20190026958A1 · Gausebeck · 2019 [cited by examiner]
US 20190065026A1 · Kiemele et al. · 2019 [cited by applicant]
US 20190354699A1 · Pekelny et al. · 2019 [cited by applicant]
US 20200020173A1 · Sharif · 2020 [cited by examiner]
US 20200175756A1 · Crowe · 2020 [cited by examiner]
US 20210005026A1 · Lesbordes · 2021 [cited by applicant]
US 20210074062A1 · Madonna · 2021 [cited by examiner]
US 20210097776A1 · Faulkner et al. · 2021 [cited by applicant]
US 20210335039A1 · Jones et al. · 2021 [cited by applicant]
US 20210343080A1 · Kim · 2021 [cited by examiner]
US 20210398351A1 · Papandreou · 2021 [cited by examiner]
US 20220068007A1 · Lafer · 2022 [cited by examiner]
US 20220222887A1 · Hundal · 2022 [cited by examiner]
US 20220284613A1 · Yin et al. · 2022 [cited by applicant]
US 20220292352A1 · Jourdan et al. · 2022 [cited by applicant]
US 20220414834A1 · Du et al. · 2022 [cited by applicant]
US 20230033956A1 · Yong et al. · 2023 [cited by applicant]
US 20230080584A1 · Zohar · 2023 [cited by examiner]
US 20230117686A1 · Jain et al. · 2023 [cited by applicant]
US 20230140460A1 · Munkberg · 2023 [cited by examiner]
US 20230143034A1 · Wu et al. · 2023 [cited by applicant]
US 20230230321A1 · Schreckenbert et al. · 2023 [cited by applicant]
US 20230239458A1 · Cantero Clares · 2023 [cited by examiner]
US 20230281925A1 · Aigerman · 2023 [cited by examiner]
US 20230290090A1 · Su · 2023 [cited by applicant]
US 20230306686A1 · Zangenehpour · 2023 [cited by examiner]
US 20230326028A1 · Zhang et al. · 2023 [cited by applicant]
US 20230368339A1 · Zheng et al. · 2023 [cited by applicant]
US 20240013462A1 · Seol · 2024 [cited by examiner]
US 20240037717A1 · Amirghodsi et al. · 2024 [cited by applicant]
US 20240087265A1 · Park · 2024 [cited by examiner]
US 20240112396A1 · Spencer · 2024 [cited by applicant]
US 20240135572A1 · Singh et al. · 2024 [cited by applicant]
US 20240144520A1 · Gori · 2024 [cited by examiner]
US 20240144586A1 · Hold-Geoffroy · 2024 [cited by examiner]
US 20240144623A1 · Gori · 2024 [cited by examiner]
US 20240161320A1 · Gadelha · 2024 [cited by examiner]
US 20240161366A1 · Mech · 2024 [cited by examiner]
US 20240161405A1 · Mech · 2024 [cited by examiner]
US 20240161406A1 · Mech · 2024 [cited by examiner]
Office Action as received in CN Application No. 2023112860680.7 dated Nov. 2, 2023. [cited by applicant]
Xie et al., SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers, 2021. [cited by applicant]
Iro Armeni, Sasha Sax, Amir R Zamir, and Silvio Savarese. Joint 2d-3d-semantic data for indoor scene understanding. arXiv preprint arXiv:1702.01105, 2017. [cited by applicant]
Olga Barinova, Victor Lempitsky, Elena Tretiak, and Push—meet Kohli. Geometric image parsing in man-made environments. In European conference on computer vision, pp. 57-70. Springer, 2010. [cited by applicant]
Shariq Farooq Bhat, Ibraheem Alhashim, and Peter Wonka. Adabins: Depth estimation using adaptive bins. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 4009-4018, 2021. [cited by applicant]
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niessner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang. Matterport3d: Learning from rgb-d data in indoor environments. International Conferen… [cited by applicant]
Jia-Ren Chang and Yong-Sheng Chen. Pyramid stereo matching network. In Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 5410-5418, 2018. [cited by applicant]
Antonio Criminisi, Ian Reid, and Andrew Zisserman. Single view metrology. International Journal of Computer Vision, 40(2):123-148, 2000. [cited by applicant]
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition, pp. 248-255. Ieee, 2009. [cited by applicant]
Patrick Denis, James H Elder, and Francisco J Estrada. Efficient edge-based methods for estimating manhattan frames in urban imagery. In European conference on computer vision, pp. 197-210. Springer, 2008. [cited by applicant]
Jonathan Deutscher, Michael Isard, and John MacCormick. Automatic camera calibration from a single manhattan image. In European Conference on Computer Vision, pp. 175-188. Springer, 2002. [cited by applicant]
Derek Hoiem, Alexei A Efros, and Martial Hebert. Putting objects in perspective. International Journal of Computer Vision, 80(1):3-15, 2008. [cited by applicant]
Yannick Hold-Geoffroy, Kalyan Sunkavalli, Jonathan Eisenmann, Matthew Fisher, Emiliano Gambaretto, Sunil Hadap, and Jean-Francois Lalonde. A perceptual measure for deep single image camera calibration. In Proceedings of… [cited by applicant]
Abhishek Kar, Shubham Tulsiani, Joao Carreira, and Jitendra Malik. Amodal completion and size constancy in natural scenes. In Proceedings of the IEEE international conference on computer vision, pp. 127-135, 2015. [cited by applicant]
Johannes Kopf, Xuejian Rong, and Jia-Bin Huang. Robust consistent video depth estimation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1611-1621, 2021. [cited by applicant]
Byeong-Uk Lee, Kyunghyun Lee, and In So Kweon. Depth completion using plane-residual representation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 13916-13925, 2021. [cited by applicant]
Hyunjoon Lee, Eli Shechtman, Jue Wang, and Seungyong Lee. Automatic upright adjustment of photographs with robust camera calibration. IEEE transactions on pattern analysis and machine intelligence, 36(5):833-844, 2013. [cited by applicant]
Jin Han Lee, Myung-Kyu Han, Dong Wook Ko, Il Hong Suh. From big to small: Multi-scale local planar guidance for monocular depth estimation. arXiv preprint arXiv:1907.10326, 2019. [cited by applicant]
Zhengqi Li, Tali Dekel, Forrester Cole, Richard Tucker, Noah Snavely, Ce Liu, and William T Freeman. Learning the depths of moving people by watching frozen people. In Proceedings of the IEEE/CVF conference on computer … [cited by applicant]
Xuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen, and Johannes Kopf. Consistent video depth estimation. ACM Transactions on Graphics (ToG), 39(4):71-1, 2020. [cited by applicant]
Fangchang Ma and Sertac Karaman. Sparse-to-dense: Depth prediction from sparse depth samples and a single image. In 2018 IEEE international conference on robotics and automation (ICRA), pp. 4796-4803. IEEE, 2018. [cited by applicant]
Ricardo Martin-Brualla, Noha Radwan, Mehdi SM Sajjadi, Jonathan T Barron, Alexey Dosovitskiy, and Daniel Duckworth. Nerf in the wild: Neural radiance fields for unconstrained photo collections. In Proceedings of the IEE… [cited by applicant]
Nikolaus Mayer, Eddy Ilg, Philip Hausser, Philipp Fischer, 942 Daniel Cremers, Alexey Dosovitskiy, and Thomas Brox. A large dataset to train convolutional networks for disparity, optical flow, and scene flow estimation.… [cited by applicant]
Ben Mildenhall, Pratul P Srinivasan, Matthew Tancik, Jonathan T Barron, Ravi Ramamoorthi, and Ren Ng. Nerf: Representing scenes as neural radiance fields for view synthesis. Communications of the ACM, 65(1):99-106, 2021. [cited by applicant]
Thomas Müller, Alex Evans, Christoph Schied, and Alexander Keller. Instant neural graphics primitives with a multiresolution hash encoding. arXiv preprint arXiv:2201.05989, 2022. [cited by applicant]
Jeong Joon Park, Peter Florence, Julian Straub, Richard Newcombe, and Steven Lovegrove. Deepsdf: Learning continuous signed distance functions for shape representation. In Proceedings of the IEEE/CVF conference on compu… [cited by applicant]
Rene' Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun. Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer. IEEE transactions on pattern analysis a… [cited by applicant]
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. U-net: Convolutional networks for biomedical image segmentation. In International Conference on Medical image computing and computer-assisted intervention, pp. 234-241… [cited by applicant]
Benjamin Ummenhofer, Huizhong Zhou, Jonas Uhrig, Nikolaus Mayer, Eddy Ilg, Alexey Dosovitskiy, and Thomas Brox. Demon: Depth and motion network for learning monocular stereo. In Proceedings of the IEEE conference on com… [cited by applicant]
Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. Pvt v2: Improved baselines with pyramid vision transformer. Computational Visual Media, 8(3):415-424, 2022. [cited by applicant]
Scott Workman, Connor Greenwell, Menghua Zhai, Ryan Baltenberger, and Nathan Jacobs. Deepfocal: A method for direct focal length estimation. In 2015 IEEE International Conference on Image Processing (ICIP), pp. 1369-137… [cited by applicant]
Scott Workman, Menghua Zhai, and Nathan Jacobs. Horizon lines in the wild. arXiv preprint arXiv:1604.02129, 2016. [cited by applicant]
Yao Yao, Zixin Luo, Shiwei Li, Tian Fang, and Long Quan. Mvsnet: Depth inference for unstructured multi-view stereo. In Proceedings of the European conference on computer vision (ECCV), pp. 767-783, 2018. [cited by applicant]
Wei Yin, Jianming Zhang, Oliver Wang, Simon Niklaus, Long Mai, Simon Chen, and Chunhua Shen. Learning to recover 3d scene shape from a single image. In Proceedings of the IEEE/CVF Conference on Computer Vision and Patte… [cited by applicant]
Menghua Zhai, Scott Workman, and Nathan Jacobs. Detecting vanishing points using global image context in a non-manhattan world. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 5657-… [cited by applicant]
Tinghui Zhou, Matthew Brown, Noah Snavely, and David G Lowe. Unsupervised learning of depth and ego-motion from video. In Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 1851-1858, 201… [cited by applicant]
Rui Zhu, Xingyi Yang, Yannick Hold-Geoffroy, Federico Perazzi, Jonathan Eisenmann, Kalyan Sunkavalli, and Manmohan Chandraker. Single view metrology in the wild. In European Conference on Computer Vision, pp. 316-333. S… [cited by applicant]
Zhe Cao, Gines Hildago, Tomas Simon, Shih-En Wei, Yaser Sheikh-OpenPose-Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields, CVPS (2019). [cited by applicant]
Kevin Lin, Lijuan Wang, Zicheng Liu—End-to-End Human Pose and Mesh Reconstruction with Transformers, CVPR (2021). [cited by applicant]
Kripasindhu Sarkar, Vladislav Golyanik, Lingjie Liu, Christian Theobalt—Style and Pose Control for Image Synthesis of Humans from a Single Monocular View, Sarkar et al., 2021. [cited by applicant]
Enric Corona, Albert Pumarola, Guillem Alenyà, Gerard Pons-Moll, Francesc Moreno-Noguer—SMPLicit: Topology-aware Generative Model for Clothed People, Corona et al., 2021. [cited by applicant]
Combined Search Report and Written Opinion as received in GB 2312456.3 dated Feb. 15, 2024. [cited by applicant]
Lili Wang et al., “Bidirectional Shadow Rendering for Interactive Mixed 360° Videos”, 2021 IEEE Virtual Reality and 3D User Interfaces (VR), p. 170-178, 2021. [cited by applicant]
Combined Search Report and Written Opinion as received in GB 2402407.7 dated Jul. 9, 2024. [cited by applicant]
U.S. Appl. No. 18/055,590, Apr. 8, 2024, Office Action. [cited by applicant]
U.S. Appl. No. 18/304,162, Apr. 25, 2024, Notice of Allowance. [cited by applicant]
U.S. Appl. No. 18/055,585, Aug. 26, 2024, Office Action. [cited by applicant]
U.S. Appl. No. 18/055,590, Sep. 17, 2024, Office Action. [cited by applicant]
U.S. Appl. No. 18/304,162, Sep. 11, 2024, Notice of Allowance. [cited by applicant]
Combined Search Report and Written Opinion as received in GB 2403090.0 dated Jan. 28, 2025. [cited by applicant]
Chung-Yi Weng et al. “Photo Wake-Up: 3D Character Animation from a Single Photo”, IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 5908-5917 (2019). [cited by applicant]
Federica Bogo et al. “Keep it SMPL: Automatic Estimation of 3D Human Pose and Shape from a Single Image”, ECCV, pp. 561-578 (2016). [cited by applicant]
Imry Kissos et al. “Beyond Weak Perspective for Monocular 3D Human Pose Estimation”, Springer Nature, pp. 541-554 (2021). [cited by applicant]
Shih-En Wei et al. “Convolutional Pose Machines”, IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 4724-4732 (2016). [cited by applicant]
U.S. Appl. No. 18/055,585, Jan. 13, 2025, Notice of Allowance. [cited by applicant]
U.S. Appl. No. 18/055,590, Feb. 3, 2025, Notice of Allowance. [cited by applicant]
U.S. Appl. No. 18/304,147, Jan. 21, 2025, Office Action. [cited by applicant]
S. Lv, X. Yang, L. Gu, X. Xing, L. Pan and M. Fang, “Delaunay Mesh Reconstruction from 3D Medical Images Based on Centroidal Voronoi Tessellations,” 2009 International Conference on Computational Intelligence and Softwa… [cited by applicant]
Shrivastava, S. (2020). Stereo Vision Based Object Detection Using V-Disparity and 3D Density-Based Clustering. In: Arai, K., Kapoor, S. (eds) Advances in Computer Vision. CVC 2019. Advances in Intelligent Systems and C… [cited by applicant]
Watson, Jamie, Oisin Mac Aodha, Daniyar Turmukhambetov, Gabriel J. Brostow and Michael Firman. University of Edinburgh. “Learning Stereo from Single Images.” European Conference on Computer Vision (2020). Aug. 2020. (Ye… [cited by applicant]
U.S. Appl. No. 18/055,594, Mail Date Mar. 5, 2025, Notice of Allowance. [cited by applicant]