IP Library Granted Patent US 12,573,153
Granted Patent B2
US 12,573,153 · App. 18/218,465 · Granted Mar 10, 2026

Determining traversable space from single images

Inventors: James Watson (London, GB); Michael David Firman (London, GB); Aaron Monszpart (London, GB); Gabriel J. Brostow (London, GB)
Assignee: Niantic Spatial, Inc.
G06T19/006G06N20/00G06T7/10G06T7/50G06T19/20G06V10/26G06V10/774G06V20/10G06T2207/20081G06T2219/2004
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,573,153
App. No.
18/218,465
Granted
Mar 10, 2026
Kind
B2
Abstract

A model predicts the geometry of both visible and occluded traversable surfaces from input images. The model may be trained from stereo video sequences, using camera poses, per-frame depth, and semantic segmentation to form training data, which is used to supervise an image to image network. In various embodiments, the model is applied to a single RGB image depicting a scene to produce information describing traversable space of the scene that includes occluded traversable. The information describing traversable space can include a segmentation mask of traversable space (both visible and occluded) and non-traversable space and a depth map indicating an estimated depth to traversable surfaces corresponding to each pixel determined to correspond to traversable space.

Claims (68)

1 . A non-transitory computer-readable storage medium storing instructions that, when executed by a computing device, cause the computing device to perform operations comprising:

receiving an image comprising a plurality of pixels, the plurality of pixels comprising information representing an object in a scene and depth information for the scene;

inputting the image into a traversability model configured to identify traversable and non-traversable space in the scene, the traversability model configured to:

identify a traversable surface in the scene based on the depth information represented by the plurality of pixels in the image, the traversable surface comprising traversable space and non-traversable space,

identify pixels in the image representing the object as object pixels, the object pixels comprising a subset of object pixels representing a footprint of the object on the traversable surface, and

determine non-traversable space and traversable space on the traversable surface based on the depth information for the object pixels, the non-traversable space including space represented by the subset of object pixels representing the footprint of the object on the traversable surface;

determining a traversable path through the scene for a virtual agent based on the determined non-traversable space and traversable space; and

displaying the virtual agent as traversing along the traversable path.

2 . The non-transitory computer-readable storage medium of claim 1 , wherein the traversability model is further configured to:

determine, based on the object pixels, the object is occluding space in the scene; and

determine, based on the depth information for the object pixels, the traversable space includes the space determined to be occluded by the object in the image.

3 . The non-transitory computer-readable storage medium of claim 2 , wherein the determined traversable path includes traversable space occluded by the object.

4 . The non-transitory computer-readable storage medium of claim 1 , wherein identifying the traversable surface in the scene based on depth information further comprises:

generating a depth map for the scene based on the depth information for the plurality of pixels;

determining traversable space in the scene based on the depth map and the footprint of the object; and

generating the traversable surface based on the depth information for the determined traversable space in the scene.

5 . The non-transitory computer-readable storage medium of claim 1 , wherein identifying the traversable surface in the scene based on depth information further comprises:

identifying pixels in the image representing a static object, the static object comprising a surface; and

determining the traversable surface is the surface of the static object.

6 . The non-transitory computer-readable storage medium of claim 1 , wherein the object pixels comprise an additional subset of object pixels positioned above a plane of the traversable surface, and the traversability model is configured to:

generate the footprint of the object on the traversable surface by projecting the additional subset of object pixels above the plane of the traversable surface onto the plane of the traversable surface.

7 . The non-transitory computer-readable storage medium of claim 1 , wherein the traversability model is configured to identify a type of the object and the footprint of the object on the traversable surface is based on the type of the object.

8 . A system comprising:

a computing device; and

a non-transitory computer-readable storage medium storing computer instructions that, when executed by the computing device, cause the computing device to perform operations comprising:

receiving an image comprising a plurality of pixels, the plurality of pixels comprising information representing an object in a scene and depth information for the scene;

inputting the image into a traversability model configured to identify traversable and non-traversable space in the scene, the traversability model configured to:

identify a traversable surface in the scene based on the depth information represented by the plurality of pixels in the image, the traversable surface comprising traversable space and non-traversable space,

identify pixels in the image representing the object as object pixels, the object pixels comprising a subset of object pixels representing a footprint of the object on the traversable surface, and

determine non-traversable space and traversable space on the traversable surface based on the depth information for the object pixels, the non-traversable space including space represented by the subset of object pixels representing the footprint of the object on the traversable surface;

determining a traversable path through the scene for a virtual agent based on the determined non-traversable space and traversable space; and

displaying the virtual agent as traversing along the traversable path.

9 . The system of claim 8 , wherein the traversability model is further configured to:

determine, based on the object pixels, the object is occluding space in the scene; and

determine, based on the depth information for the object pixels, the traversable space includes the space determined to be occluded by the object in the image.

10 . The system of claim 9 , wherein the determined traversable path includes traversable space occluded by the object.

11 . The system of claim 8 , wherein identifying the traversable surface in the scene based on depth information further comprises:

generating a depth map for the scene based on the depth information for the plurality of pixels;

determining traversable space in the scene based on the depth map and the footprint of the object; and

generating the traversable surface based on the depth information for the determined traversable space in the scene.

12 . The system of claim 8 , wherein identifying the traversable surface in the scene based on depth information further comprises:

identifying pixels in the image representing a static object, the static object comprising a surface; and

determining the traversable surface is the surface of the static object.

13 . The system of claim 8 , wherein the object pixels comprise an additional subset of object pixels positioned above a plane of the traversable surface, and the traversability model is configured to:

generate the footprint of the object on the traversable surface by projecting the additional subset of object pixels above the plane of the traversable surface onto the plane of the traversable surface.

14 . The system of claim 8 , wherein the traversability model is configured to identify a type of the object and the footprint of the object on the traversable surface is based on the type of the object.

15 . A method comprising:

receiving an image comprising a plurality of pixels, the plurality of pixels comprising information representing an object in a scene and depth information for the scene;

inputting the image into a traversability model configured to identify traversable and non-traversable space in the scene, the traversability model configured to:

identify a traversable surface in the scene based on the depth information represented by the plurality of pixels in the image, the traversable surface comprising traversable space and non-traversable space,

identify pixels in the image representing the object as object pixels, the object pixels comprising a subset of object pixels representing a footprint of the object on the traversable surface, and

determine non-traversable space and traversable space on the traversable surface based on the depth information for the object pixels, the non-traversable space including space represented by the subset of object pixels representing the footprint of the object on the traversable surface;

determining a traversable path through the scene for a virtual agent based on the determined non-traversable space and traversable space; and

displaying the virtual agent as traversing along the traversable path.

16 . The method of claim 15 , wherein the traversability model is further configured to:

determine, based on the object pixels, the object is occluding space in the scene; and

determine, based on the depth information for the object pixels, the traversable space includes the space determined to be occluded by the object in the image.

17 . The method of claim 16 , wherein the determined traversable path includes traversable space occluded by the object.

18 . The method of claim 15 , wherein identifying the traversable surface in the scene based on depth information further comprises:

generating a depth map for the scene based on the depth information for the plurality of pixels;

determining traversable space in the scene based on the depth map and the footprint of the object; and

generating the traversable surface based on the depth information for the determined traversable space in the scene.

19 . The method of claim 15 , wherein identifying the traversable surface in the scene based on depth information further comprises:

identifying pixels in the image representing a static object, the static object comprising a surface; and

determining the traversable surface is the surface of the static object.

20 . The method of claim 15 , wherein the object pixels comprise an additional subset of object pixels positioned above a plane of the traversable surface, and the traversability model is configured to:

generate the footprint of the object on the traversable surface by projecting the additional subset of object pixels above the plane of the traversable surface onto the plane of the traversable surface.

21 . The method of claim 15 , wherein the traversability model is configured to identify a type of the object and the footprint of the object on the traversable surface is based on the type of the object.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 16, 2025
From: NIANTIC, INC.
To: NIANTIC SPATIAL, INC.
Reel/Frame 071555/0833 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 16, 2023
From: WATSON, JAMES; FIRMAN, MICHAEL DAVID; MONSZPART, ARON; BROSTOW, GABRIEL J.
To: NIANTIC INTERNATIONAL TECHNOLOGY LIMITED
Reel/Frame 064614/0923 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 16, 2023
From: NIANTIC INTERNATIONAL TECHNOLOGY LIMITED
To: NIANTIC, INC.
Reel/Frame 064614/0932 →
Continuity (3)
Continuation 17193878 · Mar 5, 2021
Provisional Application 62987849 · Mar 10, 2020
Related Publication 20230360339A1 · Nov 9, 2023
References Cited (46)
US 8855408B2 · Kim et al. · 2014 [cited by applicant]
US 10977662B2 · Shaw · 2021 [cited by applicant]
US 11741675B2 · Watson · 2023 [cited by examiner]
US 20130063556A1 · Russell et al. · 2013 [cited by applicant]
US 20150231490A1 · Graepel et al. · 2015 [cited by applicant]
US 20150356365A1 · Collet et al. · 2015 [cited by applicant]
US 20190057509A1 · Lv et al. · 2019 [cited by applicant]
US 20190094875A1 · Schulter et al. · 2019 [cited by applicant]
US 20190113349A1 · Lavu et al. · 2019 [cited by applicant]
US 20190208181A1 · Rowell et al. · 2019 [cited by applicant]
US 20200054939A1 · Golden et al. · 2020 [cited by applicant]
US 20200272164A1 · Gray · 2020 [cited by applicant]
US 20200334894A1 · Long et al. · 2020 [cited by applicant]
US 20210004984A1 · Ji et al. · 2021 [cited by applicant]
US 20210073953A1 · Lee · 2021 [cited by applicant]
US 20210142497A1 · Pugh et al. · 2021 [cited by applicant]
US 20240340400A1 · Godard et al. · 2024 [cited by applicant]
CA 3100640A1 · 2019 [cited by applicant]
CN 107407567A · 2017 [cited by applicant]
CN 109215080A · 2019 [cited by applicant]
CN 110088801A · 2019 [cited by applicant]
CN 110400322A · 2019 [cited by applicant]
JP 2019008796A · 2019 [cited by applicant]
JP 2019537080A · 2019 [cited by applicant]
KR 1020180058624A · 2018 [cited by applicant]
KR 1020200020646A · 2020 [cited by applicant]
WO 2018123641A1 · 2018 [cited by applicant]
PCT International Search Report and Written Opinion, PCT Application No. PCT/IB2021/051947, Jun. 16, 2021, nine pages. [cited by applicant]
Taiwan Intellectual Property Office, Office Action, Taiwanese Patent Application No. 110108388, Mar. 23, 2022, six pages. [cited by applicant]
United States Office Action, U.S. Appl. No. 17/193,878, filed Dec. 30, 2022, 12 pages. [cited by applicant]
United States Office Action, U.S. Appl. No. 17/193,878, filed Jun. 15, 2022, 10 pages. [cited by applicant]
United States Office Action, U.S. Appl. No. 17/193,878, filed Jan. 3, 2022, 10 pages. [cited by applicant]
United States Office Action, U.S. Appl. No. 17/193,878, filed Sep. 9, 2021, 8 pages. [cited by applicant]
Echigo T: “Segmentation of a 3D scene into free areas and object surfaces by using occluded edges of trinocular stereo”, Intelligent Robots And Systems '91, IEEE, Nov. 3, 1991, pp. 863-868. [cited by applicant]
Extended European Search Report and Search Opinion received for European Application No. 21768066.9, mailed on Mar. 12, 2024, 11 pages. [cited by applicant]
Garg, R et al , “Unsupervised cnn for single view depth estimation Geometry to the rescue,” European Conference on Computer Vision, Oct. 8, 2016, pp. 740-756. [cited by applicant]
Gupta, S. et al., “Cognitive Mapping and Planning for Visual Navigation”, arxiv.org, vol. 1702.03920, No. v2, Apr. 23, 2017, pp. 1-14. [cited by applicant]
International Search Report and Written Opinion received for PCT Patent Application No. PCT/IB2021/051947, mailed on Jun. 16, 2021, 7 pages. [cited by applicant]
Office Action received for Japanese Patent Application No. 2022-554751, mailed on Jan. 7, 2025, 6 pages (3 pages of English Translation and 3 pages of Original Document). [cited by applicant]
Takemura, I., et al., “Half-DR Expression of a Blind-Spot Area Using Automatic Driving Software”, The 23rd Virtual Reality Society of Japan General Conference, Sep. 19, 2018. [cited by applicant]
Tanaka, K, et al., “Building a floor map by combining stereo vision and visual tracking of persons”, Computational Intelligence In Robotics And Automation, 2003, Proceedin Gs. 2003 Ieee International Symposium On Jul. 1… [cited by applicant]
Wang, X., et al., “Viewpoint-Predicting-Based Remote Rendering on Mobile Devices Using Multiple Depth Images,” 2015 International Conference on Virtual Reality and Visualization (ICVRV) [online], IEEE, 2015, pp. 216-223. [cited by applicant]
Yingze, B. S. et al., “Understanding the 3D layout of a cluttered room from multiple images”, IEEE Winter Conference On Applications Of Computer Vision, IEEE, Mar. 24, 2014, pp. 690-697. [cited by applicant]
Qing R. et al., “Automatic Human Body Foreground Matting Algorithm,” Journal of Computer-Aided Design & Computer Graphics, vol. 32, No. 2, Feb. 2020, pp. 277-286. [cited by applicant]
Zhenjie, S. et al., “Scene Depth Estimation for Single Digital Image in Multi-scaled Space,” Computer Technology and Development, 2013, Issue 1, Jan. 2013, pp. 1-12. [cited by applicant]
Chinese Intellectual Property Office, Office Action, Chinese Patent Application No. 202180032229.1, Dec. 5, 2025, 14 pages. [cited by applicant]