IP Library › Granted Patent US 12,705,828
Granted Patent B2
US 12,705,828 · App. 18/541,711 · Granted Aug 11, 2026

Methods and systems for scanning objects

Inventors: Charles Dasher (Lawrenceville, GA); Reda Harb (Tampa, FL)
Assignee: Adeia Guides Inc.
G06T17/00G06T7/12G06T7/20G06V10/44G06V10/764G06V2201/07
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,705,828
App. No.
18/541,711
Filed
Dec 15, 2023
Granted
Aug 11, 2026
Kind
B2
Examiner
GE, JIN
Art Unit
2619
USPC
345/419
Abstract

There are provided systems and methods for scanning objects in a virtual environment. In particular, the present disclosure pertains to the domain of three-dimensional (3D) object capturing in extended reality (XR) environments and to an optimized system and method for passively capturing 3D objects within Extended Reality (XR) environments. A selection of an object from within an XR environment by a user via a user device is detected. A 3D capture session is initiated based on the detection. Scan data corresponding to the selected object is captured during the 3D capture session, where the capturing is performed as a background process on the user device. A 3D representation or model of the object is generated in the XR environment and provided to the user via the user device.

Claims (72)

1 . A method for scanning a physical object, the method comprising:

detecting, by control circuitry, a presence of the physical object to be scanned in an extended reality environment;

initiating, by the control circuitry, a 3D scan of the physical object as a background operation of a user device;

in response to the initiation of the 3D scan, capturing, by the control circuitry, scan data corresponding to the object while the object is in a field of view of the user device;

calculating, by the control circuitry, a confidence score based on the scan data, wherein the confidence score is used to determine whether to continue the 3D scan;

applying, by the control circuitry, spatial segmentation to the scan data to identify regions of interest, and bypass segments of the physical object that have been previously captured and remain unchanged;

performing, by the control circuitry, model quantization on the scan data, wherein a degree of the quantization is adjusted based on the availability of cloud computing resources;

storing, by the control circuitry, the scan data in a database in response to the calculated confidence score exceeding a threshold; and

generating, by the control circuitry, a 3D model of the physical object in the extended reality environment based on the scan data in the database.

2 . The method of claim 1 , further comprising:

initiating a primary 3D scan of the physical object as a background operation of the user device;

based on the primary 3D scan data, determining that a 3D model related to the physical object is incomplete;

retrieving the 3D model related to the physical object from the database; and

updating the retrieved 3D model with the scan data.

3 . The method of claim 1 , wherein, in response to detecting the presence of the object, the method further comprises:

displaying an assistive user interface element around the object on a viewport of the user device to capture a plurality of angles of the object in a 3D space.

4 . The method of claim 1 , further comprising:

determining completion status data for the generation of the 3D model;

generating an overlay comprising the 3D model on the object based on the completion status data; and

displaying the overlay within the viewport of the user device.

5 . The method of claim 1 , wherein:

applying the spatial segmentation to the scan data to identify regions of interest further comprises:

distinguishing the physical object from background elements in the field of view;

extracting a set of data points corresponding to pertinent sections of the physical object to restrict an area of focus for the 3D scan; and

storing a plurality of detailed segmentations of the identified regions of interest in the database for use as a spatial reference during a subsequent encounter with the physical object.

6 . The method of claim 1 , further comprising:

receiving an indication of priority for scanning one or more objects of a plurality of objects; and

capturing the one or more objects of the plurality of objects based on the indication of priority.

7 . The method of claim 1 , further comprising:

tracking a motion of the physical object; and

capturing the scan data corresponding to the physical object based on a speed of the motion of the object, wherein the scan data is captured when the speed of the motion is within a pre-defined threshold.

8 . The method of claim 1 , wherein generating the 3D model comprises:

receiving a set of features of interest corresponding to the object; and

generating the 3D model based on the set of features of interest.

9 . The method of claim 1 , further comprising:

creating at least one classifier algorithm based on the scan data of the physical object; and

using the classifier algorithm to detect a presence of a second physical object.

10 . A system for scanning a physical object, the system comprising control circuitry configured to:

detect a presence of the physical object to be scanned in an extended reality environment;

initiate a 3D scan of the physical object as a background operation of a user device;

in response to the initiation of the 3D scan, capture scan data corresponding to the object while the object is in a field of view of the user device;

calculate a confidence score based on the scan data, wherein the confidence score is used to determine whether to continue the 3D scan;

apply spatial segmentation to the scan data to identify regions of interest and bypass segments of the physical object that have been previously captured and remain unchanged;

perform model quantization on the scan data, wherein a degree of the quantization is adjusted based on the availability of cloud computing resources;

store the scan data in a database in response to the calculated confidence score exceeding a threshold; and

generate a 3D model of the physical object in the extended reality environment based on the scan data in the database.

11 . The system of claim 10 , the control circuitry configured to:

retrieve a prior 3D model related to the physical object from the database;

determine whether the prior 3D model related to the physical object is complete; and

in response to determining the prior 3D model is not complete, update the retrieved prior 3D model with the scan data.

12 . The system of claim 10 , wherein, in response to detecting the presence of the object, the control circuitry configured to:

display an assistive user interface element around the object on a viewport of the user device to capture a plurality of angles of the object in a 3D space.

13 . The system of claim 10 , the control circuitry configured to:

determine completion status data for the generation of the 3D model;

generate an overlay comprising the 3D model on the object based on the completion status data; and

display the overlay on the viewport of the user device.

14 . The system of claim 10 , wherein the control circuitry is configured, when applying spatial segmentation to the scan data to identify regions of interest, to:

distinguish the physical object from background elements in the field of view while applying the spatial segmentation;

extract a set of data points corresponding to pertinent sections of the physical object to restrict an area of focus for the 3D scan; and

store a plurality of detailed segmentations of the identified regions of interest in the database for use as a spatial reference during a subsequent encounter with the physical object.

15 . The system of claim 11 , the control circuitry configured to:

receive an indication of priority for scanning one or more of a plurality of objects; and

capture each of the plurality of objects based on the indication of priority.

16 . The system of claim 10 , the control circuitry configured to:

track a motion of the physical object; and

capture the scan data corresponding to the physical object based on a speed of the motion of the object, wherein the scan data is captured when the speed of the motion is within a pre-defined threshold.

17 . The system of claim 10 , the control circuitry configured to:

receive a set of features of interest corresponding to the object; and

generate the 3D model based on the set of features of interest.

18 . The system of claim 10 , the control circuitry configured to:

create at least one classifier algorithm based on the scan data of the physical object; and

use the classifier algorithm to detect a presence of a second physical object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 20, 2024
From: DASHER, CHARLES; HARB, REDA
To: ADEIA GUIDES INC.
Reel/Frame 066837/0857 →
Continuity (1)
Related Publication 20250200875A1 · Jun 19, 2025
References Cited (53)
US 8818768B1 · Fan et al. · 2014 [cited by applicant]
US 9536352B2 · Anderson · 2017 [cited by applicant]
US 11436808B2 · Wang et al. · 2022 [cited by applicant]
US 11922368B1 · Wozniak · 2024 [cited by examiner]
US 12039735B2 · Feng et al. · 2024 [cited by applicant]
US 12112435B1 · Bhushan et al. · 2024 [cited by applicant]
US 20090237396A1 · Venezia et al. · 2009 [cited by applicant]
US 20130083064A1 · Geisner · 2013 [cited by examiner]
US 20130242113A1 · Tanaka · 2013 [cited by examiner]
US 20140198096A1 · Mitchell · 2014 [cited by applicant]
US 20140226858A1 · Kang · 2014 [cited by examiner]
US 20170154204A1 · Ryu et al. · 2017 [cited by applicant]
US 20170169619A1 · Mullins · 2017 [cited by examiner]
US 20180114363A1 · Rosenbaum · 2018 [cited by examiner]
US 20200035025A1 · Crocker et al. · 2020 [cited by applicant]
US 20200053253A1 · Kavallierou · 2020 [cited by applicant]
US 20200080853A1 · Tam et al. · 2020 [cited by applicant]
US 20210209339A1 · You et al. · 2021 [cited by applicant]
US 20210326777A1 · Stevens · 2021 [cited by applicant]
US 20220277543A1 · Saklatvala · 2022 [cited by examiner]
US 20230368458A1 · Dryer et al. · 2023 [cited by applicant]
US 20230419616A1 · Dudovitch et al. · 2023 [cited by applicant]
US 20240104132A1 · Banfield · 2024 [cited by examiner]
US 20240265654A1 · Tomizuka et al. · 2024 [cited by applicant]
US 20240290049A1 · Prideaux-Ghee et al. · 2024 [cited by applicant]
US 20250005867A1 · Reynolds et al. · 2025 [cited by applicant]
US 20250078401A1 · Biswas · 2025 [cited by examiner]
US 20250200885A1 · Dasher et al. · 2025 [cited by applicant]
U.S. Appl. No. 18/541,700, filed Dec. 15, 2023, Charles Dasher. [cited by applicant]
Abou-Chakra, J., “Scene Understanding, Semantic SLAM, Implicit Representations”, In Conference on Robot Learning, available online at: <https://nikosuenderhauf.github.io/projects/sceneunderstanding/>, retrieved on Nov. … [cited by applicant]
Anonymous, “Are Online Trade Schools Worth It?,” (https://www.onlineschoolscenter.com/are-online-trade-schools-worth-it/) (downloaded Aug. 14, 23; undated) (18 pages). [cited by applicant]
Anonymous, “Room scan visualization,” Microsoft, Sep. 21, 2022 (https://learn.microsoft.com/en-us/windows/mixed-reality/design/room-scan-visualization) (3 pages). [cited by applicant]
Anonymous, “Scanning and detecting 3D objects,” Sample Code, (https://developer.apple.com/documentation/arkit/arkit_in_ios/content_anchors/scanning_and_detecting_3d_objects) (downloaded Aug. 14, 2023; undated) (7 pages). [cited by applicant]
Apple Developer , “Augmented Reality—Object Capture”, Apple Developer, “Augmented Reality—Object Capture” (Downloaded Mar. 15, 24) Retrieved from https://developer.apple.com/augmented-reality/object-capture/ (2 pages). [cited by applicant]
Burgess, “Tour the British Museum on Google Street View,” Wired (Dec. 11, 2015) (https://www.wired.co.uk/article/british-museum-google-street-view) (7 pages). [cited by applicant]
Chen et al., “An Overview on Visual SLAM: From Tradition to Semantic,” Remote Sensing, MDPI, 14:3010 (May 29, 2022) (47 pages). [cited by applicant]
Chen, Zhiqin, et al., “MobileNeRF: Exploiting the Polygon Rasterization Pipeline for Efficient Neural Field Rendering on Mobile Architectures”, Chen, “MobileNeRF: Exploiting the Polygon Rasterization Pipeline for Effici… [cited by applicant]
Crawford, How to host a real estate virtual open house in 7 steps, Next (Jul. 9, 2020) (13 pages) (https://www.nextinsurance.com/blog/how-host-a-real-estate-virtual-open-house-in-7-steps). [cited by applicant]
Deci.AI, “The Ultimate Guide to Deep Learning Model Quantization and Quantization-Aware Training”, Deci.AI, “The Ultimate Guide to Deep Learning Model Quantization and Quantization-Aware Training” (Downloaded Mar. 15, 2… [cited by applicant]
Edwards, Benj , “New AI model can “cut out” any object within an image- and Meta is sharing the code”, Edwards, “New AI model can “cut out” any object within an image- and Meta is sharing the code” (2023), Retrieved fro… [cited by applicant]
Kiri Engine, “Kiri Engine App”, Kiri Engine, “Kiri Engine App”, (Downloaded Mar. 15, 2024) Retrieved from https://www.kiriengine.app (7 Pages). [cited by applicant]
Lancial, Braxton , “The Basics: Niantic Lightship”, https://lightship.dev/guides/lightship-basics/, 2022. [cited by applicant]
Mars Gadgets , “Free 3D Scanner Apps (Photogrammetry) 2022”, Mars Gadgets, “Free 30 Scanner Apps (Photogrammetry) 2022” (Sep. 26, 2021) Retrieved from https://www.youtube.com/watch?v=K3PyUMOPI1M (12 Pages). [cited by applicant]
Matterport, “Get a 2D floor plan from a 3D tour.,” (downloaded Sep. 13, 2023, https://matterport.com/how-it-works/schematic-floor-plans). [cited by applicant]
Neta Zmora et al., “Achieving FP32 Accuracy for INT8 Inference Using Quantization Aware Training with NVIDIA TensorRT”, (Jul. 20, 2021), (https://deci.ai/quantization-and-quantization-aware-training/), 8 pages. [cited by applicant]
Niantic , “Lightship VPS: Building Our 3D Map From Crowdsourced Scans”, Niantic “Lightship VPS: Building Our 3D Map From Crowdsourced Scans” (Nov. 10, 2022) Retrieved from https://nianticlabs.com/news/vps-part-2?hl=en (… [cited by applicant]
NVIDIA Game Developer, “NVIDIA's Photogrammetry ‘Live Visualizer’ using Reality Capture and Unity”, NVIDIA Game Developer, “NVIDIA's Photogrammetry ‘Live Visualizer’ using Reality Capture and Unity” (Mar. 9, 2017) Retri… [cited by applicant]
Papers With Code , “3D Semantic Segmentation”, Papers With Code, “3D Semantic Segmentation” (Downloaded Mar. 15, 2024), Retrieved from https://paperswithcode.com/task/3d-semantic-segmentation (17 Pages). [cited by applicant]
Perez, “Apple Acquires Flyby Media, Makers Of Tech That “Sees” The World Around You,” Join TechCrunch (Jan. 29, 2016) (9 pages) (Apple Acquires Flyby Media, Makers Of Tech That “Sees” The World Around You | TechCrunch). [cited by applicant]
Photogrammetry as defined by Wikipedia, https://en.wikipedia.org/wiki/Photogrammetry, retrieved from internet Oct. 3, 2023. [cited by applicant]
Roomvo, “Where imagination meets results,” (downloaded Sep. 13, 2023; https://get.roomvo.com/product-room-visualizer-flooring/) (7 pages). [cited by applicant]
Weinmann et al., “Efficient 3D Mapping and Modelling of Indoor Scenes with the Microsoft HoloLens: A Survey,” PFG, 89:319-333 (2021). [cited by applicant]
U.S. Appl. No. 18/214,776, filed Jun. 27, 2023, Jennifer Reynolds. [cited by applicant]