IP Library Granted Patent US 12,475,591
Granted Patent B2
US 12,475,591 · App. 18/801,711 · Granted Nov 18, 2025

Vision-based 6DOF camera pose estimation in bronchoscopy

Inventors: Mali Shen (Sunnyvale, CA); Menglong Ye (Mountain View, CA)
Assignee: Auris Health, Inc.
G06T7/70A61B1/2676A61B6/12G06N3/04G06T7/33G06T7/55G16H30/40G16H50/50A61B2090/3762G06T2207/10016G06T2207/10028G06T2207/10068G06T2207/10081G06T2207/20081G06T2207/20084G06T2207/30244
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,475,591
App. No.
18/801,711
Granted
Nov 18, 2025
Kind
B2
Abstract

Methods and systems provide improved navigation through tubular networks such as lung airways by providing improved estimation of location and orientation information of a medical instrument (e.g., an endoscope) within the tubular network. Various input data such as image data and CT data, are used to model the tubular networks, and the model information is used to generate a camera pose representing a specific site location within the tubular network and/or to determine navigation information including position and orientation for the medical instrument.

Claims (35)

1 . A method, comprising:

transforming video image data of an anatomical structure into a first depth map;

receiving a plurality of second depth maps based at least in part on transformed computed tomography (CT) image data, each of the second depth maps representing a virtual model of the anatomical structure; and

generating a camera pose representative of a spatial location within the anatomical structure, including:

finding a minimum of a scalar similarity function relative to a comparison of each of the plurality of second depth maps to the first depth map; and

generating the camera pose as a function of a second depth map associated with the minimum of the scalar similarity function.

2 . The method of claim 1 , further comprising referencing the video image data to a coordinate system of the CT image data prior to the transforming.

3 . The method of claim 1 , further comprising applying a convolutional neural network (CNN) to transform the video image data into the first depth map.

4 . The method of claim 1 , further comprising comparing each of the plurality of second depth maps to the first depth map for similarity of shape.

5 . The method of claim 1 , further comprising comparing each of the plurality of second depth maps to the first depth map using a normalized-cross correlation or mutual information.

6 . The method of claim 1 , further comprising generating the camera pose as a function of a second depth map associated to a given time point.

7 . The method of claim 1 , further comprising initializing the scalar similarity function with a previously generated camera pose.

8 . The method of claim 1 , further comprising generating a camera pose for each of the plurality of second depth maps.

9 . The method of claim 1 , wherein the first depth map comprises a three-dimensional virtual model of the anatomical structure.

10 . The method of claim 1 , wherein the first depth map and the plurality of second depth maps contain three-dimensional position and orientation information.

11 . The method of claim 1 , wherein the similarity function provides a minimum difference between the first depth map at a given time and one of the second depth maps at the given time.

12 . A method, comprising:

transforming video image data of a bronchial airway into a first depth map;

receiving a plurality of second depth maps based at least in part on transformed computed tomography (CT) image data, each of the second depth maps representing a virtual model of the bronchial airway; and

generating a camera pose representative of a spatial location within the bronchial airway, including:

finding a minimum of a scalar similarity function relative to a comparison of each of the plurality of second depth maps to the first depth map; and

generating the camera pose as a function of a second depth map associated with the minimum of the scalar similarity function.

13 . The method of claim 12 , further comprising generating the camera pose to include an orientation within the bronchial airway.

14 . The method of claim 12 , further comprising using an artificial neural network architecture to find the minimum of the scalar similarity function.

15 . The method of claim 12 , further comprising estimating a camera pose by passing each of the plurality of second depth maps and the first depth map into a spatial transformation network that regresses the relative transformation between each of the plurality of second depth maps and the first depth map.

16 . The method of claim 12 , wherein the generating the camera pose is an iterative or continuous process that includes using a camera pose of a previous iteration to initialize a current iteration.

17 . A medical system, comprising:

a means for transforming video image data of an anatomical structure into a first depth map;

a means for receiving a plurality of second depth maps based at least in part on transformed computed tomography (CT) image data, each of the second depth maps representing a virtual model of the anatomical structure;

a means for generating a camera pose representative of a spatial location within the anatomical structure, including:

a means for finding a minimum of a scalar similarity function relative to a comparison of each of the plurality of second depth maps to the first depth map; and

a means for generating the camera pose as a function of a second depth map associated with the minimum of the scalar similarity function.

18 . The medical system of claim 17 , further comprising a means for initializing the scalar similarity function.

19 . The medical system of claim 17 , further comprising a means for comparing each of the plurality of second depth maps to the first depth map for similarity of identified features.

20 . The medical system of claim 17 , further comprising a means for communicating the camera pose in real time, relative to the anatomical structure.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 27, 2024
From: SHEN, MALI; YE, MENGLONG
To: AURIS HEALTH, INC.
Reel/Frame 068415/0890 →
Continuity (2)
Continuation 17219804 · Mar 31, 2021
Related Publication 20240404101A1 · Dec 5, 2024
References Cited (75)
US 7388583B2 · Redert · 2008 [cited by applicant]
US 10270962B1 · Stout · 2019 [cited by applicant]
US 10380753B1 · Csordas et al. · 2019 [cited by applicant]
US 11069150B2 · Sminshisescu et al. · 2021 [cited by applicant]
US 11191423B1 · Zingaretti et al. · 2021 [cited by applicant]
US 11295460B1 · Aghdasi et al. · 2022 [cited by applicant]
US 20050129305A1 · Chen et al. · 2005 [cited by applicant]
US 20070013710A1 · Higgins et al. · 2007 [cited by applicant]
US 20070161854A1 · Alamaro et al. · 2007 [cited by applicant]
US 20090048482A1 · Hong · 2009 [cited by examiner]
US 20090088897A1 · Zhao et al. · 2009 [cited by applicant]
US 20100280365A1 · Higgins et al. · 2010 [cited by applicant]
US 20110043604A1 · Peleg et al. · 2011 [cited by applicant]
US 20110128352A1 · Higgins et al. · 2011 [cited by applicant]
US 20120056986A1 · Popovic · 2012 [cited by applicant]
US 20120069167A1 · Liu et al. · 2012 [cited by applicant]
US 20120163686A1 · Liao · 2012 [cited by examiner]
US 20120310098A1 · Popovic · 2012 [cited by applicant]
US 20130259315A1 · Angot et al. · 2013 [cited by applicant]
US 20130303883A1 · Zehavi et al. · 2013 [cited by applicant]
US 20130322717A1 · Bar-Shalev · 2013 [cited by applicant]
US 20140320629A1 · Chizeck et al. · 2014 [cited by applicant]
US 20150235408A1 · Gross et al. · 2015 [cited by applicant]
US 20150381908A1 · De Bruijn et al. · 2015 [cited by applicant]
US 20160278721A1 · Im et al. · 2016 [cited by applicant]
US 20170046833A1 · Lurie et al. · 2017 [cited by applicant]
US 20170084006A1 · Stewart · 2017 [cited by applicant]
US 20170084027A1 · Mintz et al. · 2017 [cited by applicant]
US 20180174311A1 · Kluckner et al. · 2018 [cited by applicant]
US 20190015163A1 · Abhari et al. · 2019 [cited by applicant]
US 20190060013A1 · McDowall · 2019 [cited by applicant]
US 20190087987A1 · Bae et al. · 2019 [cited by applicant]
US 20190279383A1 · Angelova et al. · 2019 [cited by applicant]
US 20200029789A1 · Hirakawa · 2020 [cited by applicant]
US 20200046436A1 · Tzeisler et al. · 2020 [cited by applicant]
US 20200051258A1 · Miao et al. · 2020 [cited by applicant]
US 20200281454A1 · Refai et al. · 2020 [cited by applicant]
US 20210090226A1 · Rauniyar et al. · 2021 [cited by applicant]
US 20210097662A1 · Wang et al. · 2021 [cited by applicant]
US 20210186619A1 · Levi et al. · 2021 [cited by applicant]
US 20210280312A1 · Freedman et al. · 2021 [cited by applicant]
US 20210406596A1 · Hoffman et al. · 2021 [cited by applicant]
US 20220074994A1 · Senn et al. · 2022 [cited by applicant]
US 20220076808A1 · Vija et al. · 2022 [cited by applicant]
US 20220113804A1 · Crowther et al. · 2022 [cited by applicant]
US 20220198693A1 · Li et al. · 2022 [cited by applicant]
US 20220230384A1 · Shin et al. · 2022 [cited by applicant]
US 20220249168A1 · Besier et al. · 2022 [cited by applicant]
US 20220257978A1 · Kamen et al. · 2022 [cited by applicant]
US 20220262023A1 · Vignard et al. · 2022 [cited by applicant]
US 20220387129A1 · Kinrot · 2022 [cited by applicant]
US 20230290042A1 · Casella et al. · 2023 [cited by applicant]
CN 111145238A · 2020 [cited by applicant]
WO 2020102584A2 · 2020 [cited by applicant]
European Communication with Extended Search Report, dated Jan. 14, 2025, from European Patent Application No. 22779246.2, pp. 1-13. [cited by applicant]
Zhao, C., Shen, M., Sun, L., & Yang, G. Z. (2019). Generative localization with uncertainty estimation through video-CT data for bronchoscopic biopsy. IEEE Robotics and Automation Letters, 5(1), 258-265. [cited by applicant]
Shen, M., Giannarou, S., & Yang, G. Z. (2015). Robust camera localisation with depth reconstruction for bronchoscopic navigation. International journal of computer assisted radiology and surgery, 10, 801-813. [cited by applicant]
Kumar, A., Wang, Y. Y., Wu, C. J., Liu, K. C., & Wu, H. S. (2014). Stereoscopic visualization of laparoscope image using depth information from 3D model. Computer methods and programs in biomedicine, 113(3), 862-868. [cited by applicant]
Nardelli, P., Jaeger, A., O'Shea, C., Khan, K. A., Kennedy, M. P., & Cantillon-Murphy, P. (2017). Pre-clinical validation of virtual bronchoscopy using 3D Slicer. International journal of computer assisted radiology and… [cited by applicant]
Non-Final Office Action, dated Feb. 24, 2023, from U.S. Appl. No. 17/219,804, pp. 1-35. [cited by applicant]
Final Office Action, dated Aug. 1, 2023, from U.S. Appl. No. 17/219,804, pp. 1-32. [cited by applicant]
Advisory Action, dated Oct. 6, 2023, from U.S. Appl. No. 17/219,804, pp. 1-3. [cited by applicant]
Non-Final Office Action, dated Nov. 6, 2023, from U.S. Appl. No. 17/219,804, pp. 1-32. [cited by applicant]
Notice of Allowance, dated May 8, 2024, from U.S. Appl. No. 17/219,804, pp. 1-10. [cited by applicant]
Brachmann, E., Krull, A., Nowozin, S., Shotton, J., Michel, F., Gumhold, S., Rother, C., Dsac—differentiable ransac for camera localization, 2017, 11 pages. [cited by applicant]
Deligianni, F., Chung, A., Yang, G.Z., Patient-specific bronchoscope simulation with pqspace-based 2D/3D registration, Jan. 6, 2010, 13 pages. [cited by applicant]
Goodfellow, I.J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A.C., Bengio, Y., Generative Adversarial Nets, 2014, 9 pages. [cited by applicant]
He, K., Zhang, X., Ren, S., Sun, J., Deep residual learning for image recognition, 2016, 9 pages. [cited by applicant]
Jun-Yan Zhu, Taesung Park, Phillip Isola, Alexei A. Efros, Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks, 2017, 10 pages. [cited by applicant]
Luo, X., Feurestein, M., Deguchi, D., Kitasaka, T., Takabatake, H., Mori, K., Development and comparison of new hybrid motion tracking for bronchoscopic navigation, 2012, 22 pages. [cited by applicant]
Rai, L., Helferty, J.P., Higgines, W.E., Combined video tracking and image-video registration for continuous bronchoscopic guidance, 2008, 15 pages. [cited by applicant]
Shen, M., Gu, Y., Liu, N., Yang, G.Z., Context-Aware Depth and Pose Estimation for Bronchoscopic Navigation, Jan. 2019, 8 pages. [cited by applicant]
International Search Report for Appl. No. PCT/IB2022/052730, dated Jun. 22, 2022, 3 pages. [cited by applicant]
Written Opinion for Appl. No. PCT/IB2022/052730, dated Jun. 22, 2022, 3 pages. [cited by applicant]
International Preliminary Report on Patentability for Appl. No. PCT/IB2022/052730, dated Oct. 3, 2023, 4 pages. [cited by applicant]