IP Library Granted Patent US 12,412,282
Granted Patent B2
US 12,412,282 · App. 17/657,366 · Granted Sep 9, 2025

Semi-supervised tracking in medical images with cycle tracking

Inventors: Abdoul Aziz Amadou (London, GB); Rui Liao (Princeton Junction, NJ); Yue Zhang (Jersey City, NJ)
Assignee: Siemens Healthineers AG
G06T7/248G06T7/74G06T2207/10016G06T2207/20081G06T2207/20084G06T2207/30004
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,412,282
App. No.
17/657,366
Granted
Sep 9, 2025
Kind
B2
Abstract

Systems and methods for tracking a location of an object of interest through a sequence of medical images are provided. First and second input medical images of a patient are received. The first input medical image comprises an annotation of a location of an object of interest. Features are extracted from the first and the second input medical images. A location of the object of interest in the second input medical image is determined using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images. The location of the object of interest in the second input medical image is output. The machine learning based location predictor network is trained based on a comparison between 1) locations of a particular object in a sequence of training images determined during a forward tracking of the particular object through the sequence of training images and 2) locations of the particular object determined during a backward tracking of the particular object through the sequence of training images.

Claims (42)

1. A computer-implemented method comprising:

receiving first and second input medical images of a patient, the first input medical image comprising an annotation of a location of an object of interest;

extracting features from the first and the second input medical images;

determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images, wherein determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images comprises:

correlating the features extracted from the first input medical image with the features extracted from the second input medical image, wherein correlating the features extracted from the first input medical image with the features extracted from the second input medical image comprises:

matching features extracted from the first input medical image for regions corresponding to the location of the object of interest with features extracted from the second input medical image, and

determining the location of the object of interest in the second input medical image based on the correlated features; and

outputting the location of the object of interest in the second input medical image,

wherein the machine learning based location predictor network is trained based on a comparison between 1) locations of a particular object in a sequence of training images determined during a forward tracking of the particular object through the sequence of training images and 2) locations of the particular object determined during a backward tracking of the particular object through the sequence of training images.

2. The computer-implemented method of claim 1 , wherein the machine learning based location predictor network determines the locations of the particular object from a first image to a last image of the sequence of training images during the forward tracking and the machine learning based location predictor network determines the locations of the particular object from the last image to the first image in the sequence of training images during the backward tracking.

3. The computer-implemented method of claim 1 , wherein the machine learning based location predictor network is further trained based on a comparison between locations of the particular object in annotated training images of the sequence of training images and annotated ground truth locations of the particular object in the annotated training images.

4. The computer-implemented method of claim 1 , wherein the machine learning based location predictor network is further trained based on a comparison between a location of the particular object in a first training image of the sequence of training images and an annotated ground truth location of the particular object in the first training image of the sequence of training images.

5. The computer-implemented method of claim 1 , wherein a first image and a last image in the sequence of training images are annotated with a ground truth location of the particular object.

6. The computer-implemented method of claim 1 , wherein the first and the second input medical images are of a sequence of medical images, the method further comprising:

repeating the receiving, the extracting, the determining, and the outputting using the second input medical image as the first input medical image and a next image of the sequence of medical images as the second input medical image.

7. The computer-implemented method of claim 1 , wherein the location of the object of interest in the first input medical image is manually annotated by a user or automatically annotated using a machine learning based network.

8. An apparatus comprising:

means for receiving first and second input medical images of a patient, the first input medical image comprising an annotation of a location of an object of interest;

means for extracting features from the first and the second input medical images;

means for determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images, wherein the means for determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images comprises:

means for correlating the features extracted from the first input medical image with the features extracted from the second input medical image, wherein the means for correlating the features extracted from the first input medical image with the features extracted from the second input medical image comprises:

means for matching features extracted from the first input medical image for regions corresponding to the location of the object of interest with features extracted from the second input medical image, and

means for determining the location of the object of interest in the second input medical image based on the correlated features; and

means for outputting the location of the object of interest in the second input medical image,

wherein the machine learning based location predictor network is trained based on a comparison between 1) locations of a particular object in a sequence of training images determined during a forward tracking of the particular object through the sequence of training images and 2) locations of the particular object determined during a backward tracking of the particular object through the sequence of training images.

9. The apparatus of claim 8 , wherein the machine learning based location predictor network determines the locations of the particular object from a first image to a last image of the sequence of training images during the forward tracking and the machine learning based location predictor network determines the locations of the particular object from the last image to the first image in the sequence of training images during the backward tracking.

10. The apparatus of claim 8 , wherein the machine learning based location predictor network is further trained based on a comparison between locations of the particular object in annotated training images of the sequence of training images and annotated ground truth locations of the particular object in the annotated training images.

11. The apparatus of claim 8 , wherein the machine learning based location predictor network is further trained based on a comparison between a location of the particular object in a first training image of the sequence of training images and an annotated ground truth location of the particular object in the first training image of the sequence of training images.

12. The apparatus of claim 8 , wherein a first image and a last image in the sequence of training images are annotated with a ground truth location of the particular object.

13. A non-transitory computer readable medium storing computer program instructions, the computer program instructions when executed by a processor cause the processor to perform operations comprising:

receiving first and second input medical images of a patient, the first input medical image comprising an annotation of a location of an object of interest;

extracting features from the first and the second input medical images;

determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images, wherein determining a location of the object of interest in the second input medical image using a machine learning based location predictor network based on the annotation of the location of the object of interest in the first input medical image and the extracted features from the first and the second input medical images comprises:

correlating the features extracted from the first input medical image with the features extracted from the second input medical image, wherein correlating the features extracted from the first input medical image with the features extracted from the second input medical image comprises:

matching features extracted from the first input medical image for regions corresponding to the location of the object of interest with features extracted from the second input medical image, and

determining the location of the object of interest in the second input medical image based on the correlated features; and

outputting the location of the object of interest in the second input medical image,

wherein the machine learning based location predictor network is trained based on a comparison between 1) locations of a particular object in a sequence of training images determined during a forward tracking of the particular object through the sequence of training images and 2) locations of the particular object determined during a backward tracking of the particular object through the sequence of training images.

14. The non-transitory computer readable medium of claim 13 , wherein the machine learning based location predictor network determines the locations of the particular object from a first image to a last image of the sequence of training images during the forward tracking and the machine learning based location predictor network determines the locations of the particular object from the last image to the first image in the sequence of training images during the backward tracking.

15. The non-transitory computer readable medium of claim 13 , wherein the first and the second input medical images are of a sequence of medical images, the operations further comprising:

repeating the receiving, the extracting, the determining, and the outputting using the second input medical image as the first input medical image and a next image of the sequence of medical images as the second input medical image.

16. The non-transitory computer readable medium of claim 13 , wherein the location of the object of interest in the first input medical image is manually annotated by a user or automatically annotated using a machine learning based network.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2023
From: SIEMENS HEALTHCARE GMBH
To: SIEMENS HEALTHINEERS AG
Reel/Frame 066267/0346 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 3, 2022
From: SIEMENS MEDICAL SOLUTIONS USA, INC.
To: SIEMENS HEALTHCARE GMBH
Reel/Frame 060092/0133 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 25, 2022
From: SIEMENS HEALTHCARE LIMITED
To: SIEMENS MEDICAL SOLUTIONS USA, INC.
Reel/Frame 060015/0600 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 28, 2022
From: AMADOU, ABDOUL AZIZ
To: SIEMENS HEALTHCARE LIMITED
Reel/Frame 059767/0066 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 4, 2022
From: LIAO, RUI; ZHANG, YUE
To: SIEMENS MEDICAL SOLUTIONS USA, INC.
Reel/Frame 059488/0385 →
Continuity (1)
Related Publication 20230316544A1 · Oct 5, 2023
References Cited (15)
US 10713794B1 · He · 2020 [cited by examiner]
US 20180253837A1 · Ghesu · 2018 [cited by examiner]
US 20190310648A1 · Yang · 2019 [cited by examiner]
US 20200184654A1 · Kim · 2020 [cited by examiner]
US 20210174518A1 · Ishikake · 2021 [cited by examiner]
US 20210366120A1 · Ito · 2021 [cited by examiner]
Andreucci et al., “Side Effects of Radiographic Contrast Media: Pathogenesis, Risk Factors, and Prevention”, BioMed Research International, vol. 2014, Article ID 741018, 2014, pp. 1-20. [cited by applicant]
Kristan et al., “The Eighth Visual Object Tracking VOT2020 Challenge Results.” ECCV Workshops, 2020, pp. 547-601. [cited by applicant]
Wang et al., “Learning Correspondence From the Cycle-Consistency of Time”, 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2019, pp. 2566-2576. [cited by applicant]
Ambrosini et al., “Fully automatic and real-time catheter segmentation in X-ray fluoroscopy”, Medical Image Computing and Computer-Assisted Intervention, 2017, pp. 577-585. [cited by applicant]
Piayda et al., “Dynamic coronary roadmapping during percutaneous coronary intervention: a feasibility study”, European Journal of Medical Research, vol. 23, 2018, pp. 1-7. [cited by applicant]
Zhang et al., “Cascade attention machine for occluded landmark detection in 2D X-Ray angiography”, IEEE WACV 2019, pp. 91-100. [cited by applicant]
Ma et al., “Dynamic coronary roadmapping via catheter tip tracking in X-ray fluoroscopy with deep learning-based Bayesian filtering”, Medical Image Analysis 61, 2020, 101634, pp. 1-29. [cited by applicant]
Ullah et al., “Synthesize and Segment: Towards Improved Catheter Segmentation via Adversarial Augmentation”, Applied Sciences, 2021, pp. 1-16. [cited by applicant]
Li et al., “SiamRPN++: Evolution of Siamese Visual Tracking With Very Deep Networks”, IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2019, pp. 4277-4286. [cited by applicant]
Cited By (1)
US 12,511,747