IP Library › Granted Patent US 11,055,847
Granted Patent B2
US 11,055,847 · App. 16/822,101 · Granted Jul 6, 2021

Adversarial and dual inverse deep learning networks for medical image analysis

Inventors: Shaohua Kevin Zhou (Princeton, NJ); Mingqing Chen (Plainsboro, NJ); Daguang Xu (Princeton, NJ); Zhoubing Xu (Plainsboro, NJ); Dong Yang (Somerset, NJ)
Assignee: Siemens Healthcare GmbH
G06T7/0012G06K9/6267G06K9/66G06N3/0454G06N3/084G06N7/005G06T7/11G06K2209/05G06T2207/10072G06T2207/10081G06T2207/10088G06T2207/10104G06T2207/10116G06T2207/10132G06T2207/20081G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,055,847
App. No.
16/822,101
Granted
Jul 6, 2021
Kind
B2
Abstract

Methods and apparatus for automated medical image analysis using deep learning networks are disclosed. In a method of automatically performing a medical image analysis task on a medical image of a patient, a medical image of a patient is received. The medical image is input to a trained deep neural network. An output model that provides a result of a target medical image analysis task on the input medical image is automatically estimated using the trained deep neural network. The trained deep neural network is trained in one of a discriminative adversarial network or a deep image-to-image dual inverse network.

Claims (36)

1. A method for automatically performing a medical image analysis task on a medical image of a patient, comprising:

receiving a medical image of a patient;

inputting the medical image to a trained deep neural network; and

automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network, wherein the trained deep neural network is trained in a deep image-to-image dual inverse network.

2. The method of claim 1 , wherein the trained deep neural network is a first deep image-to-image network trained in the deep image-to-image dual inverse network and automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network comprises:

automatically generating an output image that provides a result of the target medical image analysis task on the input medical image using the first deep image-to-image network.

3. The method of claim 2 , wherein the deep image-to-image dual inverse network includes the first deep image-to-image network trained to perform the target medical image analysis task and including an encoder that coverts an input medical image to a feature representation and a decoder that generates a predicted output image from the feature representation of the input medical image, and a second deep image-to-image network trained to perform an inverse task to the target medical image analysis task and including an encoder that coverts an output image for the target medical image analysis task to a feature representation and a decoder that generates a predicted input image from the feature representation of the output image.

4. The method of claim 3 , wherein the first deep image-to-image network and the second deep image-to-image network are trained together based on a set of training samples including ground truth input images and corresponding ground truth output images to minimize a cost function including a first loss function that calculates an error between the ground truth output images and the predicted output images generated by the first deep image-to-image network from the ground truth input images, a second loss function that calculates an error between the ground truth output images and the predicted input images generated by the second deep image-to-image network from the ground truth output images, and a third loss function that calculates and error between the feature representation of the ground truth input images generated by the encoder of the first deep image-to-image network and the feature representation of the ground truth output images generated by the encoder of the second deep image-to-image network.

5. The method of claim 4 , further comprising:

training the first deep image-to-image network and the second deep image-to-image network together to minimize the cost function by repeating the following training operations for a plurality of iterations:

with parameters of the second deep image-to-image network fixed, learning parameters of the first deep image-to-image network to minimize the first loss function and the third loss function; and

with the parameters of the first deep image-to-image network fixed, learning the parameters of the second deep image-to-image network to minimize the second loss function and the third loss function.

6. An apparatus for automatically performing a medical image analysis task on a medical image of a patient, comprising:

means for receiving a medical image of a patient;

means for inputting the medical image to a trained deep neural network; and

means for automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network, wherein the trained deep neural network is trained in a deep image-to-image dual inverse network.

7. The apparatus of claim 6 , wherein the trained deep neural network is a first deep image-to-image network trained in the deep image-to-image dual inverse network and the means for automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network comprises:

means for automatically generating an output image that provides a result of the target medical image analysis task on the input medical image using the first deep image-to-image network.

8. The apparatus of claim 7 , wherein the deep image-to-image dual inverse network includes the first deep image-to-image network trained to perform the target medical image analysis task and including an encoder that coverts an input medical image to a feature representation and a decoder that generates a predicted output image from the feature representation of the input medical image, and a second deep image-to-image network trained to perform an inverse task to the target medical image analysis task and including an encoder that coverts an output image for the target medical image analysis task to a feature representation and a decoder that generates a predicted input image from the feature representation of the output image.

9. The apparatus of claim 8 , wherein the first deep image-to-image network and the second deep image-to-image network are trained together based on a set of training samples including ground truth input images and corresponding ground truth output images to minimize a cost function including a first loss function that calculates an error between the ground truth output images and the predicted output images generated by the first deep image-to-image network from the ground truth input images, a second loss function that calculates an error between the ground truth output images and the predicted input images generated by the second deep image-to-image network from the ground truth output images, and a third loss function that calculates and error between the feature representation of the ground truth input images generated by the encoder of the first deep image-to-image network and the feature representation of the ground truth output images generated by the encoder of the second deep image-to-image network.

10. The apparatus of claim 9 , further comprising:

means for training the first deep image-to-image network and the second deep image-to-image network together to minimize the cost function by repeating the following training operations for a plurality of iterations:

with parameters of the second deep image-to-image network fixed, learning parameters of the first deep image-to-image network to minimize the first loss function and the third loss function; and

with the parameters of the first deep image-to-image network fixed, learning the parameters of the second deep image-to-image network to minimize the second loss function and the third loss function.

11. A non-transitory computer readable medium storing computer program instructions for automatically performing a medical image analysis task on a medical image of a patient, the computer program instructions when executed by a processor cause the processor to perform operations comprising:

receiving a medical image of a patient;

inputting the medical image to a trained deep neural network; and

automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network, wherein the trained deep neural network is trained in a deep image-to-image dual inverse network.

12. The non-transitory computer readable medium of claim 11 , wherein the trained deep neural network is a first deep image-to-image network trained in the deep image-to-image dual inverse network and automatically estimating an output model that provides a result of a target medical image analysis task on the input medical image using the trained deep neural network comprises:

automatically generating an output image that provides a result of the target medical image analysis task on the input medical image using the first deep image-to-image network.

13. The non-transitory computer readable medium of claim 12 , wherein the deep image-to-image dual inverse network includes the first deep image-to-image network trained to perform the target medical image analysis task and including an encoder that coverts an input medical image to a feature representation and a decoder that generates a predicted output image from the feature representation of the input medical image, and a second deep image-to-image network trained to perform an inverse task to the target medical image analysis task and including an encoder that coverts an output image for the target medical image analysis task to a feature representation and a decoder that generates a predicted input image from the feature representation of the output image.

14. The non-transitory computer readable medium of claim 13 , wherein the first deep image-to-image network and the second deep image-to-image network are trained together based on a set of training samples including ground truth input images and corresponding ground truth output images to minimize a cost function including a first loss function that calculates an error between the ground truth output images and the predicted output images generated by the first deep image-to-image network from the ground truth input images, a second loss function that calculates an error between the ground truth output images and the predicted input images generated by the second deep image-to-image network from the ground truth output images, and a third loss function that calculates and error between the feature representation of the ground truth input images generated by the encoder of the first deep image-to-image network and the feature representation of the ground truth output images generated by the encoder of the second deep image-to-image network.

15. The non-transitory computer readable medium of claim 14 , wherein the operations further comprise:

training the first deep image-to-image network and the second deep image-to-image network together to minimize the cost function by repeating the following training operations for a plurality of iterations:

with parameters of the second deep image-to-image network fixed, learning parameters of the first deep image-to-image network to minimize the first loss function and the third loss function; and

with the parameters of the first deep image-to-image network fixed, learning the parameters of the second deep image-to-image network to minimize the second loss function and the third loss function.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2023
From: SIEMENS HEALTHCARE GMBH
To: SIEMENS HEALTHINEERS AG
Reel/Frame 066267/0346 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 18, 2020
From: CHEN, MINGQING; XU, DAGUANG; XU, ZHOUBING; YANG, DONG; ZHOU, SHAOHUA KEVIN
To: SIEMENS MEDICAL SOLUTIONS USA, INC.
Reel/Frame 052148/0237 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 18, 2020
From: SIEMENS MEDICAL SOLUTIONS USA, INC.
To: SIEMENS HEALTHCARE GMBH
Reel/Frame 052148/0251 →
Continuity (3)
Division 15868062 · Jan 11, 2017
Provisional Application 62457013 · Feb 9, 2017
Related Publication 20200219259A1 · Jul 9, 2020