IP Library › Granted Patent US 12,190,556
Granted Patent B2
US 12,190,556 · App. 17/581,836 · Granted Jan 7, 2025

Learning apparatus, learning method, and learning program, graph structure extraction apparatus, graph structure extraction method, and graph structure extraction program, and learned extraction model

Inventor: Deepak Keshwani (Tokyo, JP)
Assignee: FUJIFILM Corporation
G06V10/426G06V10/774G06V10/82G06V10/98G06V2201/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,190,556
App. No.
17/581,836
Granted
Jan 7, 2025
Kind
B2
Abstract

A learning unit derives, from a target image including at least one tubular structure, in a case where an image for learning and ground-truth data of a graph structure included in the image for learning are input to an extraction model which extracts a feature vector of a plurality of nodes constituting a graph structure of the tubular structure, a loss between nodes on the graph structure included in the image for learning on the basis of an error between a feature vector distance between nodes belonging to the same graph structure and a topological distance which is a distance on a route of the graph structure between the nodes, and performs learning of the extraction model on the basis of the loss.

Claims (46)

1. A learning apparatus comprising:

at least one processor, wherein the processor is configured to:

input a learning image and ground-truth data of the learning image to an extraction model, wherein the ground-truth data of the learning image comprises an extraction result of nodes of a graph structure included in the learning image;

receive a feature map for learning outputted from the extraction model such that a feature vector distance between nodes belonging to a same graph structure included in learning image corresponds to a topological distance which is a distance on a route of the graph structure between the nodes;

derive, according to the feature map for learning and the ground-truth data for learning, a loss between the nodes on the graph structure included in the learning image on the basis of a difference between the feature vector distance and the topological distance; and

perform learning of the extraction model on the basis of the loss.

2. The learning apparatus according to claim 1 ,

wherein the processor is configured to further derive, for another learning image including at least two different graph structures, a loss such that a feature vector distance between nodes belonging to a same graph structure in the another learning image is decreased.

3. The learning apparatus according to claim 2 ,

wherein the processor is configured to further derive the loss such that a feature vector distance between nodes belonging to different graph structures in the another learning image is increased.

4. The learning apparatus according to claim 1 ,

wherein the extraction model is a fully convolutional neural network.

5. The learning apparatus according to claim 1 ,

wherein the target image and the image for learning are three-dimensional medical images.

6. The learning apparatus according to claim 1 ,

wherein the tubular structure is an artery and a vein.

7. The learning apparatus according to claim 1 ,

wherein the tubular structure is an artery, a vein, and a portal vein in a liver.

8. The learning apparatus according to claim 1 ,

wherein the tubular structure is a bronchus.

9. A graph structure extraction apparatus comprising:

at least one processor, wherein the processor is configured to:

acquire a target image including at least one tabular structure;

input the target image into the extraction model learned by the learning apparatus according to claim 1 ; and

receive an extraction result of a graph structure of the tabular structure included in the target image.

10. The graph structure extraction apparatus according to claim 9 , wherein the processor is further configured to;

label the tubular structure included in the target image according to the extraction result of the graph structure; and

display the target image in which the tubular structure is labeled, on a display.

11. A learning method comprising:

inputting a learning image and ground-truth data of the learning image to an extraction model, wherein the ground-truth data of the learning image comprises an extraction result of nodes of a graph structure included in the learning image;

receiving a feature map for learning outputted from the extraction model such that a feature vector distance between nodes belonging to a same graph structure included in learning image corresponds to a topological distance which is a distance on a route of the graph structure between the nodes;

deriving, according to the feature map for learning and the ground-truth data for learning, a loss between the nodes on the graph structure included in the learning image on the basis of a difference between the feature vector distance and the topological distance; and

performing learning of the extraction model on the basis of the loss.

12. A graph structure extraction method comprising:

acquiring a target image including at least one tabular structure;

inputting the target image into the extraction model learned by the learning method according to claim 11 ; and

receiving an extraction result of a graph structure of the tabular structure included in the target image.

13. A non-transitory computer-readable storage medium that stores a graph structure extraction program causing a computer to execute a process comprising:

acquiring a target image including at least one tabular structure;

inputting the target image into the extraction model learned by the learning method according to claim 11 ; and

receiving an extraction result of a graph structure of the tabular structure included in the target image.

14. A non-transitory computer-readable storage medium that stores a learning program causing a computer to execute a process comprising:

inputting a learning image and ground-truth data of the learning image to an extraction model, wherein the ground-truth data of the learning image comprises an extraction result of nodes of a graph structure included in the learning image;

receiving a feature map for learning outputted from the extraction model such that a feature vector distance between nodes belonging to a same graph structure included in learning image corresponds to a topological distance which is a distance on a route of the graph structure between the nodes;

deriving, according to the feature map for learning and the ground-truth data for learning, a loss between the nodes on the graph structure included in the learning image on the basis of a difference between the feature vector distance and the topological distance; and

performing learning of the extraction model on the basis of the loss.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 24, 2022
From: KESHWANI, DEEPAK
To: FUJIFILM CORPORATION
Reel/Frame 058734/0595 →
Priority Claims (1)
JP 2019-137034 · Jul 25, 2019 · national
Continuity (2)
Continuation PCTJP2020028416 · Jul 22, 2020
Related Publication 20220148286A1 · May 12, 2022
References Cited (19)
US 20100296709A1 · Ostrovsky et al. · 2010 [cited by applicant]
US 20220044767A1 · Rong · 2022 [cited by examiner]
EP 1271391A1 · 2003 [cited by examiner]
JP 2014236912 · 2014 [cited by applicant]
JP 2014236912A · 2014 [cited by examiner]
WO 2017199246 · 2017 [cited by applicant]
WO WO2017199246A1 · 2017 [cited by examiner]
Bert De Brabandere et al., “Semantic Instance Segmentation with a Discriminative Loss Function,” workshop at CVPR 2017, Aug. 2017, pp. 1-10. [cited by applicant]
Alireza Fathi et al., “Semantic Instance Segmentation via Deep Metric Learning,” Computer Vision and Pattern Recognition, Mar. 2017, pp. 1-9. [cited by applicant]
Shu Kong et al., “Recurrent Pixel Embedding for Instance Grouping,” Computer Vision and Pattern Recognition (cs.CV); Machine Learning(cs.LG); Multimedia (cs.MM), Dec. 2017, pp. 1-24. [cited by applicant]
Rossant Florence et al., “A Morphological Approach for Vessel Segmentation in Eye Fundus Images, with Quantitative Evaluation,” Journal of Medical Imaging and Health Information, vol. 1, 2011, pp. 42-49. [cited by applicant]
Andrzej Szymczak et al., “Coronary vessel trees from 3D imagery: A topological approach,” Medical Image Analysis, vol. 10, Aug. 2006, pp. 1-24. [cited by applicant]
Reza Nekovei et al., “Back Propagation Network and its Configuration for Blood Vessel Detection in Angiograms,” IEEE Transactions on Neural Networks, vol. 6, Jan. 1995, pp. 64-72. [cited by applicant]
“International Search Report (Form PCT/ISA/210) of PCT/JP2020/028416,” mailed on Sep. 15, 2020, with English translation thereof, pp. 1-6. [cited by applicant]
“Written Opinion of the International Searching Authority (Form PCT/ISA/237)” of PCT/JP2020/028416, mailed on Sep. 15, 2020, with English translation thereof, pp. 1-6. [cited by applicant]
“Search Report of Europe Counterpart Application”, issued on Jul. 18, 2022, p. 1-p. 9. [cited by applicant]
Ruben Hemelings et al., “Artery-vein segmentation in fundus images using a fully convolutional network,” Computerized Medical Imaging and Graphics, vol. 76, May 2019, pp. 1-12. [cited by applicant]
Fantin Girard et al., “Joint segmentation and classification of retinal arteries/veins from fundus images,” arXiv.org, Mar. 2019, pp. 1-15. [cited by applicant]
“Notice of Reasons for Refusal of Japan Counterpart Application”, issued on Nov. 8, 2022, with English translation thereof, p. 1-p. 4. [cited by applicant]