IP Library › Granted Patent US 12,272,065
Granted Patent B2
US 12,272,065 · App. 17/723,767 · Granted Apr 8, 2025

Apparatus and method for image segmentation

Inventors: Gyu Sang Choi (Daegu, KR); Hyun Kwang Shin (Gyeongsangbuk-do, KR)
Assignee: RESEARCH COOPERATION FOUNDATION OF YEUNGNAM UNIVERSITY
G06T7/10G06T2207/10088G06T2207/20084G06T2207/30016G06T2207/30096
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,272,065
App. No.
17/723,767
Granted
Apr 8, 2025
Kind
B2
Abstract

An apparatus for image segmentation according to an embodiment includes an acquirer configured to acquire one or more images in which an object is photographed and a segmentation performer configured to perform segmentation on the one or more images using a segmentation model which is deep learned through a plurality of images, in which the segmentation model is a U-Net-based model including a first type module based on depth-wise separable convolution (DSC) and a second type module based on global context network (GCNet).

Claims (34)

1. An apparatus for image segmentation, which is implemented with a computing device that includes one or more processors and a memory for storing one or more programs executed by the one or more processors, the apparatus comprising:

an acquirer configured to acquire one or more images in which an object is photographed; and

a segmentation performer configured to perform segmentation on the one or more images using a segmentation model which is deep learned through a plurality of images,

wherein the segmentation model is a U-Net-based model including a first type module based on depth-wise separable convolution (DSC) and a second type module based on global context network (GCNet),

wherein the first type module is configured to include a plurality of depth-wise convolution layer blocks for extracting feature information of a feature map and a plurality of point-wise convolution layer blocks including a first point-wise convolution layer block and a second point-wise convolution layer block for controlling a number of channels of the feature map,

wherein the first type module is configured to:

calculate a map for extracting feature information by repeatedly applying the depth-wise convolution layer block and the point-wise convolution layer block to an input feature map, and increase a number of output channels by using the first point-wise convolution layer block and then adjust the increased number of the output channels to the number before increasing by using the second point-wise convolution layer block;

calculate a map for controlling the number of channels by applying only the point-wise convolution layer blocks to the input feature map; and

sum the map for extracting the feature information and the map for adjusting the number of channels and output a result of the summation.

2. The apparatus of claim 1 , wherein the one or more images include a tomographic image of a brain obtained by magnetic resonance imaging (MRI).

3. The apparatus of claim 2 , wherein the segmentation performer is configured to segment the tomographic image of the brain into a plurality of sections, and determine one or more sections satisfying a preset condition among the plurality of sections as a stroke lesion.

4. The apparatus of claim 1 , wherein the segmentation model is a U-Net-based model in which at least some of a plurality of convolution layer blocks in the segmentation model are replaced with the first type module.

5. The apparatus of claim 1 , wherein the segmentation model is a U-Net-based model in which the second type module is disposed between an encoder in which down sampling is performed and a decoder in which up sampling is performed in the segmentation model.

6. The apparatus of claim 1 , wherein the second type module is configured to include a first convolution layer block for extracting feature information of a feature map, a second convolution layer block, and a global context block (GCBlock) based on the global context network.

7. The apparatus of claim 6 , wherein the second type module is configured to:

calculate a global feature map by applying the first convolution layer block, the second convolution layer block, and the global context block to an input feature map; and

sum the input feature map and the global feature map and output a result of the summation.

8. A method for image segmentation performed by a computing device that includes one or more processors and a memory for storing one or more programs executed by the one or more processors, the method comprising:

acquiring one or more images in which an object is photographed; and

performing segmentation on the one or more images using a segmentation model which is deep learned through a plurality of images,

wherein the segmentation model is a U-Net-based model including a first type module based on depth-wise separable convolution (DSC) and a second type module based on global context network (GCNet),

wherein the first type module is configured to include a plurality of depth-wise convolution layer blocks for extracting feature information of a feature map and a plurality of point-wise convolution layer blocks including a first point-wise convolution layer block and a second point-wise convolution layer block for controlling the number of channels of the feature map,

wherein the first type module is configured to:

calculate a map for extracting feature information by repeatedly applying the depth-wise convolution layer block and the point-wise convolution layer block to an input feature map, and, in calculating the map for extracting feature information, increase a number of output channels using a first point-wise convolution layer block of the plurality of point-wise convolution layer blocks and then adjust the increased number of the output channels to the number before increasing by using a second point-wise convolution layer block;

calculate a map for controlling the number of channels by applying only the point-wise convolution layer block to the input feature map; and

sum the map for extracting the feature information and the map for adjusting the number of channels and output a result of the summation.

9. The method of claim 8 , wherein the one or more images include a tomographic image of a brain obtained by magnetic resonance imaging (MRI).

10. The method of claim 9 , wherein in the performing of segmentation, the tomographic image of the brain is segmented into a plurality of sections, and one or more sections satisfying a preset condition among the plurality of sections is determined as a stroke lesion.

11. The method of claim 8 , wherein the segmentation model is a U-Net-based model in which at least some of a plurality of convolution layer blocks in the segmentation model are replaced with the first type module.

12. The method of claim 8 , wherein the segmentation model is a U-Net-based model in which the second type module is disposed between an encoder in which down sampling is performed and a decoder in which up sampling is performed in the segmentation model.

13. The method of claim 8 , wherein the second type module is configured to include a first convolution layer block for extracting feature information of a feature map, a second convolution layer block, and a global context block (GCBlock) based on the global context network.

14. The method of claim 13 , wherein the second type module is configured to:

calculate a global feature map by applying the first convolution layer block, the second convolution layer block, and the global context block to an input feature map; and

sum the input feature map and the global feature map and output a result of the summation.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 28, 2022
From: CHOI, GYU SANG; SHIN, HYUN KWANG
To: RESEARCH COOPERATION FOUNDATION OF YEUNGNAM UNIVERSITY
Reel/Frame 060331/0605 →
Priority Claims (1)
KR 10-2021-0024711 · Feb 24, 2021 · national
Continuity (1)
Related Publication 20230005152A1 · Jan 5, 2023
References Cited (11)
KR 102094320B1 · 2020 [cited by applicant]
Ronneberger, Olaf, Philipp Fischer, and Thomas Brox. “U-net: Convolutional networks for biomedical image segmentation.” MICCAI 2015: 18th international conference, Munich, Germany, Oct. 5-9, 2015, proceedings, part III … [cited by examiner]
Cao, Yue, et al. “GCNet: Non-Local Networks Meet Squeeze-Excitation Networks and Beyond.” 2019 IEEE/CVF International Conference on Computer Vision Workshop (ICCVW). IEEE, 2019. (Year: 2019). [cited by examiner]
Qi, Kehan, et al. “X-net: Brain stroke lesion segmentation based on depthwise separable convolution and long-range dependencies. ” MICCAI 2019: 22nd International Conference, Shenzhen, China, Oct. 13-17, 2019, Proceedin… [cited by examiner]
Li, Zhuoying, et al. “Memory-efficient automatic kidney and tumor segmentation based on non-local context guided 3D U-Net.” MICCAI 2020: 23rd International Conference, Lima, Peru, Oct. 4-8, 2020, Proceedings, Part IV 23… [cited by examiner]
Siddique, Nahian, et al. “U-Net and its variants for medical image segmentation: theory and applications.” arXiv preprint arXiv: 2011.01118v1 (2020). (Year: 2020). [cited by examiner]
Shin, Hyunkwang, et al. “Automated segmentation of chronic stroke lesion using efficient U-Net architecture.” Biocybernetics and Biomedical Engineering 42.1 (2022): 285-294. (Year: 2022). [cited by examiner]
Office action issued on May 18, 2023 from Korean Patent Office in a counterpart Korean Patent Application No. 10-2021-0024711 (all the cited references are listed in this IDS.) (English translation is also submitted her… [cited by applicant]
Kehan Qi et al., “X-Net: Brain Stroke Lesion Segmentation Based on Depthwise Separable Convolution and Long-Range Dependencies”, MICCAI 2019: 22nd International Conference (Oct. 10, 2019). [cited by applicant]
Jie Chen et al., “Road Crack Image Segmentation Using Global Context U-net”, Proceedings of the 2019 3rd international conference on computer science and artificial intelligence (Mar. 4, 2020). [cited by applicant]
Zhuoying Li et al., “Memory-Efficient Automatic Kidney and Tumor Segmentation Based on Non-local Context Guided 3D U-Net”, MICCAI 2020: 23rd International Conference (Sep. 29, 2020). [cited by applicant]
Cited By (1)
US 12,555,369