IP Library Granted Patent US 12,462,392
Granted Patent B2
US 12,462,392 · App. 18/090,408 · Granted Nov 4, 2025

Methods and apparatuses for auto segmentation using bounding box

Inventors: Dae Hoon Kim (Seoul, KR); Jey Yoon Ru (Seoul, KR); Luca Pimenta Medeiros (Seoul, KR)
Assignee: NUVI LABS CO., LTD.
G06T7/11G06T7/194G06T7/90G06T2207/20081G06T2207/20112
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,462,392
App. No.
18/090,408
Granted
Nov 4, 2025
Kind
B2
Abstract

Provided are a method and an apparatus for auto segmentation using a bounding box. A method for auto segmentation using a bounding box according to one embodiment of the present disclosure comprises receiving a first object image including an object labeled with a bounding box, which is a pre-learning target, learning a segmentation model by classifying an object and a background from the bounding box of the received first object image, and segmenting an object from a second object image, which is an identification target, using the learned segmentation model.

Claims (26)

1 . A method for auto segmentation executed by an apparatus for auto segmentation, the method comprising:

receiving a first object image including an object labeled with a bounding box, which is a pre-learning target;

learning a segmentation model by classifying an object and a background from the bounding box of the received first object image; and

segmenting an object from a second object image, which is an identification target, using the learned segmentation model,

wherein the learning a segmentation model calculates a mask loss by summing a first loss calculated by using a mask and a bounding box predicted in the first object image and a second loss calculated by using a mask predicted in the first object image and a color similarity map between individual pixels and their neighboring pixels within the bounding box and learns the segmentation model using the calculated mask loss.

2 . The method of claim 1 , wherein the learning a segmentation model classifies an object and a background using a color similarity map in a bounding box of the received first object image and learns the segmentation model through the classification of the object and the background.

3 . The method of claim 1 , wherein the learning a segmentation model learns the segmentation model by determining whether a pixel located in the bounding box of the received first object image belongs to one of objects to be trained or the background.

4 . The method of claim 1 , wherein the learning a segmentation model calculates a first loss so that the prediction mask is restricted to stay within the bounding box.

5 . The method of claim 1 , wherein the learning a segmentation model calculates a second loss so that an area occupied by the prediction mask contains the minimum of a background area and the maximum of an object area.

6 . The method of claim 1 , further including:

performing auto-labeling in a manner of re-training through user inspection for a bounding box exceeding a preset prediction error value.

7 . The method of claim 1 , wherein the identifying an object from a second object image identifies an object from the second object image, which is an identification target, using the learned segmentation model and a pre-learned multimodal model.

8 . An apparatus for auto segmentation using a bounding box comprising:

a memory storing one or more programs; and

a processor executing the stored one or more programs, wherein the processor is configured to:

receive a first object image including an object labeled with a bounding box, which is a pre-learning target,

learn a segmentation model by classifying an object and a background from the bounding box of the received first object image, and

segment an object from a second object image, which is an identification target, using the learned segmentation model,

wherein the processor calculates a mask loss by summing a first loss calculated by using a mask and a bounding box predicted in the first object image and a second loss calculated by using a mask predicted in the first object image and a color similarity map between individual pixels and their neighboring pixels within the bounding box and learn the segmentation model using the calculated mask loss.

9 . The apparatus of claim 8 , wherein the processor classifies an object and a background using a color similarity map in a bounding box of the received first object image and learns the segmentation model through the classification of the object and the background.

10 . The apparatus of claim 8 , wherein the processor learns the segmentation model by determining whether a pixel located in the bounding box of the received first object image belongs to one of objects to be trained or the background.

11 . The apparatus of claim 8 , wherein the processor calculates a first loss so that the prediction mask is restricted to stay within the bounding box.

12 . The apparatus of claim 8 , wherein the processor calculates a second loss so that an area occupied by the prediction mask contains the minimum of a background area and the maximum of an object area.

13 . The apparatus of claim 8 , wherein the processor performs auto-labeling in a manner of re-training through user inspection for a bounding box exceeding a preset prediction error value.

14 . The apparatus of claim 8 , wherein the processor identifies an object from the second object image, which is an identification target, using the learned segmentation model and a pre-learned multimodal model.

15 . The apparatus of claim 8 comprising a database storing the first object image including the object labeled with the bounding box, which is the pre-learning target, wherein the processor is configured to receive the first object image from the database.

Assignments (2)
CORRECTIVE ASSIGNMENT TO CORRECT THE REMOVE THE 4TH INVENTOR PREVIOUSLY RECORDED AT REEL: 062264 FRAME: 0656. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 24, 2023
From: KIM, DAE HOON; RU, JEY YOON; PIMENTA MEDEIROS, LUCA
To: NUVI LABS CO., LTD.
Reel/Frame 063430/0066 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 3, 2023
From: KIM, DAE HOON; RU, JEY YOON; PIMENTA MEDEIROS, LUCA; KIM, HONG YEOB
To: NUVI LABS CO., LTD.
Reel/Frame 062264/0656 →
Priority Claims (1)
KR 10-2022-0145994 · Nov 4, 2022 · national
Continuity (1)
Related Publication 20240161303A1 · May 16, 2024
References Cited (12)
US 10402978B1 · Kim · 2019 [cited by examiner]
US 10692002B1 · Kim · 2020 [cited by examiner]
US 20150071487A1 · Sheng · 2015 [cited by examiner]
US 20180189596A1 · Lee · 2018 [cited by examiner]
US 20220391616A1 · Tsui · 2022 [cited by examiner]
US 20230023164A1 · Parameswaran · 2023 [cited by examiner]
US 20230360238A1 · Alexander · 2023 [cited by examiner]
Tian et al., “BoxInst: High-Performance Instance Segmentation With Box Annotations”, 2021, CVPR, pp. 5443-5452 (Year: 2021). [cited by examiner]
Tian, Z. et al., “BoxInst: High-Performance Instance Segmentation With Box Annotations,” arXiv (Cornell University), (2020). [cited by applicant]
Tian, Z. et al., “Conditional Convolutions for Instance Segmentation,” arXiv (Cornell University), (2020). [cited by applicant]
Kirillov, A. et al., “Segment Anything,” arXiv.org, (2023). [cited by applicant]
Extended European Search Report dated Apr. 12, 2024 as received in Application No. 23210773.0. [cited by applicant]