IP Library Granted Patent US 12,482,268
Granted Patent B2
US 12,482,268 · App. 17/981,933 · Granted Nov 25, 2025

Data construction and learning system and method based on method of splitting and arranging multiple images

Inventors: Jin Woo Kim (Daejeon, KR); Ki Tae Kim (Daejeon, KR)
Assignee: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE
G06V20/56G06T5/50G06V10/759G06T2207/20221G06V2201/07
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,482,268
App. No.
17/981,933
Granted
Nov 25, 2025
Kind
B2
Abstract

The present disclosure relates to a data construction and learning system and method based on a method of splitting and arranging multiple images. The data construction and learning system based on a method of splitting and arranging multiple images includes an input unit configured to receive images captured by a plurality of cameras disposed in a vehicle, a memory in which a program for merging the images into a single image and estimating information on a road situation and an object has been stored, and a processor configured to execute the program. The processor merges and recognizes, as one situation, road situations and objects redundantly included in the images.

Claims (23)

1 . A data construction and learning system based on a method of splitting and arranging multiple images, the system comprising:

a processor having an input that receives images captured by a plurality of cameras disposed in a vehicle; and

a memory that stores a program for merging the images into a single image and estimating information on a road situation and an object; and

wherein the processor executes the program,

wherein the processor merges and recognizes, as one situation, road situations and objects redundantly included in the images,

wherein the input receives images having different field of views (FOVs) from the plurality of cameras,

wherein the processor constructs the single image in which the images are disposed for each section by rearranging and merging the images having the different FOVs, and

wherein the processor constructs data to be delivered in a learning level in a way to merge and manage information of an object that is partially displayed for each section and whose region of interest (ROI) is fully displayed and becomes relatively small by performing annotation on the object within the single image.

2 . The system of claim 1 , wherein:

the processor constructs structural data of objects within images by using the single image and calibrated information based on a distance sensor and a position sensor, and

the structural data comprises information of an object, a model, a class, a position, heading, a size, and a target.

3 . The system of claim 2 , wherein the processor

merges the images into the single image comprising an image main sector and an image sub-sector, and

outputs final estimation information based on only a merged image input upon real-time execution by using learning results using the structural data of the objects within the images.

4 . A data construction and learning method based on a method of splitting and arranging multiple images, which is performed by a data construction and learning system based on a method of splitting and arranging multiple images, the method comprising steps of:

(a) obtaining image data photographed by a plurality of cameras;

(b) merging the image data into a single frame having a tile array;

(c) converting, into an image domain, object information labeled on an object within the single frame; and

(d) generating learning data by merging an image converted into the image domain and labeling information of the image domain,

wherein the step (b) comprises constructing the data capable of being delivered in a learning level by merging and managing an object partially displayed in the single frame for each section, information of a region of interest (ROI) displayed in a complete form, and information of an object that is disposed at a long distance and that has a relatively small ROI.

5 . The method of claim 4 , wherein the step (a) comprises obtaining the image data from the plurality of cameras whose photographing information has been set so that the cameras have different field or views (FOVs) and resolution.

6 . The method of claim 4 , wherein the step (c) comprises converting, into the image domain, the object information comprising object coordinate information in a birds' eye view, form modeling information corresponding to an object size, and moving direction information.

7 . The method of claim 4 , wherein the step (d) comprises performing learning using structural data of objects within images comprising information of an object, a model, a class, a position, heading, a size, and a target.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 7, 2022
From: KIM, JIN WOO; KIM, KI TAE
To: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE
Reel/Frame 061677/0193 →
Priority Claims (2)
KR 10-2021-0153298 · Nov 9, 2021 · national
KR 10-2022-0099345 · Aug 9, 2022 · national
Continuity (1)
Related Publication 20230146228A1 · May 11, 2023
References Cited (13)
US 10692262B2 · Lim · 2020 [cited by examiner]
US 20120163671A1 · Choi et al. · 2012 [cited by applicant]
US 20130211656A1 · An et al. · 2013 [cited by applicant]
US 20250022144A1 · Feng · 2025 [cited by examiner]
EP 3163506A1 · 2017 [cited by examiner]
KR 1020200027425 · 2020 [cited by applicant]
KR 1020200095357 · 2020 [cited by applicant]
KR 1020210018434 · 2021 [cited by applicant]
KR 1020210128424 · 2021 [cited by applicant]
KR 102342945 · 2021 [cited by applicant]
Li, Detection of Road Objects With Small Appearance in Images for Autonomous Driving in Various Traffic Situations Using a Deep Learning Based Approach, IEEE Access, Digital Object Identifier 10.1109/ACCESS.2020.3036620… [cited by examiner]
Venkateswara, Deep-Learning Systems for Domain Adaptation in Computer Vision, IEEE Signal Processing Magazine, Nov. 2017 (Year: 2017). [cited by examiner]
Heng, Project AutoVision: Localization and 3D Scene Perception for an Autonomous Vehicle with a Multi-Camera System, 2019 International Conference on Robotics and Automation (ICRA), Palais des congres de Montreal, Montr… [cited by examiner]