IP Library › Granted Patent US 12,423,836
Granted Patent B2
US 12,423,836 · App. 18/125,739 · Granted Sep 23, 2025

Camera-radar fusion using correspondences

Inventors: Francesco Michielin (Stuttgart, DE); Oliver Vogel (Stuttgart, DE)
Assignee: Sony Group Corporation
G06T7/248G06T7/10G08G1/16G06T2207/10028G06T2207/20021G06T2207/20221G06T2207/30252
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,423,836
App. No.
18/125,739
Granted
Sep 23, 2025
Kind
B2
Abstract

A method of performing sensor fusion of data obtained from a camera ( 51 ) and a supplemental sensor ( 52 ), the method comprising performing patch tracking ( 55 ) on image data ( 53 ) provided by the camera ( 51 ) to determine tracked patches ( 56 ), and performing a fusion ( 57 ) of the image data ( 53 ) obtained from the camera ( 51 ) with supplemental data ( 54 ) provided by the supplemental sensor ( 52 ) based on the tracked patches ( 56 ).

Claims (26)

1. A method of performing sensor fusion of data obtained from a camera and a supplemental sensor, the method comprising:

performing patch tracking on image data provided by the camera to determine tracked patches;

determining, for each tracked patch, a patch scale ratio based on a change in size of the patch between two frames of the image data by calculating a determinant of a 2×2 upper left sub-matrix of an affine transformation matrix representing the change in the patch between the two frames;

determining, for each part of supplemental data provided by the supplemental sensor, an equivalent scale ratio based on a ratio of distances to an object at two different times, wherein the ratio is calculated using a radial velocity measurement and a current distance measurement;

identifying, for each part of the supplemental data, a corresponding tracked patch by matching the equivalent scale ratio of the part of supplemental data to the patch scale ratio of the tracked patch; and

performing a fusion of the image data obtained from the camera with supplemental data provided by the supplemental sensor based on the tracked patches.

2. The method of claim 1 , wherein performing patch tracking comprises performing an affine patch tracking.

3. The method of claim 1 , comprising applying part of the supplemental data provided by the supplemental sensor to a patch of the tracked patches.

4. The method of claim 1 , wherein the supplemental data is provided by the supplemental sensor in the form of a point cloud.

5. The method of claim 1 , wherein the part of the supplemental data is a point of a point cloud.

6. The method of claim 1 , wherein the supplemental sensor is a radar sensor.

7. The method of claim 1 , wherein the equivalent scale ratio is determined from distance information and radial velocity information of a point of the radar point cloud.

8. The method of claim 1 , comprising, for a point that is considered being located close to a patch of the tracked patches, comparing an equivalent scale ratio of the point with a scale ratio of the patch obtained from patch tracking to determine if the scale ratios match.

9. The method of claim 1 , comprising discarding those points of points that are considered being located close to a patch of the tracked patches, whose equivalent scale ratios do not match with a scale ratio of the patch obtained from affine patch tracking.

10. The method of claim 1 , comprising performing object segmentation on the image data captured by the camera.

11. The method of claim 1 , comprising averaging supplemental data related to patches associated with an object, and associating the averaged information with the object.

12. The method of claim 1 , wherein the supplemental data comprises Time to Collision information.

13. The method of claim 1 , wherein the sensor fusion is applied in an automotive context.

14. A device, comprising:

processing circuitry, the processing circuitry being configured to

perform patch tracking on image data provided by a camera to determine tracked patches,

determine, for each tracked patch, a patch scale ratio based on a change in size of the patch between two frames of the image data by calculating a determinant of a 2×2 upper left sub-matrix of an affine transformation matrix representing the change in the patch between the two frames,

determine, for each part of supplemental data provided by the supplemental sensor, an equivalent scale ratio based on a ratio of distances to an object at two different times, wherein the ratio is calculated using a radial velocity measurement and a current distance measurement,

identify, for each part of the supplemental data, a corresponding tracked patch by matching the equivalent scale ratio of the part of supplemental data to the patch scale ratio of the tracked patch, and

perform a fusion of the image data obtained from the camera with supplemental data provided by a supplemental sensor based on the tracked patches.

15. A system comprising a camera, a supplemental sensor, and the device of claim 14 , the device being configured to perform patch tracking on image data provided by the camera to determine tracked patches, and to perform a fusion of the image data obtained from the camera with supplemental data provided by the supplemental sensor based on the tracked patches.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 26, 2024
From: MICHIELIN, FRANCESCO; VOGEL, OLIVER
To: SONY GROUP CORPORATION
Reel/Frame 068397/0366 →
Priority Claims (1)
EP 22166015 · Mar 31, 2022 · regional
Continuity (1)
Related Publication 20230316546A1 · Oct 5, 2023
References Cited (23)
US 10593042B1 · Douillard · 2020 [cited by examiner]
US 20190337451A1 · Bacchus · 2019 [cited by examiner]
US 20190353775A1 · Kirsch · 2019 [cited by examiner]
US 20200175315A1 · Gowaikar · 2020 [cited by examiner]
US 20200180638A1 · Kanoh · 2020 [cited by examiner]
US 20200219264A1 · Brunner · 2020 [cited by examiner]
US 20210033722A1 · Søndergaard · 2021 [cited by examiner]
US 20210048521A1 · Leduc · 2021 [cited by examiner]
US 20210209785A1 · Unnikrishnan · 2021 [cited by examiner]
US 20220130109A1 · Arbabian · 2022 [cited by examiner]
US 20220153306A1 · Imran · 2022 [cited by examiner]
US 20220262129A1 · Cao · 2022 [cited by examiner]
US 20220377242A1 · Camacho · 2022 [cited by examiner]
US 20230176205A1 · Yu · 2023 [cited by examiner]
US 20230316546A1 · Michielin · 2023 [cited by examiner]
CN 108847026A · 2018 [cited by applicant]
CN 111652914A · 2020 [cited by applicant]
Feng et al., “Deep Multi-modal Object Detection and Semantic Segmentation for Autonomous Driving: Datasets, Methods, and Challenges”, arXiv:1902.07830v4 [cs.RO], Feb. 8, 2020, pp. 1-27. [cited by applicant]
Lim et al., “Radar and Camera Early Fusion for Vehicle Detection in Advanced Driver Assistance Systems”, Machine Learning for Autonomous Driving Workshop at the 33rd Conference on Neural Information Processing Systems (… [cited by applicant]
Chang et al., “Spatial Attention Fusion for Obstacle Detection Using MmWave Radar and Vision Sensor”, Sensors, vol. 20, No. 4, Available Online At: https://doi.org/10.3390/s20040956, Feb. 11, 2020, pp. 1-21. [cited by applicant]
Nobis et al., “A Deep Learning-based Radar and Camera Sensor Fusion Architecture for Object Detection”, arXiv:2005.07431v1 [cs.CV], May 15, 2020, 7 pages. [cited by applicant]
Jedoui et al., “Lecture 18: Tracking”, Computer Vision: Foundations and Applications (CS 131, 2017), Available Online At: http://vision.stanford.edu/teaching/cs131_fall1718/files/18_notes.pdf, 2017, pp. 1-6. [cited by applicant]
Molton et al., “Locally Planar Patch Features for Real-Time Structure from Motion”, In Proc. British Machine Vision Conference, Available Online At: http://citeseerx.ist.psu.edu/viewdoc/summary?doi=10.1.1.422.1885, 2004… [cited by applicant]