IP Library › Granted Patent US 12,256,050
Granted Patent B2
US 12,256,050 · App. 17/342,715 · Granted Mar 18, 2025

Image processing apparatus, image processing method, storage medium, manufacturing method of learned model, and image processing system

Inventor: Yoshinori Kimura (Tochigi, JP)
Assignee: CANON KABUSHIKI KAISHA
H04N13/111G06N3/045G06N3/08G06T7/97H04N13/282G06T2207/20081G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,256,050
App. No.
17/342,715
Granted
Mar 18, 2025
Kind
B2
Abstract

An image processing apparatus includes at least one processor or circuit configured to execute a plurality of tasks including an acquisition task configured to acquire two first images made by capturing the same object at two different viewpoints, and an image processing task configured to input the two first images into a machine learning model and to estimate a second image at one or more viewpoints different from the two viewpoints.

Claims (39)

1. An image processing apparatus comprising:

an image sensor including a plurality of pixels, with two light-receiving units only in each of the plurality of pixels, and an optical system that forms images of different viewpoints on the only two light-receiving units in each of the plurality of pixels; and

at least one processor or circuit configured to execute a plurality of tasks including:

an acquisition task configured to acquire (i) two first images made by capturing the same object at two viewpoints separated from each other in a first direction using the image sensor including the only two light-receiving units in each of the plurality of pixels and the optical system that forms the images of the different viewpoints on the only two light-receiving units in each of the plurality of pixels and (ii) an imaging condition associated with the two first images and including both of a focal length of the optical system and an ISO speed of the image sensor;

a feature amount generating task configured to generate two first feature amounts by inputting the two first images, the focal length of the optical system, and the ISO speed of the image sensor into a first neural network in a machine learning model;

a comparing task configured to generate a second feature amount by calculating corresponding points of the two first feature amounts through processing based on a matrix product of the two first feature amounts; and

an image processing task configured to estimate a second image at a viewpoint different from the two viewpoints by inputting the second feature amount into a second neural network in the machine learning model,

wherein the viewpoint different from the two viewpoints is shifted in a second direction orthogonal to the first direction, and

wherein, in the image processing task, an image at a viewpoint other than the two viewpoints is not input into the machine learning model in estimating the second image.

2. The image processing apparatus according to claim 1 , wherein the two first images are respectively input to two first neural networks having the same network architecture and weight as each other.

3. The image processing apparatus according to claim 1 , wherein the image processing task generates a third image that has an in-focus position changed from that of the second image.

4. An image processing method comprising:

an acquisition step of acquiring (i) two first images made by capturing the same object at two viewpoints separated from each other in a first direction using an image sensor including a plurality of pixels, with two light-receiving units only in each of the plurality of pixels, and an optical system that forms images of different viewpoints on the only two light-receiving units in each of the plurality of pixels and (ii) an imaging condition associated with the two first images and including both of a focal length of the optical system and an ISO speed of the image sensor;

a feature amount generating step of generating two first feature amounts by inputting the two first images, the focal length of the optical system, and the ISO speed of the image sensor into a first neural network in a machine learning model;

a comparing step of generating a second feature amount by calculating corresponding points of the two first feature amounts through processing based on a matrix product of the two first feature amounts; and

an image processing step of estimating a second image at a viewpoint different from the two viewpoints by inputting the second feature amount into a second neural network in the machine learning model,

wherein the viewpoint different from the two viewpoints is shifted in a second direction orthogonal to the first direction, and

wherein, in the image processing task, an image at a viewpoint other than the two viewpoints is not input into the machine learning model in estimating the second image.

5. A non-transitory computer-readable storage medium storing a computer program that causes a computer to execute the image processing method comprising:

an acquisition step of acquiring (i) two first images made by capturing the same object at two viewpoints separated from each other in a first direction using an image sensor including a plurality of pixels, with two light-receiving units only in each of the plurality of pixels, and an optical system that forms images of different viewpoints on the only two light-receiving units in each of the plurality of pixels and (ii) an imaging condition associated with the two first images and including both of a focal length of the optical system and an ISO speed of the image sensor;

a comparing step of generating a second feature amount by calculating corresponding points of the two first feature amounts through processing based on a matrix product of the two first feature amounts; and

an image processing step of estimating a second image at a viewpoint different from the two viewpoints by inputting the second feature amount into a second neural network in the machine learning model,

wherein the viewpoint different from the two viewpoints is shifted in a second direction orthogonal to the first direction, and

wherein, in the image processing task, an image at a viewpoint other than the two viewpoints is not input into the machine learning model in estimating the second image.

6. An image processing system including an image processing apparatus and a control apparatus that is capable of communicating with the image processing apparatus,

the image processing apparatus comprising:

an image sensor including a plurality of pixels, with two light-receiving units only in each of the plurality of pixels, and an optical system that forms images of different viewpoints on the only two light-receiving units in each of the plurality of pixels; and

at least one processor or circuit configured to execute a plurality of tasks including:

an acquisition task configured to acquire (i) two first images made by capturing the same object at two viewpoints separated from each other in a first direction using the image sensor including the only two light-receiving units in each of the plurality of pixels and the optical system that forms the images of the different viewpoints on the only two light-receiving units in each of the plurality of pixels and (ii) an imaging condition associated with the two first images and including both of a focal length of the optical system and an ISO speed of the image sensor;

a feature amount generating task configured to generate two first feature amounts by inputting the two first images, the focal length of the optical system, and the ISO speed of the image sensor into a first neural network in a machine learning model;

a comparing task configured to generate a second feature amount by calculating corresponding points of the two first feature amounts through processing based on a matrix product of the two first feature amounts; and

an image processing task configured to estimate a second image at a viewpoint different from the two viewpoints by inputting the second feature amount into a second neural network in the machine learning model,

wherein the viewpoint different from the two viewpoints is shifted in a second direction orthogonal to the first direction, and

wherein, in the image processing task, an image at a viewpoint other than the two viewpoints is not input into the machine learning model in estimating the second image;

wherein the control apparatus transmits a request for executing a process on the two first images, and

wherein the image processing apparatus executes the process on the two first images based on the request.

7. The image processing apparatus according to claim 1 ,

wherein the acquired imaging condition associated with the two first images further includes an aperture of the optical system and a pixel pitch of the image sensor, and

wherein the aperture of the optical system and the pixel pitch of the image sensor are input together with the two first images, the focal length of the optical system, and the ISO speed of the image sensor to the first neural network to generate the two first feature amounts.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 8, 2021
From: KIMURA, YOSHINORI
To: CANON KABUSHIKI KAISHA
Reel/Frame 057022/0524 →
Priority Claims (1)
JP 2020-103809 · Jun 16, 2020 · national
Continuity (1)
Related Publication 20210392313A1 · Dec 16, 2021
References Cited (19)
US 20140079297A1 · Tadayon · 2014 [cited by examiner]
US 20140098191A1 · Rime · 2014 [cited by examiner]
US 20190244379A1 · Venkataraman · 2019 [cited by applicant]
US 20200013154A1 · Jang · 2020 [cited by applicant]
US 20200134366A1 · Xu · 2020 [cited by examiner]
CN 104012088A · 2014 [cited by applicant]
CN 107729948A · 2018 [cited by applicant]
CN 108364310A · 2018 [cited by applicant]
CN 109214265A · 2019 [cited by applicant]
CN 109561296A · 2019 [cited by applicant]
CN 109978936A · 2019 [cited by applicant]
CN 110351548A · 2019 [cited by applicant]
CN 110443874A · 2019 [cited by applicant]
CN 111247559A · 2020 [cited by applicant]
JP 2018124939A · 2018 [cited by applicant]
Kalantari. “Learning-Based View Synthesis for Light Field Cameras.” SIGGRAPH Asia. 2016. 10 pages. [cited by applicant]
Office Action issued in Chinese Appln. No. 202110665990.0, mailed Sep. 4, 2023. English translation provided. [cited by applicant]
Office Action issued in Chinese Appln. No. 202110665990.0 mailed Mar. 19, 2024. English translation provided. [cited by applicant]
Schiopu et al. “Frame-Wise CNN-Based View Synthesis for Light Field Camera Arrays” IEEE. 2019. pp. 1-7. [cited by applicant]