IP Library › Granted Patent US 12,444,051
Granted Patent B2
US 12,444,051 · App. 17/428,122 · Granted Oct 14, 2025

System for OCT image translation, ophthalmic image denoising, and neural network therefor

Inventors: Arindam Bhattacharya (Dublin, CA); Warren Lewis (Yellow Springs, OH); Sophie Kubach (Menlo Park, CA); Lars Omlor (Pleasanton, CA); Mary Durbin (San Francisco, CA)
Assignees: CARL ZEISS MEDITEC, INC.; CARL ZEISS MEDITEC AG
G06T7/0014A61B3/102G06T5/50G06T5/70G06T7/37G06T11/008G16H30/40G06T2207/10084G06T2207/10101G06T2207/20081G06T2207/20084G06T2207/30041
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,444,051
App. No.
17/428,122
Granted
Oct 14, 2025
Kind
B2
Abstract

An OCT system includes a machine learning (ML) model trained to receive a single OCT scan/image and provide an image translation and/or denoise function. The ML model may be based on a neural network (NN) architecture including a series of encoding modules in a contracting path followed by a series of decoding modules in an expanding path leading to an output convolution module. An intermediate error module determines a deep error measure, e.g., between a training output image and at least one encoding module and/or decoding module, and an error from the output convolution module is combined with the deep error measure. The NN may be trained using true averaged images as ground truth, training outputs. Alternatively, the NN may be trained using randomly selected, individual OCT images/scans as training outputs.

Claims (86)

1. An optical coherence tomography (OCT) system comprising:

a light source for generating a beam of light;

a beam splitter having a beam-splitting surface for directing a first portion of the light into a reference arm and a second portion of the light into a sample arm;

optics for directing the light in the sample arm to one or more locations on a sample;

a detector for receiving light returning from the sample and reference arms and generating signals in response thereto;

a processor for converting the signals into a first image and submitting the first image to an image translation module that translates the first image to a second image characterized by one or more of decreased noise artifacts, jitter and minimized creation of fictional structures as compared to the first image; and

an output display for displaying a system output image based on the second image;

wherein the image translation module includes a machine learning module based on a neural network trained using a set of training input images and a target set of training output images,

wherein the neural network has an input module configured to receive a current training input image, a plurality of intermediate processing modules following the input module, an intermediate error module that determines an intermediate error based on an output of at least one select intermediate processing module and a current training output image, and an output module following the plurality of intermediate processing modules that determines a preliminary output error based on its output and the current training output image, and

wherein a total loss error for a current training cycle is defined based on a combination of the preliminary output error and the intermediate error.

2. The system of claim 1 , wherein the processor further combines a current second image with one or more previously obtained second images to generate the system output image.

3. The system of claim 2 , wherein the current second image is combined with the one or more previously obtained second images by one of direct averaging or weighted averaging with second images of higher image quality being weighted more heavily.

4. The system of claim 1 , wherein:

the processor defines a plurality of the first images, submits the plurality first images to the image translation module to produce a corresponding plurality of second images, and calculates motion contrast information from the plurality of second images using an OCT angiography (OCTA) processing technique; and

the system output image displays the motion contrast information.

5. The system of claim 1 , wherein the processor:

defines a plurality of the first images;

submits the plurality of first images to the image translation module to produce a corresponding plurality of second images;

applies an image registration technique to the plurality of second images to produce image alignment settings; and

aligns the plurality of first images based at least in part on the image alignment settings of the plurality of second images.

6. The system of claim 1 , wherein:

the first image is divided into a plurality of first image segments; and

the image translation module individually translates each first image segment into a corresponding second image segment, and combines the second image segments to construct the second image.

7. The system of claim 1 , wherein:

at least one of the training output images is defined as the averaging of a set of OCT test images of the same region of a test sample; and

at least a fraction of the training input images is included in the set of OCT test images.

8. The system of claim 1 , wherein:

the first image is of a first region of the sample; and

the second image has characteristics defined as the averaging of multiple hypothetical OCT scans of the first region with the first image.

9. The system of claim 1 , wherein the training of the neural network includes:

collecting a plurality of OCT test images of a target ophthalmic region;

averaging the plurality of OCT test images to define a corresponding averaged-image of the target ophthalmic region;

separately and individually inputting the OCT test images of the target ophthalmic region as training input images to the neural network, and providing their corresponding averaged-image as their individually corresponding training output image for the neural network.

10. The system of claim 9 , wherein the training of the neural network further includes:

dividing each OCT test image into a plurality of test segments;

dividing their corresponding averaged-image into a plurality of corresponding ground truth segments;

correlating test segments to corresponding ground truth segments;

separately and individually submitting the correlated test segments to the neural network as training input images and providing their correlated ground truth segments as training output images for the neural network.

11. The system of claim 9 , further including combining a currently inputted OCT test image with a corresponding current output of the neural network to define a combination network output, and comparing the combination network output with the corresponding training output image to determine the total loss error for a current training cycle.

12. The system of claim 1 , wherein the training input images and training output images include a mixture of images of healthy eyes and images of diseased eyes.

13. The system of claim 1 , wherein:

the first image is of a first imaging modality; and

the second image simulates a second imaging modality different than the first imaging modality.

14. The system of claim 13 , wherein the first and second modalities are a mixture including one or more of time domain OCT, spectral domain OCT, swept source OCT, and adaptive optics OCT (AO-OCT).

15. The system of claim 1 , wherein:

the OCT system is of a first modality;

the machine learning module is trained using third images taken with a first OCT device of the first modality as the set of training input images and fourth images taken with a second OCT device of a second modality as the target set of training output images, the second modality being different than the first modality; and

the second image has features characteristic of an image generated by an OCT system of the second modality.

16. The system of claim 15 , wherein:

the first OCT device of the first modality is of a non-adaptive optics OCT type;

the second OCT device of the second modality is of an adaptive optics OCT type;

the third images obtained by the first OCT device are bigger than the fourth images obtained by the second OCT device, the third images are divided into third image segments of similar size as the fourth images and each third image segment is correlated to a corresponding fourth image; and

the correlated third segments are separately and individually submitted to the neural network as training input images and their correspondingly correlated fourth images are provided as training output images for the neural network.

17. The system of claim 1 , wherein

the plurality of intermediate processing modules form a contracting path and an expanding path, the contracting path following the input module and including a plurality of encoding modules, each encoding module having a convolution stage, an activation function, and a max pooling operation,

the expanding path following the contracting path and having a plurality of decoding modules, each decoding module concatenates its current value with that of a corresponding encoding module;

the output module is an output convolution module excluding a pooling layer and a sigmoid layer activation function, the output convolution module receiving the output from the last decoding module in the expanding path and producing the preliminary output error; and

the intermediate error module determines the intermediate error for at least one encoding module and/or one decoding module.

18. The system of claim 17 , wherein during training of the neural network:

the intermediate error module determines the error measure as an error between the current training output image and the current value of the intermediate error module's corresponding encoding module and/or decoding module; and

the preliminary output error from the output convolution module is based on the current training output image and the current value of the output convolution module.

19. An ophthalmic imaging system comprising:

a processor for acquiring a first image, and submitting the first image to an image modification module that defines a second image based on the first image; and

an output display for displaying an output image based on the second image;

wherein the image modification module includes a machine model based on a neural network trained using a set of training input images and a target set of training output images, the neural network having:

a) an input module for receiving a current training input image;

b) a contracting path following the input module, the contracting path including a plurality of encoding modules, each encoding module having a convolution stage, an activation function, and a max pooling operation;

c) an expanding path following the contracting path, the expanding path having a plurality of decoding modules, each decoding module concatenating its current value with that of a corresponding encoding module;

d) an output convolution module excluding a pooling layer and an activation function, the output convolution module receiving the output from the last decoding module in the expanding path and producing a preliminary output error; and

e) an intermediate error module that determines an error measure for at least one encoding module and/or one decoding module; and

during training of the neural network, the preliminary output error of the output convolution module is combined with the output error of the intermediate error module.

20. The system of claim 19 , wherein the activation function is a rectifier linear unit or sigmoid layer.

21. The system of claim 19 , wherein the preliminary output error is based on a current training output image and the error measure is based on a current output of the at least one encoding module and/or one decoding module and the current training output image.

22. The system of claim 21 , wherein

a training cycle error for a current training cycle is based on the combined errors from the output convolution module and the intermediate error module.

23. The system of claim 19 , wherein the error measure is based on a square loss function.

24. The system of claim 19 , wherein the ophthalmic imaging system is an optical coherence tomography (OCT) angiography system, and the first image is a vasculature image.

25. The system of claim 19 , wherein the ophthalmic imaging system is an optical coherence tomography (OCT) system of a first modality, and training of the neural network includes:

collecting one or more of third images of different target ophthalmic regions using a first OCT system of the first modality;

collecting one or more of fourth images of the same target ophthalmic regions using a second OCT system of a second modality different than the first modality;

defining one or more of the training output images from the one or more fourth images; and

defining one or more of the training input images from the one or more third images, wherein each input training image has a corresponding training output image.

26. The system of claim 25 , wherein:

the first OCT system of a non-adaptive optics OCT type; and

the second OCT system is of an adaptive optics OCT type.

27. The system of claim 19 , wherein one of the set of training input images or target set of training output images are obtained by an optical coherence tomography (OCT) system, and the other of the set of training input images or target set of training output images are obtained by a fundus imaging system.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2025
From: BHATTACHARYA, ARINDAM; LEWIS, WARREN; KUBACH, SOPHIE; OMLOR, LARS; DURBIN, MARY K.
To: CARL ZEISS MEDITEC AG
Reel/Frame 071909/0774 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 3, 2022
From: BHATTACHARYA, ARINDAM; LEWIS, WARREN; KUBACH, SOPHIE; OMLOR, LARS; DURBIN, MARY K.
To: CARL ZEISS MEDITEC, INC.
Reel/Frame 060708/0148 →
Continuity (2)
Provisional Application 62805835 · Feb 14, 2019
Related Publication 20220058803A1 · Feb 24, 2022
References Cited (40)
US 6549801B1 · Chen et al. · 2003 [cited by applicant]
US 6741359B2 · Wei et al. · 2004 [cited by applicant]
US 7301644B2 · Knighton et al. · 2007 [cited by applicant]
US 8967806B2 · Bublitz et al. · 2015 [cited by applicant]
US 8998411B2 · Tumlinson et al. · 2015 [cited by applicant]
US 9332902B2 · Tumlinson et al. · 2016 [cited by applicant]
US 9700206B2 · An et al. · 2017 [cited by applicant]
US 9706915B2 · Everett et al. · 2017 [cited by applicant]
US 9759544B2 · An et al. · 2017 [cited by applicant]
US 10426442B1 · Schnorr · 2019 [cited by examiner]
US 11763549B1 · Shu · 2023 [cited by examiner]
US 20050171438A1 · Chen et al. · 2005 [cited by applicant]
US 20100027857A1 · Wang · 2010 [cited by applicant]
US 20120277579A1 · Sharma et al. · 2012 [cited by applicant]
US 20120307014A1 · Wang · 2012 [cited by applicant]
US 20150208915A1 · Schallek · 2015 [cited by examiner]
US 20160317027A1 · Goto · 2016 [cited by examiner]
US 20170023887A1 · Nakamura et al. · 2017 [cited by applicant]
US 20170256054A1 · Matsumura · 2017 [cited by examiner]
US 20190005603A1 · Chen et al. · 2019 [cited by applicant]
US 20190130217A1 · Wu · 2019 [cited by examiner]
US 20200037872A1 · Shiba · 2020 [cited by examiner]
EP 3404611A1 · 2018 [cited by applicant]
JP 2005525893 · 2005 [cited by applicant]
JP 2008528954 · 2011 [cited by applicant]
WO 2018200493 · 2018 [cited by applicant]
C. Song, M. Ahn, D. Gweon and H. Cho, “High-resolution high-speed spectral domain optical coherence tomography,” 2009 International Symposium on Optomechatronic Technologies, Istanbul, Turkey, 2009, pp. 11-15, doi: 10.1… [cited by examiner]
M. K. S. Yadav and N. P. Singh, “A Survey on Optical Coherence Tomography,” 2023 IEEE International Conference on Computer Vision and Machine Intelligence (CVMI), Gwalior, India, 2023, pp. 1-6, doi: 10.1109/CVMI59935.20… [cited by examiner]
Japan Patent Office; Japan Office Action dated Jan. 25, 2024 in Application No. 2021-547257. [cited by applicant]
Blazkiewicz et al., (2005). “Signal-to-noise ratio study of full-field Fourier-domain optical coherence tomography,” Applied Optics, 44(36):7722-7729. [cited by applicant]
Devalla et al., (2018). “A Deep Learning Approach to Denoise Optical Coherence Tomography Images of the Optic Nerve Head,” arXiv:1809.10589, 19 pages. [cited by applicant]
Halupka et al., (2018). “Retinal optical coherence tomography image enhancement via deep learning”, Biomedical Optics Express, 9(12):6205-6221. [cited by applicant]
Hillmann et al., (2011). “Holoscopy—holographic optical coherence tomography,” Optics Letters, 36(13):2390-2392. [cited by applicant]
International Search Report and Written Opinion received for International Patent Application No. PCT/EP2020/053515 mailed on Mar. 25, 2020, 15 pages. [cited by applicant]
Jin et al., (2017). “Deep Convolutional Neural Network for Inverse Problems in Imaging,” IEEE Transactions On Image Processing, 26(9):4509-4522. [cited by applicant]
Lee et al., (2018). “Generating retinal flow maps from structural optical coherence tomography with artificial intelligence”, bioRxiv, 25, 26 pages. [cited by applicant]
Ma et al., (2018). “Speckle noise reduction in optical coherence tomography images based on edge-sensitive cGAN”, Biomedical Optics Express, 9(11):5129-5146. [cited by applicant]
Nakamura et al., (2007). “High-Speed three dimensional human retinal imaging by line field spectral domain optical coherence tomography,” Optics Express, 15(12):7103-7116. [cited by applicant]
Xu et al., (2018). “Improving the Resolution of Retinal OCT with Deep Learning”, MIUA2018, 046(v2):1-8. [cited by applicant]
Japan Patent Office; Japan Office Action dated May 24, 2024 in Application No. 2021-547257. [cited by applicant]