IP Library › Granted Patent US 12,705,475
Granted Patent B2
US 12,705,475 · App. 17/639,054 · Granted Aug 11, 2026

Simultaneous measurements of gradients in optical networks

Inventors: Shanhui Fan (Stanford, CA); Tyler William Hughes (San Diego, CA); David A.B. Miller (Stanford, CA); Sunil K. Pai (Stanford, CA); Olav Solgaard (Stanford, CA); Ian A.D. Williamson (Palo Alto, CA)
Assignee: The Board of Trustees of the Leland Stanford Junior University
G06N3/067
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,705,475
App. No.
17/639,054
Granted
Aug 11, 2026
Kind
B2
Abstract

Improved training of optical neural networks is provided. In one example: 1) we choose input and target vectors, we program those into an input vector generator and a measurement unit, respectively, we turn on the optical input source power, and we monitor the electrical signal representing the cost function. 2) we can then modulate two or more controllable elements inside the optical network at different frequencies and look for the size and sign of the corresponding distinct AC variations in the measured cost function, simultaneously giving us the gradients with respect to each element.

Claims (12)

1 . A method of training a photonic neural network, the method comprising:

providing an optical network having two or more optical inputs, two or more optical outputs and two or more control inputs, wherein control signals provided to the control inputs determine an input-output relation between the optical inputs and the optical outputs;

providing one or more predetermined input training patterns to the optical inputs of the optical network;

providing an adjustable output analyzer connected to the optical outputs of the optical network and configured to provide a cost function output;

simultaneously measuring two or more derivatives of the cost function with respect to the control signals as part of training the photonic neural network with the one or more predetermined input training patterns;

wherein the simultaneously measuring two or more derivatives of the cost function with respect to the control signals comprises dithering two or more of the control signals at two or more distinct dither frequencies and measuring corresponding distinct frequency components in the cost function output.

2 . The method of claim 1 , wherein two or more predetermined input training patterns are provided to the optical inputs of the optical network at various times, whereby the two or more derivatives of the cost function are analog time averages over the two or more predetermined input training patterns.

3 . The method of claim 1 , wherein two or more predetermined input training patterns are provided to the optical inputs of the optical network at two or more distinct wavelengths, whereby the two or more derivatives of the cost function are analog wavelength averages over the two or more predetermined input training patterns.

4 . The method of claim 1 , wherein the one or more predetermined input training patterns are provided as modulated input training patterns, whereby the corresponding distinct frequency components in the cost function output resulting from the two or more distinct dither frequencies are heterodyne shifted away from the two or more distinct dither frequencies.

5 . The method of claim 1 , further comprising adjusting the control signals to optimize the cost function with an optimization method that makes use of the two or more derivatives of the cost function with respect to the control signals, whereby the photonic neural network is trained according to the one or more predetermined input training patterns.

6 . The method of claim 1 , wherein the optical network includes two or more meshes of linear optical components connected in alternating series via one or more nonlinearity units, and wherein the control inputs include at least inputs to each of the two or more meshes of linear optical components.

7 . The method of claim 1 , wherein the optical network includes at least one optical element having a compound control input, wherein the compound control input includes a first input and a second input, wherein the first input has a lower bandwidth than the second input, and wherein a dither of the compound control input is delivered via the second input.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 28, 2022
From: FAN, SHANHUI; HUGHES, TYLER WILLIAM; MILLER, DAVID A.B.; PAI, SUNIL K.; SOLGAARD, OLAV; WILLIAMSON, IAN A.D.
To: THE BOARD OF TRUSTEES OF THE LELAND STANFORD JUNIOR UNIVERSITY
Reel/Frame 059118/0265 →
Continuity (2)
Provisional Application 62897657 · Sep 9, 2019
Related Publication 20220327369A1 · Oct 13, 2022
References Cited (13)
US 7062166B2 · Jacobowitz · 2006 [cited by examiner]
US 10042190B2 · Liu · 2018 [cited by examiner]
US 10338319B2 · Miller · 2019 [cited by applicant]
US 10496069B2 · Nazarathy · 2019 [cited by examiner]
US 11460753B2 · Hughes · 2022 [cited by examiner]
US 12026615B2 · Hughes · 2024 [cited by examiner]
US 20170351293A1 · Carolan · 2017 [cited by applicant]
US 20180061344A1 · Kurokawa · 2018 [cited by applicant]
US 20180259707A1 · Zalevsky · 2018 [cited by applicant]
US 20200250532A1 · Shen · 2020 [cited by examiner]
“High-speed, model-free adaptive control using parallel synchronous detection” by Loizos et al., 20th Symposium on Integrated Circuits and System Design (SBCCI07), pp. 224-229 (Year: 2007). [cited by examiner]
“Parallel Dither and Dropout for Regularising Deep Neural Networks” by Simpson, arXiv preprint arXiv:1508.07130 (Year: 2015). [cited by examiner]
“The multidither principle in adaptive optics” by O'Meara, J. Opt. Soc. Am., vol. 67, No. 3, pp. 306-315 (Year: 1977). [cited by examiner]