IP Library Granted Patent US 12,468,976
Granted Patent B1
US 12,468,976 · App. 17/332,606 · Granted Nov 11, 2025

Probabilistic inference in machine learning using a quantum oracle

Inventors: Nan Ding (Los Angeles, CA); Masoud Mohseni (Redondo Beach, CA); Hartmut Neven (Malibu, CA)
Assignee: Google LLC
G06N20/00G06N7/01G06N10/60
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,468,976
App. No.
17/332,606
Granted
Nov 11, 2025
Kind
B1
Abstract

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for using a quantum oracle to make inference in complex machine learning models that is capable of solving artificial intelligent problems. Input to the quantum oracle is derived from the training data and the model parameters, which maps at least part of the interactions of interconnected units of the model to the interactions of qubits in the quantum oracle. The output of the quantum oracle is used to determine values used to compute loss function values or loss function gradient values or both during a training process.

Claims (160)

1 . A method performed by a system of one or more computers for probabilistic inference in a model for use in machine learning, the model comprising interconnected units, the method comprising:

receiving data for training the model, the data comprising observed data for training and validating the model;

deriving input to a quantum information processor using the received data and a state of the model, wherein the input maps at least some interactions of different interconnected units of the model to interactions of qubits in the quantum information processor;

providing the input to the quantum information processor for learning part of the inference in the model; and

receiving, from the quantum information processor, data representing the learned inference.

2 . The method of claim 1 , wherein learning part of the inference in the model comprises learning values of variables in a hidden layer of the model.

3 . The method of claim 1 , further comprising computing a value of a loss function or a value of a gradient of the loss function using the data representing the learned inference for training or validating the model.

4 . The method of claim 3 , wherein the loss function is given by

L

T

=

-

log

p

~

(

x

^

,

y

^

)

=

-

F

T

(

log

p

(

x

^

,

s

,

y

^

)

)

+

log

x

,

y

exp

(

F

T

(

log

p

(

x

,

s

,

y

)

)

)

where T represents temperature, x represents variables corresponding to the observed data for use as input data during the training of the model, y represents variables corresponding to the observed data used as labels during the training of the model, s represents variables in a hidden layer of the model, p represents a probability distribution, and F represents a tempered log-partition function.

5 . The method of claim 3 , wherein computing a value of a gradient of the loss function comprises performing a contrastive divergence algorithm, wherein learning part of the inference in the model comprises computing values of variables in a hidden layer of the model as part of the contrastive divergence algorithm.

6 . The method of claim 1 , wherein the observed data comprises variables in visible layers of the model.

7 . A method performed by a quantum information processor, the method comprising:

receiving, as input, data for learning part of an inference in a model comprising interconnected units, wherein

the input is derived using i) data for training the model, the data comprising observed data for training and validating the model, and ii) a state of the model, and

the input maps at least some interactions of different interconnected units of the model to interactions of qubits in the quantum information processor;

processing the input to learn the part of the inference in the model; and

providing, as output, data representing the learned part of the inference.

8 . The method of claim 7 , wherein processing the input to learn part of the inference in the model comprises processing the input to learn values of variables in a hidden layer of the model.

9 . The method of claim 7 , wherein processing the input to learn part of the inference in the model comprises computing values of variables in a hidden layer of the model as part of a contrastive divergence algorithm.

10 . The method of claim 7 , wherein the observed data comprises variables in visible layers of the model.

11 . A system for probabilistic inference in a model for use in machine learning, the model comprising interconnected units, the system comprising one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving data for training the model, the data comprising observed data for training and validating the model;

deriving input to a quantum information processor using the received data and a state of the model, wherein the input maps at least some interactions of different interconnected units of the model to interactions of qubits in the quantum information processor;

providing the input to the quantum information processor for learning part of the inference in the model; and

receiving, from the quantum information processor, data representing the learned inference.

12 . The system of claim 11 , wherein learning part of the inference in the model comprises learning values of variables in a hidden layer of the model.

13 . The system of claim 11 , wherein the operations further comprise computing a value of a loss function or a value of a gradient of the loss function using the data representing the learned inference for training or validating the model.

14 . The system of claim 13 , wherein the loss function is given by

L

T

=

-

log

p

~

(

x

^

,

y

^

)

=

-

F

T

(

log

p

(

x

^

,

s

,

y

^

)

)

+

log

x

,

y

exp

(

F

T

(

log

p

(

x

,

s

,

y

)

)

)

where T represents temperature, x represents variables corresponding to the observed data for use as input data during the training of the model, y represents variables corresponding to the observed data used as labels during the training of the model, s represents variables in a hidden layer of the model, p represents a probability distribution, and F represents a tempered log-partition function.

15 . The system of claim 13 , wherein computing a value of a gradient of the loss function comprises performing a contrastive divergence algorithm, wherein learning part of the inference in the model comprises computing values of variables in a hidden layer of the model as part of the contrastive divergence algorithm.

16 . The system of claim 11 , wherein the observed data comprises variables in visible layers of the model.

17 . A quantum system, comprising a quantum information processor having hardware connections comprising couplers that connect qubits included in the quantum information processor, wherein the quantum system is configured to perform operations comprising:

receiving, as input, data for learning part of an inference in a model comprising interconnected units, wherein

the input is derived using i) data for training the model, the data comprising observed data for training and validating the model, and ii) a state of the model, and

the input maps at least some interactions of different interconnected units of the model to interactions of qubits in the quantum information processor;

processing the input to learn the part of the inference in the model; and

providing, as output, data representing the learned part of the inference.

18 . The system of claim 17 , wherein processing the input to learn part of the inference in the model comprises processing the input to learn values of variables in a hidden layer of the model.

19 . The system of claim 17 , wherein processing the input to learn part of the inference in the model comprises computing values of variables in a hidden layer of the model as part of a contrastive divergence algorithm.

20 . The system of claim 17 , wherein the observed data comprises variables in visible layers of the model.

Assignments (3)
CORRECTIVE ASSIGNMENT TO CORRECT THE CORRECT THE CONVEYING PARTY EXECUTION DATE FROM 09/30/2017 TO 09/29/2017 PREVIOUSLY RECORDED AT REEL: 56493 FRAME: 244. ASSIGNOR(S) HEREBY CONFIRMS THE CHANGE OF NAME. Recorded Apr 7, 2025
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 070761/0876 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 4, 2021
From: DING, NAN; MOHSENI, MASOUD; NEVEN, HARTMUT
To: GOOGLE INC.
Reel/Frame 056445/0086 →
CHANGE OF NAME Recorded Jun 4, 2021
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 056493/0244 →
Continuity (3)
Continuation 16413273 · May 15, 2019
Continuation 14484039 · Sep 11, 2014
Provisional Application 61876744 · Sep 11, 2013
References Cited (34)
US 20090075825A1 · Rose · 2009 [cited by applicant]
US 20120084242A1 · Levin · 2012 [cited by applicant]
US 20120254586A1 · Amin · 2012 [cited by applicant]
US 20130144925A1 · Macready · 2013 [cited by applicant]
US 20140187427A1 · Macready et al. · 2014 [cited by applicant]
US 20140223224A1 · Berkley · 2014 [cited by applicant]
US 20140337612A1 · Williams · 2014 [cited by applicant]
US 20140344322A1 · Ranjbar · 2014 [cited by applicant]
US 20150006443A1 · Rose et al. · 2015 [cited by applicant]
US 20150052092A1 · Tang · 2015 [cited by applicant]
US 20150111754A1 · Harris et al. · 2015 [cited by applicant]
US 20150161524A1 · Hamze · 2015 [cited by applicant]
US 20150317558A1 · Adachi et al. · 2015 [cited by applicant]
US 20150363708A1 · Amin · 2015 [cited by applicant]
US 20160071021A1 · Raymond · 2016 [cited by applicant]
US 20160132785A1 · Amin et al. · 2016 [cited by applicant]
US 20160162798A1 · Marandi et al. · 2016 [cited by applicant]
US 20160321559A1 · Rose et al. · 2016 [cited by applicant]
US 20170017894A1 · Lanting · 2017 [cited by applicant]
Bian et al., QUBOs for Machine Learning, 2009, 29 pages. [cited by applicant]
Danchev et al “Totally Corrective Boosting with Cardinality Penalization,” arXiv 1504.01446v1, Apr. 7, 2015, 16 pages. [cited by applicant]
Denchev et al. “Robust Classification with Adiabatic Quantum Optimization,” arXiv 1205.1148v2, May 30, 2012, 17 pages. [cited by applicant]
Denil et al, Toward the Implementation of a Quantum RBM, 2011, 9 pages. [cited by applicant]
Dong et al, “Quantum Reinforcement Learning,” IEEE Transactions on Systems, Man, and Cybernetics, Jul. 2008, 13 pages. [cited by applicant]
Hoehn et al. “A Quantum Post—“Lock and Key” Model of Neuroreceptor Activation,” Scientific Reports, Dec. 1, 2014, 31 pages. [cited by applicant]
Liu et al., “The Algorithm and Application of Quantum Wavelet Neural Networks,” 2010 Chinese Control and Decision Conference, May 2010, 5 pages. [cited by applicant]
Macready et al. “Structured Classification using Binary Quadratic Programmes,” D-Wave Systems, SFU and Google, Nov. 25, 2008, 20 pages. [cited by applicant]
Neven et al. “Image recognition with an adiabatic quantum computer I. Mapping to quadratic unconstrained binary optimization,” arXiv 0804.4457v1, Apr. 28, 2008, 7 pages. [cited by applicant]
Neven et al. “QBoost: Large Scale Classifier Training with Adiabatic Quantum Optimization,” JMLR Workshop and Conference Proceedings, 2012, 16 pages. [cited by applicant]
Neven et al. “Training a Binary Classifier with the Quantum Adiabatic Algorithm,” arXiv0811.0416v1, Nov. 4, 2008, 11 pages. [cited by applicant]
Neven et al, “Training a Binary Classifier with the Quantum Adabiatic Algorithm,” arXiv preprint arXiv:0811.0416, Nov. 2008, 11 pages. [cited by applicant]
Neven et al. “Training a Large Scale Classifier with the Quantum Adiabatic Algorithm,” arXiv 0912.0779v1, Dec. 2009, 14 pages. [cited by applicant]
Pudenz et al., “Quantum adiabatic machine learning,” Quantum information processing, May 2013, 21 pages. [cited by applicant]
Smelyanskiy et al, A Near-Term Quantum Computing Approach for Hard Computational Problems in Space Exploration, Apr. 2012, 69 pages. [cited by applicant]