IP Library › Granted Patent US 12,314,846
Granted Patent B2
US 12,314,846 · App. 16/284,322 · Granted May 27, 2025

Answering cognitive queries from sensor input signals

Inventors: Giovanni Cherubini (Rueschlikon, CH); Evangelos Stavros Eleftheriou (Rueschlikon, CH)
Assignee: International Business Machines Corporation
G06N3/08G06F7/582G06N3/04G06N3/045
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,314,846
App. No.
16/284,322
Granted
May 27, 2025
Kind
B2
Abstract

A computer-implemented method for answering a cognitive query from sensor input signals may be provided. The method comprises feeding sensor input signals to an input layer of an artificial neural network comprising a plurality of hidden neuron layers and an output neural layer, determining hidden layer output signals from each of the plurality of hidden neuron layers and output signals from the output neural layer, and generating a set of pseudo-random bit sequences by applying a set of mapping functions using the output signals of the output layer and the hidden layer output signals of one of the hidden neuron layers as input data for one mapping function. Furthermore, the method comprises determining a hyper-vector using the set of pseudo-random bit sequences, and storing the hyper-vector in an associative memory, in which a distance between different hyper-vectors is determinable.

Claims (58)

1. A method for answering a cognitive query from sensor input signals, said method comprising:

feeding sensor input signals to an input layer of an artificial neural network comprising a plurality of hidden neuron layers and an output neural layer;

determining hidden layer output signals from each of said plurality of hidden neuron layers and output signals from said output neural layer;

generating a set of pseudo-random bit sequences by applying a plurality of mapping functions by at least feeding said output signals of said output layer and feeding said hidden layer output signals of at least one of said hidden neuron layers directly from said at least one of said hidden neuron layers as input data to one mapping function, the one mapping function receiving the output signals of the output layer and the hidden layer output signals of at least one of the hidden neuron layers additional to receiving the output signals of the output layer, the at least one of the hidden neuron layers including at least one hidden neuron layer other than a hidden neuron layer immediately preceding the output layer, wherein the one mapping function generates at least one pseudo-random bit sequence of the set of pseudo-random bit sequences based on the received output signals of said output layer and intermediate information of the artificial neural network represented by the hidden layer output signals received additionally to the output signals of said output layer, each of the plurality of mapping functions being fed as input, at least one selected from the group consisting of: signals received directly from at least two of the plurality of hidden neuron, and signals received directly from the output neural layer and the at least one of the plurality of hidden neuron layers, wherein each one of the plurality of hidden neuron layers is used in direct feeding in generating the set of pseudo-random bit sequences;

determining a hyper-vector using said set of pseudo-random bit sequences by concatenating the set of pseudo-random bit sequences generated by the plurality of mapping functions; and

storing said hyper-vector in an associative memory, in which a distance between different hyper-vectors is determinable.

2. The method according to claim 1 , also comprising

combining different hyper-vectors stored in said associative memory for deriving said cognitive query and candidate answers, each of which is obtained from said sensor input signals.

3. The method according to claim 2 , also comprising

measuring a distance between said hyper-vector representing said cognitive query and said hyper-vectors representing said candidate answers.

4. The method according to claim 3 , also comprising

selecting said hyper-vector relating to said candidate answers having a minimum distance from said hyper-vector representing said cognitive query.

5. The method according to claim 2 , wherein said feeding sensor input data to an input layer of an artificial neural network also comprises

feeding said sensor input data to input layers of a plurality of artificial neural networks.

6. The method according to claim 1 , wherein said generating a set of pseudo-random bit sequences also comprises

using said output signals of said output layer and said hidden layer output signals of a plurality of said hidden neuron layers as input data for one mapping function.

7. The method according to claim 1 , wherein said neural network is a convolutional neural network.

8. The method according to claim 1 , also comprising

generating a plurality of random hyper-vectors as role vectors at an end of a training of said neural network to be stored in said associative memory.

9. The method according to claim 5 , also comprising

generating continuously a plurality of pseudo-random hyper-vectors as filler vectors based on output data of said plurality of neural networks for each new set of sensor input data, said pseudo-random hyper-vectors to be stored in said associative memory.

10. The method according to claim 2 , wherein said combining comprises

binding different hyper-vectors by a vector-element-wise binary XOR operation.

11. The method according to claim 2 , wherein said combining comprises

bundling different hyper-vectors by a vector-element-wise binary average.

12. The method according to claim 8 , wherein said role vectors are stored in said associative memory at said end of a training of said artificial neural network if a stochastic gradient descent value determined between two sets of input vectors of said sensor input signals to said neural network remain below a predefined threshold value.

13. A machine-learning system for answering a cognitive query from sensor input signals, said machine-learning system comprising:

a sensor adapted for feeding sensor input signals to an input layer of an artificial neural network comprising a plurality of hidden neuron layers and an output neural layer;

at least one computer processor adapted for determining hidden layer output signals from each of said plurality of hidden neuron layers and output signals from said output neural layer;

the at least one computer processor further for generating a set of pseudo-random bit sequences by applying a plurality of mapping functions by at least feeding said output signals of said output layer and feeding said hidden layer output signals of at least one of said hidden neuron layers directly from said at least one of said hidden neuron layers as input data to one mapping function, the one mapping function receiving the output signals of the output layer and the hidden layer output signals of at least one of the hidden neuron layers additional to receiving the output signals of the output layer, the at least one of the hidden neuron layers including at least one hidden neuron layer other than a hidden neuron layer immediately preceding the output layer, wherein the one mapping function generates at least one pseudo-random bit sequence of the set of pseudo-random bit sequences based on the received output signals of said output layer and intermediate information of the artificial neural network represented by the hidden layer output signals received additionally to the output signals of said output layer, each of the plurality of mapping functions being fed as input, at least one selected from the group consisting of: signals received directly from at least two of the plurality of hidden neuron, and signals received directly from the output neural layer and the at least one of the plurality of hidden neuron layers, wherein each one of the plurality of hidden neuron layers is used in direct feeding in generating the set of pseudo-random bit sequences;

the at least one computer processor further adapted for determining a hyper-vector using said set of pseudo-random bit sequences by concatenating the set of pseudo-random bit sequences generated by the plurality of mapping functions; and

a storage module adapted for storing said hyper-vector in an associative memory, in which a distance between different hyper-vectors is determinable.

14. The machine-learning system according to claim 13 , wherein the at least one computer processor is further

for combining different hyper-vectors stored in said associative memory for deriving said cognitive query and candidate answers, each of which is obtained from said sensor input signals.

15. The machine-learning system according to claim 14 , wherein the at least one computer processor is further

adapted for measuring a distance between said hyper-vector representing said cognitive query and said hyper-vectors representing said candidate answers.

16. The machine-learning system according to claim 15 , wherein the at least one computer processor is further

adapted for selecting said hyper-vector relating to said candidate answers having a minimum distance from said hyper-vector representing said cognitive query.

17. The machine-learning system according to claim 14 , wherein said sensor for feeding sensor input data to an input layer of an artificial neural network is also adapted for

feeding said sensor input data to input layers of a plurality of artificial neural networks.

18. The machine-learning system according to claim 13 , wherein said generating a set of pseudo-random bit sequences also comprises

using said output signals of said output layer and said hidden layer output signals of a plurality of said hidden neuron layers as input data for one mapping function.

19. The machine-learning system according to claim 13 , wherein said neural network is a convolutional neural network.

20. The machine-learning system according to claim 13 , wherein the at least one computer processor is further

adapted for generating a plurality of random hyper-vectors as role vectors at an end of a training of said neural network to be stored in said associative memory.

21. The machine-learning system according to claim 17 , wherein the at least one computer processor is further

adapted generating continuously a plurality of pseudo-random hyper-vectors as filler vectors based on output data of said plurality of neural networks for each new set of sensor input data, said pseudo-random hyper-vectors to be stored in said associative memory.

22. The machine-learning system according to claim 14 , wherein the at least one computer processor is further adapted for

binding different hyper-vectors by a vector-element-wise binary XOR operation.

23. The machine-learning system according to claim 14 , wherein the at least one computer processor is further adapted for

bundling different hyper-vectors by a vector-element-wise binary average.

24. The machine-learning system according to claim 20 , wherein said role vectors are stored in said associative memory at said end of a training of said artificial neural network if a stochastic gradient descent value determined between two sets of input vectors of said sensor input signals to said neural network remain below a predefined threshold value.

25. A computer program product for answering a cognitive query from sensor input signals, said computer program product comprising a computer readable storage medium having program instructions embodied therewith, said program instructions being executable by a computing system to cause the computing system to:

feed sensor input signals to an input layer of an artificial neural network comprising a plurality of hidden neuron layers and an output neural layer;

determine hidden layer output signals from each of said plurality of hidden neuron layers and output signals from said output neural layer;

generate a set of pseudo-random bit sequences by applying a plurality of mapping functions by at least feeding said output signals of said output layer and feeding said hidden layer output signals of at least one of said hidden neuron layers directly from said at least one of said hidden neuron layers as input data to one mapping function, the one mapping function receiving the output signals of the output layer and the hidden layer output signals of at least one of the hidden neuron layers additional to receiving the output signals of the output layer, the at least one of the hidden neuron layers including at least one hidden neuron layer other than a hidden neuron layer immediately preceding the output layer, wherein the one mapping function generates at least one pseudo-random bit sequence of the set of pseudo-random bit sequences based on the received output signals of said output layer and intermediate information of the artificial neural network represented by the hidden layer output signals received additionally to the output signals of said output layer, each of the plurality of mapping functions being fed as input, at least one selected from the group consisting of: signals received directly from at least two of the plurality of hidden neuron, and signals received directly from the output neural layer and the at least one of the plurality of hidden neuron layers, wherein each one of the plurality of hidden neuron layers is used in direct feeding in generating the set of pseudo-random bit sequences;

determine a hyper-vector using said set of pseudo-random bit sequences by concatenating the set of pseudo-random bit sequences generated by the plurality of mapping functions; and

store said hyper-vector in an associative memory, in which a distance between different hyper-vectors is determinable.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2019
From: CHERUBINI, GIOVANNI; ELEFTHERIOU, EVANGELOS STAVROS
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 048426/0194 →
Continuity (1)
Related Publication 20200272895A1 · Aug 27, 2020
References Cited (46)
US 7034670B2 · Kennedy et al. · 2006 [cited by applicant]
US 7991714B2 · Widrow et al. · 2011 [cited by applicant]
US 8521669B2 · Knoblauch · 2013 [cited by applicant]
US 10909450B2 · Chen · 2021 [cited by examiner]
US 20150120706A1 · Hoffman et al. · 2015 [cited by applicant]
US 20150310858A1 · Li et al. · 2015 [cited by applicant]
US 20160034811A1 · Paulik et al. · 2016 [cited by applicant]
US 20170091615A1 · Liu et al. · 2017 [cited by applicant]
US 20170178666A1 · Yu · 2017 [cited by applicant]
US 20180089556A1 · Zeiler et al. · 2018 [cited by applicant]
US 20180181860A1 · Verbist et al. · 2018 [cited by applicant]
US 20180211128A1 · Hotson · 2018 [cited by examiner]
US 20180232637A1 · Caicedo Fernandez et al. · 2018 [cited by applicant]
US 20180253645A1 · Burr · 2018 [cited by applicant]
US 20190164056A1 · Hoshizuki · 2019 [cited by examiner]
CN 106127301A · 2016 [cited by applicant]
CN 107810508A · 2018 [cited by applicant]
CN 108604313A · 2018 [cited by applicant]
CN 108694200A · 2018 [cited by applicant]
JP 2001052175A · 2001 [cited by applicant]
JP 2005190429A · 2005 [cited by applicant]
JP 2018165926A · 2018 [cited by applicant]
JP 2018536244A · 2018 [cited by applicant]
WO 2016206765A1 · 2016 [cited by applicant]
WO 2017112466A1 · 2017 [cited by applicant]
WO 2017096396A1 · 2017 [cited by applicant]
WO 2018188240A1 · 2018 [cited by applicant]
Imani, Mohsen, et al. “Low-power sparse hyperdimensional encoder for language recognition.” IEEE Design & Test 34.6 (2017): 94-101. (Year: 2017). [cited by examiner]
Bernardi, Marcello De, M. H. R. Khouzani, and Pasquale Malacaria. “Pseudo-random number generation using generative adversarial networks.” Joint European Conference on Machine Learning and Knowledge Discovery in Databas… [cited by examiner]
Mukhlason, Ahmad, Ahmad Kamil Mahmood, and Noreen Izza Arshad. “Knowledge Management Driven eLearning System: Developing An Automatic Question Answering System With Expert Locator.” Journal of Knowledge Management Pract… [cited by examiner]
Neubert, Peer, Stefan Schubert, and Peter Protzel. Learning vector symbolic architectures for reactive robot behaviours. Universitätsbibliothek Chemnitz, 2017. (Year: 2017). [cited by examiner]
Das, Abhishek, et al. “Embodied question answering.” Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 2018. (Year: 2018). [cited by examiner]
Imani, Mohsen, et al. “Voicehd: Hyperdimensional computing for efficient speech recognition.” 2017 IEEE international conference on rebooting computing (ICRC). IEEE, 2017. (Year: 2017). [cited by examiner]
Karras, D. A., and V. Zorkadis. “Overfitting in multilayer perceptrons as a mechanism for (pseudo) random number generation in the design of secure electronic commerce systems.” IEEE/AFCEA EUROCOMM 2000. Information Sys… [cited by examiner]
Vakhshiteh, Fatemeh, and Farshad Almasganj. “Exploration of properly combined audiovisual representation with the entropy measure in audiovisual speech recognition.” Circuits, Systems, and Signal Processing 38 (2019): 2… [cited by examiner]
Amendment submitted to the UK Intellectual Property Office in counterpart UK Patent Application No. 2112651.1 dated Dec. 20, 2021, 6 pages. [cited by applicant]
Emruli, B., et al., “Analogical Mapping with Sparse Distributed Memory: A Simple Model that Learns to Generalize from Examples”, Cognitive Computation, Mar. 2014, pp. 74-88, vol. 6, Issue 1. [cited by applicant]
Gong, Y., et al., “Compressing Deep Convolutional Networks using Vector Quantization”, arXiv:1412.6115 v1, Dec. 18, 2014, 10 pages. [cited by applicant]
Slomin, N., “The Information Bottleneck: Theory and Applications,” Thesis submitted for the degree “Doctor of Philosophy” to the Senate of the Hebrew University, 2002, 157 pages. [cited by applicant]
Schwartz-Ziv, et al., “Opening the black box of Deep Neural Networks via Information,” arXiv:1703.00810v3, Apr. 29, 2017, 19 pages. [cited by applicant]
Emruli, B., et al., “Analogical mapping and inference with binary spatter codes and sparse distributed memory”, The 2013 International Joint Conference on Neural Networks (IJCNN 2013), Aug. 2013, 8 pages. [cited by applicant]
Kanerva, P., “Computing with 10,000-Bit Words”, 2014 52nd Annual Allerton Conference on Communication, Control, and Computing, Sep. 30-Oct. 3, 2014, 8 pages. [cited by applicant]
International Search Report dated Jun. 4, 2020, issued in PCT/IB2020/051259, 9 pages. [cited by applicant]
UK Examination Report dated Nov. 5, 2021 issued in GB2112651.1, 7 pages. [cited by applicant]
Notice of Reasons for Refusal dated Sep. 26, 2023 issued in JP Application No. 2021-547178, 4 pages. [cited by applicant]
Decision to Grant a Patent dated Jan. 9, 2024 issued in JP Application No. 2021-547178, 5 pages. [cited by applicant]