IP Library › Granted Patent US 12,387,581
Granted Patent B2
US 12,387,581 · App. 17/613,674 · Granted Aug 12, 2025

Method and device for detecting smoke

Inventor: Andreas Wellhausen (Hannover, DE)
Assignee: Robert Bosch GmbH
G08B17/125G06T7/0002G06T7/73G06T2207/20081G06T2207/20084G06T2207/30232
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,387,581
App. No.
17/613,674
Granted
Aug 12, 2025
Kind
B2
Abstract

The invention relates to a machine learning system ( 10 ) which is configured, on the basis of a plurality of images captured in succession, to detect smoke ( 12 ) within the images. The machine learning system ( 10 ) comprises a convolutional recurrent neural network. The invention also relates to a method for detecting smoke by means of this machine learning system.

Claims (45)

1. A machine learning system comprising:

a convolutional recurrent neural network, wherein the machine learning system is parameterized, on the basis of successive images of a sequence of images, to detect smoke in the images, and wherein the convolutional recurrent neural network is configured as a plurality of concatenated convolutional LongShortTermMemory (convolutional LSTM) modules connected to each other in a predetermined sequence;

wherein a first convolutional LSTM module in the predetermined sequence is configured to receive an input variable from the sequence of images,

wherein a number of filters that each respective convolutional LSTM module of the plurality of concatenated convolutional LSTM modules comprises increases with increasing depth of a position of each respective convolutional LSTM module,

wherein the depth of the position is a position within a sequence of each respective convolutional LSTM module of the machine learning system relative to the first convolutional LSTM module of the plurality of concatenated convolutional LSTM modules,

wherein a pooling layer is positioned between each of the convolutional LSTM modules, and

wherein a second convolutional LSTM module of the plurality of the concatenated convolutional LSTM modules is connected to a fully connected neural network, which is parameterized to output an output variable of the machine learning system.

2. The machine learning system as claimed in claim 1 , which is further parameterized to output an output variable that characterizes whether smoke is represented within the images.

3. The machine learning system as claimed in claim 1 , which is parameterized, on the basis of each individual image of the images, to output a matrix as an output variable,

wherein elements of the matrix are each assigned to a section of a predetermined plurality of sections of a respective image of the images, and

wherein the elements of the matrix characterize whether smoke is represented within this assigned section of the respective image.

4. The machine learning system as claimed in claim 1 , further comprising an input, wherein the input is configured to provide the sequence of images captured by a camera to the machine learning system.

5. The machine learning system as claimed in claim 1 , wherein the machine learning system is further parameterized to determine an amount of the detected smoke based on the successive images of the sequence of images, wherein the amount of the detected smoke is at least one selected from a group consisting of: an intensity and a density.

6. The machine learning system as claimed in claim 1 , wherein the convolutional LSTM module is configured to extract global information and temporal information of the successive images of the sequence of images together.

7. The machine learning system as claimed in claim 1 , wherein the convolutional LSTM module is configured to process global information and temporal information of the successive images of the sequence of images together.

8. The machine learning system as claimed in claim 1 ,

wherein the images are processed successively by the plurality of layers of the machine learning system,

wherein a resolution of the processed image decreases with increasing depth of a position of each respective convolutional LSTM module, wherein the depth of the position is a position within a sequence of each respective convolutional LSTM module of the machine learning system relative to the first convolutional LSTM module of the plurality of concatenated convolutional LSTM modules.

9. The machine learning system as claimed in claim 1 , wherein each of the plurality of convolutional LSTM modules are configured to process temporal and global information simultaneously.

10. A method for detecting smoke within an image by means of a machine learning system, which is parameterized, on the basis of a plurality of images captured in direct succession to detect smoke in the images,

wherein the machine learning system comprises a plurality of layers connected in a specified sequence and at least one of the plurality of layers comprises a convolutional recurrent neural network, wherein the convolutional recurrent neural network is a plurality of concatenated convolutional LongShortTermMemory (convolutional LSTM) modules,

wherein a number of filters that each respective convolutional LSTM module of the plurality of concatenated convolutional LSTM modules comprises increases with increasing depth of a position of each respective convolutional LSTM module,

wherein the depth of the position is a position within a sequence of each respective convolutional LSTM module of the machine learning system relative to a first convolutional LSTM module of the plurality of concatenated convolutional LSTM modules,

wherein a pooling layer is positioned between each of the convolutional LSTM modules, and

wherein a second convolutional LSTM module of the plurality of the concatenated convolutional LSTM modules is connected to a fully connected neural network, which is parameterized to output an output variable of the machine learning system, said method comprising the following steps:

obtaining a plurality of images captured in direct succession;

receiving an input variable at a first layer in the specified sequence, the first layer comprising the first convolutional LSTM module;

propagating the plurality of the captured images through the machine learning system in succession,

wherein during the propagation, the images are processed successively by the plurality of layers of the machine learning system and a final layer of the sequence of layers outputs an output variable.

11. The method as claimed in claim 10 , wherein the machine learning system is trained on the basis of a plurality of training data, comprising training images (x) and respectively assigned training output variables (y s ),

wherein the training images (x) are propagated through the machine learning system during the training and, on the basis of the determined output variables of the machine learning system and the respectively assigned training output variables of the training images, a parameterization of the machine learning system is adjusted in such a way that a deviation between the determined output variables and the training output variables becomes a minimum.

12. The method as claimed in claim 10 , wherein a smoke detector is activated on the basis of the determined output variable of the machine learning system.

13. The method of claim 10 , wherein a resolution of the processed image decreases with increasing depth of a position of each respective layer, wherein the depth of the position is a position within a sequence of each respective layer of the machine learning system relative to the first layer.

14. The method of claim 10 , further comprising processing temporal and global information simultaneously at each layer.

15. A non-transitory, computer-readable medium that contains instructions that when executed on a computer cause said computer to detect smoke within an image by obtaining a plurality of images captured in direct succession;

propagating the plurality of the captured images through a machine learning system in succession;

wherein the machine learning system comprises a plurality of layers connected in a specified sequence and at least one of the layers comprises a convolutional recurrent neural network,

wherein the convolutional recurrent neural network is a plurality of concatenated convolutional LongShortTermMemory (convolutional LSTM) modules,

wherein a number of filters that each respective convolutional LSTM module of the plurality of concatenated convolutional LSTM modules comprises increases with increasing depth of a position of each respective convolutional LSTM module,

wherein the depth of the position is a position within a sequence of each respective convolutional LSTM module of the machine learning system relative to a first convolutional LSTM module of the plurality of concatenated convolutional LSTM modules,

wherein a pooling layer is positioned between each of the convolutional LSTM modules,

wherein a second convolutional LSTM module of the plurality of the concatenated convolutional LSTM modules is connected to a fully connected neural network, which is parameterized to output an output variable of the machine learning system, and

wherein during the propagation, the images are processed successively by the layers of the machine learning system, wherein a first layer of the plurality of layers is the first convolutional LSTM module configured to receive an input variable from the plurality of images and a final layer of the sequence of layers outputs an output variable.

16. The non-transitory, computer-readable medium of claim 15 , wherein each of the plurality of convolutional LSTM modules are configured to process temporal and global information simultaneously.

17. The non-transitory, computer-readable medium of claim 15 , wherein a resolution of the processed image decreases with increasing depth of a position of each respective layer, wherein the depth of the position is a position within a sequence of each respective layer of the machine learning system relative to the first layer.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 1, 2021
From: WELLHAUSEN, ANDREAS
To: ROBERT BOSCH GMBH
Reel/Frame 058259/0914 →
Priority Claims (1)
DE 10 2019 207 711.0 · May 27, 2019 · national
Continuity (1)
Related Publication 20220230519A1 · Jul 21, 2022
References Cited (27)
US 10013640B1 · Angelova · 2018 [cited by examiner]
US 10929996B2 · Angelova · 2021 [cited by examiner]
US 11275989B2 · Tschernezki · 2022 [cited by examiner]
US 20180253839A1 · Zur · 2018 [cited by examiner]
US 20180260951A1 · Yang · 2018 [cited by examiner]
US 20180349477A1 · Jaech · 2018 [cited by examiner]
US 20190331832A1 · Chandra · 2019 [cited by examiner]
US 20200388028A1 · Agus · 2020 [cited by examiner]
CN 107451552A · 2017 [cited by applicant]
CN 108830305A · 2018 [cited by applicant]
CN 109582794A · 2019 [cited by applicant]
DE 102014216644A1 · 2016 [cited by applicant]
JP 2018101317A · 2018 [cited by applicant]
JP 2018101416A · 2018 [cited by applicant]
Hu et al., Real-Time Fire Detection Based on Deep Convolutional Long-Recurrent Networks and Optical Flow Method, Proceedings of the 37th Chinese Control Conference (Year: 2018). [cited by examiner]
Hu et al., “Real-Time Fire Detection Based on Deep Convolutional Long-Recurrent Networks and Optical Flow Method,” Proceedings of the 37th Chinese Control Conference, Jul. 25-27, 2018, Wuhan, China (Year: 2018). [cited by examiner]
Medel et al., “Anomaly Detection in Video Using Predictive Convolutional Long Short-Term Memory Networks,” Rochester Institute of Technology, Rochester, New York (Year: 2016). [cited by examiner]
Translation of International Search Report for Application No. PCT/EP2020/063957 dated Sep. 7, 2020 (3 pages). [cited by applicant]
Luo et al., “A Slight Smoke Perceptual Network”, IEEE Access, vol. 7, 2019, pp. 42889-42896. [cited by applicant]
Filonenko et al., “Smoke Detection on Video Sequence Using Convolutional and Recurent Neural Networks”, Annual International Conference on the Theory and Applications of Cryptographic Techniques, 2017, pp. 558-566. [cited by applicant]
Hu et al., “Real-Time Fire Detection Based on Deep Convolutional Long-Recurrent Networks and Opitcal Flow Method”, Proceedings of the 37th Chinese Control Conference, 2018, pp. 9061-9066. [cited by applicant]
Han et al., “Predictive Models of Fire via Deep learning Exploiting Colorific Variation”, IEEE International Conference on Artificial Intelligence in Information and Communication, 2019, pp. 579-581. [cited by applicant]
Shi et al., “Convolutional LSTM Network: A Machine Learning Approach for Precipitation Nowcasting”, arXiv 2015, pp. 1-11. [cited by applicant]
Yin et al., “Recurrent Convolutional Network for Video-based Smoke Detection,” Multimedia Tools Applications, 2019, vol. 78, pp. 237-256. [cited by applicant]
Guo et al., “Hierarchical LSTM for Sign Language Translation,” 32nd AAAI Conference on Artificial Intelligence, 2018, pp. 6845-6852. [cited by applicant]
Chen et al., “Image Identification of Forest Fire Smoke Based on Convolutional Neural Network Algorithm,” Instrumentation Technology, 2019, Issue 5, 8 pages including English Translation. [cited by applicant]
Zapata-Impata et al., “Learning Spatio Temporal Tactile Features with a ConvLSTM for the Direction of Slip Detection,” Sensors, 2019, vol. 19, 16 pages. [cited by applicant]