IP Library › Granted Patent US 12,198,032
Granted Patent B2
US 12,198,032 · App. 17/286,982 · Granted Jan 14, 2025

Electronic device and control method therefor

Inventors: Chiyoun Park (Suwon-si, KR); Jaedeok Kim (Suwon-si, KR); Hyunjoo Jung (Suwon-si, KR)
Assignee: Samsung Electronics Co., Ltd.
G06N3/04G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,198,032
App. No.
17/286,982
Granted
Jan 14, 2025
Kind
B2
Abstract

An electronic device and a control method therefor are provided. The electronic device may comprise: a memory for storing at least one instruction; and a processor connected to the memory so as to control the electronic device, wherein the processor: by executing the at least one instruction, appends a second layer including a learnable function to a first layer in an artificial neural network including a plurality of layers; updates a parameter value included in the second layer by learning of the artificial neural network; acquires a function value by inputting the updated parameter value to the learnable function; and eliminates at least one channel among a plurality of channels included in the first layer on the basis of the acquired function value so as to achieve update to a third layer.

Claims (37)

1. An electronic device comprising:

a memory configured to store at least one instruction; and

a processor connected to the memory and configured to control the electronic device,

wherein the processor, by executing the at least one instruction, is further configured to:

append a second layer including a trainable function to a first layer in an artificial neural network including a plurality of layers, wherein the trainable function is obtained by adding a first function outputting 0 or 1 to a differentiable second function and the differentiable second function is obtained by multiplying a differentiable function by a function having a predetermined gradient,

obtain a first function value by inputting parameter value of the second layer to the trainable function,

obtain output data of the second layer by multiplying output data of the first layer by the first function value channel-wise,

generate a loss function based on the output data of the first layer and the output data of the second layer,

obtain parameter value that outputs minimum function value of the loss function,

update the parameter value of the second layer to the obtained parameter value that outputs the minimum function value of the loss function,

obtain a second function value by inputting the updated parameter value to the trainable function, and

update the first layer to a third layer by eliminating at least one channel among a plurality of channels included in the first layer based on the obtained second function value,

wherein the loss function is a function of adding an additional loss function indicating a size or computational complexity of a layer of the artificial neural network to be obtained after compression of the artificial neural network to a loss function using a sum of output values of the trainable function.

2. The electronic device of claim 1 , wherein the processor is further configured to:

based on the updated parameter value being negative, obtain a function value of 0 by inputting a negative parameter value to the trainable function, and

based on the updated parameter value being positive, obtain a function value of 1 by inputting a positive parameter value to the trainable function.

3. The electronic device of claim 2 , wherein the processor is further configured to:

based on the obtained function value of 0, eliminate a channel of the first layer corresponding to the parameter value input to the trainable function; and

based on the obtained function value of 1, maintain the channel of the first layer corresponding to the parameter value input to the trainable function.

4. The electronic device of claim 1 , wherein the processor is further configured to update the first layer to a third layer by eliminating the second layer and eliminating the channel of the first layer based on an obtained function value.

5. The electronic device of claim 1 , wherein the processor is further configured to:

obtain a function value by inputting the updated parameter value to the trainable function;

change a weight of a first kernel of the first layer based on the obtained function value; and

update the first kernel of the first layer to a second kernel including the changed weight.

6. The electronic device of claim 5 , wherein the processor is further configured to:

based on the obtained function value being 0, change a weight of the first kernel of the first layer corresponding to the parameter value input to the trainable function to 0; and

based on the obtained function value being 1, maintain the weight of the first kernel of the first layer corresponding to the parameter value input to the trainable function.

7. A control method of an electronic device, the method comprising:

appending a second layer including a trainable function to a first layer in an artificial neural network including a plurality of layers, wherein the trainable function is obtained by adding a first function outputting 0 or 1 to a differentiable second function and the differentiable second function is obtained by multiplying a differentiable function by a function having a predetermined gradient;

obtaining a first function value by inputting parameter value of the second layer to the trainable function;

obtaining output data of the second layer by multiplying output data of the first layer by the first function value channel-wise;

generating a loss function based on the output data of the first layer and the output data of the second layer;

obtaining parameter value that outputs minimum function value of the loss function;

updating the parameter value of the second layer to the obtained parameter value that outputs the minimum function value of the loss function;

obtaining a second function value by inputting the updated parameter value to the trainable function; and

updating the first layer to a third layer by eliminating at least one channel among a plurality of channels included in the first layer based on the obtained second function value,

wherein the loss function is a function of adding an additional loss function indicating a size or computational complexity of a layer of the artificial neural network to be obtained after compression of the artificial neural network to a loss function using a sum of output values of the trainable function.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 20, 2021
From: PARK, CHIYOUN; KIM, JAEDEOK; JUNG, HYUNJOO
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 055976/0381 →
Priority Claims (1)
KR 10-2019-0118918 · Sep 26, 2019 · national
Continuity (2)
Provisional Application 62793497 · Jan 17, 2019
Related Publication 20210383190A1 · Dec 9, 2021
References Cited (67)
US 5253327A · Yoshihara · 1993 [cited by applicant]
US 8086052B2 · Toth et al. · 2011 [cited by applicant]
US 8463722B2 · Knoblauch · 2013 [cited by applicant]
US 10380484B2 · Goel et al. · 2019 [cited by applicant]
US 10819844B2 · Kim et al. · 2020 [cited by applicant]
US 11055320B2 · Chandna et al. · 2021 [cited by applicant]
US 11132621B2 · Botea et al. · 2021 [cited by applicant]
US 20030174872A1 · Chalana et al. · 2003 [cited by applicant]
US 20030174881A1 · Simard et al. · 2003 [cited by applicant]
US 20150332171A1 · Nakagawa · 2015 [cited by applicant]
US 20160358070A1 · Brothers et al. · 2016 [cited by applicant]
US 20170169326A1 · Diamos et al. · 2017 [cited by applicant]
US 20170293757A1 · Rosenman et al. · 2017 [cited by applicant]
US 20170330076A1 · Valpola · 2017 [cited by applicant]
US 20180046915A1 · Sun et al. · 2018 [cited by applicant]
US 20180107925A1 · Choi et al. · 2018 [cited by applicant]
US 20180114110A1 · Han et al. · 2018 [cited by applicant]
US 20180114114A1 · Molchanov et al. · 2018 [cited by applicant]
US 20180189950A1 · Norouzi et al. · 2018 [cited by applicant]
US 20180247193A1 · Holtham · 2018 [cited by applicant]
US 20180336468A1 · Kadav et al. · 2018 [cited by applicant]
US 20180357538A1 · Hwang et al. · 2018 [cited by applicant]
US 20180373975A1 · Yu · 2018 [cited by examiner]
US 20190080241A1 · Guo et al. · 2019 [cited by applicant]
US 20190114391A1 · Jaganathan et al. · 2019 [cited by applicant]
US 20190114462A1 · Jang · 2019 [cited by examiner]
US 20190197406A1 · Darvish Rouhani et al. · 2019 [cited by applicant]
US 20190279089A1 · Wang · 2019 [cited by applicant]
US 20190325267A1 · Chen · 2019 [cited by applicant]
US 20200034661A1 · Kim et al. · 2020 [cited by applicant]
US 20200042796A1 · Kim et al. · 2020 [cited by applicant]
US 20200059551A1 · Kim et al. · 2020 [cited by applicant]
US 20200302303A1 · Chen et al. · 2020 [cited by applicant]
US 20210117651A1 · Kotake · 2021 [cited by applicant]
CN 101414351A · 2009 [cited by applicant]
CN 106548234A · 2017 [cited by applicant]
JP 4258268B2 · 2009 [cited by applicant]
JP 6760318B2 · 2020 [cited by applicant]
KR 1020170092595A · 2017 [cited by applicant]
KR 1020180045635A · 2018 [cited by applicant]
KR 1020180075368A · 2018 [cited by applicant]
KR 1020180134740A · 2018 [cited by applicant]
KR 1020190094133A · 2019 [cited by applicant]
KR 1020190103084A · 2019 [cited by applicant]
KR 1020190106861A · 2019 [cited by applicant]
KR 102124171B1 · 2020 [cited by applicant]
KR 102163498B1 · 2020 [cited by applicant]
WO 2016083657A1 · 2016 [cited by applicant]
Gao, Xitong, et al. “Dynamic channel pruning: Feature boosting and suppression.” arXiv preprint arXiv:1810.05331 (2018). (Year: 2018). [cited by examiner]
Hahn, Sangchul, and Heeyoul Choi. “Gradient acceleration in activation functions.” (2018). (Year: 2018). [cited by examiner]
Hua, et al. “Channel Gating Neural Networks” arXiv preprint arXiv:1805.12549v1 (2018). (Year: 2018). [cited by examiner]
Wen, Wei, et al. “Learning structured sparsity in deep neural networks.” Advances in neural information processing systems 29 (2016). (Year: 2016). [cited by examiner]
Srivastava, Rupesh K., et al. “Compete to compute.” Advances in neural information processing systems 26 (2013). (Year: 2013). [cited by examiner]
Loffe, Sergey, and Christian Szegedy. “Batch normalization: Accelerating deep network training by reducing internal covariate shift.” International conference on machine learning. pmlr, 2015. (Year: 2015). [cited by examiner]
Ghosh et al.; Trusted Neural Networks for Safety-Constrained Autonomous Control; Cornell University; arXiv.org>cs>arXiv:1085.07075v1; May 18, 2018. [cited by applicant]
International Search Report dated May 13, 2021; International Appln. No. PCT/KR2021/001041. [cited by applicant]
Mladenov et al.; Solving Sudoku puzzles by using Hopfield neural networks; ResearchGate; Conference Paper, https://www.researchgate.net/publication/262170343; May 2011. [cited by applicant]
Yue et al.; Sudoku Solver by Q'tron Neural Networks; Dept. of Computer Science and Engineering, Tatung University Taipe; D.-S. Huang, K. Li, and G.W. Irwin (Eds.): ICIC 2006, LNCS 4113, pp. 943-952, 2006; Springer-Verla… [cited by applicant]
European Examination Report dated Jun. 16, 2023, issued in European Patent Application No. 19910158.5. [cited by applicant]
European Search Report dated Oct. 10, 2022; European Appln. No. 21747753.8—1203/4022526 PCT/KR2021001041. [cited by applicant]
Indian Office Action dated Dec. 5, 2022; Indian Appln. No. 202217029971. [cited by applicant]
Hua et al.; Channel Gating Neural Networks; XP080884351; arXiv:1805.12549v1 [cs.LG]; May 29, 2018. [cited by applicant]
European Search Report dated Nov. 19, 2021; European Appln. No. 19910158.5—1203/3852017 PCT/KR2019016235. [cited by applicant]
Indian Office Action dated May 17, 2024, issued in Indian Patent Application No. 202217029971. [cited by applicant]
European Office Action dated May 29, 2024, issued in European Patent Application No. 19 910 158.5—1203. [cited by applicant]
U.S. Office Action dated Aug. 5, 2024, issued by the U.S. Patent and Trademark Office U.S. Appl. No. 17/158,561. [cited by applicant]
Chinese Office Action dated Oct. 30, 2024, issued in Chinese Application No. 202180011639.8. [cited by applicant]