IP Library › Granted Patent US 12,443,829
Granted Patent B2
US 12,443,829 · App. 16/537,752 · Granted Oct 14, 2025

Neural network processing method and apparatus based on nested bit representation

Inventors: Seohyung Lee (Seoul, KR); Youngjun Kwak (Seoul, KR); Jinwoo Son (Seoul, KR); Changyong Son (Anyang-si, KR); Sangil Jung (Suwon-si, KR); Chang Kyu Choi (Seongnam-si, KR); Jaejoon Han (Seoul, KR)
Assignee: Samsung Electronics Co., Ltd.
G06N3/063G06N3/045G06N3/08
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,443,829
App. No.
16/537,752
Granted
Oct 14, 2025
Kind
B2
Abstract

A neural network processing method and apparatus based on nested bit representation is provided. The processing method includes obtaining first weights for a first layer of a source model of a first layer of a neural network, determining a bit-width for the first layer of the neural network, obtaining second weights for the first layer of the neural network by extracting at least one bit corresponding to the determined bit-width from each of the first weights for the first layer of a source model corresponding to the first layer of the neural network, and processing input data of the first layer of the neural network by executing the first layer of the neural network based on the obtained second weights.

Claims (37)

1. A processor-implemented training method, comprising:

executing an iterative process, by one or more processors, with training data and an in-training first neural network, configured to perform a first task, to generate a trained first neural network, configured to perform the first task and a second task different from the first task, that has a trained first layer including trained first weights that each have a first bit-width corresponding to a first precision, the iterative process including:

quantizing in-training first weights, having the first bit-width of an in-training first layer of the in-training first neural network to generate second weights of a first layer of a second neural network, that have a second bit-width that is less than the first bit-width;

executing the second neural network using the second weights, including applying the training data to the first layer of the second neural network and determining loss values, corresponding to the second task, of the first layer of the second neural network;

updating the in-training first weights of the in-training first layer of the in-training first neural network based on the determined loss values; and

performing, for each of the updated in-training first weights, a quantization of a corresponding updated in-training first weight of the updated in-training first weights to generate a corresponding first weight of the trained first weights that includes a nested second weight having the second bit-width that shares bits with the corresponding first weight,

wherein the updating of the in-training first weights comprises updating the in-training first weights of the first bit-width based on statistical information of loss gradients corresponding to the determined loss values,

wherein the updating of the in-training first weights further comprises calculating the statistical information by assigning a high weighted value to a loss gradient corresponding to a weight for which a high priority is set among the second weights of the second bit-width, and

wherein the nested second weight is nested in the corresponding first weight and stored in a same memory space.

2. The method of claim 1 , wherein the executing further includes executing the in-training first neural network using the in-training first weights, including applying the training data to the first layer of the in-training first neural network and determining other loss values, corresponding to the first task, of the first layer of the in-training first neural network, and

wherein the updating of the in-training first weights is further based on the determined other loss values.

3. The method of claim 1 , wherein the quantizing of the in-training first weights includes extracting less that all bits from each of the in-training first weights to generate the second weights.

4. The method of claim 1 , wherein the quantizing of the in-training first weights includes:

determining, for each of the in-training first weights, a corresponding second weight of the second weights to be a corresponding bit group of bits of the second bit-width, included in a corresponding in-training first weight of the in-training first weights, that have lower significance than at least one other bit of the corresponding in-training first weight.

5. The method of claim 2 , wherein the updating of the in-training first weights comprises:

updating the in-training first weights based on respective statistical information of respective loss gradients corresponding to the determined loss values and the determined other loss values.

6. A processor-implemented training method, comprising:

executing an iterative process, by one or more processors, with training data to generate a first neural network, with first layer higher bit-width weights, using lower-precision weights corresponding to an other neural network that are nested within the higher bit-width weights, including:

quantizing weights of a high bit-width, corresponding to a first layer of the first neural network, to generate weights of a low bit-width corresponding to a first layer of the other neural network;

executing the other neural network, including applying the training data to the first layer of the other neural network and determining loss values of the first layer of the other neural network corresponding to the weights of the low bit-width; and

updating the weights of the high bit-width of the first neural network based on the determined loss values,

wherein the updating of the weights of the high bit-width comprises updating the weights of the high bit-width based on statistical information of loss gradients corresponding to the determined loss values,

wherein the updating of the weights of the high bit-width further comprises calculating the statistical information by assigning a high weighted value to a loss gradient corresponding to a weight for which a high priority is set among the weights of the low bit-width, and

wherein the weights of the low bit-width are nested in corresponding higher bit-width and stored in a same memory space.

7. A neural network training apparatus, comprising:

one or more processors configured to execute instructions; and

a memory storing the instructions, that when executed by the one or more processors configure the one or more processors to:

execute an iterative process with training data and an in-training first neural network, configured to perform a first task, to generate a trained first neural network, configured to perform the first task and a second task different from the first task, that has a trained first layer including trained first weights that each have a first bit-width corresponding to a first precision, the iterative including:

a quantizing of in-training first weights, having the first bit-width of an in-training first layer of the in-training first neural network to generate second weights of a first layer of a second neural network, that have a second bit-width that is less than the first bit-width;

an executing of the second neural network using the second weights, including applying the training data to the first layer of the second neural network and determining loss values of the first layer of the second neural network; and

an updating of the in-training first weights of the in-training first layer of the in-training first neural network based on the determined loss values;

a performance, for each of the updated in-training first weights, a quantization of a corresponding updated in-training first weight of the updated in-training first weights to generate a corresponding first weight of the trained first weights that includes a nested second weight having the second bit-width that shares bits with the corresponding first weight,

wherein the updating of the in-training first weights comprises updating the in-training first weights of the first bit-width based on statistical information of loss gradients corresponding to the determined loss values,

wherein the updating of the in-training first weights further comprises calculating the statistical information by assigning a high weighted value to a loss gradient corresponding to a weight for which a high priority is set among the second weights of the second bit-width,

wherein the nested second weight is nested in the corresponding first weight and stored in a same memory space.

8. The training apparatus of claim 7 , wherein the quantizing of the in-training first weights includes extracting less that all bits from each of the in-training first weights to generate the second weights.

9. The training apparatus of claim 7 , wherein the quantizing of the in-training first weights includes determining, for each of the in-training first weights, a corresponding second weight of the second weights to be a corresponding bit group of bits of the second bit-width, included in a corresponding in-training first weight of the in-training first weights, that have lower significance than at least one other bit of the corresponding in-training first weight.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 12, 2019
From: LEE, SEOHYUNG; KWAK, YOUNGJUN; SON, JINWOO; SON, CHANGYONG; JUNG, SANGIL; CHOI, CHANG KYU; HAN, JAEJOON
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 050022/0539 →
Priority Claims (1)
KR 10-2018-0165585 · Dec 19, 2018 · national
Continuity (1)
Related Publication 20200202199A1 · Jun 25, 2020
References Cited (39)
US 9699484B2 · Su et al. · 2017 [cited by applicant]
US 11423311B2 · Brothers et al. · 2022 [cited by applicant]
US 20160328647A1 · Lin et al. · 2016 [cited by applicant]
US 20170270408A1 · Shi et al. · 2017 [cited by applicant]
US 20180032866A1 · Son et al. · 2018 [cited by applicant]
US 20190251436A1 · Son et al. · 2019 [cited by applicant]
CN 107451659A · 2017 [cited by applicant]
KR 1020040015613A · 2004 [cited by applicant]
KR 1020160140394A · 2016 [cited by applicant]
KR 1020160143548A · 2016 [cited by applicant]
KR 1020180052990A · 2018 [cited by applicant]
KR 1020180077533A · 2018 [cited by applicant]
Micikevicius, P., Narang, S., Alben, J., Diamos, G., Elsen, E., Garcia, D., Ginsburg, B., Houston, M., Kuchaiev, O., Venkatesh, G., & Wu, H.. (2017). Mixed Precision Training. (Year: 2017). [cited by examiner]
Jia, X., Song, S., He, W., Wang, Y., Rong, H., Zhou, F., Xie, L., Guo, Z., Yang, Y., Yu, L., Chen, T., Hu, G., Shi, S., & Chu, X.. (2018). Highly Scalable Deep Learning Training System with Mixed-Precision: Training Ima… [cited by examiner]
Wen, H., Zhou, S., Liang, Z., Zhang, Y., Feng, D., Zhou, X., & Yao, C.. (2016). Training Bit Fully Convolutional Network for Fast Semantic Segmentation. (Year: 2016). [cited by examiner]
Jason Sachs, 2016 “Round Round Get Around: Why Fixed-Point Right Shifts Are Just Fine” archived by https://web.archive.org/web/20170706132623/https://www.embeddedrelated.com/showarticle/1015.php (Year: 2016). [cited by examiner]
Kim, D., Lee, M., Choi, D., & Song, B. (2017). Multi-Modal Emotion Recognition Using Semi-Supervised Learning and Multiple Neural Networks in the Wild. In Proceedings of the 19th ACM International Conference on Multimod… [cited by examiner]
J. Lee, C. Kim, S. Kang, D. Shin, S. Kim and H.- J. Yoo, “UNPU: An Energy-Efficient Deep Neural Network Accelerator With Fully Variable Weight Bit Precision,” in IEEE Journal of Solid-State Circuits, vol. 54, No. 1, pp.… [cited by examiner]
Lin, X., Zhao, C., & Pan, W. (2017). Towards Accurate Binary Convolutional Neural Network. In Proceedings of the 31st International Conference on Neural Information Processing Systems (pp. 344-352). Curran Associates In… [cited by examiner]
Vukotić, V., Raymond, C., & Gravier, G. (2016). Bidirectional Joint Representation Learning with Symmetrical Deep Neural Networks for Multimodal and Crossmodal Applications. In Proceedings of the 2016 ACM on Internation… [cited by examiner]
Artit Wangperawong, Cyrille Brun, Olav Laudy, & Rujikorn Pavasuthipaisit. (2016). Churn analysis using deep convolutional neural networks and autoencoders. (Year: 2016). [cited by examiner]
Zhaoqi Li, Yu Ma, Catalina Vajiac, & Yunkai Zhang. (2018). Exploration of Numerical Precision in Deep Neural Networks. (Year: 2018). [cited by examiner]
W. Choi, K. Choi and J. Park, “Low Cost Convolutional Neural Network Accelerator Based on Bi-Directional Filtering and Bit-Width Reduction,” in IEEE Access, vol. 6, pp. 14734-14746, 2018, doi: 10.1109/ACCESS.2018.281601… [cited by examiner]
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar and L. Fei-Fei, “Large-Scale Video Classification with Convolutional Neural Networks,” 2014 IEEE Conference on Computer Vision and Pattern Recognition, Columb… [cited by examiner]
Sheng Chen, Yang Liu, Xiang Gao, & Zhen Han. (2018). MobileFaceNets: Efficient CNNs for Accurate Real-Time Face Verification on Mobile Devices. (Year: 2018). [cited by examiner]
Zhengping Ji, et al. (2015) Reducing weight precision of convolutional neural networks towards large-scale on-chip image recognition, Proceedings vol. 9496, Independent Component Analyses, Compressive Sampling, Large Da… [cited by examiner]
Anwar, Saji et al., “Fixed Point Optimization of Deep Convolutional Neural Networks for Object Recognition.” [cited by applicant]
Ma, Xiaolong et al., “An Area and Energy Efficient Design of Domain-Wall Memory-Based Deep Convolutional Neural Networks using Stochastic Computing.” [cited by applicant]
Chung, Jinil et al., “Domain Wall Memory-Based Design of Deep Neural Network Convolutional Layers.” [cited by applicant]
Partial European Search Report issued on May 15, 2020 for the corresponding European Patent Application No. 19204904.7 (17 pages in English). [cited by applicant]
Kim, Eunwoo, et al., “NestedNet: Learning Nested Sparse Structures in Deep Neural Networks”, [cited by applicant]
Sachs, Jason “Round Round Get Around: Why Fixed-Point Right-Shifts Are Just Fine” EmbeddedRelated.com, Nov. 22, 2016, (10 pages). [cited by applicant]
Wen, He, et al. “Training Bit Fully Convolutional Network for Fast Semantic Segmentation.” arXiv preprint arXiv:1612.00212 (2016)., (8 pages). [cited by applicant]
Micikevicius, Paulius, et al. “Mixed Precision Training.” arXiv preprint arXiv:1710.03740 (2017)., (14 pages). [cited by applicant]
Jia, Xianyan, et al. “Highly Scalable Deep Learning Training System with Mixed-Precision: Training Imagenet in Four Minutes. CoRR abs/1807.11205 (2018).” arXiv preprint arXiv:1807.11205 (2018)., (9 pages). [cited by applicant]
Chinese Office Action issued on Jul. 14, 2024, in counterpart Chinese Patent Application No. 201910778533.5 (23 pages in English, 14 pages in Chinese). [cited by applicant]
Howard, Andrew G., et al., “MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications”, Apr. 2017, pp. 1-9. [cited by applicant]
Kim, Eunwoo, et al., “NestedNet: Learning Nested Sparse Structures in Deep Neural Networks”, [cited by applicant]
Korean Office Action issued on Dec. 17, 2024 in corresponding Korean Application No. 10-2018-0165585. (5pages in English, 9pages in Korean). [cited by applicant]