IP Library Granted Patent US 12,223,417
Granted Patent B2
US 12,223,417 · App. 18/322,988 · Granted Feb 11, 2025

Efficient convolution in machine learning environments

Inventor: Dhawal Srivastava (Scottsdale, AZ)
Assignee: INTEL CORPORATION
G06N3/063G06F18/2113G06N3/044G06N3/045G06N3/08G06N3/084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,223,417
App. No.
18/322,988
Filed
May 24, 2023
Granted
Feb 11, 2025
Kind
B2
Art Unit
2621
USPC
706/25
Abstract

A mechanism is described for facilitating smart convolution in machine learning environments. An apparatus of embodiments, as described herein, includes one or more processors including one or more graphics processors, and detection and selection logic to detect and select input images having a plurality of geometric shapes associated with an object for which a neural network is to be trained. The apparatus further includes filter generation and storage logic (“filter logic”) to generate weights providing filters based on the plurality of geometric shapes, where the filter logic is further to sort the filters in filter groups based on common geometric shapes of the plurality of geographic shapes, and where the filter logic is further to store the filter groups in bins based on the common geometric shapes, wherein each bin corresponds to a geometric shape.

Claims (37)

1. An apparatus comprising:

processor circuitry coupled to a memory, the processor circuitry to:

initialize geometric shape-based training of a filter group based on one or more values obtained from a bin, wherein the bin is identified based on the filter group and selected based on a geometric shape of the object; and

initiate geometric shape-specific training of a neural network based on the trained filter group.

2. The apparatus of claim 1 , wherein the processor circuitry is further to:

detect and select input images having geometric shapes associated with the object for which the neural network is to be trained, wherein the bin includes the filter group associated with the geometric shape of the object;

generate weights providing filters based on the geometric shapes;

sort the filters in filter groups based on common geometric shapes of the geographic shapes; and

store the filter groups in bins based on the common geometric shapes, wherein one or more bins correspond to one or more geometric shapes.

3. The apparatus of claim 1 , wherein the processor circuitry is further to detect layers of the neural network, wherein the layers include higher layers and lower layers; and detect and identify existing convolution filters associated with the lower-level layers.

4. The apparatus of claim 3 , wherein the processor circuitry is further to:

separate the existing convolution filters of the neural network into pairs of new convolution filters, where a new convolution filter is half in size of an existing convolution filter; and

train the neural network based on the pairs of new convolution filters, wherein the processor circuitry includes graphics processor circuitry co-located with application processor circuitry on a common semiconductor package.

5. A method comprising:

initializing, by one or more processors of a computing device, geometric shape-based training of a filter group based on one or more values obtained from a bin, wherein the bin is identified based on the filter group and selected based on a geometric shape of the object; and

initiating geometric shape-specific training of a neural network based on the trained filter group.

6. The method of claim 5 , further comprising:

detecting and selecting input images having geometric shapes associated with the object for which the neural network is to be trained, wherein the bin includes the filter group associated with the geometric shape of the object;

generating weights providing filters based on the geometric shapes;

sorting the filters in filter groups based on common geometric shapes of the geographic shapes; and

storing the filter groups in bins based on the common geometric shapes, wherein one or more bins correspond to one or more geometric shapes.

7. The method of claim 5 , further comprising detecting layers of the neural network, wherein the layers include higher layers and lower layers; and detect and identify existing convolution filters associated with the lower-level layers.

8. The method of claim 7 , further comprising:

separating the existing convolution filters of the neural network into pairs of new convolution filters, where a new convolution filter is half in size of an existing convolution filter; and

training the neural network based on the pairs of new convolution filters, wherein the one or more processors include one or more graphics processors co-located with one or more application processors on a common semiconductor package.

9. At least one non-transitory computer-readable medium having stored thereon instructions which, when executed, cause a computing device to facilitate operations comprising:

initializing geometric shape-based training of a filter group based on one or more values obtained from a bin, wherein the bin is identified based on the filter group and selected based on a geometric shape of the object; and

initiating geometric shape-specific training of a neural network based on the trained filter group.

10. The non-transitory computer-readable medium of claim 9 , wherein the operations further comprise:

detecting and selecting input images having geometric shapes associated with the object for which the neural network is to be trained, wherein the bin includes the filter group associated with the geometric shape of the object;

generating weights providing filters based on the geometric shapes;

sorting the filters in filter groups based on common geometric shapes of the geographic shapes; and

storing the filter groups in bins based on the common geometric shapes, wherein one or more bins correspond to one or more geometric shapes.

11. The non-transitory computer-readable medium of claim 9 , wherein the operations further comprise detecting layers of the neural network, wherein the layers include higher layers and lower layers; and detect and identify existing convolution filters associated with the lower-level layers.

12. The non-transitory computer-readable medium of claim 11 , wherein the operations further comprise:

separating the existing convolution filters of the neural network into pairs of new convolution filters, where a new convolution filter is half in size of an existing convolution filter; and

training the neural network based on the pairs of new convolution filters, wherein the computing device comprises one or more processors including one or more graphics processors co-located with one or more application processors on a common semiconductor package.

Continuity (2)
Continuation 15859487 · Dec 30, 2017
Related Publication 20230419090A1 · Dec 28, 2023
References Cited (25)
US 7219085B2 · Buck · 2007 [cited by applicant]
US 7873812B1 · Mimar · 2011 [cited by applicant]
US 10528864B2 · Dally et al. · 2020 [cited by applicant]
US 10860922B2 · Dally et al. · 2020 [cited by applicant]
US 10891538B2 · Dally et al. · 2021 [cited by applicant]
US 11710028B2 · Srivastava · 2023 [cited by examiner]
US 20070047802A1 · Puri · 2007 [cited by applicant]
US 20160062947A1 · Chetlur et al. · 2016 [cited by applicant]
US 20170103309A1 · Chang et al. · 2017 [cited by applicant]
US 20180046906A1 · Dally et al. · 2018 [cited by applicant]
CN 109993278A · 2019 [cited by applicant]
EP 3509017A · 2019 [cited by applicant]
EP 4036809A · 2022 [cited by applicant]
Extended European Search Report for EP Application No. 22156704.3, mailed Jun. 22, 2022, 10 pages. [cited by applicant]
Summons to Attend oral proceedings pursuant to Rule 115(1) EPC, EP Application No. 18209351.8, Reference 113435EP, mailed Feb. 6, 2021, 8 pages. [cited by applicant]
Paul-Louise Prove, “An Introduction to different Types of Convolutions in Deep Learning”, Jun. 22, 2017 (Jun. 22, 2017), XP055592739. Retrieved from the Internet: https://towardsdatascience.com/types-of-convolutions-in-… [cited by applicant]
Xingyi, Li et al., “Filter Shaping For Convolutional Neural Networks”, ICLR 2017: 5th International Conference on Learning Representations, Apr. 26, 2017 (Apr. 26, 2017), XP055592764. [cited by applicant]
EP Communication Pursuant to Article 94(3) EPC for EP Application No. 18209351.8, Reference 113435EP, mailed Feb. 3, 2020, 6 pages. [cited by applicant]
Extended European Search Report for EP Application No. 18209351.8, mailed Jun. 6, 2019, 9 pages. [cited by applicant]
Goodfellow, et al. “Adaptive Computation and Machine Learning Series”, Book, Nov. 18, 2016, pp. 98-165, Chapter 5, The MIT Press, Cambridge, MA. [cited by applicant]
Ross, et al. “Intel Processor Graphics: Architecture & Programming”, Power Point Presentation, Aug. 2015, 78 pages, Intel Corporation, Santa Clara, CA. [cited by applicant]
Shane Cook, “CUDA Programming”, Book, 2013, pp. 37-52, Chapter 3, Elsevier Inc., Amsterdam Netherlands. [cited by applicant]
Nicholas Wilt, “The CUDA Handbook; A Comprehensive Guide to GPU Programming”, Book, Jun. 22, 2013, pp. 41-57, Addison-Wesley Professional, Boston, MA. [cited by applicant]
Stephen Junkins, “The Compute Architecture of Intel Processor Graphics Gen9”, paper, Aug. 14, 2015, 22 pages, Version 1.0, Intel Corporation, Santa Clara, CA. [cited by applicant]
EP Communication pursuant to Article 94(3) EPC issued in EP Application No. 22 156 704.3, mailed Apr. 23, 2024, 4 pages. [cited by applicant]
Cited By (1)
US 12,718,068