IP Library › Granted Patent US 12,067,472
Granted Patent B2
US 12,067,472 · App. 15/942,298 · Granted Aug 20, 2024

Defect resistant designs for location-sensitive neural network processor arrays

Inventors: Rathinakumar Appuswamy (San Jose, CA); John V. Arthur (Mountain View, CA); Andrew S. Cassidy (San Jose, CA); Pallab Datta (San Jose, CA); Steven K. Esser (San Jose, CA); Myron D. Flickner (San Jose, CA); Jennifer Klamo (San Jose, CA); Dharmendra S. Modha (San Jose, CA); Hartmut Penner (San Jose, CA); Jun Sawada (Austin, TX); Brian Taba (Cupertino, CA)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
G06N3/04G06N3/063
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,067,472
App. No.
15/942,298
Filed
Mar 30, 2018
Granted
Aug 20, 2024
Kind
B2
Art Unit
2123
USPC
706/29
Abstract

Defect resistant designs for location-sensitive neural network processor arrays are provided. In various embodiments, plurality of neural network processor cores are arrayed in a grid. The grid has a plurality of rows and a plurality of columns. A network interconnects at least those of the plurality of neural network processor cores that are adjacent within the grid. The network is adapted to bypass a defective core of the plurality of neural network processor cores by providing a connection between two non-adjacent rows or columns of the grid, and transparently routing messages between the two non-adjacent rows or columns, past the defective core.

Claims (47)

1. A system comprising:

a plurality of neural network processor cores arrayed in a grid, the grid having a plurality of rows and a plurality of columns; and

a network interconnecting at least those of the plurality of neural network processor cores that are adjacent within the grid, wherein:

the network comprises a switchable bypass between each adjacent core along each direction of two dimensions of the grid, each switchable bypass being directly connected, by a wire, to an input of one of the plurality of neural network processor cores and an output of that one of the plurality of neural network processor cores, and to an input of an adjacent one of the plurality of neural network processor cores, and

the network is adapted to bypass a defective core of the plurality of neural network processor cores by

applying the switchable bypass to provide a connection between two of the plurality of neural network processor cores in non-adjacent rows or columns of the grid, and

transparently routing messages between the two non-adjacent rows or columns, past the defective core, via the switchable bypass.

2. The system of claim 1 , wherein:

the network comprises a loop connecting all cores within a row or column.

3. The system of claim 1 , wherein:

the network comprises a loop connecting all cores within two adjacent rows or columns.

4. The system of claim 1 , wherein the bypass comprises a multiplexor.

5. The system of claim 1 , wherein the bypass precedes each core, and is adapted to select between an input from an adjacent core and a non-adjacent core.

6. The system of claim 1 , wherein the bypass succeeds each core, and is adapted to select between an output of each core and of a preceding core.

7. The system of claim 6 , wherein the bypass is connected to a succeeding core without any additional connection between the core and the succeeding core.

8. The system of claim 1 , wherein the plurality of neural network processor cores are adapted to operate at a reduced clock frequency when bypassing a defective core.

9. The system of claim 1 , wherein each of the plurality of neural network processor cores is identified by a physical address.

10. The system of claim 9 , wherein the network is adapted to map a logical address to each of the cores, and wherein bypassing the defective core comprises remapping the logical addresses.

11. The system of claim 1 , wherein the network is adapted to deliver message between cores, based on relative addresses.

12. A system comprising:

a plurality of neural network processor cores arrayed in a grid, the grid having a plurality of rows and a plurality of columns; and

a network interconnecting at least those of the plurality of neural network processor cores that are adjacent within the grid, wherein:

the network comprises a switchable bypass between each adjacent core along each direction of two dimensions of the grid, each switchable bypass being directly connected, by a wire, to an input of one of the plurality of neural network processor cores and an output of that one of the plurality of neural network processor cores, and to an input of an adjacent one of the plurality of neural network processor cores, and

the network is adapted to bypass a defective core of the plurality of neural network processor cores by

applying the switchable bypass to provide a connection between two of the plurality of neural network processor cores in non-adjacent cores, and

transparently routing messages between the two non-adjacent cores, past the defective core, via the switchable bypass.

13. The system of claim 12 , wherein each of the plurality of neural network processor cores is identified by a physical address.

14. The system of claim 13 , wherein the network is adapted to map a logical address to each of the cores, and wherein bypassing the defective core comprises remapping the logical addresses.

15. The system of claim 12 , wherein the network is adapted to deliver message between cores, based on relative addresses.

16. A method of bypassing a defective core of a plurality of neural network processor cores, the cores being arrayed in a grid, the grid having a plurality of rows and a plurality of columns, a network interconnecting at least those of the plurality of neural network processor cores that are adjacent within the grid, wherein:

the network comprises a switchable bypass between each adjacent core, the method comprising:

directly connecting each switchable bypass, by a wire, to an input of one of the plurality of neural network processor cores and an output of that one of the plurality of neural network processor cores, and to an input of an adjacent one of the plurality of neural network processor cores;

applying the switchable bypass to provide a connection between two of the plurality of neural network processor cores in non-adjacent rows or columns of the grid along each direction of two dimensions of the grid; and

transparently routing messages between the two non-adjacent rows or columns, past the defective core, via the switchable bypass.

17. The method of claim 16 , wherein:

the network comprises a loop connecting all cores within a row or column.

18. The method of claim 16 , wherein:

the network comprises a loop connecting all cores within two adjacent rows or columns.

19. The method of claim 16 , wherein the bypass comprises a multiplexor.

20. The method of claim 16 , wherein the bypass precedes each core, and is adapted to select between an input from an adjacent core and a non-adjacent core.

21. The method of claim 16 , wherein the bypass succeeds each core, and is adapted to select between an output of each core and of a preceding core.

22. The method of claim 21 , wherein the bypass is connected to a succeeding core without any additional connection between the core and the succeeding core.

23. A method of bypassing a defective core of a plurality of neural network processor cores, the cores being arrayed in a grid, the grid having a plurality of rows and a plurality of columns, a network interconnecting at least those of the plurality of neural network processor cores that are adjacent within the grid, wherein:

the network comprises a switchable bypass between each adjacent core, the method comprising:

directly connecting each switchable bypass, by a wire, to an input of one of the plurality of neural network processor cores and an output of that one of the plurality of neural network processor cores, and to an input of an adjacent one of the plurality of neural network processor cores;

applying the switchable bypass to provide a connection between two non-adjacent cores along each direction of two dimensions of the grid; and

transparently routing messages between the two non-adjacent cores, past the defective core, via the switchable bypass.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 2, 2018
From: APPUSWAMY, RATHINAKUMAR; ARTHUR, JOHN V.; CASSIDY, ANDREW S.; DATTA, PALLAB; ESSER, STEVEN K.; FLICKNER, MYRON D.; KLAMO, JENNIFER; MODHA, DHARMENDRA S.; PENNER, HARTMUT; SAWADA, JUN; TABA, BRIAN
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 045811/0683 →
Continuity (1)
Related Publication 20190303741A1 · Oct 3, 2019
Cited By (1)
US 12,711,095