IP Library Granted Patent US 10,445,451
Granted Patent B2
US 10,445,451 · App. 15/640,535 · Granted Oct 15, 2019

Processors, methods, and systems for a configurable spatial accelerator with performance, correctness, and power reduction features

Inventors: Kermin Fleming (Hudson, MA); Kent D. Glossop (Merrimack, NH); Simon C. Steely, Jr. (Hudson, NH); Ping Tak Peter Tang (Edison, NJ)
Assignee: Intel Corporation
G06F17/505G06F12/0802G06F15/7867G06F15/8015G11C8/12H03K19/1778H03K19/17736H03K19/17756H03K19/17764H03K19/17776
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,445,451
App. No.
15/640,535
Granted
Oct 15, 2019
Kind
B2
Abstract

Systems, methods, and apparatuses relating to a configurable spatial accelerator are described. In one embodiment, a processor includes a plurality of processing elements; and an interconnect network between the plurality of processing elements to receive an input of a dataflow graph comprising a plurality of nodes, wherein the dataflow graph is to be overlaid into the interconnect network and the plurality of processing elements with each node represented as a dataflow operator in the plurality of processing elements, and the plurality of processing elements is to perform an operation when an incoming operand set arrives at the plurality of processing elements. At least one of the plurality of processing elements includes a plurality of control inputs.

Claims (5)

1. A processor comprising:

a plurality of processing elements, wherein at least one of the processing elements is to perform a floating-point operation with selectable precision control; and

an interconnect network between the plurality of processing elements to receive an input of a dataflow graph comprising a plurality of nodes, wherein the dataflow graph is to be overlaid into the interconnect network and the plurality of processing elements with each node represented as a dataflow operator in the plurality of processing elements, and the plurality of processing elements is to perform an operation when an incoming operand set arrives at the plurality of processing elements.

2. The processor of claim 1 , wherein:

the at least one of the plurality of processing elements is configurable as a floating-point multiplier, at least an other of the plurality of processing elements is configurable as a floating-point adder, and the at least one and the at least an other of the plurality of processing elements are coupled together to perform a fused multiply-add.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 4, 2017
From: FLEMING, KERMIN; GLOSSOP, KENT D.; STEELY, SIMON C., JR.; TANG, PING TAK PETER
To: INTEL CORPORATION
Reel/Frame 043208/0054 →
Continuity (1)
Related Publication 20190005161A1 · Jan 3, 2019
Cited By (3)
US 12,386,602 US 12,568,344 US 12,705,282