IP Library Granted Patent US 11,669,418
Granted Patent B1
US 11,669,418 · App. 16/674,083 · Granted Jun 6, 2023

Simultaneous multi-processor apparatus applicable to achieving exascale performance for algorithms and program systems

Inventors: Earle Jennings (Somerset, CA); George Landers (Tigard, OR)
Assignee: QSIGMA, INC.
G06F11/2028G06F11/10G06F11/1443G06F11/1625G06F11/1679G06F15/161G06F15/17362G06F11/1423G06F11/202
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,669,418
App. No.
16/674,083
Granted
Jun 6, 2023
Kind
B1
Abstract

Apparatus adapted for exascale computers are disclosed. The apparatus includes, but is not limited to at least one of: a system, data processor chip (DPC), Landing module (LM), chips including LM, anticipator chips, simultaneous multi-processor (SMP) cores, SMP channel (SMPC) cores, channels, bundles of channels, printed circuit boards (PCB) including bundles, floating point adders, accumulation managers, QUAD Link Anticipating Memory (QUADLAM), communication networks extended by coupling links of QUADLAM, log2 calculators, exp2 calculators, logALU, Non-Linear Accelerator (NLA), and stairways. Methods of algorithm and program development, verification and debugging are also disclosed. Collectively, embodiments of these elements disclose a class of supercomputers that obsolete Amdahl's Law, providing cabinets of petaflop performance and systems that may meet or exceed an exaflop of performance for Block LU Decomposition (Linpack).

Claims (20)

1. An apparatus, comprising:

A. an anticipator chip adapted to respond to a system performance requirement of Kpref exaflop by a system for a sustained runtime for an algorithm and an incremental state of said algorithm received by said anticipator;

B. said anticipator chip is adapted to respond to said incremental state by creating an anticipated requirement; and

C. said anticipator chip is adapted to respond to said anticipated requirement by directing at least part of said system to achieve said system performance requirement

D. wherein said anticipator further includes a state table adapted for configuration to integrate said incremental states of said algorithm to update said state table to account for said anticipated requirement; and said anticipator responds to a successor incremental state based upon said state table in order to generate a successor anticipated requirement;

E. wherein said state table is adapted to integrate said incremental states of said algorithm to update said state table to account for said anticipated requirement, for each of said incremental states;

F. wherein said algorithm includes a form of Block LU Decomposition with partial pivoting of a matrix A including at least N rows and at least N columns of double precision floating point numbers, where said N is a member of a second group consisting of said number multiplied by K of said rows and said columns resulting in said Kpref sustained performance, where said K is 1024;

G. wherein said number is a member of the second group consisting of 1/4 to 16;

H. wherein said sustained runtime is at least one and no more than 8 hours; wherein said incremental state includes a pivot decision for one of said columns of said matrix A; and

J. wherein said Kpref exaflop is a member of a first group consisting of ¼ exaflop to 1 exaflop.

2. The apparatus of claim 1 , wherein said anticipated requirement, includes

A. an anticipated future memory transfer requirement of at least one memory unit array as an associated large memory to said anticipator chip; and

B. an anticipated future transfer requirement of at least one Landing Module (LM) chip as at least one associated communication node chip to said anticipator chip.

3. The apparatus of claim 2 , wherein said anticipator adapted to respond to said anticipated requirement includes said anticipator configured to perform

A. said anticipator scheduling memory transfers of said associated memory unit array to fulfill said anticipated future memory transfer requirement;

B. said anticipator configuring at least one of said associated communication node chips to fulfill said anticipated future transfer requirement.

4. The apparatus of claim 2 , wherein said anticipated requirement further includes:

A. an anticipated internal transfer requirement for a Data Processor Chip (DPC) as an associated DPC to said anticipator chip; and

B. said anticipator configuring at most one of said associated DPC to respond to said anticipated internal transfer requirement of said associated DPC with any coupled said associated communication node chips so that said performance requirement is met in the average over said sustained runtime.

5. The apparatus of claim 2 , wherein said system performance requirement includes said system performing said Kpref exaflop for said sustained runtime directed by said algorithm.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 23, 2022
From: QSIGMA INC. (CA)
To: QSIGMA INC
Reel/Frame 062214/0202 →
Continuity (7)
Division 15844740 · Dec 18, 2017
Division 15695939 · Sep 5, 2017
Provisional Application 62328470 · Apr 27, 2016
Provisional Application 62261836 · Dec 1, 2015
Provisional Application 62243885 · Oct 20, 2015
Provisional Application 62233547 · Sep 28, 2015
Provisional Application 62207432 · Aug 20, 2015