IP Library Patent Application 18748826
Patent Application
App. No. 18/748,826

ARTIFICIAL INTELLIGENCE TRAINING SYSTEM

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
18/748,826
Abstract

A computing system is provided that includes at least one processing unit, at least one high bandwidth memory (HBM) unit, and at least one high bandwidth flash (HBF) unit. The HBM and HBF units are all in electrical communication with the at least one processing unit. The computing system also includes control circuitry that is configured to train a large language model according to a low-rank adaptation (LoRA) technique. The control circuitry is configured to store a full-weight matrix in the at least one HBF unit and to store at least one low-rank matrix in the at least one HBM unit.

Claims (33)

1 . A method of training a large language model using a low rank adaptation (LoRA) technique, comprising the steps of:

preparing a computing system that includes at least one processing unit and at least one high bandwidth memory (HBM) unit in electrical communication with the at least one processing unit and at least one high bandwidth flash (HBF) unit in electrical communication with the at least one processing unit;

storing a full-weight matrix in the at least one HBF unit; and

storing at least one low-rank matrix in the at least one HBM unit.

2 . The method as set forth in claim 1 , further including the step of generating the at least one low-rank matrix from the full-weight matrix.

3 . The method as set forth in claim 2 , wherein the step of generating the at least one low rank matrix from the full-weight matrix includes the step of generating a pair of low-rank matrices from the full-weight matrix.

4 . The method as set forth in claim 3 , further including the step of adjusting the low-rank matrices based on an input.

5 . The method as set forth in claim 4 , wherein after the step of adjusting the low-rank matrices based on the input, the method further includes the step of adjusting the full-weight matrix based on the adjusted low-rank matrices.

6 . The method as set forth in claim 1 , wherein the at least one HBF unit includes a plurality of HBF units that do not allow random access, and

wherein the plurality of HBF units have arrays of memory cells that are arranged in a plurality of word lines and memory holes.

7 . The method as set forth in claim 1 , wherein the at least one HBM unit includes a plurality of HBM units that allow random access.

8 . The method as set forth in claim 7 , wherein the plurality of HBM units are dynamic random access memory (DRAM).

9 . The method as set forth in claim 1 , wherein the at least one HBF unit has a bandwidth of at least 3 TB/s.

10 . A computing system, comprising:

at least one processing unit, at least one high bandwidth memory (HBM) unit in electrical communication with the at least one processing unit, and at least one high bandwidth flash (HBF) unit in electrical communication with the at least one processing unit;

control circuitry that is configured to train a large language model according to a low-rank adaptation (LoRA) technique, the control circuitry being configured to;

store a full-weight matrix in the at least one HBF unit, and

store at least one low-rank matrix in the at least one HBM unit.

11 . The computing system as set forth in claim 10 , wherein the control circuitry is configured to generate the at least one low-rank matrix from the full-weight matrix.

12 . The computing system as set forth in claim 10 , wherein the at least one low-rank matrix includes a pair of low-rank matrices.

13 . The computing system as set forth in claim 12 , wherein the control circuitry is configured to adjust the low-rank matrices based on an input.

14 . The computing system as set forth in claim 13 , wherein after adjusting the low-rank matrices based on the input, the control circuitry is configured to adjust the full-weight matrix based on the adjusted low-rank matrices.

15 . The computing system as set forth in 10 wherein the at least one HBF unit includes a plurality of HBF units that do not allow random access, and

wherein the plurality of HBF units have arrays of memory cells that are arranged in a plurality of word lines and memory holes.

16 . The computing system as set forth in claim 10 , wherein the at least one HBM unit includes a plurality of HBM units that allow random access.

17 . The computing system as set forth in claim 16 , wherein the plurality of HBM units are dynamic random access memory (DRAM).

18 . The computing system as set forth in claim 10 , wherein the at least one HBF unit has a bandwidth of at least 3 TB/s.

19 . An apparatus, comprising:

at least one processing unit, at least one high bandwidth memory (HBM) unit that is volatile and is in electrical communication with the at least one processing unit, and at least one high bandwidth flash (HBF) unit that is non-volatile and is in electrical communication with the at least one processing unit;

an artificial intelligence training means for training a large language model according to a low-rank adaptation (LoRA) technique, the artificial intelligence training means being configured to;

store a full-weight matrix in the at least one HBF unit, and

store at least one low-rank matrix in the at least one HBM unit.

20 . The apparatus as set forth in claim 19 , wherein the artificial intelligence training means is configured to adjust the at least one low-rank matrix based on an input and then adjust the full-weight matrix based on the adjusted at least one low-rank matrix.

Assignments (4)
SECURITY AGREEMENT Recorded Apr 25, 2025
From: SANDISK TECHNOLOGIES, INC.
To: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
Reel/Frame 071050/0001 →
PARTIAL RELEASE OF SECURITY INTERESTS Recorded Apr 25, 2025
From: JPMORGAN CHASE BANK, N.A., AS AGENT
To: SANDISK TECHNOLOGIES, INC.
Reel/Frame 071382/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 31, 2024
From: SANDISK TECHNOLOGIES LLC
To: SANDISK TECHNOLOGIES, INC.
Reel/Frame 069796/0423 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 23, 2024
From: GUYOT, CYRIL; MATEESCU, ROBERT; QIN, MINGHAI
To: SANDISK TECHNOLOGIES LLC
Reel/Frame 068668/0096 →