IP Library › Granted Patent US 11,854,661
Granted Patent B2
US 11,854,661 · App. 17/948,937 · Granted Dec 26, 2023

Copy data in a memory system with artificial intelligence mode

Inventor: Alberto Troia (Munich, DE)
Assignee: Micron Technology, Inc.
G11C7/22G06F11/1448G06F12/0246G06F18/214G06N3/08G11C11/409
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,854,661
App. No.
17/948,937
Granted
Dec 26, 2023
Kind
B2
Abstract

The present disclosure includes apparatuses and methods related to copying data in a memory system with an artificial intelligence (AI) mode. An apparatus can receive a command indicating that the apparatus operate in an artificial intelligence (AI) mode, a command to perform AI operations using an AI accelerator based on a status of a number of registers, and a command to copy data between memory devices that are performing AI operations. The memory system can copy neural network data, activation function data, bias data, input data, and/or output data from a first memory device to a second memory device, such that that the first memory device can use the neural network data, activation function data, bias data, input data, and/or output data in a first AI operation and the second memory device can use the neural network data, activation function data, bias data, input data, and/or output data in a second AI operation.

Claims (37)

1. An apparatus, comprising:

a controller; and

a number of memory devices coupled to the controller, wherein each of the number of memory devices is configured as part of a neural network and includes a number of memory arrays and wherein the controller is configured to:

provide input information associated with data stored in a first memory device of the number of memory devices for a training operation, an inference operation, or both;

receive a command comprising a first address identifying the first memory device and a second address identifying a second memory device of the number of memory devices as a target for the stored data;

transmit the stored data from the first memory device to the second memory device in response to the command and the input information; and

execute the training operation, the inference operation, or both, at the second memory device using the stored data.

2. The apparatus of claim 1 , wherein the input information comprises a set of data providing information associated with the stored data.

3. The apparatus of claim 1 , wherein the command instructs the second memory device to receive the stored data and write the stored data to a memory array in the second memory array.

4. The apparatus of claim 1 , wherein the command is issued by a host device.

5. The apparatus of claim 1 , wherein the input information is stored in the second memory device.

6. The apparatus of claim 1 , wherein the input information is stored at a host device.

7. The apparatus of claim 1 , wherein the controller configured to transmit the stored data from the first memory device to the second memory device comprises copying the input information for the training operation, the inference operation, or both.

8. A system, comprising:

a controller; and

a number of memory devices coupled to the controller, wherein each of the number of memory devices are configured as part of a neural network and include a number of memory arrays and wherein and the is controller configured to:

receive a first command to execute a first training or inference operation on a first memory device of the number of memory devices;

receive a second command to copy data from the first memory device to a second memory device of the number of memory device and to execute a second training or inference operation on the second memory device, wherein the command comprises a first address and a second address, the first address identifying the first memory device storing data associated with the first training or inference operation and the second address identifying the second memory device as a target for the data;

transmit data from the first memory device to the second memory device in response to the first command, the second command, and input information comprising a set of data providing information associated with the copied data; and

execute the training or inference operation at the second memory device using the copied data.

9. The system of claim 8 , wherein the copied data comprises data associated with the training or inference operation.

10. The system of claim 8 , wherein neural network data is copied from the first memory device to the second memory device.

11. The system of claim 10 , wherein the neural network data is used by the first memory device to perform the first training or inference operation and the second memory device to perform the second training or inference operation.

12. The system of claim 8 , wherein the command is issued by a host device, and the input information is stored at the host device.

13. The system of claim 8 , wherein the command is issued by a host device, and the input information is stored at the second memory device.

14. The system of claim 8 , wherein activation function data is copied from the first memory device to the second memory device.

15. A method, comprising:

receiving, at a first memory device of a number of memory devices, input information associated with stored data for a training operation, an inference operation, or both;

receiving, from a host device and at the first memory device, a command comprising a first address and a second address, the first address identifying the first memory device storing the data for the training operation, the inference operation,

or both, and the second address identifying a second memory device of the number of memory devices as a target for the stored data;

transmitting the stored data from the first memory device to the second memory device in response to the command and the input information, wherein the stored data comprises the input information for the training operation, the inference operation, or both; and

executing the training operation, the inference operation, or both, at the second memory device using the stored data.

16. The method of claim 15 , wherein transmitting the stored data from the first memory device to the second memory device comprises copying neural network data for the training operation, the inference operation, or both.

17. The method of claim 15 , wherein receiving the command comprises receiving a command that selects the first memory device and the second memory device to transfer data on a bus shared by the number of memory devices.

18. The method of claim 15 , wherein receiving the command comprises receiving a command that enables the second memory device to enter an artificial intelligence (AI) mode to perform an AI operation.

19. The method of claim 15 , further comprising copying the stored data from the first memory device to a third memory device of the number of memory devices in response to receiving another command that comprises the first address and a third address, the first address identifying the first memory device storing data for the training or inference operation and the third address identifying the second memory device as the target for the stored data.

20. The method of claim 19 , further including performing another training or inference operation on the third memory device using the data copied from the first memory device.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 20, 2022
From: TROIA, ALBERTO
To: MICRON TECHNOLOGY, INC.
Reel/Frame 061488/0782 →
Continuity (4)
Continuation 17856099 · Jul 1, 2022
Continuation 17328751 · May 24, 2021
Continuation 16554924 · Aug 29, 2019
Related Publication 20230017697A1 · Jan 19, 2023