IP Library Granted Patent US 12688146
Granted Patent B2
US 12688146 · App. 18/626,775 · Granted Jul 21, 2026

Multi-tile memory management for detecting cross tile access providing multi-tile inference scaling and providing page migration

Inventors: Lakshminarayanan Striramassarma (Folsom, CA); Prasoonkumar Surti (Folsom, CA); Varghese George (Folsom, CA); Ben Ashbaugh (Folsom, CA); Aravindh Anantaraman (Folsom, CA); Valentin Andrei (San Jose, CA); Abhishek Appu (El Dorado Hills, CA); Nicolas Galoppo Von Borries (Portland, OR); Altug Koker (El Dorado Hills, CA); Mike Macpherson (Portland, OR); Subramaniam Maiyuran (Gold River, CA); Nilay Mistry (Bangalore, IN); Elmoustapha Ould-Ahmed-Vall (Chandler, AZ); Selvakumar Panneer (Portland, OR); Vasanth Ranganathan (El Dorado Hills, CA); Joydeep Ray (Folsom, CA); Ankur Shah (Folsom, CA); Saurabh Tangri (Folsom, CA)
Assignee: INTEL CORPORATION
G06F15/7839G06F7/5443G06F7/575G06F7/588G06F9/3001G06F9/30014G06F9/30036G06F9/3004G06F9/30043G06F9/30047G06F9/30065G06F9/30079G06F9/3887G06F9/3888G06F9/5011G06F9/5077G06F12/0215G06F12/0238G06F12/0246G06F12/0607G06F12/0802G06F12/0804G06F12/0811G06F12/0862G06F12/0866G06F12/0871G06F12/0875G06F12/0882G06F12/0888G06F12/0891G06F12/0893G06F12/0895G06F12/0897G06F12/1009G06F12/128G06F13/1626G06F15/8046G06F17/16G06F17/18G06T1/20G06T1/60H03M7/46G06F9/3802G06F9/3818G06F9/3867G06F2212/1008G06F2212/1021G06F2212/1044G06F2212/302G06F2212/401G06F2212/455G06F2212/60G06N3/08G06T15/06
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12688146
App. No.
18/626,775
Granted
Jul 21, 2026
Kind
B2
Abstract

Multi-tile Memory Management for Detecting Cross Tile Access, Providing Multi-Tile Inference Scaling with multicasting of data via copy operation, and Providing Page Migration are disclosed herein. In one embodiment, a graphics processor for a multi-tile architecture includes a first graphics processing unit (GPU) having a memory and a memory controller, a second graphics processing unit (GPU) having a memory and a cross-GPU fabric to communicatively couple the first and second GPUs. The memory controller is configured to determine whether frequent cross tile memory accesses occur from the first GPU to the memory of the second GPU in the multi-GPU configuration and to send a message to initiate a data transfer mechanism when frequent cross tile memory accesses occur from the first GPU to the memory of the second GPU.

Claims (20)

1 . An apparatus comprising:

graphics processing circuitry comprising a first graphics processing unit (GPU) including a first local memory and a memory controller, and a second GPU including a second local memory, wherein the graphics processing circuitry to:

facilitate the memory controller to send a message to a graphics driver to initiate data transfer when frequent cross tile memory accesses occur between the first and second GPUs, wherein content of the data transfer is accessed by the first GPU via the first local memory and by the second GPU via the second local memory; and

perform a page allocation to the first local memory when a first access to a page occurs in the second local memory, wherein the graphics processing circuitry is further to enable split frame between the first GPU and the second GPU.

2 . The apparatus of claim 1 , wherein, based on the contents of the data transfer, the first GPU to handle rendering for a first portion of a display and the second GPU to handle rendering for a second portion of the display, wherein the second portion is different from the first portion, wherein the graphics processing circuitry is further to facilitate the first GPU to determine whether frequent cross tile memory accesses occur between the first and second GPUs.

3 . The apparatus of claim 1 , wherein the data transfer between the first and second GPUs is initiated in response to the graphics driver receiving the message from the memory controller.

4 . The apparatus of claim 1 , wherein the data transfer allows for accessing a page table to provide a translation of virtual addresses to physical addresses.

5 . A method comprising:

facilitating, by graphics processing circuitry of a computing device, the memory controller to send a message to a graphics driver to initiate data transfer when frequent cross tile memory accesses occur between a first graphics processing unit (GPU) and a second GPU having a first local memory and a second local memory, respectively, wherein content of the data transfer is accessed by the first GPU via the first local memory and by the second GPU via the second local memory, wherein the first GPU includes a memory controller; and

performing a page allocation to the first local memory when a first access to a page occurs in the second local memory, wherein the graphics processing circuitry is further to enable split frame between the first GPU and the second GPU.

6 . The method of claim 5 , wherein, based on the contents of the data transfer, the first GPU to handle rendering for a first portion of a display and the second GPU to handle rendering for a second portion of the display, wherein the second portion is different from the first portion, wherein the graphics processing circuitry is further to facilitate the first GPU to determine whether frequent cross tile memory accesses occur between the first and second GPUs.

7 . The method of claim 5 , wherein the data transfer between the first and second GPUs is initiated in response to the graphics driver receiving the message from the memory controller.

8 . The method of claim 5 , wherein the data transfer allows for accessing a page table to provide a translation of virtual addresses to physical addresses.

9 . At least one non-transitory computer-readable medium having stored thereon instructions which, when executed, cause a computing device to perform operations comprising:

facilitating, by graphics processing circuitry of the computing device, the memory controller to send a message to a graphics driver to initiate data transfer when frequent cross tile memory accesses occur between a first graphics processing unit (GPU) and a second GPU having a first local memory and a second local memory, respectively, wherein content of the data transfer is accessed by the first GPU via the first local memory and by the second GPU via the second local memory, wherein the first GPU includes a memory controller; and

performing a page allocation to the first local memory when a first access to a page occurs in the second local memory, wherein the graphics processing circuitry is further to enable split frame between the first GPU and the second GPU.

10 . The non-transitory computer-readable medium of claim 9 , wherein, based on the contents of the data transfer, the first GPU to handle rendering for a first portion of a display and the second

GPU to handle rendering for a second portion of the display, wherein the second portion is different from the first portion, wherein the graphics processing circuitry is further to facilitate the first GPU to determine whether frequent cross tile memory accesses occur between the first and second GPUs.

11 . The non-transitory computer-readable medium of claim 9 , wherein the data transfer between the first and second GPUs is initiated in response to the graphics driver receiving the message from the memory controller.

12 . The method of claim 5 , wherein the data transfer allows for accessing a page table to provide a translation of virtual addresses to physical addresses.