IP Library Granted Patent US 12,259,816
Granted Patent B2
US 12,259,816 · App. 17/809,475 · Granted Mar 25, 2025

Composable infrastructure enabled by heterogeneous architecture, delivered by CXL based cached switch SOC

Inventors: Shreyas Shah (San Jose, CA); George Apostol, Jr. (Los Gatos, CA); Nagarajan Subramaniyan (San Jose, CA); Jack Regula (Durham, NC); Jeffrey S. Earl (San Jose, CA)
Assignee: AVAGO TECHNOLOGIES INTERNATIONAL SALES PTE., LIMITED
G06F12/0868G06F12/0646G06F12/0815G06F12/0837G06F12/0862G06F12/1466G06F13/1642G06F13/1668G06F13/1673G06F13/4022G06F13/4221G06F2213/0026G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,259,816
App. No.
17/809,475
Granted
Mar 25, 2025
Kind
B2
Abstract

Described herein are systems, methods, and products utilizing a cache coherent switch on chip. The cache coherent switch on chip may utilize Compute Express Link (CXL) interconnect open standard and allow for multi-host access and the sharing of resources. The cache coherent switch on chip provides for resource sharing between components while independent of a system processor, removing the system processor as a bottleneck. Cache coherent switch on chip may further allow for cache coherency between various different components. Thus, for example, memories, accelerators, and/or other components within the disclose systems may each maintain caches, and the systems and techniques described herein allow for cache coherency between the different components of the system with minimal latency.

Claims (45)

1. A system comprising:

a first server device comprising:

a first memory device; and

a first cache coherent switch on chip, communicatively coupled to the first memory device via a Compute Express Link (CXL) protocol; and

a second server device, communicatively coupled to the first server device via a data connection, the second server device comprising:

a second memory device; and

a second cache coherent switch on chip, communicatively coupled to the second memory device via the CXL protocol and communicatively coupled to the first cache coherent switch on chip by the data connection via the CXL protocol, wherein second cache coherent switch on chip is configured to:

determine that the second memory device is to be shared with the first server device;

generate an access key for the first server device to access the second memory device; and

provide, to the first server device via the data connection, the access key; and

wherein the first cache coherent switch is configured to communicate CXL protocol memory commands to the second cache coherent switch, each of the CXL protocol memory commands including an instruction and the access key for the first server device to access the second memory device.

2. The system of claim 1 , wherein the first cache coherent switch on chip is configured to:

receive the access key; and

provide, to the second cache coherent switch on chip, read or write instructions for the second memory device, wherein the read or write instructions comprise the access key and a request for reading or writing first data.

3. The system of claim 2 , wherein the second cache coherent switch on chip is further configured to:

receive the read or write instructions;

read or write, in response to receiving the read or write instructions, the first data; and

provide, to the first cache coherent switch on chip and based on the reading or writing the first data, a read or write acknowledgement.

4. The system of claim 3 , wherein the read or write acknowledgement is configured to maintain cache coherency within the system.

5. The system of claim 1 , wherein the first cache coherent switch on chip and the second cache coherent switch on chip are further configured to share cached memory within the first memory device and the second memory device.

6. The system of claim 5 , wherein the first cache coherent switch on chip is configured to:

receive a request for recall of cached first data;

determine that the cached first data is not available within the first memory device; and

provide, to the second cache coherent switch on chip, a request for the cached first data.

7. The system of claim 6 , wherein the second cache coherent switch on chip is further configured to:

determine that the cached first data is stored within the second memory device; and

provide, to the first cache coherent switch on chip, the cached first data.

8. The system of claim 6 , wherein the second server device further comprises:

a microprocessor; and

a third memory device communicatively coupled to the microprocessor, wherein the second cache coherent switch on chip is further configured to:

determine that the cached first data is not stored within the second memory device; and

provide, to the microprocessor, the request for the cached first data from the third memory device;

receive the cached first data from the microprocessor; and

provide, to the first cache coherent switch on chip, the cached first data.

9. The system of claim 1 , further comprising:

a third server device, communicatively coupled to the first server device and the second server device via the data connection, the third server device comprising:

a third memory device; and

a third cache coherent switch on chip, communicatively coupled to the third memory device via the CXL protocol and communicatively coupled to the first cache coherent switch on chip and the second cache coherent switch on chip by the data connection via the CXL protocol, wherein the first memory device, the second memory device, and the third memory device are shared between the first cache coherent switch on chip, the second cache coherent switch on chip, and the third cache coherent switch on chip.

10. The system of claim 9 , wherein the first cache coherent switch on chip is configured to:

determine that a requested cached first data is not stored within the first memory device; and

provide, to the second cache coherent switch on chip and the third cache coherent switch on chip, a request for the cached first data.

11. The system of claim 9 , wherein the first cache coherent switch on chip is configured to:

receive, based on the request for the cached first data, erasure data; and

construct the cached first data by replacing missing data blocks with the erasure data.

12. The system of claim 1 , wherein the first memory device and the second memory device are DRAM or persistent memory.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 3, 2023
From: ELASTICS.CLOUD, INC.
To: AVAGO TECHNOLOGIES INTERNATIONAL SALES PTE. LIMITED
Reel/Frame 065104/0547 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 1, 2022
From: SHAH, SHREYAS; APOSTOL, GEORGE, JR.; SUBRAMANIYAN, NAGARAJAN; REGULA, JACK; EARL, JEFFREY S.
To: ELASTICS.CLOUD, INC.
Reel/Frame 060423/0956 →
Continuity (2)
Provisional Application 63223045 · Jul 18, 2021
Related Publication 20230017643A1 · Jan 19, 2023
References Cited (41)
US 11388268B1 · Siva et al. · 2022 [cited by applicant]
US 11573898B2 · Passint et al. · 2023 [cited by applicant]
US 20150169452A1 · Persson et al. · 2015 [cited by applicant]
US 20160299860A1 · Harriman · 2016 [cited by applicant]
US 20160381176A1 · Cherubini et al. · 2016 [cited by applicant]
US 20190042518A1 · Marolia et al. · 2019 [cited by applicant]
US 20200192798A1 · Natu · 2020 [cited by applicant]
US 20200322287A1 · Connor · 2020 [cited by examiner]
US 20200341930A1 · Cannata et al. · 2020 [cited by applicant]
US 20210011755A1 · Shah · 2021 [cited by applicant]
US 20210075633A1 · Sen et al. · 2021 [cited by applicant]
US 20210117244A1 · Herdrich et al. · 2021 [cited by applicant]
US 20210132999A1 · Haywood · 2021 [cited by examiner]
US 20210240655A1 · Das Sharma · 2021 [cited by applicant]
US 20210311643A1 · Shanbhogue · 2021 [cited by examiner]
US 20210311646A1 · Malladi et al. · 2021 [cited by applicant]
US 20210311739A1 · Malladi et al. · 2021 [cited by applicant]
US 20210318976A1 · Zhang et al. · 2021 [cited by applicant]
US 20210320866A1 · Le et al. · 2021 [cited by applicant]
US 20210374056A1 · Malladi · 2021 [cited by examiner]
US 20210382838A1 · Mittal et al. · 2021 [cited by applicant]
US 20220124038A1 · Leguay et al. · 2022 [cited by applicant]
US 20220147476A1 · Nam et al. · 2022 [cited by applicant]
US 20220164288A1 · Ramagiri et al. · 2022 [cited by applicant]
US 20220292026A1 · Hornung et al. · 2022 [cited by applicant]
US 20220350767A1 · Mcgraw et al. · 2022 [cited by applicant]
US 20220398207A1 · Norrie et al. · 2022 [cited by applicant]
US 20220405212A1 · Kakaiya · 2022 [cited by examiner]
US 20230012822A1 · Shah et al. · 2023 [cited by applicant]
US 20230017583A1 · Shah et al. · 2023 [cited by applicant]
US 20230017643A1 · Shah et al. · 2023 [cited by applicant]
US 20230409302A1 · Kodama et al. · 2023 [cited by applicant]
“FlexPod Datacenter with Citrix VDI and VMware vSphere 7 for up to 2500 Seats”, Cisco, Published Apr. 2022, http://www.cisco.com/go/designzone, 497 pages. [cited by applicant]
Amir Roozbeh, “Realizing Next-Generation Data Centers via Software-Defined “Hardware” Infrastructures and Resource Disaggregation”, Doctoral Thesis KTH Royal Institute of Technology, 227 pages. [cited by applicant]
Davide Giri et al, “NoC-Based Support of Heterogeneous Cache-Coherence Models for Accelerators”, 2018 Twelfth IEEE/ACM International Symposium on Networks-on-Chip (NOCS), IEEE Oct. 4, 2018, pp. 1-8, Section III and figu… [cited by applicant]
Int'l Application Serial No. PCT/US22/73233, ISR/WO mailed 10/14/229 pgs. [cited by applicant]
Kshitij Bhardwaj et al, “Determining Optimal Coherence Interface for Many-Accelerator SoC's Using Bayesian Optimization”, IEEE Computer Architecture Letters, IEEE Sep. 16, 2019, pp. 119-123 Section 3.1; and figure 2. [cited by applicant]
Kshitij Bhardwaj, et al., “A Comprehensive Methodology to Determine Optimal Coherence Interfaces for Many-Accelerator SoC's”, ISLPED '20 Proceedings of the ACM/IEEE International Symposium on Low Power Electronics and D… [cited by applicant]
Prateek Shantharama, et al., “Hardware Accelerated Platforms and Infrastructures for Network Functions: A Survey of Enabling Technologies and Research Studies”.IEEE Jul. 9, 2020, Digital Object Identifier 10.1109/ACCESS… [cited by applicant]
Yakun Sophia Shao, et al. “Co-Designing Accelerators and SoC Interfaces using gem5-Aladdin”, 2016 49th Annual IEEE/ACM International Symposium on Microarchitecture (MICRO). IFEE, Oct. 15, 2016, pp. 1-12, pp. 3-5 and fig… [cited by applicant]
Zuckerman, et al. “Cohmeleon: Learning-Based Orchestration ofAccelerator Coherence in Heterogeneous SoCs”, Columbia University, New York, New York, arXiv:2109.06382v1 [cs.AR] Sep. 14, 2021, 14 pages. [cited by applicant]