IP Library › Granted Patent US 12,430,370
Granted Patent B2
US 12,430,370 · App. 19/062,115 · Granted Sep 30, 2025

Method and system for multi-level artificial intelligence supercomputer design

Inventors: Vijay Madisetti (Alpharetta, GA); Arshdeep Bahga (Chandigarh, IN)
Assignee: Vijay Madisetti
G06F16/3329G06F40/284
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,430,370
App. No.
19/062,115
Filed
Feb 25, 2025
Granted
Sep 30, 2025
Kind
B2
Examiner
YEN, ERIC L
Art Unit
2658
USPC
704/9
Abstract

A method for generating a merged large language model (LLM) from input data including receiving derived prompts generated from a user prompt and relevant contexts, receiving the relevant contexts related to the generation of the plurality of derived prompts, providing derived prompts and the relevant contexts to one or more merged LLMs, generating a plurality of results, and sending the plurality of results to an output broker.

Claims (60)

1. A method for generating a merged large language model (LLM) from input data comprising:

receiving a plurality of derived prompts generated from a user prompt and one or more relevant contexts;

sending the plurality of derived prompts to one or more prompt embedding models;

receiving one or more knowledge documents responsive to sending the plurality of derived prompts to the one or more prompt embedding models;

receiving a plurality of updated derived prompts that were generated based on the plurality of derived prompts and the one or more knowledge documents;

providing the plurality of updated derived prompts to one or more merged LLMs, generating a plurality of results; and

sending the plurality of results to an output broker.

2. The method of claim 1 wherein the one or more merged LLMs comprises a plurality of merged LLMs.

3. The method of claim 1 wherein:

the one or more merged LLMs are the result of merging a first LLM trained on a first subset of input data for a first specialized task and a second LLM trained on a second subset of input data for a second specialized task;

the first subset of the input data has a first data type; and

the second subset of the input data has a second data type.

4. The method of claim 3 wherein:

the first data type is one of text data, image data, audio data, video data, and code data; and

the second data type is one of text data, image data, audio data, video data, and code data and is different from the first data type.

5. The method of claim 3 wherein the one or more merged LLMs has at least one of an improved precision and an improved accuracy relative to at least one of the first LLM and the second LLM.

6. The method of claim 4 wherein the one or more merged LLMs has at least one of an improved precision and an improved accuracy relative to each of the first LLM and the second LLM.

7. The method of claim 3 wherein the first LLM is trained in parallel with the second LLM.

8. The method of claim 3 wherein:

the first LLM is a first family of LLMs (h-LLM);

the second LLM is a second h-LLM; and

the one or more merged LLMs is one or more merged h-LLMs.

9. The method of claim 3 wherein the one or more merged LLMs are produced by a merging process that is configured to produce a merged LLM that is specialized to perform a specific task.

10. The method of claim 1 further comprising filtering the plurality of results for at least one of accuracy or service level assurance.

11. The method of claim 1 wherein:

the one or more merged LLMs is added to a plurality of merged LLMs; and

providing the plurality of updated derived prompts to the one or more merged LLMs comprises sending the plurality of updated derived prompts to the plurality of merged LLMs, generating the plurality of results.

12. A method for generating a merged large language model (LLM) from input data comprising:

training a first LLM on a first subset of the input data for a first specialized task;

training a second LLM on a second subset of the input data for a second specialized task;

performing a merging process on the first LLM and the second LLM to produce one or more merged LLMs, the one or more merged LLMs being operable to:

receive a plurality of updated derived prompts generated from a plurality of derived prompts and one or more relevant contexts; and

generate a plurality of results based on the plurality of updated derived prompts.

13. The method of claim 12 wherein the one or more merged LLMs comprises a plurality of merged LLMs.

14. The method of claim 12 wherein:

the first subset of the input data has a first data type; and

the second subset of the input data has a second data type.

15. The method of claim 14 wherein:

the first data type is one of text data, image data, audio data, video data, and code data; and

the second data type is one of text data, image data, audio data, video data, and code data and is different from the first data type.

16. The method of claim 12 wherein the one or more merged LLMs has at least one of an improved precision and an improved accuracy relative to at least one of the first LLM and the second LLM.

17. The method of claim 12 wherein the first LLM is trained in parallel with the second LLM.

18. The method of claim 12 wherein:

the first LLM is a first family of LLMs (h-LLM);

the second LLM is a second h-LLM; and

the one or more merged LLMs is one or more merged h-LLMs.

19. The method of claim 12 wherein the one or more merged LLMs are produced by a merging process that is configured to produce a merged LLM that is specialized to perform a specific task.

20. The method of claim 12 further comprising filtering the plurality of results for at least one of accuracy or service level assurance.

21. The method of claim 12 wherein:

the one or more merged LLMs is added to a plurality of merged LLMs; and

providing the plurality of updated derived prompts to the one or more merged LLMs comprises sending the plurality of updated derived prompts to the plurality of merged LLMs, generating the plurality of results.

22. A system for generating a merged large language model (LLM) from input data comprising:

a processor;

a network communication device positioned in communication with the processor and operable to communicate across a computer network; and

a non-transitory computer-readable storage medium positioned in communication with the processor and having stored thereon software that, when executed by the processor is operable to:

receive a plurality of derived prompts generated from a user prompt and one or more relevant contexts;

receive the one or more relevant contexts;

receive a plurality of updated derived prompts from the plurality of derived prompts and one or more relevant contexts;

provide the plurality of updated derived prompts to one or more merged LLMs, generating a plurality of results; and

send the plurality of results to an output broker.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 6, 2025
From: BAHGA, ARSHDEEP
To: MADISETTI, VIJAY
Reel/Frame 070420/0143 →
Continuity (6)
Continuation 18826342 · Sep 6, 2024
Continuation 18470487 · Sep 20, 2023
Continuation 18348692 · Jul 7, 2023
Provisional Application 63469571 · May 30, 2023
Provisional Application 63463913 · May 4, 2023
Related Publication 20250209098A1 · Jun 26, 2025
References Cited (27)
US 11960983B1 · Gelfenbeyn · 2024 [cited by examiner]
US 12135937B1 · Mancuso · 2024 [cited by examiner]
US 20220292266A1 · Dey · 2022 [cited by examiner]
US 20240001228A1 · Gilg · 2024 [cited by examiner]
US 20240184834A1 · McCann · 2024 [cited by examiner]
US 20240202225A1 · Siebel · 2024 [cited by examiner]
US 20240202452A1 · Schillace · 2024 [cited by examiner]
US 20240202539A1 · Poirier · 2024 [cited by examiner]
US 20240281472A1 · LaRhette · 2024 [cited by examiner]
US 20240281663A1 · Chen · 2024 [cited by examiner]
US 20240290331A1 · Lu · 2024 [cited by examiner]
US 20240320444A1 · Maschmeyer · 2024 [cited by examiner]
US 20240346162A1 · Luitjens · 2024 [cited by examiner]
US 20240346256A1 · Qin · 2024 [cited by examiner]
US 20240346566A1 · Qin · 2024 [cited by examiner]
US 20240362212A1 · Sun · 2024 [cited by examiner]
US 20240370658A1 · Zhang · 2024 [cited by examiner]
US 20240403194A1 · Hawes · 2024 [cited by examiner]
US 20240403341A1 · Berglund · 2024 [cited by examiner]
US 20240404632A1 · Ho · 2024 [cited by examiner]
US 20240406125A1 · Hine · 2024 [cited by examiner]
US 20240411797A1 · Blum · 2024 [cited by examiner]
US 20240412029A1 · Yang · 2024 [cited by examiner]
US 20240414108A1 · Sun · 2024 [cited by examiner]
US 20240414191A1 · Humphrey · 2024 [cited by examiner]
US 20240420012A1 · Austin · 2024 [cited by examiner]
US 20250014605A1 · Chen · 2025 [cited by examiner]
Cited By (2)
US 12,639,346 US 12,743,451