IP Library › Granted Patent US 12,596,736
Granted Patent B2
US 12,596,736 · App. 18/962,656 · Granted Apr 7, 2026

Systems and methods for using prompt dissection for large language models

Inventors: Vineeth Chinmaya Murthy (Bengaluru, IN); Rafal Powalski (Warsaw, PL)
Assignee: Instabase, Inc.
G06F16/3347G06F40/30
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,596,736
App. No.
18/962,656
Granted
Apr 7, 2026
Kind
B2
Abstract

Systems and methods for using prompt dissection to prompts one or more machine learning models for a set of one or more documents are disclosed. Exemplary implementations may: store a set of document segments from the set of one or more documents; present a user interface to obtain a compound query that includes a first subquery and a second subquery; create subquery vectors that semantically represent the subqueries; determine a subset of a set of semantic vectors based on comparisons with the subquery vectors; create a combination of the individual document segments that are associated with the subset; provide one or more prompts to the one or more machine learning models, using the combination of the individual document segments as context; present replies from the one or more machine learning models, and/or perform other steps.

Claims (33)

1 . A system configured to use prompt dissection to prompt one or more machine learning models for a set of one or more documents, the system comprising:

electronic storage configured to electronically store information, wherein the stored information includes a set of document segments, wherein individual document segments are included in the set of one or more documents, wherein the stored information further includes a set of semantic vectors, wherein individual semantic vectors are associated with individual document segments; and

one or more hardware processors configured by machine-readable instructions to:

effectuate a presentation of a user interface, the user interface being configured to obtain a compound query from a user, wherein the compound query includes two or more subqueries including a first subquery and a second subquery;

create, using the one or more machine learning models, a first subquery vector representing the first subquery;

determine a subset of the set of semantic vectors, wherein the determination is based on comparisons, wherein the comparisons include a first comparison using the first subquery vector;

create a combination of the individual document segments that are associated with the subset of the set of semantic vectors as determined;

provide one or more prompts to the one or more machine learning models that prompt the one or more machine learning models to generate one or more replies, using the created combination of the individual document segments as context for the one or more machine learning models, wherein the one or more prompts are based on the compound query, wherein the second subquery constrains the one or more replies, wherein an individual machine learning model from the one or more machine learning models is limited to a predetermined numbers of tokens as the context for the one or more prompts, and wherein the combination of the individual document segments is created such that the predetermined number of tokens is not exceeded; and

present to the user, through the user interface, the one or more replies obtained from the one or more machine learning models in reply to the one or more prompts, wherein the one or more replies have been constrained by the second subquery.

2 . The system of claim 1 , wherein the first comparison compares similarity between the first subquery vector and individual ones of the set of semantic vectors.

3 . The system of claim 1 , wherein the first comparison compares dissimilarity between the first subquery vector and individual ones of the set of semantic vectors.

4 . The system of claim 1 , wherein the second subquery is a limitation for the one or more replies.

5 . The system of claim 1 , wherein the second subquery is a filter that constrains the one or more replies.

6 . The system of claim 1 , wherein the second subquery is related to formatting of the one or more replies.

7 . The system of claim 1 , wherein the second subquery is related to modifying the one or more replies prior to presentation to the user.

8 . The system of claim 1 , wherein the individual machine learning model is a large language model.

9 . The system of claim 8 , wherein the large language model is based on or derived from Generative Pre-trained Transformer 3 (GPT3) or a successor of Generative Pre-trained Transformer 3 (GPT3).

10 . A method of using prompt dissection to prompt one or more machine learning models for a set of one or more documents, the method comprising:

electronically storing information, wherein the stored information includes a set of document segments, wherein individual document segments are included in the set of one or more documents, wherein the stored information further includes a set of semantic vectors, wherein individual semantic vectors are associated with individual document segments; and

effectuating a presentation of a user interface that obtains a compound query from a user, wherein the compound query includes two or more subqueries including a first subquery and a second subquery;

creating, using the one or more machine learning models, a first subquery vector representing the first subquery;

determining a subset of the set of semantic vectors, wherein the determination is based on comparisons, wherein the comparisons include a first comparison using the first subquery vector;

creating a combination of the individual document segments that are associated with the subset of the set of semantic vectors as determined;

providing one or more prompts to the one or more machine learning models that prompt the one or more machine learning models to generate one or more replies, using the created combination of the individual document segments as context for the one or more machine learning models, wherein the one or more prompts are based on the compound query, wherein the second subquery constrains the one or more replies, wherein an individual machine learning model from the one or more machine learning models is limited to a predetermined numbers of tokens as the context for the one or more prompts, and wherein the combination of the individual document segments is created such that the predetermined number of tokens is not exceeded; and

presenting to the user, through the user interface, the one or more replies obtained from the one or more machine learning models in reply to the one or more prompts, wherein the one or more replies have been constrained by the second subquery.

11 . The method of claim 10 , wherein the first comparison compares similarity between the first subquery vector and individual ones of the set of semantic vectors.

12 . The method of claim 10 , wherein the first comparison compares dissimilarity between the first subquery vector and individual ones of the set of semantic vectors.

13 . The method of claim 10 , wherein the second subquery is a limitation for the one or more replies.

14 . The method of claim 10 , wherein the second subquery is a filter that constrains the one or more replies.

15 . The method of claim 10 , wherein the second subquery is related to formatting of the one or more replies.

16 . The method of claim 10 , wherein the second subquery is related to modifying the one or more replies prior to presentation to the user.

17 . The method of claim 10 , wherein the individual machine learning model is a large language model.

18 . The method of claim 17 , wherein the large language model is based on or derived from Generative Pre-trained Transformer 3 (GPT3) or a successor of Generative Pre-trained Transformer 3 (GPT3).

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 27, 2024
From: MURTHY, VINEETH CHINMAYA; POWALSKI, RAFAL
To: INSTABASE, INC.
Reel/Frame 069425/0164 →
Continuity (2)
Continuation 18358780 · Jul 25, 2023
Related Publication 20250086216A1 · Mar 13, 2025
References Cited (51)
US 7620976B2 · Low · 2009 [cited by applicant]
US 8881307B2 · Nun · 2014 [cited by applicant]
US 10242212B2 · Tegegne · 2019 [cited by applicant]
US 10614345B1 · Tecuci · 2020 [cited by applicant]
US 10642832B1 · Neumann · 2020 [cited by examiner]
US 12182125B1 · Buniatyan · 2024 [cited by applicant]
US 12405985B1 · Kanagovi · 2025 [cited by applicant]
US 20150317486A1 · Muller · 2015 [cited by applicant]
US 20190340949A1 · Meisner · 2019 [cited by applicant]
US 20200159848A1 · Yeo · 2020 [cited by examiner]
US 20200311349A1 · Balasubramanian · 2020 [cited by examiner]
US 20210034621A1 · Patel · 2021 [cited by examiner]
US 20220164346A1 · Mitra · 2022 [cited by examiner]
US 20230205824A1 · Jablokov · 2023 [cited by applicant]
US 20230315731A1 · Xu · 2023 [cited by examiner]
US 20230385261A1 · Siddiqui · 2023 [cited by examiner]
US 20240096125A1 · Yebes Torres · 2024 [cited by applicant]
US 20240202539A1 · Poirier · 2024 [cited by examiner]
US 20240221007A1 · Hormati · 2024 [cited by examiner]
US 20240256965A1 · Chung · 2024 [cited by applicant]
US 20240289559A1 · Gajek · 2024 [cited by applicant]
US 20240296279A1 · Gardner · 2024 [cited by applicant]
US 20240311407A1 · Barron · 2024 [cited by applicant]
US 20240338361A1 · Hazel · 2024 [cited by applicant]
US 20250045314A1 · Madnani · 2025 [cited by applicant]
US 20250045445A1 · Srinivasan · 2025 [cited by examiner]
US 20250077527A1 · Vaughn · 2025 [cited by applicant]
US 20250086190A1 · Azarmi · 2025 [cited by applicant]
US 20250111167A1 · McIntyre · 2025 [cited by applicant]
US 20250111237A1 · Krishnamurthy · 2025 [cited by applicant]
US 20250117605A1 · De Wynter · 2025 [cited by applicant]
US 20250148020A1 · Padmashali · 2025 [cited by applicant]
US 20250165714A1 · Krabach · 2025 [cited by examiner]
US 20250272507A1 · Shaul · 2025 [cited by applicant]
US 20250348741A1 · Rimchala · 2025 [cited by applicant]
US 20250384206A1 · Baird · 2025 [cited by applicant]
US 20260057181A1 · Murthy · 2026 [cited by applicant]
CN 117951274A · 2024 [cited by applicant]
CN 118332072A · 2024 [cited by applicant]
CN 118656482A · 2024 [cited by applicant]
CN 118939782A · 2024 [cited by applicant]
Li, M., Zhao, Y., Yu, B., Song, F., Li, H., Yu, H., & Li, Y. (2023). Api-bank: A comprehensive benchmark for tool-augmented llms. arXiv preprint arXiv:2304.08244. (15 pages) (Year: 2023). [cited by applicant]
Schick, T., Dwivedi-Yu, J., Dessì, R., Raileanu, R., Lomeli, M., Hambro, E., & Scialom, T. (2023). Toolformer: Language models can teach themselves to use tools. Advances in Neural Information Processing Systems, 36, 68… [cited by applicant]
Qiao, S., Gui, H., Lv, C., Jia, Q., Chen, H., & Zhang, N. (2023). Making language models better tool learners with execution feedback. arXiv preprint arXiv:2305.13068. (10 pgs) (Year: 2023). [cited by applicant]
Paranjape, B., Lundberg, S., Singh, S., Hajishirzi, H., Zettlemoyer, L., & Ribeiro, M. T. (2023). Art: Automatic multi-step reasoning and tool-use for large language models. arXiv preprint arXiv:2303.09014. (26 pgs) (Ye… [cited by applicant]
Liu, Yilun, et al., “What Do You Want? User-Centric Prompt Generation for Text-to-image Synthesis via Multi-turn Guidance”, arXiv, Cornell University repository, document No. arXiv:2408.12910v1 [cs.AI], Aug. 23, 2024, d… [cited by applicant]
Zmigrod, Ran, et al., “What is the value of {templates}?” Rethinking Document Information Extraction Datasets for LLMs, arXiv, Cornell University repository, document No. arXiv:2410.15484v1 [cs.CL], Oct. 20, 2024, downl… [cited by applicant]
Song, Yunpeng, et al., “VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning”, UIST '24, Pittsburgh, PA, Oct. 13-16, 2024, 17 pages. [cited by applicant]
Neumann, Mark, et al., “PAWLS: PDF Annotation With Labels and Structure”, arXiv, Cornell University repository, document No. arXiv:2101.10281v1 [cs.CL], Jan. 25, 2021, downloaded from: https://arxiv.org/abs/2408.12910v1… [cited by applicant]
Cheng, Kanzhi, et al., “SeeClick: Hamessing GUI Grounding for Advanced Visual GUI Agents”, Proc. of the 62nd Annual Meeting of the Association for Computational Linguistics, vol. I: Long Papers, Aug. 11-16, 2024, pp. 93… [cited by applicant]
Almohaimeed, Saleh, et al., “GAT-SQL: An Advanced Prompt Engineering Approach for Effective Text-to-SQL Interactions”, 2024IEEE Congress on Evolutionary Computation (CEC), Yokohama, Japan, Jun. 30, 2024 Jul. 5, 2024, 10… [cited by applicant]