IP Library Granted Patent US 12,602,539
Granted Patent B2
US 12,602,539 · App. 18/390,675 · Granted Apr 14, 2026

Proactive assistance via a cascade of LLMS

Inventors: Victor Carbune (Zürich, CH); Matthew Sharifi (Kilchberg, CH)
Assignee: Google LLC
G06F40/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,602,539
App. No.
18/390,675
Granted
Apr 14, 2026
Kind
B2
Abstract

A method for providing proactive assistance includes obtaining, by a digital assistant, a contextual event associated with a user of a user device. The method includes determining, using a local large language model (LLM) executing on the user device, a remote LLM prompt confidence. The method includes determining that the remote LLM prompt confidence satisfies a threshold. Based on determining that the remote LLM prompt confidence satisfies the threshold, the method includes generating a remote LLM prompt for a remote LLM executing remote from the user device. The method includes transmitting, to the remote LLM, the remote LLM prompt. The method includes receiving, at the digital assistant, from the remote LLM, response content providing the proactive assistance associated with the contextual event. The method includes providing, for output from the user device, presentation content based on the response content received from the remote LLM.

Claims (60)

1 . A computer-implemented method executed by data processing hardware that causes the data processing hardware to perform operations comprising:

obtaining, by a digital assistant, a contextual event associated with a user of a user device;

generating, using a local large language model (LLM) executing on the user device, initial response content associated with the contextual event, the initial response content providing an offer for the digital assistant to interact with a remote LLM to perform an action on the user's behalf based on the contextual event;

providing, for output from the user device, initial presentation content based on the initial response content, the initial presentation content prompting the user to consent to the offer for the digital assistant to interact with the remote LLM to perform the action on the user's behalf;

receiving an initial presentation content interaction indicating user interaction with the initial presentation content:

determining, using the local LLM, a remote LLM prompt confidence based on the received initial presentation content, the remote LLM prompt confidence indicating a likelihood of prompting a remote LLM for proactive assistance associated with the contextual event;

determining that the remote LLM prompt confidence satisfies a threshold;

based on determining that the remote LLM prompt confidence satisfies the threshold, generating a remote LLM prompt for the remote LLM executing remote from the user device;

transmitting, to the remote LLM, the remote LLM prompt;

receiving, at the digital assistant, from the remote LLM, response content providing the proactive assistance associated with the contextual event; and

providing, for output from the user device, presentation content based on the response content received from the remote LLM.

2 . The method of claim 1 , wherein the remote LLM prompt confidence comprises a probability generated by the local LLM.

3 . The method of claim 1 , wherein

the remote LLM prompt confidence comprises the initial presentation content interaction.

4 . The method of claim 1 , wherein the initial presentation content interaction comprises user consent for transmitting the remote LLM prompt to the remote LLM.

5 . The method of claim 1 , wherein the remote LLM prompt is based on output from the local LLM.

6 . The method of claim 5 , wherein the output comprises a summary of the contextual event.

7 . The method of claim 5 , wherein:

the contextual event comprises personal identification information associated with the user; and

generating the remote LLM prompt comprises redacting, using the output from the local LLM, the personal identification information.

8 . The method of claim 1 , wherein the contextual event comprises at least one of:

sensor data captured by a sensor of the user device; or

application-specific data generated by another application executing on the user device.

9 . The method of claim 1 , wherein:

the contextual event comprises non-textual data; and

the operations further comprise transforming the non-textual data into textual data.

10 . The method of claim 1 , wherein:

the operations further comprise batching a plurality of contextual events together; and

determining the remote LLM prompt confidence is further based on the batched plurality of contextual events.

11 . A system comprising:

data processing hardware; and

memory hardware in communication with the data processing hardware, the memory hardware storing instructions that when executed on the data processing hardware cause the data processing hardware to perform operations comprising:

obtaining, by a digital assistant, a contextual event associated with a user of a user device;

generating, using a local large language model (LLM) executing on the user device, initial response content associated with the contextual event, the initial response content providing an offer for the digital assistant to interact with a remote LLM to perform an action on the user's behalf based on the contextual event;

providing, for output from the user device, initial presentation content based on the initial response content, the initial presentation content prompting the user to consent to the offer for the digital assistant to interact with the remote LLM to perform the action on the user's behalf;

receiving an initial presentation content interaction indicating user interaction with the initial presentation content;

determining, using the local LLM, a remote LLM prompt confidence based on the received initial presentation content, the remote LLM prompt confidence indicating a likelihood of prompting a remote LLM for proactive assistance associated with the contextual event;

determining that the remote LLM prompt confidence satisfies a threshold;

based on determining that the remote LLM prompt confidence satisfies the threshold, generating a remote LLM prompt for the remote LLM executing remote from the user device;

transmitting, to the remote LLM, the remote LLM prompt;

receiving, at the digital assistant, from the remote LLM, response content providing the proactive assistance associated with the contextual event; and

providing, for output from the user device, presentation content based on the response content received from the remote LLM.

12 . The system of claim 11 , wherein the remote LLM prompt confidence comprises a probability generated by the local LLM.

13 . The system of claim 11 , wherein

the remote LLM prompt confidence comprises the initial presentation content interaction.

14 . The system of claim 11 , wherein the initial presentation content interaction comprises user consent for transmitting the remote LLM prompt to the remote LLM.

15 . The system of claim 11 , wherein the remote LLM prompt is based on output from the local LLM.

16 . The system of claim 15 , wherein the output comprises a summary of the contextual event.

17 . The system of claim 15 , wherein:

the contextual event comprises personal identification information associated with the user; and

generating the remote LLM prompt comprises redacting, using the output from the local LLM, the personal identification information.

18 . The system of claim 11 , wherein the contextual event comprises at least one of:

sensor data captured by a sensor of the user device; or

application-specific data generated by another application executing on the user device.

19 . The system of claim 11 , wherein:

the contextual event comprises non-textual data; and

the operations further comprise transforming the non-textual data into textual data.

20 . The system of claim 11 , wherein:

the operations further comprise batching a plurality of contextual events together; and

determining the remote LLM prompt confidence is further based on the batched plurality of contextual events.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2023
From: CARBUNE, VICTOR; SHARIFI, MATTHEW
To: GOOGLE LLC
Reel/Frame 065921/0777 →
Continuity (1)
Related Publication 20250209261A1 · Jun 26, 2025
References Cited (10)
US 11195534B1 · Shen · 2021 [cited by examiner]
US 20230037085A1 · Biadsy et al. · 2023 [cited by applicant]
US 20240126794A1 · Cook · 2024 [cited by examiner]
US 20240362468A1 · Lee · 2024 [cited by examiner]
US 20250077237A1 · Sachindran · 2025 [cited by examiner]
US 20250150474A1 · Jones · 2025 [cited by examiner]
US 20250165752A1 · Lovric · 2025 [cited by examiner]
US 20250190872A1 · Kissane · 2025 [cited by examiner]
Yun Zhu et al: “Towards an On-device Agent for Text Rewriting”, arxiv.org, Cornell University Library, 201 Olin Library Cornell University Ithaca, NY 14853, Aug. 22, 2023 (Aug. 22, 2023), XP091598325. [cited by applicant]
International Search Report and Written Opinion issued in related PCT Application No. PCT/US2024/053134, dated Jan. 29, 2025. [cited by applicant]