IP Library › Granted Patent US 12,609,120
Granted Patent B2
US 12,609,120 · App. 19/057,541 · Granted Apr 21, 2026

Method and system for providing assistance for cognitively impaired users by utilizing artificial intelligence

Inventor: Leigh M. Rothschild (Miami, FL)
Assignee: Ariel Inventions, LLC
G10L15/22G10L25/93
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,609,120
App. No.
19/057,541
Granted
Apr 21, 2026
Kind
B2
Abstract

In an embodiment, the disclosure relates to a device for assisting a respondent in a conversation. The device includes a microphone configured to detect a voice input, and a transmitter communicatively coupled to a server and configured to transmit the voice input to the server. The server is to generate vectors associated with the voice input, feed the vectors associated with the voice input to an Artificial Intelligence utilizing a trained Machine Learning (ML) model, and obtain, from the trained ML model, an output corresponding to the vectors. The device further includes a receiver communicatively coupled to the server, and configured to receive from the server, the output generated by the ML model. A speaker is communicatively coupled with the receiver and is configured to generate a voice-based response based on the output, for assisting the respondent in responding to the conversation.

Claims (62)

1 . A device for assisting a respondent in a conversation between a querier and the respondent, the device comprising:

a controller communicatively coupled to a microphone, wherein the controller is configured to:

fetch a voice input from the microphone;

detect a silent period during a vocal conversation between the querier and the respondent, wherein the silent period has a duration;

compare the duration of the silent period with a threshold time period; and

trigger a transmitter to transmit the voice input to a server upon detecting that the duration of the silent period is greater than the threshold time period, the server being configured to:

generate an output corresponding to the voice input received by the server, wherein the output comprises at least one token as a response to an excerpt from the voice input; and,

transmit the output to a receiver; and

a speaker communicatively coupled with the receiver and configured to generate a voice-based response based on the output, for assisting the respondent in responding to the conversation.

2 . The device of claim 1 , further comprising:

a wireless module communicatively coupled with the receiver and to a mobile device, wherein the wireless module is configured to:

receive, from the receiver, the generated output; and

transmit the output to the mobile device,

wherein the mobile device is configured to generate and display a text-based response based on the output for assisting the respondent in responding to the conversation.

3 . The device of claim 1 , wherein the device is an ear-worn device.

4 . The device of claim 1 , wherein the device is a stationary speaker device.

5 . The device of claim 1 , wherein the output is one of: a text-based output and a voice-based output.

6 . The device of claim 1 , wherein the output comprises at least one of:

a rephrasing of an excerpt from the vocal conversation between the querier and the respondent; and

an answer to a query associated with the vocal conversation between the querier and the respondent.

7 . The device of claim 1 , further comprising:

an imaging device communicatively coupled to the controller, wherein the imaging device is configured to obtain at least one image during the conversation between the querier and the respondent, and wherein the controller is further configured to:

receive the at least one image from the imaging device, and

determine at least one of: an identity of the querier, and an identification of an object captured in the at least one image.

8 . The device of claim 1 , wherein the device is communicatively coupled to a sensor configured to assist in determining gestures of a user present within a predetermined range, and generating commands for a remote device based the determined gestures.

9 . A method of assisting a respondent in a conversation, the method comprising:

receiving, by a controller, a voice input from a microphone, wherein the voice input comprises an excerpt from a vocal conversation between a querier and the respondent;

detecting, by the controller, a silent period during the vocal conversation between the querier and the respondent, wherein the silent period has a duration;

comparing, by the controller, the duration of the silent period with a threshold time period;

triggering, by the controller, a transmitter to transmit the voice input to a server upon detecting that the duration of the silent period is greater than the threshold time period, wherein the server operates to:

generate an output corresponding to the voice input received by the server, wherein the output comprises at least one token as a response to the excerpt from the voice input;

receiving, by the controller, the output from the server via a receiver;

generating, by the controller, a voice-based response based on the output; and

transmitting, by the controller, the voice-based response to a speaker for playing the voice-based response.

10 . The method of claim 9 , further comprising:

generating, by the controller, a text-based response based on the output for assisting the respondent in responding to the conversation; and

transmitting, by the controller, the text-based response to a mobile device, via a wireless module, wherein the mobile device is configured to display the text-based response for assisting the respondent in responding to the conversation.

11 . The method of claim 9 , wherein the output is one of: a text-based output and a voice-based output.

12 . The method of claim 9 , wherein the output comprises at least one of:

a rephrasing of the excerpt from the vocal conversation between the querier and the respondent; and

an answer to a query associated with the vocal conversation between the querier and the respondent.

13 . The method of claim 9 , further comprising:

receiving, by the controller, at least one image from an imaging device, wherein the imaging device operates to obtain the at least one image during the conversation between the querier and the respondent; and

determining, by the controller, at least one of: an identity of the querier and an identification of an object captured in the at least one image.

14 . The method of claim 8 , the voice input is processed by a device communicatively coupled to the controller, wherein the device operates to reduce or exclude ambient noise from the voice input.

15 . A method of generating voice-assistance for a respondent, the method comprising:

receiving, by a controller, a voice input from a microphone, wherein the voice input comprises an excerpt from a speech by the respondent;

detecting, by the controller, a silent period based on the voice input, wherein the silent period has a duration;

comparing, by the controller, the duration of the silent period with a threshold time period;

triggering, by the controller, a transmitter to transmit the voice input to a server upon detecting that the duration of the silent period is greater than the threshold time period, wherein the server operates to generate an output corresponding to the voice input received by the server, wherein the output comprises at least one token as a response to the excerpt from the speech in the voice input;

receiving, by the controller, the output from the server;

generating, by the controller, a voice-based response based on the output; and

transmitting, by the controller, the voice-based response to a speaker for playing the voice-based response.

16 . The method of claim 15 , further comprising:

generating, by the controller, a text-based response based on the output; and

transmitting, by the controller, the text-based response to a mobile device, via a wireless module, wherein the mobile device is configured to display the text-based response for providing voice-assistance to the respondent.

17 . The method of claim 15 , wherein the output is one of: a text-based output and a voice-based output.

18 . The method of claim 15 , wherein the output comprises an answer to a query associated with the speech by the respondent.

19 . The method of claim 15 , further comprising:

receiving, by the controller, at least one image from an imaging device, wherein the imaging device operates to obtain the at least one image during the speech by the respondent; and

determining, by the controller, at least one of: an identity of the querier and an identification of an object captured in the at least one image.

20 . The method of claim 15 , the voice input is processed by a device communicatively coupled to the controller, wherein the device operates to reduce or exclude ambient noise from the voice input.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 30, 2026
From: ARIEL INVENTIONS, LLC
To: HORIZON IP TECHNOLOGIES LLC
Reel/Frame 074528/0873 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 28, 2026
From: ROTHSCHILD, LEIGH M., MR.
To: ARIEL INVENTIONS, LLC
Reel/Frame 074506/0970 →
Continuity (2)
Continuation 18656244 · May 6, 2024
Related Publication 20250342833A1 · Nov 6, 2025
References Cited (8)
US 8537980B2 · Frazier · 2013 [cited by examiner]
US 9491573B2 · Khare · 2016 [cited by examiner]
US 11302320B2 · Deros · 2022 [cited by examiner]
US 20200043479A1 · Mont-Reynaud · 2020 [cited by examiner]
US 20210256046A1 · Newell · 2021 [cited by examiner]
US 20210352560A1 · Cai · 2021 [cited by examiner]
US 20240160298A1 · Nongpiur · 2024 [cited by examiner]
JP 7187212B2 · 2022 [cited by examiner]