IP Library › Granted Patent US 11,687,737
Granted Patent B2
US 11,687,737 · App. 17/932,162 · Granted Jun 27, 2023

Method for multi-channel audio synchronization for task automation

Inventors: Andrei Papancea (New York, NY); Vlad Papancea (New York, NY)
Assignee: NLX Inc.
G06F40/58G10L15/005H04L67/12
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,687,737
App. No.
17/932,162
Granted
Jun 27, 2023
Kind
B2
Abstract

A method for coordinating actions between an audio channel and a synchronized non-audio channel includes receiving an indication of a start of a session associated with a user and having an audio channel that is synchronized with a non-audio channel. Thereafter, repeated determinations are made as to whether a prompt on the non-audio channel has been received from the user. In response to each determination that the prompt on the non-audio channel has not been received from the user, a signal is sent to cause an inaudible output on the audio channel to the user. In response to a determination that the prompt on the non-audio channel has been received from the user, an audible output is selected based on an activity by the user on the non-audio channel, and a signal is sent to cause the audible output to be output on the audio channel.

Claims (69)

1. An apparatus, comprising:

a processor; and

a memory operably coupled to the processor, the memory storing instructions to cause the processor to:

receive an indication of a start of a session associated with a user and having an audio channel that is synchronized with a non-audio channel;

repeatedly determine, after the receiving, whether a prompt on the non-audio channel has been received from the user; and

send a signal to cause an inaudible output on the audio channel to the user in response to each determination that the prompt on the non-audio channel has not been received from the user.

2. The apparatus of claim 1 , wherein the memory further stores instructions to cause the processor to:

in response to a determination that the prompt on the non-audio channel has been received from the user, select an audible output based on an activity by the user on the non-audio channel;

select, at a first time, a first language from a plurality of languages; and

select, at a second time after the first time, a second language from the plurality of languages,

the selecting the audible output being based on the second language.

3. The apparatus of claim 1 , wherein:

the audio channel is associated with a first device type from a plurality of device types,

the non-audio channel is associated with a second device type from the plurality of device types, and

the plurality of device types includes a phone, a smart speaker, an earphone and an Internet of Things (IoT) device.

4. The apparatus of claim 1 , wherein the memory further stores instructions to cause the processor to:

in response to a determination that the prompt on the non-audio channel has been received from the user, select an audible output based on an activity by the user on the non-audio channel, and

receive, via an application programming interface (API), a signal from a device for the non-audio channel,

the selecting the audible output being based on the signal from the device for the non-audio channel.

5. The apparatus of claim 1 , wherein:

during a first time period, the non-audio channel is associated with and the selecting is performed with respect to a first digital non-audio channel, and

during a second time period after the first time period, the non-audio channel is associated with and the selecting is performed with respect to a second digital non-audio channel different from the first digital non-audio channel.

6. The apparatus of claim 1 , wherein the repeatedly determining and the sending the signal to cause the inaudible output being repeated until an end of the session, the method further comprising:

after the start of the session and before the end of the session, performing at least one of:

determine that a prompt on the audio channel received from the user includes an indication that the user would like to discontinue the non-audio channel, or

determine that the prompt on the non-audio channel includes an indication that the user would like to discontinue the non-audio channel;

terminate the non-audio channel of the session, in response to the indication that the user would like to discontinue the non-audio channel; and

send, after the terminating, a signal to connect a communication device of the user with a communication device of a live agent.

7. The apparatus of claim 1 , wherein the non-audio channel is associated with a communication device of the user, the communication device of the user having a plurality of output modes.

8. A non-transitory, processor-readable medium storing instructions to cause a processor to:

initiate a request for a session associated with a user to cause an audio channel associated with the session to synchronize with a non-audio channel associated with the session;

repeatedly determine whether a prompt on the non-audio channel has been received from the user; and

cause an inaudible output on the audio channel to the user in response to each determination that the prompt on the non-audio channel has not been received from the user.

9. The non-transitory, processor-readable medium of claim 8 , wherein the audio channel is configured to ignore audible input from the user during the session.

10. The non-transitory, processor-readable medium of claim 8 , wherein the instructions includes instructions to cause the processor to:

cause an audible output to be output on the audio channel in response to a determination that the prompt on the non-audio channel has been received from the user,

the audible output includes a first portion associated with a first voice and a second portion associated with a second voice different than the first voice.

11. The non-transitory, processor-readable medium of claim 8 , wherein the instructions includes instructions to cause the processor to:

receive an indication from the user to end the session; and

connect to a compute device associated with at least one of a live chat or a live agent.

12. The non-transitory, processor-readable medium of claim 8 , wherein the audio channel is associated with a first compute device, and the non-audio channel is associated with a second compute device different than the first compute device.

13. The non-transitory, processor-readable medium of claim 8 , wherein:

the initiating of the request, the repeatedly determining, and the causing of the inaudible output is performed by a first compute device, and

the initiating of the request includes calling, via the first compute device, a phone number associated with a second compute device to cause the second compute device to generate the session.

14. The non-transitory, processor-readable medium of claim 8 , wherein the initiating of the request, the repeatedly determining, and the causing of the inaudible output are performed by a voice assistant device, the instructions includes instructions to cause the processor to:

receive, by the voice assistant device, a voice command from the user that includes an indication of the request, the initiating of the request performed automatically in response to the receiving of the voice command.

15. A non-transitory, processor-readable medium storing instructions to cause a processor to:

receive a representation of a request from a compute device associated with a user to complete a task;

cause an audio channel associated with the user to synchronize with at least one non-audio channel associated with the user;

send a first signal to cause a first audible output to be output by the audio channel;

repeatedly determine whether a prompt on the at least one non-audio channel has been received from the user;

send a second signal to cause an inaudible output on the audio channel to the user in response to each determination that the prompt on the at least one non-audio channel has not been received from the user; and

in response to a determination that the prompt on the at least one non-audio channel has been received from the user:

select a second audible output based on a determination that the prompt is in accordance with the task,

select a third audible output based on a determination that the prompt is not in accordance with the task, and

send a third signal to cause one of the second audible output or the third audible output to be output on the audio channel.

16. The non-transitory, processor-readable medium of claim 15 , wherein the prompt is a first prompt, the instructions further includes instructions to cause the processor to:

repeatedly determine whether a second prompt on the at least one non-audio channel has been received from the user;

send a fourth signal to cause the inaudible output on the audio channel to the user in response to each determination that the second prompt on the at least one non-audio channel has not been received from the user; and

in response to the determination that the second prompt on the at least one non-audio channel has been received from the user:

select a fourth audible output based on an activity by the user on the at least one non-audio channel, and

send a fourth signal to cause the fourth audible output to be output on the audio channel.

17. The non-transitory, processor-readable medium of claim 15 , wherein the compute device is a mobile device, the instructions further including instructions to cause the processor to:

transmit a hyperlink to the mobile device via at least one of a text message or an email,

the causing of the audio channel associated with the user to synchronize with the at least one non-audio channel associated with the user performed automatically in response to the user selecting the hyperlink.

18. The non-transitory, processor-readable medium of claim 15 , wherein the audio channel is associated with a first device type from a plurality of device types and the at least one non-audio channel is associated with a second device type from the plurality of device types, the plurality of device types includes a phone, a smart speaker, a speaker, an earphone and an Internet of Things (IoT) device.

19. The non-transitory, processor-readable medium of claim 15 , wherein at least one of the first audible output, the second audible output, or the third audible output include a first portion output in a first language during a first time, and a second portion output in a second language different than the first language during a second time after the first time.

20. The non-transitory, processor-readable medium of claim 15 , wherein the compute device is a first compute device, the instructions further including instructions to cause the processor to:

connect to a second compute device associated with at least one of a live chat or a live agent in response to an indication from the user to connect with at least one of the live chat or the live agent.

Assignments (3)
CHANGE OF NAME Recorded May 18, 2026
From: NLX INC.
To: NLX LLC
Reel/Frame 075582/0229 →
BILL OF SALE Recorded May 6, 2026
From: NLX LLC
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 075567/0710 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 9, 2022
From: PAPANCEA, ANDREI; PAPANCEA, VLAD
To: NLX INC.
Reel/Frame 062044/0241 →
Continuity (3)
Continuation 17532662 · Nov 22, 2021
Provisional Application 63116952 · Nov 23, 2020
Related Publication 20230004731A1 · Jan 5, 2023
Cited By (1)
US 12,190,416