IP Library Granted Patent US 11,804,231
Granted Patent B2
US 11,804,231 · App. 17/305,270 · Granted Oct 31, 2023

Information exchange on mobile devices using audio

Inventor: Ian Fitzgerald (Arlington Heights, IL)
Assignee: Capital One Services, LLC
G10L19/167G06F3/0482G06F3/165G10L19/0204G10L19/0216G10L19/038G10L25/18
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,804,231
App. No.
17/305,270
Granted
Oct 31, 2023
Kind
B2
Abstract

In some implementations, a user device may receive input that triggers transmission of information via sound. The user device may select an audio clip based on a setting associated with the device, and may modify a digital representation of the selected audio clip using an encoding algorithm and based on data associated with a user of the device. The user device may transmit, to a remote server, an indication of the selected audio clip, an indication of the encoding algorithm, and the data associated with the user. The user device may use a speaker to play audio, based on the modified digital representation, for recording by other devices. Accordingly, the user device may receive, from the remote server and based on the speaker playing the audio, a confirmation that users associated with the other devices have performed an action based on the data associated with the user of the device.

Claims (69)

1. A system for conveying information to one or more other devices using audio, the system comprising:

one or more memories; and

one or more processors, communicatively coupled to the one or more memories, configured to:

receive at least a first portion of information and a second portion of information for transmission to the one or more other devices, wherein the first portion of information includes an identifier associated with a user of the system;

select an audio clip based on a preference associated with the user;

modify a digital representation of the audio clip to encode the first portion of information and the second portion of information within the digital representation, wherein the digital representation includes a spectrogram associated with the audio clip;

transmit, to a remote server, a copy of the modified digital representation, a copy of the first portion of information and the second portion of information, and an indicator of an association between the modified digital representation and the first portion of information and the second portion of information;

use at least one speaker associated with the system to play audio, based on the modified digital representation, for recording by the one or more other devices; and

receive, from the remote server and based on the at least one speaker playing the audio, a confirmation that users associated with the one or more other devices have performed an action based on the first portion of information and the second portion of information.

2. The system of claim 1 , wherein the one or more processors are further configured to:

receive, from the remote server, an indication of a plurality of audio clips;

display, to the user, visual indicators associated with the plurality of audio clips;

receive, from the user, a selection of the audio clip from the plurality of audio clips;

store the selection as the preference associated with the user; and

transmit, to the remote server, an indication of the selected audio clip.

3. The system of claim 1 , wherein the one or more processors are further configured to:

receive, from the user, at least one file encoding the selected audio clip;

store, as the preference, an indicator associated with the at least one file; and

transmit, to the remote server, an indication of the selected audio clip.

4. The system of claim 1 , wherein the one or more processors, to select the audio clip, are configured to:

transmit, to the remote server, a request for the preference associated with the user, wherein the request includes at least one credential associated with the user; and

receive, from the remote server and based on the request, an indication of the preference associated with the user.

5. The system of claim 1 , wherein the one or more processors, to modify the digital representation, are configured to:

apply a wavelet transform to obtain the spectrogram associated with the audio clip;

select one or more subbands within the spectrogram that satisfy a threshold; and

embed a vector encoding the first portion of information and the second portion of information within the one or more subbands.

6. The system of claim 1 , wherein the action includes a transaction that is associated with the user of the system and the users of the one or more other devices and that is based at least in part on the second portion of information.

7. A method of conveying information to one or more other devices using audio, comprising:

receiving, at a device, input that triggers transmission of information via sound;

selecting an audio clip from a plurality of stored audio clips, based on a setting associated with the device;

modifying a digital representation of the selected audio clip using at least one encoding algorithm and based on data associated with a user of the device;

transmitting, to a remote server, an indication of the selected audio clip, an indication of the at least one encoding algorithm, and the data associated with the user of the device;

using at least one speaker associated with the device to play audio, based on the modified digital representation, for recording by the one or more other devices; and

receiving, from the remote server and based on the at least one speaker playing the audio, a confirmation that users associated with the one or more other devices have performed an action based on the data associated with the user of the device.

8. The method of claim 7 , wherein the data associated with the user of the device is based, at least in part, on the input that triggers transmission.

9. The method of claim 7 , further comprising:

displaying visual indicators associated with the plurality of stored audio clips;

receiving, at the device, a selection of the audio clip from the plurality of stored audio clips; and

storing, on the device, the selection as the setting associated with the device.

10. The method of claim 7 , wherein the at least one encoding algorithm comprises:

a tone insertion algorithm associated with inserting a tone within the selected audio clip, wherein the tone encodes the data associated with the user of the device, and wherein a frequency range associated with the tone is not within a range of frequencies associated with human hearing;

a phase coding algorithm associated with shifting one or more initial phases corresponding to one or more segments of the selected audio clip, wherein one or more additional phases included in each segment are shifted to maintain a relative difference with the initial phase for the segment; or

a discrete wavelet transform associated with embedding one or more vectors, that encode the data associated with the user of the device, within one or more subbands associated with the selected audio clip.

11. The method of claim 7 , wherein using the at least one speaker comprises:

outputting, to a driver, the modified digital representation of the selected audio clip,

wherein the modified digital representation is fed to a digital-to-analog converter by the driver.

12. The method of claim 7 , wherein the data associated with the user of the device includes an identifier associated with the user and an amount associated with a request from the user.

13. The method of claim 7 , wherein the indication of the selected audio clip and at least a portion of the data associated with the user of the device are transmitted earlier than the indication of the at least one encoding algorithm and a remainder of the data associated with the user of the device.

14. The method of claim 7 , further comprising:

displaying one or more visual indicators associated with the confirmation,

wherein the confirmation indicates one or more identifiers associated with one or more users of the one or more other devices.

15. A non-transitory computer-readable medium storing a set of instructions for conveying information to at least one other device using audio, the set of instructions comprising:

one or more instructions that, when executed by one or more processors of a device, cause the device to:

receive data associated with a user of the device;

modify a digital representation of an audio clip to encode the data within the digital representation;

transmit, to a remote server, a copy of the modified digital representation, a copy of the data, and an indicator of an association between the modified digital representation and the data;

use at least one speaker associated with the device to play audio, based on the modified digital representation, for recording by the at least one other device; and

receive, from the remote server and based on the at least one speaker playing the audio, a confirmation that at least one user associated with the at least one other device has performed an action based on the data.

16. The non-transitory computer-readable medium of claim 15 , wherein the digital representation includes at least one of a spectrogram, a phase matrix, or a discretized signal.

17. The non-transitory computer-readable medium of claim 15 , wherein the one or more instructions, when executed by the one or more processors, further cause the device to:

authenticate the device with the remote server before transmitting to the remote server.

18. The non-transitory computer-readable medium of claim 15 , wherein the one or more instructions, when executed by the one or more processors, further cause the device to:

receive an instruction to halt playing of the audio;

use the at least one speaker to halt playing of the audio based on the instruction; and

transmit, to the remote server, an indication to remove the association between the modified digital representation and the data.

19. The non-transitory computer-readable medium of claim 15 , wherein the one or more instructions, that cause the device to modify the digital representation, cause the device to:

generate a plurality of bits, wherein a quantity of the plurality of bits is selected based at least in part on a setting associated with the remote server; and

modify the digital representation to encode the plurality of bits, wherein the plurality of bits are used to identify the data associated with the user.

20. The non-transitory computer-readable medium of claim 15 , wherein the data associated with the user of the device includes an identifier associated with the user and an amount associated with a request from the user.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 2, 2021
From: FITZGERALD, IAN
To: CAPITAL ONE SERVICES, LLC
Reel/Frame 056744/0774 →
Continuity (1)
Related Publication 20230005491A1 · Jan 5, 2023
Cited By (1)
US 12,469,508