IP Library Granted Patent US 10,855,676
Granted Patent B2
US 10,855,676 · App. 16/398,940 · Granted Dec 1, 2020

Audio verification

Inventors: Manjana Chandrasekharan (Los Angeles, CA); Keiko Horiguchi (Palo Alto, CA); Amanda Joy Stent (Chatham, NJ); Ricardo Alberto Baeza-Yates (Palo Alto, CA); Jeffrey Kuwano (San Jose, CA); Achint Oommen Thomas (Milpitas, CA); Yi Chang (Milpitas, CA)
Assignee: Oath Inc.
H04L63/083G06F3/165G06F3/167G06F21/31G10L17/06G10L21/003G10L25/51H04L9/3226G06F2221/2133
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,855,676
App. No.
16/398,940
Granted
Dec 1, 2020
Kind
B2
Abstract

One or more techniques and/or systems are provided for audio verification. An audio signal, comprising a code for user verification, may be identified. A second audio signal is created comprising speech. The audio signal and the second audio signal may be altered to comprise a same or similar volume, pitch, amplitude, and/or speech rate. The audio signal and the second audio signal may be combined to generate a verification audio signal. The verification audio signal may be presented to a user for the user verification. Verification may be performed to determine whether the user has access to content or a service based upon user input, obtained in response to the user verification audio signal, matching the code within the user verification audio signal. In an example, the user verification may comprise verifying that the user is human.

Claims (96)

1. A system of audio verification comprising:

a processor; and

memory comprising processor-executable instructions that when executed by the processor cause implementation of an audio generation component configured to:

identify an audio signal comprising a code for user verification;

create a second audio signal based upon one or more audio segments;

identify a pitch and a volume of the audio signal at a first time;

identify a second pitch and a second volume of the second audio signal at a second time;

determine an average pitch between the pitch of the audio signal and the second pitch of the second audio signal;

determine an average volume between the volume of the audio signal and the second volume of the second audio signal;

alter the pitch of the audio signal and the second pitch of the second audio signal to be the average pitch at a third time;

alter the volume of the audio signal and the second volume of the second audio signal to be the average volume at a fourth time;

combine the audio signal and the second audio signal to generate a verification audio signal in response to:

determining that the pitch of the audio signal at the third time and the second pitch of the second audio signal at the third time are both the average pitch so that a bot is unable to discern a difference between the pitch of the audio signal and the second pitch of the second audio signal when the audio signal and the second audio signal are combined; and

determining that the volume of the audio signal at the fourth time and the second volume of the second audio signal at the fourth time are both the average volume so that a bot is unable to discern a difference between the volume of the audio signal and the second volume of the second audio signal when the audio signal and the second audio signal are combined;

present the verification audio signal to a user for the user verification, the user verification comprising verifying that the user is human; and

verify whether the user has access to content or a service based upon user input, obtained in response to the verification audio signal, matching the code within the verification audio signal.

2. The system of claim 1 , the audio generation component configured to:

identify a speaking rate of the audio signal; and

identify a second speaking rate of the second audio signal.

3. The system of claim 2 , the audio generation component configured to:

alter at least one of the speaking rate of the audio signal or the second speaking rate of the second audio signal until the speaking rate and the second speaking rate are within a threshold speaking rate similarity.

4. The system of claim 1 , the audio generation component configured to:

identify an amplitude of the audio signal; and

identify a second amplitude of the second audio signal.

5. The system of claim 4 , the audio generation component configured to:

alter at least one of the amplitude of the audio signal or the second amplitude of the second audio signal until the amplitude and the second amplitude are within a threshold amplitude similarity.

6. The system of claim 1 , the audio generation component configured to:

create the second audio signal utilizing a first audio segment and a second audio segment.

7. The system of claim 6 , the audio generation component configured to at least one of:

extract at least one of the first audio segment or the second audio segment from an audio content database; or

generate at least one of the first audio segment or the second audio segment utilizing a random speech generator.

8. The system of claim 6 , the audio generation component configured to:

randomly extract one or more portions from at least one of the first audio segment or the second audio segment; and

randomly stitch the one or more portions together to create the second audio signal.

9. The system of claim 6 , the audio generation component configured to:

randomly extract one or more portions from at least one of the first audio segment or the second audio segment;

randomly layer the one or more portions over each other to create a layered segment and a second layered segment; and

stitch the layered segment and the second layered segment together to create the second audio signal.

10. The system of claim 6 , the audio generation component configured to:

randomly extract one or more portions from at least one of the first audio segment or the second audio segment;

randomly stitch the one or more portions together to create an initial second audio signal; and

reverse the initial second audio signal to create the second audio signal.

11. The system of claim 1 , wherein the second audio signal comprises computer generated speech.

12. The system of claim 1 , the audio generation component configured to:

provide the user with an option to enter the user input audibly;

responsive to the user entering the user input audibly, identify acoustic features that are indicative of a human voice; and

responsive to the acoustic features indicating that the user input was spoken by the human voice, verify the user access to the content or the service.

13. The system of claim 1 , the audio generation component configured to:

provide the user an option to enter the user input audibly;

responsive to the user entering the user input audibly, identify acoustic features that are indicative of a human voice; and

responsive to the acoustic features indicating the user input was not spoken by a human voice, deny the user access to the content or the service.

14. A method of audio verification comprising:

identifying an audio signal comprising a code for user verification;

creating a second audio signal based upon one or more audio segments;

identifying a speaking rate and an amplitude of the audio signal at a first time;

identifying a second speaking rate and a second amplitude of the second audio signal at a second time;

altering the speaking rate of the audio signal and the second speaking rate of the second audio signal by altering the speaking rate be more similar to the second speaking rate at the second time and altering the second speaking rate to be more similar to the speaking rate at the first time;

altering the amplitude of the audio signal and the second amplitude of the second audio signal by altering the amplitude be more similar to the second amplitude at the second time and altering the second amplitude to be more similar to the amplitude at the first time;

combining the audio signal and the second audio signal to generate a verification audio signal in response to:

determining that a bot is unable to discern a difference between the speaking rate of the audio signal and the second speaking rate of the second audio signal when the audio signal and the second audio signal are combined; and

determining that a bot is unable to discern a difference between the amplitude of the audio signal and the second amplitude of the second audio signal when the audio signal and the second audio signal are combined;

presenting the verification audio signal to a user for the user verification, the user verification comprising verifying that the user is human; and

verifying whether the user has access to content or a service based upon user input, obtained in response to the verification audio signal, matching the code within the verification audio signal.

15. The method of claim 14 , comprising:

identifying a pitch of the audio signal;

identifying a second pitch of the second audio signal; and

altering at least one of the pitch of the audio signal or the second pitch of the second audio signal until the pitch and the second pitch are within a threshold pitch similarity.

16. The method of claim 14 , comprising:

identifying a volume of the audio signal;

identifying a second volume of the second audio signal; and

altering at least one of the volume of the audio signal or the second volume of the second audio signal until the volume and the second volume are within a threshold volume similarity.

17. The method of claim 14 , comprising:

creating the second audio signal utilizing a first audio segment and a second audio segment.

18. The method of claim 17 , comprising at least one of:

extracting at least one of the first audio segment or the second audio segment from an audio content database; or

generating at least one of the first audio segment or the second audio segment utilizing a random speech generator.

19. A system of audio verification comprising:

a processor; and

memory comprising processor-executable instructions that when executed by the processor cause implementation of an audio generation component configured to:

identify an audio signal comprising a code for user verification;

create a second audio signal based upon one or more audio segments;

identify at least two of a pitch, an amplitude, a volume or a speaking rate of the audio signal;

identify at least two of a second pitch, a second amplitude, a second volume or a second speaking rate of the second audio signal;

at least two of:

alter at least one of the pitch of the audio signal or the second pitch of the second audio signal;

alter at least one of the volume of the audio signal or the second volume of the second audio signal;

alter at least one of the amplitude of the audio signal or the second amplitude of the second audio signal; or

alter at least one of the speaking rate of the audio signal or the second speaking rate of the second audio signal;

combine the audio signal and the second audio signal to generate a verification audio signal in response to at least two of:

determining that a bot is unable to discern a difference between the pitch of the audio signal and the second pitch of the second audio signal when the audio signal and the second audio signal are combined;

determining that a bot is unable to discern a difference between the volume of the audio signal and the second volume of the second audio signal when the audio signal and the second audio signal are combined;

determining that a bot is unable to discern a difference between the amplitude of the audio signal and the second amplitude of the second audio signal when the audio signal and the second audio signal are combined; or

determining that a bot is unable to discern a difference between the speaking rate of the audio signal and the second speaking rate of the second audio signal when the audio signal and the second audio signal are combined;

present the verification audio signal to a user for the user verification, the user verification comprising verifying that the user is human; and

verify whether the user has access to content or a service based upon user input, obtained in response to the verification audio signal, matching the code within the verification audio signal.

20. The system of claim 19 , the one or more audio segments extracted in real-time from an on-going audio stream corresponding to at least one of a news show, a radio show or a talk show.

Assignments (3)
PATENT SECURITY AGREEMENT (FIRST LIEN) Recorded Sep 29, 2022
From: YAHOO ASSETS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 061571/0773 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 16, 2021
From: YAHOO AD TECH LLC (FORMERLY VERIZON MEDIA INC.)
To: YAHOO ASSETS LLC
Reel/Frame 058982/0282 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 26, 2020
From: OATH INC.
To: VERIZON MEDIA INC.
Reel/Frame 054258/0635 →