IP Library Granted Patent US 9,117,212
Granted Patent B2
US 9,117,212 · App. 14/167,793 · Granted Aug 25, 2015

System and method for authentication using speaker verification techniques and fraud model

Inventors: John F. Sheets (San Francisco, CA); Kim R. Wagner (Sunnyvale, CA)
Assignee: Visa International Service Association
G06Q20/4014G06Q20/3821G06Q20/4016G06Q20/40145G10L17/00G10L17/04G10L17/24G10L17/22
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,117,212
App. No.
14/167,793
Granted
Aug 25, 2015
Kind
B2
Abstract

Embodiments of the invention provide for speaker verification on a communication device without requiring a user to go through a formal registration process with the issuer or network. Certain embodiments allow the use of a captured voice sample attempting to reproduce a word string having a random element to authenticate the user. Authentication of the user is based on both a match score indicating how closely the captured voice samples match to previously stored voice samples of the user and a pass or fail response indicating whether the voice sample is an accurate reproduction of the word string. The processing network maintains a history of the authenticated transactions and voice samples.

Claims (28)

1. A method for authenticating a user for a transaction, comprising:

providing, by a device, a word string that comprises a random element;

transmitting an audio segment, to a server computer, wherein the audio segment originated from the user and wherein the server computer authenticates the user for the transaction based at least in part on the transmitted audio segment; and

receiving, from the server computer, an indication that the user is authenticated for the transaction, wherein the server computer holds the audio segment in a queue for a predetermined period of time, and delays updating of a fraud model with the audio segment being held in the queue until after the predetermined period of time has elapsed and when no fraud has been reported for the predetermined period of time.

2. The method of claim 1 further comprising receiving a match score indicative of a comparison between the transmitted audio segment and the fraud model.

3. The method of claim 1 wherein the audio segment is a reproduction, by the user, of the provided word string.

4. The method of claim 1 wherein the indication is indicative of whether the received audio segment is an accurate reproduction of the provided word string.

5. The method of claim 1 wherein the word string is seven words or less.

6. The method of claim 1 wherein the server computer is a voice biometric matching server.

7. The method of claim 1 wherein the fraud model is based on a plurality of previously transmitted audio segments.

8. The method of claim 1 further comprising displaying, by the device, the word string to the user.

9. The method of claim 1 wherein the random element is preceded by a fixed element of the word string.

10. The method of claim 9 wherein the fixed element is greater in length than the random element.

11. A device, comprising:

a processor; and

a non-transitory computer-readable storage medium, comprising code executable by the processor for implementing a method for authenticating a user for a transaction, the method comprising:

providing, by the device, a word string that comprises a random element;

transmitting an audio segment, to a server computer, wherein the audio segment originated from the user and wherein the server computer authenticates the user for the transaction based at least in part on the transmitted audio segment; and

receiving, from the server computer, an indication that the user is authenticated for the transaction, wherein the server computer holds the audio segment in a queue for a predetermined period of time, and delays updating of a fraud model with the audio segment being held in the queue until after the predetermined period of time has elapsed and when no fraud has been reported for the predetermined period of time.

12. The device of claim 11 wherein the method further comprises receiving a match score indicative of a comparison between the transmitted audio segment and the fraud model.

13. The device of claim 12 wherein the audio segment is a reproduction, by the user, of the provided word string.

14. The device of claim 11 wherein the indication is indicative of whether the received audio segment is an accurate reproduction of the provided word string.

15. The device of claim 11 wherein the word string is seven words or less.

16. The device of claim 11 wherein the server computer is a voice biometric matching server.

17. The device of claim 11 wherein the fraud model is based on a plurality of previously transmitted audio segments.

18. The device of claim 11 wherein the device further comprises a display configured to display the word string to the user.

19. The device of claim 11 wherein the random element is preceded by a fixed element of the word string.

20. The device of claim 19 wherein the fixed element is greater in length than the random element.

Continuity (3)
Continuation 13899470 · May 21, 2013
Provisional Application 61761155 · Feb 5, 2013
Related Publication 20140222678A1 · Aug 7, 2014