IP Library Granted Patent US 10,510,337
Granted Patent B2
US 10,510,337 · App. 15/949,145 · Granted Dec 17, 2019

Method and device for voice recognition training

Inventors: Michael E. Gunn (Barrington, IL); Boris Bekkerman (Highwood, IL); Mark A. Jasiuk (Chicago, IL); Pratik M. Kamdar (Gurnee, IL); Jeffrey A. Sierawski (Wauconda, IL)
Assignee: Google LLC
G10L15/063G06F3/04842G10L15/20G10L17/04G10L17/20G10L25/84G10L2015/0638H04W88/02
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,510,337
App. No.
15/949,145
Granted
Dec 17, 2019
Kind
B2
Abstract

A method on a mobile device for voice recognition training is described. A voice training mode is entered. A voice training sample for a user of the mobile device is recorded. The voice training mode is interrupted to enter a noise indicator mode based on a sample background noise level for the voice training sample and a sample background noise type for the voice training sample. The voice training mode is returned to from the noise indicator mode when the user provides a continuation input that indicates a current background noise level meets an indicator threshold value.

Claims (62)

1. A method comprising:

executing, by a processor of a mobile device, a first mode of the mobile device, the first mode configured to:

display on a screen in communication with the processor a first graphical user interface including a prompt instructing a user associated with the mobile device to speak a designated phrase for training a voice recognition system of the mobile device, the voice recognition system configured to recognize a voice of the user;

receive a first voice training sample corresponding to the user speaking the designated phrase; and

determine whether a noise level for the received first voice training sample exceeds a predetermined threshold; and

in response to determining that the noise level for the received first voice training sample exceeds the predetermined threshold, executing, by the processor, a second mode of the mobile device, the second mode configured to display on the screen of the mobile device a second graphical user interface comprising:

a notification that recommends an environment conducive to voice training; and

a graphical element,

wherein the second mode is further configured to enable the graphical element of the second graphical user interface for selection by the user when a background noise level does not exceed the predetermined threshold, the enabled graphical element, when selected by the user, causes the processor to transition from executing in the second mode back to executing in the first mode.

2. The method of claim 1 , further comprising, in response to determining that the noise level for the first voice training sample exceeds the predetermined threshold:

ceasing, by the processor, execution of the first mode of the mobile device; and

rejecting, by the processor, the first voice training sample from use in training the voice recognition system.

3. The method of claim 1 , wherein the second mode is further configured to not process any voice samples spoken by the user during execution of the second mode of the mobile device.

4. The method of claim 1 , wherein the first mode is further configured to, when the noise level for the received first voice training sample does not exceed the predetermined threshold, process the first voice training sample for use in training the speech recognition system.

5. The method of claim 4 , wherein the first mode is further configured to display again, in the first graphical user interface, the prompt instructing the user to speak the designated phrase.

6. The method of claim 1 , wherein the designated phrase comprises a trigger phrase including one or more words.

7. A mobile device comprising:

a processor;

a screen in communication with the processor; and

memory hardware in communication with the processor and storing instructions, that when executed by the processor, cause the processor to perform one or more operations comprising:

executing a first mode of the mobile device, the first mode configured to:

display on the screen a first graphical user interface including a prompt instructing a user associated with the mobile device to speak a designated phrase for training a voice recognition system of the mobile device, the voice recognition system configured to recognize a voice of the user;

receive a first voice training sample corresponding to the user speaking the designated phrase; and

determine whether a noise level for the received first voice training sample exceeds a predetermined threshold; and

in response to determining that the noise level for the received first voice training sample exceeds the predetermined threshold, executing a second mode of the mobile device, the second mode configured to display on the screen a second graphical user interface comprising:

a notification that recommends an environment conducive to voice training; and

a graphical element,

wherein the second mode is further configured to enable the graphical element of the second graphical user interface for selection by the user when a background noise level does not exceed the predetermined threshold, the enabled graphical element, when selected by the user, causes the processor to transition from executing in the second mode back to executing in the first mode.

8. The mobile device of claim 7 , wherein the operations further comprise, in response to determining that the noise level for the first voice training sample exceeds the predetermined threshold:

ceasing execution of the first mode of the mobile device; and

rejecting the first voice training sample from use in training the voice recognition system.

9. The mobile device of claim 7 , wherein the second mode is further configured to not process any voice samples spoken by the user during execution of the second mode of the mobile device.

10. The mobile device of claim 7 , wherein the first mode is further configured to, when the noise level for the received first voice training sample does not exceed the predetermined threshold, process the first voice training sample for use in training the speech recognition system.

11. The mobile device of claim 10 , wherein the first mode is further configured to display again, in the first graphical user interface, the prompt instructing the user to speak the designated phrase.

12. The mobile device of claim 7 , wherein the designated phrase comprises a trigger phrase including one or more words.

13. A method of voice recognition training, the method comprising:

receiving, at a processor of a mobile device, a voice training sample corresponding to a user speaking a designated phrase while the mobile device is displaying a first user interface on a screen of the mobile device, the first user interface prompting the user to speak the designated phrase for training voice recognition software configured to recognize a voice of the user;

determining, by the processor, whether a noise level for the received voice training sample exceeds a predetermined threshold; and

in response to determining that the noise level for the received voice training sample exceeds the predetermined threshold, displaying, by the processor, a second user interface on the screen of the mobile device, the second user interface configured to display:

a notification that recommends an environment conducive to voice training; and

a graphical element,

wherein the second mode is further configured to enable the graphical element of the second graphical user interface for selection by the user when a background noise level does not exceed the predetermined threshold, the enabled graphical element, when selected by the user, causes the processor to transition from executing in the second mode back to executing in the first mode.

14. The method of claim 13 , further comprising, in response to determining that the noise level for the received voice training sample exceeds the predetermined threshold, rejecting, by the mobile device, the voice training sample from use in training the voice recognition software.

15. The method of claim 13 , further comprising, when displaying the second user interface, preventing, by the processor, the mobile device from accepting voice samples spoken by the user.

16. The method of claim 13 , further comprising, when the determined noise level for the received voice training sample does not exceed the predetermined threshold, processing, by the processor, the voice training sample for use in training the speech recognition software.

17. The method of claim 13 , wherein the designated phrase comprises a trigger phrase including one or more words.

18. The method of claim 13 , further comprising executing, by the processor, the voice recognition software on the mobile device.

19. A mobile device comprising:

a processor;

a screen in communication with the processor; and

memory hardware in communication with the processor and storing instructions, that when executed by the processor, cause the processor to perform one or more operations comprising:

receiving a voice training sample corresponding to a user speaking a designated phrase while the mobile device is displaying a first user interface on the screen, the first user interface prompting the user to speak the designated phrase for training voice recognition software configured to recognize a voice of the user;

determining whether a noise level for the received voice training sample exceeds a predetermined threshold; and

in response to determining that the noise level for the received voice training sample exceeds the predetermined threshold, displaying a second user interface on the screen, the second user interface configured to display:

a notification that recommends an environment conducive to voice training; and

a graphical element,

wherein the second mode is further configured to enable the graphical element of the second graphical user interface for selection by the user when a background noise level does not exceed the predetermined threshold, the enabled graphical element, when selected by the user, causes the processor to transition from executing in the second mode back to executing in the first mode.

20. The mobile device of claim 19 , wherein the operations further comprise, in response to determining that the noise level for the received voice training sample exceeds the predetermined threshold, rejecting the voice training sample from use in training the voice recognition software.

21. The mobile device of claim 19 , wherein the operations further comprise, when displaying the second user interface, preventing the mobile device from accepting voice samples spoken by the user.

22. The mobile device of claim 19 , wherein the operations further comprise, when the determined noise level for the received voice training sample does not exceed the predetermined threshold, processing the voice training sample for use in training the speech recognition software.

23. The mobile device of claim 19 , wherein the designated phrase comprises a trigger phrase including one or more words.

24. The mobile device of claim 19 , wherein the operations further comprise executing the voice recognition software on the mobile device.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 10, 2018
From: GUNN, MICHAEL E.; BEKKERMAN, BORIS; JASIUK, MARK A.; KAMDAR, PRATIK M.; SIERAWSKI, JEFFREY A.
To: MOTOROLA MOBILITY LLC
Reel/Frame 045488/0732 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 10, 2018
From: MOTOROLA MOBILITY LLC
To: GOOGLE TECHNOLOGY HOLDINGS LLC
Reel/Frame 045880/0230 →
Continuity (5)
Continuation 15466448 · Mar 22, 2017
Continuation 14142210 · Dec 27, 2013
Provisional Application 61892527 · Oct 18, 2013
Provisional Application 61857696 · Jul 23, 2013
Related Publication 20180301142A1 · Oct 18, 2018