IP Library › Granted Patent US 11,595,514
Granted Patent B2
US 11,595,514 · App. 17/118,387 · Granted Feb 28, 2023

Handling calls on a shared speech-enabled device

Inventors: Vinh Quoc Ly (Sunnyvale, CA); Raunaq Shah (San Francisco, CA); Okan Kolak (Sunnyvale, CA); Deniz Binay (San Fransisco, CA); Tianyu Wang (Los Altos, CA)
Assignee: GOOGLE LLC
H04M3/42008G06F3/167G10L15/1822G10L15/22G10L15/30G10L17/00H04L61/4594H04L65/1096H04M3/42059G10L2015/223G10L2015/225G10L2015/227
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,595,514
App. No.
17/118,387
Granted
Feb 28, 2023
Kind
B2
Abstract

In some implementations, a determination that a first party has spoken a query for a voice-enabled virtual assistant during a voice call between the first party and a second party is made, in response to the determination that the first party has spoken the query for the voice-enabled virtual assistant during the voice call between the first party and the second party, the voice call between the first party and the second party is placed on hold, a determination that the voice-enabled virtual assistant has resolved the query is made, and in response to the determination that the voice-enabled virtual assistant has handled the query, the voice call between the first party and the second party is resumed from hold.

Claims (80)

1. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving, by a speech-enabled device, an utterance of a first user that requests a voice call with a second user;

determining, by the speech-enabled device and before the voice call is initiated, whether the second user corresponds to a particular type of user;

when the second user corresponds to the particular type of user:

determining, by the speech-enabled device, a recipient voice call number for the second user;

requesting, by the speech-enabled device and from a server, a temporary voice call number for the speech-enabled device; and

initiating, by the speech-enabled device, the voice call to the recipient voice call number using the temporary voice call number for the speech-enabled device; and

when the second user does not correspond to the particular type of user:

classifying, by the speech-enabled device and before the voice call is initiated, the utterance as spoken by a particular known user;

determining, by the speech-enabled device, the recipient voice call number for the second user;

determining, by the speech-enabled device and before the voice call is initiated, whether a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance; and

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance, initiating, by the speech-enabled device, the voice call to the recipient voice call number with the voice number used to place voice calls as the particular known user instead of with another voice number.

2. The system of claim 1 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether speech in the utterance matches speech corresponding to the particular known user.

3. The system of claim 1 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether a visual image of at least a portion of a speaker of the utterance matches visual information corresponding to the particular known user.

4. The system of claim 1 , wherein determining, by the speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance comprises:

accessing a user account profile of the particular known user;

determining that the user account profile indicates a voice call device of the particular known user; and

determining that voice number used to place voice calls as the particular known user is connected with the speech-enabled device through a local wireless connection.

5. The system of claim 4 , wherein initiating the voice call comprises:

initiating the voice call through the phone connected with the speech-enabled device.

6. The system of claim 1 , wherein initiating, by the speech-enabled device, the voice call to the recipient voice call number with the voice number used to place voice calls as the particular known user comprises:

initiating the voice call through a Voice over Internet Protocol call provider.

7. The system of claim 1 , the operations further comprising:

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is not known for the particular known user classified as having spoken the utterance:

initiating, by the speech-enabled device, the voice call to the recipient voice call number with an anonymous voice call number used to place voice calls anonymously.

8. A method comprising:

receiving, by a speech-enabled device, an utterance of a first user that requests a voice call with a second user;

determining, by the speech-enabled device and before the voice call is initiated, whether the second user corresponds to a particular type of user;

when the second user corresponds to the particular type of user:

determining, by the speech-enabled device, a recipient voice call number for the second user;

requesting, by the speech-enabled device and from a server, a temporary voice call number for the speech-enabled device; and

initiating, by the speech-enabled device, the voice call to the recipient voice call number using the temporary voice call number for the speech-enabled device; and

when the second user does not correspond to the particular type of user:

classifying, by the speech-enabled device and before the voice call is initiated, the utterance as spoken by a particular known user;

determining, by the speech-enabled device, the recipient voice call number for the second user;

determining, by the speech-enabled device and before the voice call is initiated, whether a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance; and

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance, initiating, by the speech-enabled device, the voice call to the recipient voice call number with the voice number used to place voice calls as the particular known user instead of with another voice number.

9. The method of claim 8 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether speech in the utterance matches speech corresponding to the particular known user.

10. The method of claim 8 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether a visual image of at least a portion of a speaker of the utterance matches visual information corresponding to the particular known user.

11. The method of claim 8 , wherein determining, by the speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance comprises:

accessing a user account profile of the particular known user;

determining that the user account profile indicates a voice call device of the particular known user; and

determining that voice number used to place voice calls as the particular known user is connected with the speech-enabled device through a local wireless connection.

12. The method of claim 11 , wherein initiating the voice call comprises:

initiating the voice call through the phone connected with the speech-enabled device.

13. The method of claim 8 , wherein initiating, by the speech-enabled device, the voice call to the recipient voice call number with the voice number used to place voice calls as the particular known user comprises:

initiating the voice call through a Voice over Internet Protocol call provider.

14. The method of claim 8 , further comprising:

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is not known for the particular known user classified as having spoken the utterance:

initiating, by the speech-enabled device, the voice call to the recipient voice call number with an anonymous voice call number used to place voice calls anonymously.

15. A non-transitory computer-readable storage medium comprising instructions that, when executed, configure one or more processors of a computing system to perform operations comprising:

receiving, by a speech-enabled device, an utterance of a first user that requests a voice call with a second user;

determining, by the speech-enabled device and before the voice call is initiated, whether the second user corresponds to a particular type of user;

when the second user corresponds to the particular type of user:

determining, by the speech-enabled device, a recipient voice call number for the second user;

requesting, by the speech-enabled device and from a server, a temporary voice call number for the speech-enabled device; and

initiating, by the speech-enabled device, the voice call to the recipient voice call number using the temporary voice call number for the speech-enabled device; and

when the second user does not correspond to the particular type of user:

classifying, by the speech-enabled device and before the voice call is initiated, the utterance as spoken by a particular known user;

determining, by the speech-enabled device, the recipient voice call number for the second user;

determining, by the speech-enabled device and before the voice call is initiated, whether a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance; and

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance, initiating, by the speech-enabled device, the voice call to the recipient voice call number with the voice number used to place voice calls as the particular known user instead of with another voice number.

16. The non-transitory computer-readable storage medium of claim 15 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether speech in the utterance matches speech corresponding to the particular known user.

17. The non-transitory computer-readable storage medium of claim 15 , wherein classifying the utterance as spoken by a particular known user comprises:

determining whether a visual image of at least a portion of a speaker of the utterance matches visual information corresponding to the particular known user.

18. The non-transitory computer-readable storage medium of claim 15 , wherein determining, by the speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is known for the particular known user classified as having spoken the utterance comprises:

accessing a user account profile of the particular known user;

determining that the user account profile indicates a voice call device of the particular known user; and

determining that voice number used to place voice calls as the particular known user is connected with the speech-enabled device through a local wireless connection.

19. The non-transitory computer-readable storage medium of claim 18 , wherein initiating the voice call comprises:

initiating the voice call through the phone connected with the speech-enabled device.

20. The non-transitory computer-readable storage medium of claim 15 , the operations further comprising:

in response to determining, by speech-enabled device and before the voice call is initiated, that a voice number used to place voice calls as the particular known user is not known for the particular known user classified as having spoken the utterance:

initiating, by the speech-enabled device, the voice call to the recipient voice call number with an anonymous voice call number used to place voice calls anonymously.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 7, 2021
From: LY, VINH QUOC; SHAH, RAUNAQ; KOLAK, OKAN; BINAY, DENIZ; WANG, TIANYU
To: GOOGLE LLC
Reel/Frame 054846/0452 →
Continuity (3)
Continuation 15980836 · May 16, 2018
Provisional Application 62506805 · May 16, 2017
Related Publication 20210092225A1 · Mar 25, 2021