IP Library › Granted Patent US 12,046,241
Granted Patent B2
US 12,046,241 · App. 18/312,580 · Granted Jul 23, 2024

Device leadership negotiation among voice interface devices

Inventors: Kenneth Mixter (Los Altos Hills, CA); Diego Melendo Casado (Mountain View, CA); Alexander H. Gruenstein (Mountain View, CA); Terry Tai (New York, NY); Christopher Thaddeus Hughes (Redwood City, CA); Matthew Nirvan Sharifi (Kilchberg, CH)
Assignee: Google LLC
G10L15/22G10L15/32G10L2015/088G10L2015/223G10L25/60
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,046,241
App. No.
18/312,580
Granted
Jul 23, 2024
Kind
B2
Abstract

The various implementations described herein include methods and systems for determining device leadership among voice interface devices. In one aspect, a method is performed at a first electronic device of a plurality of electronic devices, each having microphones, a speaker, processors, and memory storing programs for execution by the processors. The first device detects a voice input. It determines a device state and a relevance of the voice input. It identifies a subset of electronic devices from the plurality to which the voice input is relevant. In accordance with a determination that the subset includes the first device, the first device determines a first score of a criterion associated with the voice input and receives second scores of the criterion from other devices in the subset. In accordance with a determination that the first score is higher than the second scores, the first device responds to the detected input.

Claims (46)

1. A computer-implemented method when executed on data processing hardware of a first voice assistant device causes the data processing hardware to perform operations comprising:

receiving a voice input comprising a hotword and a voice request subsequent to the hotword, the voice input captured by the first voice assistant device and a second voice assistant device, the first voice assistant device and the second voice assistant device each communicatively coupled to a local network implemented at a network interface and configured to respond to voice requests that are subsequent to the hotword;

detecting, in the voice input, the hotword;

based on detecting the hotword, processing the voice input to determine that the voice request comprises:

a media transfer request to transfer playback of media content to a group of one or more media output devices; and

a user voice designation of the group of the one or more media output devices, the user voice designation comprising a description of a destination of the group of the one or more media output devices;

sending, from the first voice assistant device via the local network, a multicast message received by the second voice assistant device, the multicast message received by the second voice assistant device causing the second voice assistant device to not respond to the voice request despite the second voice assistant device capturing the voice input of the hotword and the voice request subsequent to the hotword; and

based on determining that the voice input comprises the media transfer request, causing, using the user voice designation of the group of the one or more media output devices, each media output device in the group of the one or more media output devices to playback the media content.

2. The method of claim 1 , wherein:

the media output devices in the group of the one or more media output devices comprise speakers; and

the media content played back by each media output device in the group of the one or more media output devices comprises music audibly played back by each media output device in the group of the one or more media output devices.

3. The method of claim 1 , wherein the description of the destination of the group of the one or more media output devices comprises a particular room within a house where the group of the one or more media output devices are located.

4. The method of claim 1 , wherein the description of the destination of the group of the one or more media output devices comprises a particular space within a house where the group of the one or more media output devices are located.

5. The method of claim 1 , wherein causing each media output device in the group of the one or more media output devices to playback the media content comprises causing each media output device to playback the media content streamed from a remote content source.

6. The method of claim 1 , wherein the first voice assistant device and the each media output device in the group of the one or more media output devices are communicatively coupled to the local network implemented at the network interface.

7. The method of claim 6 , wherein the first voice assistant device is configured to communicate with at least one media output device in the group of the one or more media output devices through the local network.

8. The method of claim 1 , wherein the operations further comprise displaying, via an array of light emitting diodes (LEDs) of the first voice assistant device, a visual pattern on the LEDs while processing the voice input.

9. The method of claim 1 , wherein the voice input is captured by a microphone implemented by the first voice assistant device.

10. The method of claim 1 , wherein the operations further comprise audibly outputting, from a speaker of the first voice activated device, a voice message response to the voice request confirming that the voice request has been fulfilled.

11. A first voice assistant device comprising:

data processing hardware; and

memory hardware in communication with the data processing hardware and storing instructions that when executed on the data processing hardware causes the data processing hardware to perform operations comprising:

receiving a voice input comprising a hotword and a voice request subsequent to the hotword, the voice input captured by the first voice assistant device and a second voice assistant device, the first voice assistant device and the second voice assistant device each communicatively coupled to a local network implemented at a network interface and configured to respond to voice requests that are subsequent to the hotword;

detecting, in the voice input, the hotword;

based on detecting the hotword, processing the voice input to determine that the voice request comprises:

a media transfer request to transfer playback of media content to a group of one or more media output devices; and

a user voice designation of the group of the one or more media output devices, the user voice designation comprising a description of a destination of the group of the one or more media output devices;

sending, from the first voice assistant device via the local network, a multicast message received by the second voice assistant device, the multicast message received by the second voice assistant device causing the second voice assistant device to not respond to the voice request despite the second voice assistant device capturing the voice input of the hotword and the voice request subsequent to the hotword; and

based on determining that the voice input comprises the media transfer request, causing, using the user voice designation of the group of the one or more media output devices, each media output device in the group of the one or more media output devices to playback the media content.

12. The first voice assistant device of claim 11 , wherein:

the media output devices in the group of the one or more media output devices comprise speakers; and

the media content played back by each media output device in the group of the one or more media output devices comprises music audibly played back by each media output device in the group of the one or more media output devices.

13. The first voice assistant device of claim 11 , wherein the description of the destination of the group of the one or more media output devices comprises a particular room within a house where the group of the one or more media output devices are located.

14. The first voice assistant device of claim 11 , wherein the description of the destination of the group of the one or more media output devices comprises a particular space within a house where the group of the one or more media output devices are located.

15. The first voice assistant device of claim 11 , wherein causing each media output device in the group of the one or more media output devices to playback the media content comprises causing each media output device to playback the media content streamed from a remote content source.

16. The first voice assistant activated device of claim 11 , wherein the first voice assistant device and the each media output device in the group of the one or more media output devices are communicatively coupled to the local network implemented at the network interface.

17. The first voice assistant device of claim 16 , wherein the first voice assistant device is configured to communicate with at least one media output device in the group of the one or more media output devices through the local network.

18. The first voice assistant device of claim 11 , further comprising:

an array of light emitting diodes (LEDs),

wherein the operations further comprise displaying a visual pattern on the LEDs while processing the voice input.

19. The first voice assistant device of claim 11 , further comprising:

a microphone,

wherein the voice input is captured by the microphone.

20. The first voice assistant device of claim 11 , further comprising:

a speaker,

wherein the operations further comprise audibly outputting, from the speaker, a voice message response to the voice request confirming that the voice request has been fulfilled.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 4, 2023
From: MIXTER, KENNETH; CASADO, DIEGO MELENDO; GRUENSTEIN, ALEXANDER HOUSTON; TAI, TERRY; HUGHES, CHRISTOPHER THADDEUS; SHARIFI, MATTHEW NIRVAN
To: GOOGLE INC.
Reel/Frame 063544/0940 →
CHANGE OF NAME Recorded May 4, 2023
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 063548/0916 →
Continuity (9)
Continuation 17242273 · Apr 27, 2021
Continuation 16786943 · Feb 10, 2020
Continuation 16159339 · Oct 12, 2018
Continuation 15788658 · Oct 19, 2017
Continuation 15284483 · Oct 3, 2016
Continuation In Part 15088477 · Apr 1, 2016
Continuation 14675932 · Apr 1, 2015
Provisional Application 62061830 · Oct 9, 2014
Related Publication 20230274741A1 · Aug 31, 2023