IP Library Granted Patent US 11,758,232
Granted Patent B2
US 11,758,232 · App. 17/713,261 · Granted Sep 12, 2023

Presentation and management of audio and visual content across devices

Inventors: Michael Lee Loritsch (Seattle, WA); John Martin Miller (Seattle, WA); Paul Anthony Kotas (Seattle, WA); Ross Tucker (Seattle, WA)
Assignee: Amazon Technologies, Inc.
H04N21/47202G06F3/165G06F3/167G10L15/08G10L15/22H04N21/4524G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,758,232
App. No.
17/713,261
Granted
Sep 12, 2023
Kind
B2
Abstract

Systems, methods, and computer-readable media are disclosed for systems and methods of presentation and management of audio and visual content across devices. Example methods may include causing presentation of first audio content at a speaker device, causing presentation of a first audio notification indicative of visual content available for presentation, causing presentation of second audio content after the first audio notification, and sending first visual content to a first display device for presentation during presentation of the second audio content.

Claims (84)

1. A device comprising:

a display;

a microphone;

memory that stores computer-executable instructions; and

at least one processor configured to access the memory and execute the computer-executable instructions to:

receive first voice input;

determine that audio content is playing when the first voice input is received;

cause playback of the audio content to be paused;

cause analysis of the first voice input to determine that both a trigger word and a request are present in the first voice input;

determine first visual content associated with the request; and

cause the first visual content to be sent to a first display device for presentation, wherein the first display device is associated with the user account.

2. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

cause presentation of first audio content at a speaker device in response to the first voice input.

3. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine a user account associated with the device;

cause presentation of a first audio notification indicative of visual content available at the first display device;

determine, while the first visual content is presented at the first display device, a first user interaction with a second display device; and

send the first visual content to the second display device.

4. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine that the first voice input is complete; and

cause playback of the audio content to resume while the first voice input is being analyzed.

5. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine selection of an audio playback option at a user interface of the first display device;

cause presentation of second audio content at a speaker device;

receive second voice input; and

cause presentation of third audio content at the speaker device.

6. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

receive second voice input indicating a request for audio playback of the first visual content;

cause presentation of second audio content at a speaker device, wherein the second audio content is a text-to-speech presentation of a first portion of the first visual content;

receive third voice input; and

cause presentation of third audio content at the speaker device, wherein the third audio content is a text-to-speech presentation of a second portion of the first visual content.

7. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine that a user is at a first location physically closest to the first display device;

determine that the user has moved to a second location physically closest to a second display device; and

send the first visual content to the second display device.

8. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine a second user interaction with the first visual content at the first display device;

send second visual content to the first display device;

determine a third user interaction with a second display device; and

send the second visual content to the second display device.

9. The device of claim 1 , wherein the device is further configured to execute the computer-executable instructions to:

determine a second user interaction with a third display device; and

send the first visual content to the third display device.

10. The device of claim 1 , wherein the device is a displayless device.

11. A method comprising:

receiving, by a device associated with a user account, first voice input;

determining that audio content is playing when the first voice input is received;

causing playback of the audio content to be paused;

causing analysis of the first voice input to determine that both a trigger word and a request are present in the first voice input;

determining first visual content associated with the request; and

causing the first visual content to be sent to a first display device for presentation, wherein the first display device is associated with the user account.

12. The method of claim 11 , further comprising:

causing presentation of first audio content at a speaker device in response to the first voice input.

13. The method of claim 11 , further comprising:

determining a user account associated with the device;

causing presentation of a first audio notification indicative of visual content available at the first display device;

determining, while the first visual content is presented at the first display device, a first user interaction with a second display device; and

sending the first visual content to the second display device.

14. The method of claim 11 , further comprising:

determining that the first voice input is complete; and

causing playback of the audio content to resume while the first voice input is being analyzed.

15. The method of claim 11 , further comprising:

determining selection of an audio playback option at a user interface of the first display device;

causing presentation of second audio content at a speaker device;

receiving second voice input; and

causing presentation of third audio content at the speaker device.

16. The method of claim 11 , further comprising:

receiving second voice input indicating a request for audio playback of the first visual content;

causing presentation of second audio content at a speaker device, wherein the second audio content is a text-to-speech presentation of a first portion of the first visual content;

receiving third voice input; and

causing presentation of third audio content at the speaker device, wherein the third audio content is a text-to-speech presentation of a second portion of the first visual content.

17. The method of claim 11 , further comprising:

determining that a user is at a first location physically closest to the first display device;

determining that the user has moved to a second location physically closest to a second display device; and

sending the first visual content to the second display device.

18. The method of claim 11 , further comprising:

determining a second user interaction with the first visual content at the first display device;

sending second visual content to the first display device;

determining a third user interaction with a second display device; and

sending the second visual content to the second display device.

19. The method of claim 11 , further comprising:

determining a second user interaction with a third display device; and

sending the first visual content to the third display device.

20. The method of claim 11 , wherein the device is a displayless device.

Continuity (3)
Continuation 16695513 · Nov 26, 2019
Continuation 15710911 · Sep 21, 2017
Related Publication 20220303630A1 · Sep 22, 2022
Cited By (2)
US 12,334,078 US 12,701,375