IP Library › Granted Patent US 12,003,821
Granted Patent B2
US 12,003,821 · App. 16/853,451 · Granted Jun 4, 2024

Techniques for enhanced media experience

Inventors: Douglas A. Fidaleo (Santa Clarita, CA); Michael P. Goslin (Sherman Oaks, CA)
Assignee: DISNEY ENTERPRISES, INC.
H04N21/478H04N21/4316H04N21/84
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,003,821
App. No.
16/853,451
Granted
Jun 4, 2024
Kind
B2
Abstract

The present disclosure sets forth a technique for providing supplemental content along with source media. The technique includes receiving source media and content metadata associated with the source media, generating supplemental content based on the content metadata and preferences of one or more media consumers, and delivering the supplement content along with the source media to the one or more media consumers.

Claims (46)

1. A computer-implemented method for providing supplemental content along with source media, the method comprising:

receiving source media and content metadata associated with the source media;

generating supplemental content based on a subset of the content metadata that is associated with a first portion of the source media and preferences of one or more media consumers, wherein the generated supplemental content does not include the first portion of the source media;

delivering the generated supplemental content along with the first portion of the source media to the one or more media consumers;

receiving video sensor data captured by one or more video sensors that indicates a first reaction of the one or more media consumers to at least one of the first portion of the source media or the supplemental content;

determining at least one of a first emotional state of the one or more media consumers, a first sentiment state of the one or more media consumers, or a first level of engagement of the one or more media consumers based on the video sensor data indicating the first reaction; and

generating additional supplemental content based on the at least one of the first emotional state, the first sentiment state, or the first level of engagement.

2. The computer-implemented method of claim 1 , wherein the source media comprises audio media or audio-visual media.

3. The computer-implemented method of claim 1 , further comprising:

modifying the preferences of the one or more media consumers based on the at least one of the first emotional state, the first sentiment state, or the first level of engagement;

generating the additional supplemental content based on the content metadata and the modified preferences; and

delivering the additional supplemental content to the one or more media consumers along with the source media.

4. The computer-implemented method of claim 1 , wherein the generated supplemental content comprises one or more audio elements, video elements, or holographic elements.

5. The computer-implemented method of claim 1 , wherein the generated supplemental content is customized based on an identity of at least one of the one or more media consumers.

6. The computer-implemented method of claim 1 , wherein the generated supplemental content comprises a virtual companion that consumes the source media along with the one or more media consumers.

7. The computer-implemented method of claim 6 , wherein the virtual companion reacts to one or more of the source media, another virtual companion, or the one or more media consumers.

8. The computer-implemented method of claim 6 , wherein the virtual companion interacts with one or more of the media consumers or another virtual companion.

9. The computer-implemented method of claim 1 , wherein delivering the generated supplemental content along with the first portion of the source media comprises ducking or pausing the first portion of the source media when the generated supplemental content is delivered.

10. One or more non-transitory computer-readable storage media including instructions that, when executed by one or more processors, cause the one or more processors to provide supplemental content along with source media by performing steps comprising:

receiving source media and content metadata associated with the source media;

generating supplemental content based on a subset of the content metadata that is associated with a first portion of the source media and preferences of one or more media consumers, wherein the generated supplemental content does not include the first portion of the source media;

delivering the generated supplemental content along with the first portion of the source media to the one or more media consumers;

receiving video sensor data captured by one or more video sensors that indicates a first reaction of the one or more media consumers to at least one of the first portion of the source media or the supplemental content;

determining at least one of a first emotional state of the one or more media consumers, a first sentiment state of the one or more media consumers, or a first level of engagement of the one or more media consumers based on the video sensor data indicating the first reaction; and

generating additional supplemental content based on the at least one of the first emotional state, the first sentiment state, or the first level of engagement.

11. The one or more non-transitory computer-readable storage media of claim 10 , wherein the generated supplemental content comprises a virtual companion that consumes the first portion of the source media along with the one or more media consumers.

12. The one or more non-transitory computer-readable storage media of claim 10 , wherein the content metadata includes at least one of global content metadata or time-aligned content metadata.

13. The one or more non-transitory computer-readable storage media of claim 10 , wherein the preferences of the one or more media consumers include at least one of historical preferences or transient preferences of the one or more media consumers.

14. The one or more non-transitory computer-readable storage media of claim 13 , wherein generating the supplemental content comprises giving greater weight to the transient preferences than to the historical preferences.

15. The one or more non-transitory computer-readable storage media of claim 13 , further comprising determining the transient preferences based on one or more of a world state associated with a location where the source media is being consumed, local metadata determined from the one or more video sensors at the location, a sentiment of the one or more media consumers, an engagement of the one or more media consumers, or an emotion of the one or more media consumers.

16. The one or more non-transitory computer-readable storage media of claim 10 , wherein the steps further comprise:

analyzing the source media; and

determining the content metadata based on the analyzing.

17. The one or more non-transitory computer-readable storage media of claim 16 , wherein determining the content metadata is further based on at least one of metadata retrieved from one or more data sources or input from one or more test media consumers.

18. A computing device, comprising:

a memory; and

one or more processors coupled to the memory;

wherein the one or more processors are configured to:

receive source media and content metadata associated with the source media;

generate a virtual companion based on the content metadata and preferences of one or more media consumers;

generate supplemental content to be delivered by the virtual companion, wherein the supplemental content is generated based on a subset of the content metadata that is associated with a first portion of the source media and does not include the first portion of the source media;

have the virtual companion deliver the generated supplemental content to the one or media consumers along with the first portion of the source media;

receive video sensor data captured by one or more video sensors that indicates a first reaction of the one or more media consumers to at least one of the first portion of the source media or the generated supplemental content;

determine at least one of a first emotional state of the one or more media consumers, a first sentiment state of the one or more media consumers, or a first level of engagement of the one or more media consumers based on the video sensor data indicating the first reaction, wherein the first level of engagement comprises a value on an engagement spectrum; and

generate additional supplemental content based on the at least one of the first emotional state, the first sentiment state, or the value on the engagement spectrum that is determined based on the video sensor data indicating the first reaction.

19. The computing device of claim 18 , wherein to generate the supplemental content, the one or more processors are configured to select a first supplemental content element from labeled supplemental content or customize a second supplemental content element from the labeled supplemental content.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 16, 2020
From: FIDALEO, DOUGLAS A.; GOSLIN, MICHAEL P.
To: DISNEY ENTERPRISES, INC.
Reel/Frame 054380/0986 →
Continuity (1)
Related Publication 20210329342A1 · Oct 21, 2021
Cited By (2)
US 12,531,824 US 12,604,062