IP Library › Granted Patent US 12,205,211
Granted Patent B2
US 12,205,211 · App. 17/506,054 · Granted Jan 21, 2025

Emotion-based sign language enhancement of content

Inventors: Marc Brandon (Westlake Village, CA); Mark Arana (Agoura Hills, CA)
Assignee: Disney Enterprises, Inc.
G06T13/00G06F3/1423G06F40/20G06F40/58G06T11/00G06V20/40G06V20/41G09B21/009G10L15/22G10L21/055G10L25/57G10L25/63H04N21/242H04N21/488
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,205,211
App. No.
17/506,054
Filed
Oct 20, 2021
Granted
Jan 21, 2025
Kind
B2
Art Unit
2615
USPC
345/474
Abstract

A content enhancement system includes a computing platform having processing hardware and a system memory storing software code. The processing hardware is configured to execute the software code to receive audio-video (A/V) content, to execute at least one of a visual analysis or an audio analysis of the A/V content, and to determine, based on executing the at least one of the visual analysis or the audio analysis, an emotional aspect of the A/V content. The processing hardware is further configured to execute the software code to generate, using the emotional aspect of the A/V content, a sign language translation of the A/V content, the sign language translation including one or more of a gesture, a posture, or a facial expression conveying the emotional aspect.

Claims (50)

1. A content enhancement system comprising:

a computing platform including a processing hardware and a system memory storing a character profile database and a software code, the character profile database including a first emotive profile associated with a first animated character for conveying an emotion using one or more of a first gesture, a first posture, or a first facial expression and a second emotive profile associated with a second animated character for conveying the emotion using one or more of a second gesture, a second posture, or a second facial expression different from the first gesture, the first posture, or the first facial expression, respectively;

the processing hardware configured to execute the software code to:

receive video content;

execute a visual analysis of the video content;

determine, based on executing the visual analysis of the video content, an emotional aspect of the video content being the emotion;

identify one of the first animated character or the second animated character for performing the sign language enhancement;

obtain, from the character profile database, one of the first emotive profile or the second emotive profile corresponding to the identified one of the first animated character or the second animated character; and

generate, using the emotional aspect of the video content and the obtained one of the first emotive profile or the second emotive profile, a sign language enhancement conveying the emotion using one or more of the first gesture, the first posture, or the first facial expression or one or more of the second gesture, the second posture, or the second facial expression of the obtained one of the first emotive profile or the second emotive profile, respectively.

2. The content enhancement system of claim 1 , further comprising:

a display;

wherein the processing hardware is further configured to execute the software code to:

render the video content on the display; and

render the sign language enhancement on the display concurrently with rendering the video content corresponding to the sign language enhancement.

3. The content enhancement system of claim 1 , wherein the identified one of the first animated character or the second animated character is selectable by a user based on a subject matter of the video content.

4. The content enhancement system of claim 1 , further comprising:

a display;

wherein the processing hardware is further configured to execute the software code to:

render the video content on the display for viewing by all of a plurality of users concurrently viewing the video content on the display; and

contemporaneously with rendering the video content on the display, render the sign language enhancement such that the sign language enhancement is visible to at least one of the plurality of users on another display, but not all of the plurality of users concurrently viewing the video content on the display.

5. The content enhancement system of claim 1 , wherein the video content is further accompanied with at least one of audio content or text, and wherein the processing hardware is further configured to execute the software code to determine the emotional aspect further based on analyzing the at least one of the audio content or the text.

6. The content enhancement system of claim 1 , wherein the video content includes a plurality of video frames, and wherein the processing hardware is further configured to execute the software code to execute the visual analysis on a frame-by-frame basis.

7. The content enhancement system of claim 1 , wherein the processing hardware is further configured to execute the software code to:

synchronize the sign language enhancement with a timecode of the video content to produce an enhanced video content; and

record the enhanced video content, or broadcast or stream the enhanced video content to a user system.

8. The content enhancement system of claim 1 , further comprising at least one machine learning (ML) model-based emotion analyzer, wherein the emotional aspect of the video content is determined using the at least one ML model-based emotion analyzer.

9. The content enhancement system of claim 1 , wherein the first emotive profile defines the first animated character that is more excitable than the second animated character defined by the second emotive profile, wherein the first animated character conveys the emotion with one or more exaggerated facial expressions, emphatic gestures, or frequent changes in posture than the second animated character.

10. The content enhancement system of claim 1 , wherein the first emotive profile defines the first animated character that is more stoic than the second animated character defined by the second emotive profile, wherein the first animated character conveys the emotion with one or more subdued facial expressions, subdued gestures, or less body movements than the second animated character.

11. A method for use by a content enhancement system including a computing platform having a processing hardware and a system memory storing a character profile database and a software code, the character profile database including a first emotive profile associated with a first animated character for conveying an emotion using one or more of a first gesture, a first posture, or a first facial expression and a second emotive profile associated with a second animated character for conveying the emotion using one or more of a second gesture, a second posture, or a second facial expression different from the first gesture, the first posture, or the first facial expression, respectively, the method comprising:

receiving, by the software code executed by the processing hardware, video content;

executing, by the software code executed by the processing hardware, a visual analysis of the video content;

determining, by the software code executed by the processing hardware, based on executing the visual analysis of the video content, an emotional aspect of the video content being the emotion;

identifying, by the software code executed by the processing hardware, one of the first animated character or the second animated character for performing the sign language enhancement;

obtaining, by the software code executed by the processing hardware, from the character profile database, one of the first emotive profile or the second emotive profile corresponding to the identified one of the first animated character or the second animated character; and

generating, by the software code executed by the processing hardware, using the emotional aspect of the video content and the obtained one of the first emotive profile or the second emotive profile, a sign language enhancement conveying the emotion using one or more of the first gesture, the first posture, or the first facial expression or one or more of the second gesture, the second posture, or the second facial expression of the obtained one of the first emotive profile or the second emotive profile, respectively.

12. The method of claim 11 , wherein the content enhancement system further comprises a display, the method further comprising:

rendering, by the software code executed by the processing hardware, the video content on the display; and

rendering, by the software code executed by the processing hardware, the sign language enhancement on the display concurrently with rendering the video content corresponding to the sign language enhancement.

13. The method of claim 11 , wherein the identified one of the first animated character or the second animated character is selectable by a user based on a subject matter of the video content.

14. The method of claim 11 , wherein the content enhancement system further comprises a display, the method further comprising:

rendering, by the software code executed by the processing hardware, the video content on the display for viewing by all of a plurality of users concurrently viewing the video content on the display; and

contemporaneously with rendering the video content on the display, rendering, by the software code executed by the processing hardware, the sign language enhancement such that the sign language enhancement is visible to at least one of the plurality of users on another display, but not all of the plurality of users concurrently viewing the video content on the display.

15. The method of claim 11 , wherein the video content is further accompanied with at least one of audio content or text, and wherein the emotional aspect is determined further based on analyzing, by the software code executed by the processing hardware, the at least one of the audio content or the text.

16. The method of claim 11 , wherein the video content includes a plurality of video frames, and wherein the software code, executed by the processing hardware, executes the visual analysis on a frame-by-frame basis.

17. The method of claim 11 , further comprising:

synchronizing, by the software code executed by the processing hardware, the sign language enhancement with a timecode of the video content to produce an enhanced video content; and

recording, by the software code executed by the processing hardware, the enhanced video content, or broadcasting or streaming the enhanced video content to a user system.

18. The method of claim 11 , wherein the content enhancement system further comprises at least one machine learning (ML) model-based emotion analyzer, and wherein the emotional aspect of the video content is determined using the at least one ML model-based emotion analyzer.

19. The method of claim 11 , wherein the first emotive profile defines the first animated character that is more excitable than the second animated character defined by the second emotive profile, wherein the first animated character conveys the emotion with one or more exaggerated facial expressions, emphatic gestures, or frequent changes in posture than the second animated character.

20. The method of claim 11 , wherein the first emotive profile defines the first animated character that is more stoic than the second animated character defined by the second emotive profile, wherein the first animated character conveys the emotion with one or more subdued facial expressions, subdued gestures, or less body movements than the second animated character.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 20, 2021
From: BRANDON, MARC; ARANA, MARK
To: DISNEY ENTERPRISES, INC.
Reel/Frame 057850/0206 →
Continuity (2)
Provisional Application 63184692 · May 5, 2021
Related Publication 20220358701A1 · Nov 10, 2022
References Cited (65)
US 6327272B1 · Van Steenbrugge · 2001 [cited by applicant]
US 6483532B1 · Girod · 2002 [cited by applicant]
US 6545685B1 · Dorbie · 2003 [cited by applicant]
US 7827547B1 · Sutherland et al. · 2010 [cited by applicant]
US 7827574B1 · Hendricks et al. · 2010 [cited by applicant]
US 8566075B1 · Bruner · 2013 [cited by applicant]
US 9215514B1 · Kline · 2015 [cited by applicant]
US 10061817B1 · Frenkel · 2018 [cited by examiner]
US 10375237B1 · Williams et al. · 2019 [cited by applicant]
US 10514766B2 · Gates · 2019 [cited by examiner]
US 11176484B1 · Dorner · 2021 [cited by examiner]
US 11315602B2 · Wu · 2022 [cited by examiner]
US 20020104083A1 · Hendricks et al. · 2002 [cited by applicant]
US 20050097593A1 · Raley et al. · 2005 [cited by applicant]
US 20060018254A1 · Sanders · 2006 [cited by applicant]
US 20090141793A1 · Gramelspacher et al. · 2009 [cited by applicant]
US 20090262238A1 · Hope et al. · 2009 [cited by applicant]
US 20100254408A1 · Kuno · 2010 [cited by applicant]
US 20110096232A1 · Dewa et al. · 2011 [cited by applicant]
US 20110157472A1 · Keskinen · 2011 [cited by applicant]
US 20110162021A1 · Lee · 2011 [cited by applicant]
US 20130141551A1 · Kim · 2013 [cited by applicant]
US 20140046661A1 · Bruner · 2014 [cited by examiner]
US 20140242955A1 · Kang et al. · 2014 [cited by applicant]
US 20150163545A1 · Freed et al. · 2015 [cited by applicant]
US 20150317304A1 · An et al. · 2015 [cited by applicant]
US 20150317307A1 · Mahkovec et al. · 2015 [cited by applicant]
US 20150350139A1 · Speer et al. · 2015 [cited by applicant]
US 20160098850A1 · Shintani et al. · 2016 [cited by applicant]
US 20160191958A1 · Nauseef · 2016 [cited by examiner]
US 20160198214A1 · Levy et al. · 2016 [cited by applicant]
US 20160294714A1 · Persson et al. · 2016 [cited by applicant]
US 20170006248A1 · An et al. · 2017 [cited by applicant]
US 20170111670A1 · Ducloux et al. · 2017 [cited by applicant]
US 20170132828A1 · Zelenin · 2017 [cited by examiner]
US 20180063325A1 · Wilcox · 2018 [cited by examiner]
US 20180075659A1 · Browy et al. · 2018 [cited by applicant]
US 20190052473A1 · Soni et al. · 2019 [cited by applicant]
US 20190096407A1 · Lambourne et al. · 2019 [cited by applicant]
US 20190213401A1 · Kuang · 2019 [cited by examiner]
US 20190251344A1 · Menefee et al. · 2019 [cited by applicant]
US 20200294525A1 · Santos · 2020 [cited by examiner]
US 20210241309A1 · Wolf, Jr. · 2021 [cited by examiner]
US 20210352380A1 · Duncan · 2021 [cited by examiner]
US 20220141547A1 · Plunkett, Jr. · 2022 [cited by examiner]
US 20220171960A1 · Nelson · 2022 [cited by examiner]
US 20220327309A1 · Carlock · 2022 [cited by examiner]
US 20220335971A1 · Gruszka et al. · 2022 [cited by applicant]
US 20220343576A1 · Marey · 2022 [cited by examiner]
WO 2018052901 · 2018 [cited by applicant]
WO 2019157344 · 2019 [cited by applicant]
International Search Report & Written Opinion for International Application PCT/US2022/025123, dated Jul. 4, 2022. [cited by applicant]
“Guidelines for positioning of sign language interpreters in conference, including web-streaming” Sign Language Network, Dec. 21, 2015, pp. 1-6. [cited by applicant]
“Sign Language Video Encoding for Digital Cinema” ISDCF Document 13, Jul. 18, 2018, pp. 1-6. [cited by applicant]
International Search Report & Written Opinion for International Application PCT/US2022/027717, dated Jul. 15, 2022. [cited by applicant]
International Search Report & Written Opinion for International Application PCT/US2022/027716, dated Jul. 13, 2022. [cited by applicant]
International Search Report & Written Opinion for International Application PCT/US2022/027719, dated Jul. 15, 2022. [cited by applicant]
International Search Report & Written Opinion for International Application PCT/US2022/027713, dated Aug. 9, 2022. [cited by applicant]
Daniel Jones “Demystifying Audio Watermarking, Fingerprinting and Modulation.” Published Jan. 19, 2017, 10 pgs. [cited by applicant]
Tiago, Maritan U. de Arahjo, et al. “An Approach to Generate and Embed Sign Language Video Tracks Into Multimedia Contents” Information Sciences vol. 281, Oct. 10, 2014, 7 Pgs. [cited by applicant]
ISDCF Doc4-16-Channel Audio Packaging Guide obtained from https://files.isdcf.com/papers/ISDCF-Doc4-Audio-channel-recommendations.pdf (2017). [cited by applicant]
File History of U.S. Appl. No. 17/735,907, filed May 3, 2022. [cited by applicant]
File History of U.S. Appl. No. 17/735,920, filed May 3, 2022. [cited by applicant]
File History of U.S. Appl. No. 17/735,926, filed May 3, 2022. [cited by applicant]
File History of U.S. Appl. No. 17/735,935, filed May 3, 2022. [cited by applicant]
Cited By (1)
US 12,475,910