IP Library Granted Patent US 12,198,470
Granted Patent B2
US 12,198,470 · App. 17/351,252 · Granted Jan 14, 2025

Server device, terminal device, and display method for controlling facial expressions of a virtual character

Inventor: Akihiko Shirai (Tokyo, JP)
Assignee: GREE, INC.
G06V40/174A63F13/213A63F13/2145A63F13/215A63F13/424A63F13/426A63F13/52G06F3/0482G06F3/04845G06F3/04883G06T7/20G06V40/172G10L15/18G10L15/22G10L25/63G06T2207/20081G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,198,470
App. No.
17/351,252
Granted
Jan 14, 2025
Kind
B2
Abstract

A non-transitory computer readable medium storing computer executable instructions which, when executed by processing circuitry, cause the processing circuitry to acquire, from a first sensor, data related to a face of a performer; provide first data, that is generated based on the data, to a classifier; receive, from the classifier, specific facial expression data indicating one specific facial expression among a plurality of predetermined specific facial expressions on the basis of the first data; and select, for display output, a specific facial expression corresponding to the specific facial expression data received from the classifier.

Claims (43)

1. A non-transitory computer readable medium storing computer executable instructions which, when executed by processing circuitry, cause the processing circuitry to:

acquire, from a first sensor, data related to a face of a performer;

provide first data to a classifier, the first data being generated based on the data;

receive, from the classifier, specific facial expression data based on the first data, wherein the specific facial expression data indicates one specific facial expression among a plurality of specific facial expressions, the classifier having selected the one specific facial expression based on any of processing of the first data and a manual input indicating the one specific facial expression, and the manual input is a direction of a swipe operation performed by the performer over a touch panel and an amount of movement of the swipe operation;

select, for display output, a specific facial expression corresponding to the specific facial expression data received from the classifier; and

operate in accordance with an algorithm in which, in a case where the specific facial expression data is received from the classifier in response to the first data, when specific facial expression designation data that designates the one specific facial expression among the plurality of specific facial expressions is acquired from the performer via a user interface in response to the first data, a specific facial expression corresponding to the specific facial expression designation data is selected as the specific facial expression for the display output.

2. The non-transitory computer readable medium according to claim 1 , wherein the first data includes data related to an amount of movement of a specific point in the face of the performer.

3. The non-transitory computer readable medium according to claim 2 , wherein

the processing circuitry is further caused to acquire, from a second sensor, audio data related to sound of the performer,

the processing circuitry provides second data, that is generated based on the audio data, to the classifier together with the first data, and

the specific facial expression data received by the processing circuitry from the classifier is based on the first data and the second data.

4. The non-transitory computer readable medium according to claim 3 , wherein the second data includes data related to volume, sound pressure, speech rate, and/or formant of the audio data of the performer.

5. The non-transitory computer readable medium according to claim 3 , wherein the second data includes data related to a word, a word ending, and/or an exclamation, and the data is obtained by the processing circuitry through natural language processing performed on the audio data.

6. The non-transitory computer readable medium according to claim 1 , wherein the processing circuitry provides the specific facial expression designation data to the classifier as teacher data for the first data and/or the second data.

7. The non-transitory computer readable medium according to claim 1 , wherein the processing circuitry is further caused to receive the manual input, indicating the one specific facial expression, and provide the one specific facial expression to the classifier.

8. The non-transitory computer readable medium according to claim 3 , wherein the classifier executes a principal component analysis on the first data and/or the second data to generate a learning model.

9. The non-transitory computer readable medium according to claim 8 , wherein when specific facial expression designation data that designates one specific facial expression among the plurality of specific facial expressions is input from the performer to the processing circuitry via a user interface in response to the first data and/or the second data, the classifier executes principal component analysis on the specific facial expression designation data, in addition to the first data and/or the second data, to generate a learning model.

10. The non-transitory computer readable medium according to claim 2 , wherein the processing circuitry is further caused to:

control a display to display a script that instructs the performer to show a facial expression related to any one specific facial expression among the plurality of specific facial expressions; and

provide third data, indicating the one specific facial expression, to the classifier as teacher data for the first data and/or the second data in association with the script.

11. The non-transitory computer readable medium according to claim 1 , wherein the processing circuitry is further caused to:

receive a learning model from a server via a network; and

provide the learning model to the classifier.

12. The non-transitory computer readable medium according to claim 1 , wherein the processing circuitry is further caused to control storage of a learning model generated by the classifier in a memory in association with the performer.

13. The non-transitory computer readable medium according to claim 1 , wherein the processing circuitry is further caused to control transmission of the data to a server via a network.

14. The non-transitory computer readable medium according to claim 1 , wherein the plurality of specific facial expressions include:

first facial expressions that express emotion,

second facial expressions in which a face shape is unrealistically deformed, and/or

third facial expressions in which a face is given a symbol, a shape, and/or a color.

15. The non-transitory computer readable medium according to claim 14 , wherein the first facial expressions that are shown based on a user interface mapped into a language- and culture-independent psychological space, the user interface including a Wheel of Emotions.

16. A display method, comprising:

acquiring, by processing circuitry from a first sensor, data related to a face of a performer;

providing first data to a classifier, the first data being generated based on the data;

receiving, from the classifier, specific facial expression data based on the first data, wherein the specific facial expression data indicates one specific facial expression among a plurality of specific facial expressions, the classifier having selected the one specific facial expression based on any processing of the first data and a manual input indicating the one specific facial expression, and the manual input is a direction of a swipe operation performed by the performer over a touch panel and an amount of movement of the swipe operation;

selecting, by the processing circuitry for display output, a specific facial expression corresponding to the specific facial expression data received from the classifier; and

operating by the processing circuitry in accordance with an algorithm in which, in a case where the specific facial expression data is received from the classifier in response to the first data, when specific facial expression designation data that designates the one specific facial expression among the plurality of specific facial expressions is acquired from the performer via a user interface in response to the first data, a specific facial expression corresponding to the specific facial expression designation data is selected as the specific facial expression for the display output.

17. A device, comprising:

processing circuitry configured to:

acquire, from a first sensor, data related to a face of a performer;

provide first data to a classifier, the first data being generated based on the data;

receive, from the classifier, specific facial expression data based on the first data, wherein the specific facial expression data indicates one specific facial expression among a plurality of specific facial expressions, the classifier having selected the one specific facial expression based on any of processing of the first data and a manual input indicating the one specific facial expression, and the manual input is a direction of a swipe operation performed by the performer over a touch panel and an amount of movement of the swipe operation;

select, for display output, a specific facial expression corresponding to the specific facial expression data received, from the classifier; and

operate in accordance with an algorithm in which in a case where the specific facial expression data is received from the classifier in response to the first data, when specific facial expression designation data that designates the one specific facial expression among the plurality of specific facial expressions is acquired from the performer via a user interface in response to the first data, a specific facial expression corresponding to the specific facial expression designation data is selected as the specific facial expression for the display output.

Assignments (3)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY STREET ADDRESS AND ATTORNEY DOCKET NUMBER PREVIOUSLY RECORDED AT REEL: 71308 FRAME: 765. ASSIGNOR(S) HEREBY CONFIRMS THE CHANGE OF NAME. Recorded Jun 10, 2025
From: GREE, INC.
To: GREE HOLDINGS, INC.
Reel/Frame 071611/0252 →
CHANGE OF NAME Recorded May 16, 2025
From: GREE, INC.
To: GREE HOLDINGS, INC.
Reel/Frame 071308/0765 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 18, 2021
From: SHIRAI, AKIHIKO
To: GREE, INC.
Reel/Frame 056581/0924 →
Priority Claims (1)
JP 2018-236543 · Dec 18, 2018 · national
Continuity (2)
Continuation PCTJP2019049342 · Dec 17, 2019
Related Publication 20210312167A1 · Oct 7, 2021
References Cited (67)
US 5647834A · Ron · 1997 [cited by applicant]
US 9350951B1 · Rowe · 2016 [cited by examiner]
US 10586368B2 · Cao · 2020 [cited by examiner]
US 10719968B2 · Cao · 2020 [cited by examiner]
US 20060290699A1 · Dimtrva · 2006 [cited by examiner]
US 20070139512A1 · Hada et al. · 2007 [cited by applicant]
US 20080301557A1 · Kotlyar · 2008 [cited by examiner]
US 20100141663A1 · Becker · 2010 [cited by examiner]
US 20110310237A1 · Wang et al. · 2011 [cited by applicant]
US 20130235045A1 · Corazza · 2013 [cited by examiner]
US 20150042662A1 · Latorre-Martinez · 2015 [cited by examiner]
US 20160328533A1 · Kawai et al. · 2016 [cited by applicant]
US 20160357402A1 · Matas · 2016 [cited by examiner]
US 20170160813A1 · Divakaran et al. · 2017 [cited by applicant]
US 20170311863A1 · Matsunaga · 2017 [cited by applicant]
US 20180144761A1 · Amini · 2018 [cited by examiner]
US 20180336714A1 · Stoyles et al. · 2018 [cited by applicant]
US 20190005309A1 · Hyun · 2019 [cited by examiner]
US 20190215482A1 · Sathya · 2019 [cited by examiner]
US 20190325633A1 · Miller, IV · 2019 [cited by examiner]
US 20200265627A1 · Seo et al. · 2020 [cited by applicant]
US 20210141663A1 · Deshpande · 2021 [cited by examiner]
US 20220070385A1 · Van Os · 2022 [cited by examiner]
JP 10271470A · 1998 [cited by applicant]
JP 2001126077A · 2001 [cited by applicant]
JP 2002315966A · 2002 [cited by applicant]
JP 2005323340A · 2005 [cited by applicant]
JP 200671936A · 2006 [cited by applicant]
JP 2006202188A · 2006 [cited by applicant]
JP 2007256502A · 2007 [cited by applicant]
JP 2008299430A · 2008 [cited by applicant]
JP 2009153692A · 2009 [cited by applicant]
JP 201259107A · 2012 [cited by applicant]
JP 2013020365A · 2013 [cited by applicant]
JP 2014211719A · 2014 [cited by applicant]
JP 2017156854A · 2017 [cited by applicant]
JP 2018092635A · 2018 [cited by applicant]
JP 2018116589A · 2018 [cited by applicant]
WO 2016021235A1 · 2016 [cited by applicant]
Katsunori Nakamura et al., Development of facial image generation system based on hybrid system, Information Processing Society of Japan 81st National Convention, Japan, Information Processing Society of Japan, Mar. 14,… [cited by applicant]
Notice of Reasons for Refusal issued in corresponding Japanese application 2020-561452, 6 pp. [cited by applicant]
International Search Report and Written Opinion dated Jan. 28, 2020, received for PCT Application No. PCT/JP2019/049342, Filed on Dec. 17, 2019, 11 pages including English Translation. [cited by applicant]
Miyazaki et al., “An expression synthesis in caricaturing”, Transactions of the society of Instrument and control engineers, vol. 36, No. 5, May 31, 2000, pp. 448-455. [cited by applicant]
Kondo et al., “Facial Expression recognition using gabor wavelet in face image”, Proceedings of the 1999 IEICE General Conference Foundation and Boundary, Mar. 1999, pp. 341. [cited by applicant]
Aoshima et al., “Sharing a sense of affinity by emotion view of 600 Thousand People”, Proceedings of Interaction, Mar. 11, 2011, 18 pages. [cited by applicant]
Office Action dated Aug. 30, 2022, in corresponding U.S. Appl. No. 17/077,135, 33 pages. [cited by applicant]
Advisory Action dated Jan. 20, 2023 in co-pending U.S. Appl. No. 17/077,135, 4 pages. [cited by applicant]
Office Action dated Jul. 4, 2023, in corresponding Japanese patent Application No. 2022-084302, 6 pages. [cited by applicant]
Non-final Office Action dated Apr. 10, 2023 in co-pending U.S. Appl. No. 17/077,135, 35 pages. [cited by applicant]
International Search Report and Written Opinion dated Jun. 16, 2020, corresponding PCT/JP2020/018556, 11 pages. [cited by applicant]
Office Action dated Feb. 16, 2021, in corresponding Japanese patent Application No. 2019-239318, 8 pages. [cited by applicant]
Apple Japan Inc., “iPhone X iko de animoji wo tsukau (Using animoji on iPhone X and later)”, total 4 pages, Oct. 24, 2018, searched Nov. 12, 2018, [online] URL: https://support.apple.com/ja-jp/HT208190. [cited by applicant]
Dwango Co., Ltd., “kasutamu kyasuto (custom cast)”, total 8 pages, Oct. 3, 2018, searched Nov. 12, 2018, [online] URL: https://customcast.jp/. [cited by applicant]
Japanese Office Action dated Oct. 25, 2022, in Japanese Patent Application No. 2020-561452, 5 pp. [cited by applicant]
Non-final Office Action dated Apr. 1, 2022, in U.S. Appl. No. 17/077,135. [cited by applicant]
Office Action issued on Aug. 10, 2023, in corresponding U.S. Appl. No. 17/077,135, 8 pages. [cited by applicant]
Japanese Office Action issued Feb. 6, 2024, in corresponding Japanese Patent Application No. 2022-166878, 25 pages. [cited by applicant]
Toshimitsu Miyajima et al., “Avatar facial expression control method using fundamental frequency and sound pressure in voice chat system”, Journal of Human Interface Society, Nov. 25, 2007, vol. 9, No. 4, p. 503-512. [cited by applicant]
U.S. Advisory Action issued Dec. 6, 2023, in corresponding U.S. Appl. No. 17/077.135, 4 pages. [cited by applicant]
Office Action issued Mar. 1, 2024, in co-pending U.S. Appl. No. 17/077,135, 35 pages. [cited by applicant]
Office Action issued on Jul. 4, 2023, in corresponding Japanese patent Application No. 2022-084302, 6 pages. [cited by applicant]
“Avatar technology for making a big difference in live broadcasting”, [online] Dec. 5, 2016, [searched Jun. 27, 2023], Internet <https://technote.qualiarts.jp/article/3>, total 23 pages. [cited by applicant]
Japanese Office Action issued Apr. 9, 2024, in corresponding Japanese Patent Application No. 2022-166878, 15pp. [cited by applicant]
Office Action issued May 14, 2024 in Japanese Patent Application No, 2023-077309. [cited by applicant]
Office Action issued in Jun. 18, 2024 in related Japanese Patent Application No. 2022-166878. [cited by applicant]
Office Action issued in Oct. 8, 2024 in related Japanese Patent Application No. 2023-077309. [cited by applicant]
Office Action issued in Nov. 19, 2024 in related Japanese Patent Application No. 2023-214338. [cited by applicant]