IP Library Granted Patent US 12,524,965
Granted Patent B2
US 12,524,965 · App. 18/473,768 · Granted Jan 13, 2026

Generation of three-dimensional (3D) blend-shapes from 3D scans using neural network

Inventor: Kohei Miyamoto (San Jose, CA)
Assignees: SONY GROUP CORPORATION; SONY CORPORATION OF AMERICA
G06T17/205G06T5/70G06T7/11G06T2207/10028G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,524,965
App. No.
18/473,768
Granted
Jan 13, 2026
Kind
B2
Abstract

An electronic device and a method for generation of three-dimensional (3D) blend-shapes from 3D scans using neural network are disclosed. The electronic device acquires a set of 3D scans including a body portion of an object. The electronic device determines a set of segments of the body portion from each 3D scan. The electronic device applies a neural network model on the acquired set of 3D scans. The electronic device determines a set of vertex difference vectors. Each vector of the determined set of vertex difference vectors corresponds to a 3D blend-shape. Each segment of the determined set of segments moves independently in the 3D blend-shape. The electronic device reconstructs a 3D mesh sequence. The electronic device re-trains the neural network model. The re-trained neural network model determines a set of 3D blend-shapes based on a set of input 3D scans.

Claims (69)

1 . An electronic device, comprising:

circuitry configured to:

acquire a set of three-dimensional (3D) scans including a body portion of an object;

determine a set of segments of the body portion from each 3D scan of the acquired set of 3D scans;

apply, based on the determined set of segments, a neural network model on the acquired set of 3D scans, wherein the neural network model includes an encoder model and a decoder model;

determine, by the encoder model, based on the acquired set of 3D scans, a set of weights associated with the determined set of segments;

determine, by the decoder model, a set of vertex difference vectors associated with the determined set of segments, wherein

each vector of the determined set of vertex difference vectors corresponds to a 3D blend-shape associated with the determined set of segments, and

each segment of the determined set of segments is configured to move independently in the 3D blend-shape;

determine, based on the determined set of weights and the determined set of vertex difference vectors, a regularization function associated with the determined set of segments;

reconstruct a 3D mesh sequence based on the determined set of vertex difference vectors; and

re-train the neural network model based on the acquired set of 3D scans, the determined regularization function, and the reconstructed 3D mesh sequence, wherein

the re-trained neural network model is configured to determine a set of 3D blend-shapes based on a set of input 3D scans, and

the set of 3D blend-shapes includes the 3D blend-shape.

2 . The electronic device according to claim 1 , wherein the body portion of the object corresponds to a face of a person.

3 . The electronic device according to claim 1 , wherein the determination of the set of segments of the body portion is based on at least one of a clustering technique or a specific input.

4 . The electronic device according to claim 1 , wherein the determined regularization function is configured to reduce a number of the set of 3D blend-shapes.

5 . The electronic device according to claim 1 , wherein the determined regularization function corresponds to a Lasso (L1) regression function.

6 . The electronic device according to claim 1 , wherein the circuitry is further configured to:

determine a smoothening function associated with the determined set of segments, wherein

the determined smoothening function is configured to smoothen boundaries of the determined set of segments, and

the re-trained neural network model is further based on the determined smoothening function.

7 . The electronic device according to claim 6 , wherein the determined smoothening function corresponds to a Laplacian boundary-smoothening function.

8 . The electronic device according to claim 1 , wherein the each vector of the determined set of vertex difference vectors corresponds to at least one of:

a region-based blend-shape of a segment of the determined set of segments, or

mask information associated with the segment.

9 . The electronic device according to claim 1 , wherein

the determined set of vertex difference vectors includes a first vector and a second vector,

the first vector has a first valid area,

the second vector has a second valid area,

the first valid area overlaps with the second valid area, and

the overlap between the first valid area and the second valid area hide boundaries between segments associated with the first vector and the second vector.

10 . The electronic device according to claim 9 , wherein

the overlap between the first valid area and the second valid area is smoothened with a smoothening function, and

the re-trained neural network model is based on the smoothening function.

11 . A method, comprising:

in an electronic device:

acquiring a set of three-dimensional (3D) scans including a body portion of an object;

determining a set of segments of the body portion from each 3D scan of the acquired set of 3D scans;

applying, based on the determined set of segments, a neural network model on the acquired set of 3D scans, wherein the neural network model includes an encoder model and a decoder model;

determining, by the encoder model, based on the acquired set of 3D scans, a set of weights associated with the determined set of segments;

determining, by the decoder model, a set of vertex difference vectors associated with the determined set of segments, wherein

each vector of the determined set of vertex difference vectors corresponds to a 3D blend-shape associated with the determined set of segments, and

each segment of the determined set of segments is configured to move independently in the 3D blend-shape;

determining, based on the determined set of weights and the determined set of vertex difference vectors, a regularization function associated with the determined set of segments;

reconstructing a 3D mesh sequence based on the determined set of vertex difference vectors; and

re-training the neural network model based on the acquired set of 3D scans, the determined regularization function, and the reconstructed 3D mesh sequence, wherein

the re-trained neural network model is configured to determine a set of 3D blend-shapes based on a set of input 3D scans, and

the set of 3D blend-shapes includes the 3D blend-shape.

12 . The method according to claim 11 , wherein the determined regularization function is configured to reduce a number of the set of 3D blend-shapes.

13 . The method according to claim 11 , wherein the determined regularization function corresponds to a Lasso (L1) regression function.

14 . The method according to claim 11 , further comprising:

determining a smoothening function associated with the determined set of segments, wherein

the determined smoothening function is configured to smoothen boundaries of the determined set of segments, and

the re-training of the neural network model is further based on the determined smoothening function.

15 . The method according to claim 14 , wherein the determined smoothening function corresponds to a Laplacian boundary-smoothening function.

16 . A non-transitory computer-readable medium having stored thereon, computer-executable instructions that when executed by an electronic device, cause the electronic device to execute operations, the operations comprising:

acquiring a set of three-dimensional (3D) scans including a body portion of an object;

determining a set of segments of the body portion from each 3D scan of the acquired set of 3D scans;

applying, based on the determined set of segments, a neural network model on the acquired set of 3D scans, wherein the neural network model includes an encoder model and a decoder model;

determining, by the encoder model, based on the acquired set of 3D scans, a set of weights associated with the determined set of segments;

determining, by the decoder model, a set of vertex difference vectors associated with the determined set of segments, wherein

each vector of the determined set of vertex difference vectors corresponds to a 3D blend-shape associated with the determined set of segments, and

each segment of the determined set of segments is configured to move independently in the 3D blend-shape;

determining, based on the determined set of weights and the determined set of vertex difference vectors, a regularization function associated with the determined set of segments;

reconstructing a 3D mesh sequence based on the determined set of vertex difference vectors; and

re-training the neural network model based on the acquired set of 3D scans, the determined regularization function, and the reconstructed 3D mesh sequence, wherein

the re-trained neural network model is configured to determine a set of 3D blend-shapes based on a set of input 3D scans, and

the set of 3D blend-shapes includes the 3D blend-shape.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 25, 2023
From: MIYAMOTO, KOHEI
To: SONY GROUP CORPORATION; SONY CORPORATION OF AMERICA
Reel/Frame 065012/0288 →
Continuity (1)
Related Publication 20250104358A1 · Mar 27, 2025
References Cited (34)
US 10181213B2 · Bullivant · 2019 [cited by examiner]
US 10783690B2 · Sagar · 2020 [cited by examiner]
US 10939143B2 · Iyer · 2021 [cited by examiner]
US 11354844B2 · Sagar · 2022 [cited by examiner]
US 11836837B2 · Hu · 2023 [cited by examiner]
US 11893671B2 · Sagar · 2024 [cited by examiner]
US 11941737B2 · Wang · 2024 [cited by examiner]
US 11995752B2 · Park · 2024 [cited by examiner]
US 12094043B2 · Gagne · 2024 [cited by examiner]
US 12131407B2 · Wang · 2024 [cited by examiner]
US 12159346B2 · Selvet · 2024 [cited by examiner]
US 12169900B2 · Khakhulin · 2024 [cited by examiner]
US 12182922B2 · Valentin · 2024 [cited by examiner]
US 12184923B2 · Zhang · 2024 [cited by examiner]
US 12198374B2 · Sun · 2025 [cited by examiner]
US 12277670B2 · Noh · 2025 [cited by examiner]
US 20040095344A1 · Dojyun · 2004 [cited by examiner]
US 20170243387A1 · Li · 2017 [cited by examiner]
US 20200184721A1 · Ge · 2020 [cited by examiner]
US 20220164921A1 · Noh · 2022 [cited by examiner]
US 20220405996A1 · Shirai · 2022 [cited by examiner]
US 20230062756A1 · Takeda et al. · 2023 [cited by applicant]
US 20230222721A1 · Chen · 2023 [cited by examiner]
US 20230237753A1 · Bradley · 2023 [cited by examiner]
US 20230260184A1 · Chen · 2023 [cited by examiner]
US 20240062467A1 · Sarkis · 2024 [cited by examiner]
US 20240078356A1 · Tiwari · 2024 [cited by examiner]
US 20240087230A1 · Shetty · 2024 [cited by examiner]
US 20240221270A1 · Leyton · 2024 [cited by examiner]
US 20240355051A1 · Ahn · 2024 [cited by examiner]
US 20250061634A1 · Huang · 2025 [cited by examiner]
US 20250104358A1 · Miyamoto · 2025 [cited by examiner]
Regateiro, et al., “Deep4D: A Compact Generative Representation for Volumetric Video”, Frontiers in Virtual Reality, vol. 2, Article 739010, Nov. 1, 2021, 17 pages. [cited by applicant]
Casas, et al., “4D Video Textures for Interactive Character Appearance”, vol. 33, No. 2, 2014, 10 pages. [cited by applicant]