IP Library Granted Patent US 12,573,129
Granted Patent B2
US 12,573,129 · App. 18/096,972 · Granted Mar 10, 2026

Method and device for representing rendered scenes

Inventors: Seokhwan Jang (Suwon-si, KR); Nahyup Kang (Suwon-si, KR); Jiyeon Kim (Suwon-si, KR); Hyewon Moon (Suwon-si, KR); Donghoon Sagong (Suwon-si, KR); Minjung Son (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G06T15/08G06N3/045G06N3/08G06N3/09G06T7/40G06T7/50G06T15/005G06T15/06G06T2200/04G06T2207/20081G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,573,129
App. No.
18/096,972
Granted
Mar 10, 2026
Kind
B2
Abstract

Disclosed are a method and device for representing rendered scenes. A data processing method of training a neural network model includes obtaining spatial information of sampling data, obtaining one or more volume-rendering parameters by inputting the spatial information of the sampling data to the neural network model, obtaining a regularization term based on a distribution of the volume-rendering parameters, performing volume rendering based on the volume-rendering parameters, and training the neural network model to minimize a loss function determined based on the regularization term and based on a difference between a ground truth image and an image that is estimated according to the volume rendering.

Claims (55)

1 . A method of training a neural network model for scene representation, the method comprising:

obtaining spatial information of sampling data, the spatial information including information about points sampled from a ray or from a three-dimensional model;

obtaining volume-rendering parameters by inputting the spatial information of the sampling data to the neural network model, which generates the volume-rendering parameters, the volume-rendering parameters corresponding to the points, respectively;

obtaining a regularization term quantifying a statistical distribution of the volume-rendering parameters;

performing volume rendering based on the volume-rendering parameters; and

training the neural network model to minimize a loss function, wherein the loss function is determined based on the regularization term and is determined based on a difference between a ground truth image and an image that is estimated according to the volume rendering.

2 . The method of claim 1 , wherein the training the neural network model comprises:

training the neural network model such that the distribution of the volume-rendering parameters has a predetermined feature.

3 . The method of claim 2 , wherein the training the neural network model comprises:

training the neural network model such that the distribution of the volume-rendering parameters is clustered on a surface of a scene.

4 . The method of claim 1 , wherein the obtaining the regularization term comprises:

obtaining the regularization term based on a metric quantifying a feature of the distribution of the volume-rendering parameters.

5 . The method of claim 1 , wherein the obtaining the regularization term comprises:

obtaining an entropy measure corresponding to the distribution of the volume-rendering parameters; and

obtaining an information potential corresponding to the volume-rendering parameters, based on the entropy measure.

6 . The method of claim 5 , wherein the training the neural network model is performed such that the information potential is maximized.

7 . The method of claim 1 , wherein

the loss function is determined by adding a second loss function to a first loss function, wherein

the first loss function is determined based on the difference between the ground truth image and the image is estimated through the volume rendering and the second loss function is determined based on the regularization term.

8 . The method of claim 1 , wherein the obtaining the regularization term comprises:

obtaining the distribution of the volume-rendering parameters corresponding to sample point of sampling data included in a set of sample points in a predetermined area;

obtaining a statistical value of the distribution of the volume-rendering parameters corresponding to the sample points; and

determining the statistical value to be the regularization term.

9 . The method of claim 1 , wherein the obtaining the spatial information of the sampling data comprises obtaining spatial information of a ray and obtaining sampling information.

10 . A scene representation method comprising:

obtaining spatial information of sampling data, the sampling data sampled from a three-dimensional (3D) model, the spatial information including information about points sampled from a ray or from a three-dimensional model; and

performing volume rendering of the 3D model by inputting the spatial information of the sampling data to a neural network model that generates volume rendering parameters, wherein

the spatial information of the sampling data is determined based on a quantification of an amount of information in the volume rendering parameters.

11 . The method of claim 10 , wherein

the neural network model is trained to transform the distribution of the volume rendering parameters to perform the volume rendering.

12 . The method of claim 10 , wherein

the spatial information of sampling data comprises at either position information or information on a number of sample points in the sampling data.

13 . The method of claim 10 , wherein

the obtaining the spatial information of sampling data further comprises obtaining a depth map corresponding to a scene of the 3D model.

14 . The method of claim 10 , wherein

the obtaining the spatial information of sampling data further comprises obtaining information on a surface of the 3D model.

15 . The method of claim 10 , wherein

the obtaining the spatial information of the sampling data comprises obtaining spatial information of a ray and obtaining sampling information.

16 . A non-transitory computer-readable storage medium storing instructions that, when executed by a processor, cause the processor to perform the method of claim 1 .

17 . An electronic device comprising:

one or more processors;

memory storing instructions configured to, when executed by the one or more processors, cause the one or more processors to:

obtain spatial information of sampling data, the spatial information including information about points sampled from a ray or from a three-dimensional model,

obtain volume-rendering parameters by inputting the spatial information of the sampling data to a neural network model, which generates the volume-rendering parameters, the volume-rendering parameters corresponding to the points, respectively,

obtain a regularization term quantifying an amount of information in the volume-rendering parameters,

perform volume rendering based on the volume-rendering parameters, and

train the neural network model to minimize a loss function, wherein the loss function is determined based on the regularization term and is determined based on a difference between a ground truth image and an image that is estimated according to the volume rendering.

18 . The electronic device of claim 17 , wherein the instructions are further configured to cause the one or more processors to train the neural network model such that the distribution of the volume-rendering parameters has a predetermined feature.

19 . The electronic device of claim 17 , wherein the instructions are further configured to cause the one or more processors to train the neural network model such that the distribution of the volume-rendering parameters is clustered on a surface of a scene.

20 . The electronic device of claim 17 , wherein the instructions are further configured to cause the one or more processors to obtain the regularization term based on a metric for quantifying a feature of the distribution of the volume-rendering parameters.

21 . An electronic device comprising:

one or more processors configured to

obtain spatial information of sampling data, the sampling data sampled from a ray or a three-dimensional (3D) model, and

perform volume rendering of the 3D model by inputting the spatial information of the sampling data to a neural network model that generates volume rendering parameters, wherein

the spatial information of the sampling data is determined based on a regularization term quantifying an amount of information in the volume rendering parameters.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 13, 2023
From: JANG, SEOKHWAN; KANG, NAHYUP; KIM, JIYEON; MOON, HYEWON; SAGONG, DONGHOON; SON, MINJUNG
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 062375/0095 →
Priority Claims (1)
KR 10-2022-0099247 · Aug 9, 2022 · national
Continuity (1)
Related Publication 20240054716A1 · Feb 15, 2024
References Cited (14)
US 20090219287A1 · Wang · 2009 [cited by examiner]
US 20220036602A1 · Duckworth et al. · 2022 [cited by applicant]
CN 113888689A · 2022 [cited by applicant]
Author: Wang et al.; Title: Image and Distribution Based vol. Rendering for Large Data Sets; Publication: IEEE; Source: https://ieeexplore.ieee.org/stamp/stamp.jsp?tp=&arnumber=8365973 (Year: 2018). [cited by examiner]
Author: Wang et al.; Title: Statistical Visualization and Analysis of Large Data Using a Value-based Spatial Distribution; Publication: IEEE; Source: https://ieeexplore.ieee.org/stamp/stamp.jsp?tp=&arnumber=8031590 (Yea… [cited by examiner]
Author: Sakhaee et al.; Title: A Statistical Direct vol. Rendering Framework for Visualization of Uncertain Data; Publication: IEEE; Source: https://ieeexplore.ieee.org/stamp/stamp.jsp?tp=&arnumber=7778257 (Year: 2016). [cited by examiner]
Acu, Ana-Maria, et al. “Information potential for some probability density functions.” Applied Mathematics and Computation 389 (2021): 125578, (15 pages). [cited by applicant]
Chang, Di, et al. “RC-MVSNet: Unsupervised Multi-View Stereo with Neural Rendering.” European Conference on Computer Vision. Cham: Springer Nature Switzerland, arXiv:2203.03949v3 [cs.CV] Jul. 13, 2022, (24 pages). [cited by applicant]
Oechsle, Michael, et al. “UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View Reconstruction.” Proceedings of the IEEE/CVF International Conference on Computer Vision. 2021, (11 pages). [cited by applicant]
Extended European search report issued on Dec. 20, 2023, in counterpart European Patent Application No. 23164183.8 (12 pages). [cited by applicant]
Arandjelovic et al. “Nerf in detail: Learning to sample for view synthesis.” [cited by applicant]
Fang et al. “Neusample: Neural sample field for efficient view synthesis.” [cited by applicant]
Mildenhall et al. “Nerf: Representing scenes as neural radiance fields for view synthesis.” [cited by applicant]
Ahn et al. “PANeRF: Pseudo-view Augmentation for Improved Neural Radiance Fields Based on Few-shot Inputs.” [cited by applicant]