IP Library › Granted Patent US 12,579,734
Granted Patent B2
US 12,579,734 · App. 18/641,421 · Granted Mar 17, 2026

Method for rendering viewpoints and electronic device

Inventor: Xuan Sun (Beijing, CN)
Assignee: BOE Technology Group Co., Ltd.
G06T15/20G06T5/50G06T5/60G06T2207/20084G06T2207/20221
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,579,734
App. No.
18/641,421
Granted
Mar 17, 2026
Kind
B2
Abstract

A method for rendering viewpoints is provided. The method for rendering viewpoints of the present disclosure includes: acquiring a first initial feature of a two-dimensional image and a second initial feature of a depth image corresponding to the two-dimensional image by performing initial feature extraction on the two-dimensional image and the depth image; acquiring a first initial dimension reduction feature by splicing the first initial feature and the second initial feature in a channel dimension and performing channel dimension reduction; acquiring a fusion feature by performing image distortion and restoration on the first initial dimension reduction feature for multiple times; and generating a plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature.

Claims (63)

1 . A method for rendering viewpoints, comprising:

acquiring a first initial feature of a two-dimensional image and a second initial feature of a depth image corresponding to the two-dimensional image by performing initial feature extraction on the two-dimensional image and the depth image;

acquiring a first initial dimension reduction feature by splicing the first initial feature and the second initial feature in a channel dimension and performing channel dimension reduction;

acquiring a fusion feature by performing image distortion and restoration on the first initial dimension reduction feature for multiple times; and

generating a plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature;

wherein after generating the plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature, the method further comprises;

acquiring a third initial feature and a fourth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a second initial dimension reduction feature and a third initial dimension reduction feature by performing width and height dimension down-sampling on the third initial feature and the fourth initial feature, reducing the width and height dimensions to 1, and reducing the channel dimension to 1; and

discriminating the second initial dimension reduction feature and the third initial dimension reduction feature based on the label image.

2 . The method for rendering viewpoints according to claim 1 , wherein after generating the plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature, the method further comprises:

acquiring a fifth initial feature and a sixth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a fourth initial dimension reduction feature and a fifth initial dimension reduction feature by performing width and height dimension down-sampling on the fifth initial feature and the sixth initial feature and reducing the channel dimension to 1; and

discriminating the fourth initial dimension reduction feature and the fifth initial dimension reduction feature based on the label image.

3 . The method for rendering viewpoints according to claim 2 , wherein a number of groups of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature is a plurality of groups; wherein

each group of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature corresponds to discriminated pixel blocks from large to small.

4 . The method for rendering viewpoints according to claim 1 , wherein before acquiring the first initial feature of the two-dimensional image and the second initial feature of the depth image corresponding to the two-dimensional image by performing initial feature extraction on the two-dimensional image and the depth image, the method further comprises:

generating the depth image corresponding to the two-dimensional image by inputting the two-dimensional image into a monocular depth estimation model.

5 . The method for rendering viewpoints according to claim 1 , wherein after generating the plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature, the method further comprises:

acquiring a synthesized viewpoint image by splicing the plurality of the viewpoint images in a width dimension; and

generating a three-dimensional image by interleaving the synthesized viewpoint image.

6 . An electronic device, comprising:

at least one processor; and

a memory stores one or more computer programs executable by the at least one processor,

wherein the at least one processor, when loading and executing the one or more computer programs, is caused to perform:

acquiring a first initial feature of a two-dimensional image and a second initial feature of a depth image corresponding to the two-dimensional image by performing initial feature extraction on the two-dimensional image and the depth image;

acquiring a first initial dimension reduction feature by splicing the first initial feature and the second initial feature in a channel dimension and performing channel dimension reduction;

acquiring a fusion feature by performing image distortion and restoration on the first initial dimension reduction feature for multiple times; and

generating a plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature;

wherein the at least one processor, when loading and executing the one or more computer programs, is caused to perform:

acquiring a third initial feature and a fourth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a second initial dimension reduction feature and a third initial dimension reduction feature by performing width and height dimension down-sampling on the third initial feature and the fourth initial feature, reducing the width and height dimensions to 1, and reducing the channel dimension to 1; and

discriminating the second initial dimension reduction feature and the third initial dimension reduction feature based on the label image.

7 . The electronic device according to claim 6 , wherein the at least one processor, when loading and executing the one or more computer programs, is caused to perform:

acquiring a fifth initial feature and a sixth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a fourth initial dimension reduction feature and a fifth initial dimension reduction feature by performing width and height dimension down-sampling on the fifth initial feature and the sixth initial feature and reducing the channel dimension to 1; and

discriminating the fourth initial dimension reduction feature and the fifth initial dimension reduction feature based on the label image.

8 . The electronic device according to claim 7 , wherein a number of groups of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature is a plurality of groups; wherein

each group of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature corresponds to discriminated pixel blocks from large to small.

9 . The electronic device according to claim 6 , wherein the at least one processor, when loading and executing the one or more computer programs, is caused to perform:

generating the depth image corresponding to the two-dimensional image by inputting the two-dimensional image into a monocular depth estimation model.

10 . The electronic device according to claim 6 , wherein the at least one processor, when loading and executing the one or more computer programs, is caused to perform:

acquiring a synthesized viewpoint image by splicing the plurality of the viewpoint images in a width dimension; and

generating a three-dimensional image by interleaving the synthesized viewpoint image.

11 . A non-transitory computer-readable storage medium storing at least one computer program therein, wherein the at least one computer program, when executed by a processor, causes the processor to perform:

acquiring a first initial feature of a two-dimensional image and a second initial feature of a depth image corresponding to the two-dimensional image by performing initial feature extraction on the two-dimensional image and the depth image;

acquiring a first initial dimension reduction feature by splicing the first initial feature and the second initial feature in a channel dimension and performing channel dimension reduction;

acquiring a fusion feature by performing image distortion and restoration on the first initial dimension reduction feature for multiple times; and

generating a plurality of viewpoint images by performing fusion and channel dimension reduction on the fusion feature;

wherein the at least one computer program, when executed by a processor, causes the processor to perform:

acquiring a third initial feature and a fourth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a second initial dimension reduction feature and a third initial dimension reduction feature by performing width and height dimension down-sampling on the third initial feature and the fourth initial feature, reducing the width and height dimensions to 1, and reducing the channel dimension to 1; and

discriminating the second initial dimension reduction feature and the third initial dimension reduction feature based on the label image.

12 . The non-transitory computer-readable storage medium according to claim 11 , wherein the at least one computer program, when executed by a processor, causes the processor to perform:

acquiring a fifth initial feature and a sixth initial feature by performing initial feature extraction on the viewpoint image and a corresponding label image;

acquiring a fourth initial dimension reduction feature and a fifth initial dimension reduction feature by performing width and height dimension down-sampling on the fifth initial feature and the sixth initial feature and reducing the channel dimension to 1; and

discriminating the fourth initial dimension reduction feature and the fifth initial dimension reduction feature based on the label image.

13 . The non-transitory computer-readable storage medium according to claim 12 , wherein a number of groups of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature is a plurality of groups; wherein

each group of the fourth initial dimension reduction feature and the fifth initial dimension reduction feature corresponds to discriminated pixel blocks from large to small.

14 . The non-transitory computer-readable storage medium according to claim 11 , wherein the at least one computer program, when executed by a processor, causes the processor to perform:

generating the depth image corresponding to the two-dimensional image by inputting the two-dimensional image into a monocular depth estimation model.

15 . The non-transitory computer-readable storage medium according to claim 11 , wherein the at least one computer program, when executed by a processor, causes the processor to perform:

acquiring a synthesized viewpoint image by splicing the plurality of the viewpoint images in a width dimension; and

generating a three-dimensional image by interleaving the synthesized viewpoint image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 21, 2024
From: SUN, XUAN
To: BOE TECHNOLOGY GROUP CO., LTD.
Reel/Frame 067174/0745 →
Continuity (2)
Continuation PCTCN2023112869 · Aug 14, 2023
Related Publication 20250061643A1 · Feb 20, 2025
References Cited (26)
US 12423786B2 · Zhang · 2025 [cited by examiner]
US 20200160546A1 · Gu · 2020 [cited by examiner]
US 20210264632A1 · Tankovich · 2021 [cited by examiner]
US 20220277472A1 · Birchfield · 2022 [cited by examiner]
US 20230092248A1 · Xiong · 2023 [cited by examiner]
US 20230154055A1 · Besenbruch · 2023 [cited by examiner]
US 20230360241A1 · Sayed · 2023 [cited by examiner]
US 20240061075A1 · Popov · 2024 [cited by examiner]
US 20240104345A1 · Peng · 2024 [cited by examiner]
US 20240217538A1 · Guizilini · 2024 [cited by examiner]
US 20240419382A1 · Dekel · 2024 [cited by examiner]
US 20250095302A1 · Wetmore · 2025 [cited by examiner]
US 20250131647A1 · Talegaonkar · 2025 [cited by examiner]
US 20250148633A1 · Yasarla · 2025 [cited by examiner]
CN 102609974A · 2012 [cited by applicant]
CN 107945282A · 2018 [cited by applicant]
CN 109462747A · 2019 [cited by applicant]
CN 110246146A · 2019 [cited by applicant]
CN 110689599A · 2020 [cited by applicant]
CN 111325693A · 2020 [cited by applicant]
CN 111669564A · 2020 [cited by applicant]
CN 112738495A · 2021 [cited by applicant]
CN 113487664A · 2021 [cited by applicant]
DE 102018130230A1 · 2020 [cited by applicant]
Viswanath et al.; “Dimensionality reduction-based fusion approaches for imaging and non-imaging biomedical data: concepts, workflow, and use-cases”; BMC Medical Imaging (Year: 2017). [cited by examiner]
PCT/CN2023/112869 international search report dated Dec. 23, 2023. [cited by applicant]