IP Library Granted Patent US 12,541,883
Granted Patent B2
US 12,541,883 · App. 18/031,037 · Granted Feb 3, 2026

Method and an apparatus for reconstructing an occupancy map of a point cloud frame

Inventors: Celine Guede (Cesson Sevigne, FR); Julien Ricard (Plouër-sur-Rance, FR); Pierre Andrivon (Liffre, FR); Jean-Eudes Marvie (Betton, FR)
Assignee: InterDigital CE Patent Holdings, SAS
G06T9/001G06T3/4046G06T9/002
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,541,883
App. No.
18/031,037
Granted
Feb 3, 2026
Kind
B2
Abstract

At least one embodiment relates to a method and an apparatus for reconstructing an occupancy map comprising occupancy data of a volumetric content, wherein reconstructing the occupancy map comprises: —decoding the occupancy map at a first resolution, —determining a scale factor as a function of the first resolution, —upscaling the occupancy map by the scale factor, using a neural network.

Claims (36)

1 . A method, comprising:

signaling, in a bitstream, data representative of a volumetric content, the volumetric content being represented as a set of projections onto one or more atlases, the one or more atlases being video-based encoded, the one or more atlases comprising at least one attribute frame, one geometry frame and one occupancy map, and

signaling, in the bitstream, an information indicating a method among at least two methods for upscaling at least one of the one or more atlases when reconstructing the volumetric content,

wherein the at least one of the one or more atlases is an occupancy map, and

wherein one of the at least two methods is a neural network-based upscaling.

2 . The method of claim 1 , wherein the information further comprises at least one of an indicator representative of a neural network model, a size of block used as input of a neural network or a number of iterations according to which upscaling of the occupancy map is iterated.

3 . An apparatus, comprising one or more processors, wherein the one or more processors are configured to:

signal, in a bitstream, data representative of a volumetric content, the volumetric content being represented as a set of projections onto one or more atlases, the one or more atlases being video-based encoded, the one or more atlases comprising at least one attribute frame, one geometry frame and one occupancy map, and

signal, in the bitstream, an information indicating a method among at least two methods for upscaling at least one of the one or more atlases when reconstructing the volumetric content,

wherein the at least one of the one or more atlases is an occupancy map, and

wherein one of the at least two methods is a neural network-based upscaling.

4 . The apparatus of claim 3 , wherein the information further comprises at least one of an indicator representative of a neural network model, a size of block used as input of a neural network or a number of iterations according to which upscaling of the occupancy map is iterated.

5 . A method, comprising reconstructing a volumetric content from a bitstream, the volumetric content being represented as a set of projections onto one or more atlases, the one or more atlases being video-based encoded, the one or more atlases comprising at least one attribute frame, one geometry frame and one occupancy map, wherein reconstructing the volumetric content comprises:

decoding an information indicating a method among at least two methods for upscaling at least one of the one or more atlases, wherein one of the at least two methods is a neural network-based upscaling,

wherein the at least one of the one or more atlases is an occupancy map,

decoding the occupancy map from the bitstream, and

upscaling the occupancy map.

6 . The method of claim 5 , wherein the information indicates use of a neural network for upscaling the occupancy map.

7 . The method of claim 6 , wherein reconstructing the volumetric content further comprises obtaining the neural network based on at least one of an indicator representative of a neural network model, a size of block used as input of the neural network or a scale factor by which the occupancy map is upscaled.

8 . The method of claim 7 , wherein the scale factor is determined based on a resolution of the occupancy map and a nominal resolution of the volumetric content.

9 . The method of claim 5 , wherein the information further comprises an indicator representative of a neural network model.

10 . The method of claim 5 , wherein the information further comprises a size of block used as input of a neural network.

11 . The method of claim 5 , wherein a neural network is associated to a second scale factor, and responsive to a determination that the second scale factor does not match a scale factor to which the occupancy map is to be upscaled, upscaling the occupancy map comprises iterating upscaling of the occupancy map using the neural network.

12 . The method of claim 11 , wherein the information further indicates a number of iterations according to which upscaling of the occupancy map is iterated.

13 . The method of claim 5 , wherein the information is signaled at a patch level, or at an image level, or at a sequence level, or in a supplemental enhancement information message, or in a unit header of the bitstream.

14 . The method of claim 5 , wherein the occupancy map comprises at least one value indicating whether at least one sample in a geometry or attribute frame corresponds to at least one associated sample in the volumetric content.

15 . A non-transitory computer readable storage medium having stored thereon instructions for causing one or more processors to perform the method of claim 5 .

16 . An apparatus, comprising one or more processors, wherein the one or more processors are configured to reconstruct a volumetric content from a bitstream, the volumetric content being represented as a set of projections onto one or more atlases, the one or more atlases being video- based encoded, the one or more atlases comprising at least one attribute frame, one geometry frame and one occupancy map, wherein reconstructing the volumetric content comprises:

decoding an information indicating a method among at least two methods for upscaling at least one of the one or more atlases, wherein one of the at least two methods is a neural network-based upscaling,

wherein the at least one of the one or more atlases is an occupancy map,

decoding the occupancy map from the bitstream, and

upscaling the occupancy map.

17 . The apparatus of claim 16 , comprising at least one of (i) an antenna configured to receive a signal, the signal including data representative of at least one part of a volumetric content, (ii) a band limiter configured to limit the signal to a band of frequencies that includes the data representative of the at least one part of the volumetric content, or (iii) a display configured to display the at least one part of the volumetric content.

18 . The apparatus according to claim 17 , comprising a TV, a cell phone, a tablet, a Set Top Box or a Head Mounted Display.

19 . The apparatus of claim 16 , wherein the information indicates use of a neural network for upscaling the occupancy map.

20 . The apparatus of claim 16 , wherein the information further comprises at least one of an indicator representative of a neural network model, a size of block used as input of a neural network or a number of iterations according to which upscaling of the occupancy map is iterated.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 11, 2023
From: RICARD, JULIEN; ANDRIVON, PIERRE; MARVIE, JEAN-EUDES
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 063282/0068 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 11, 2023
From: GUEDE, CELINE
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 063282/0087 →
Priority Claims (1)
EP 20306184 · Oct 9, 2020 · regional
Continuity (1)
Related Publication 20230377204A1 · Nov 23, 2023
References Cited (38)
US 10242484B1 · Cernigliaro · 2019 [cited by examiner]
US 10818069B2 · Cernigliaro · 2020 [cited by examiner]
US 11158109B2 · Cernigliaro · 2021 [cited by examiner]
US 11170556B2 · Oh · 2021 [cited by examiner]
US 11341687B2 · Oh · 2022 [cited by examiner]
US 11380019B2 · Oh · 2022 [cited by examiner]
US 11394979B2 · Oh · 2022 [cited by examiner]
US 11418564B2 · Oh · 2022 [cited by examiner]
US 11601634B2 · Oh · 2023 [cited by examiner]
US 11818190B2 · Oh · 2023 [cited by examiner]
US 12021910B2 · Lee · 2024 [cited by examiner]
US 12273405B2 · Oh · 2025 [cited by examiner]
US 20190114821A1 · Cernigliaro · 2019 [cited by examiner]
US 20190114822A1 · Cernigliaro · 2019 [cited by examiner]
US 20210005016A1 · Oh · 2021 [cited by examiner]
US 20210209806A1 · Oh · 2021 [cited by examiner]
US 20210227232A1 · Oh · 2021 [cited by examiner]
US 20210320962A1 · Oh · 2021 [cited by examiner]
US 20220060529A1 · Oh · 2022 [cited by examiner]
US 20240048604A1 · Oh · 2024 [cited by examiner]
Anonymous, “Common Test Conditions for Point Cloud Compression”, International Organization for Standardization (ISO), Coding of Moving Pictures and Audio, ISO/IEC JTC 1/SC 29/WG 11, Document: N19324, Alpbach, Austria, … [cited by applicant]
Anonymous, “V-PCC Codec Description”, International Organization for Standardization (ISO), Coding of moving pictures and audio, ISO/IEC JTC 1/SC 29WG 11, Document: N19332, Jun. 17, 2020, 73 pages. [cited by applicant]
Zhou et al., “A Review of Deep Learning for Single Image Super-Resolution”, Institute of Electronics and Electrical Engineers (IEEE), 2019 International Conference on Intelligent Informatics and Biomedical Sciences (ICI… [cited by applicant]
Yang et al., “Deep learning for Single Image Super-Resolution: A Brief Review”, Institute of Electronics and Electrical Engineers (IEEE), IEEE Transactions on Multimedia, vol. 21, Issue No. 12, Dec. 2019, 15 pages. [cited by applicant]
Anonymous, “High Efficiency Video Coding”, ITU-T Telecommunication Standardization Sector of ITU, Series H: Audiovisual and Multimedia Systems, Infrastructure of audiovisual services—Coding of moving video, Recommendati… [cited by applicant]
Quach et al., “Learning Convolutional Transforms for Lossy Point Cloud Geometry Compression”, Institute of Electrical and Electronics Engineers (IEEE), 2019 IEEE International Conference on Image Processing (ICIP), Taip… [cited by applicant]
Smith et al., “3D Object Super-Resolution”, ARXIV Library, Cornell University, Feb. 27, 2018, 9 pages. [cited by applicant]
Anonymous, “V-PCC performance evaluation and anchor results”, International Organization for Standardization (ISO), Coding of Moving Pictures and Audio, ISO/IEC JTC 1/SC29 WG11, Document: w19087, Brussels, Belgium, Jan.… [cited by applicant]
Qian et al., “PU-GCN: Point Cloud Upsampling using Graph Convolutional Networks”, ARXIV Library, Cornell University, Mar. 28, 2020, 27 pages. [cited by applicant]
Jia et al., “Convolutional Neural Network-based Occupancy Map Accuracy Improvement for Video-based Point Cloud Compression”, Institute of Electronics and Electrical Engineers (IEEE), IEEE Transactions on Multimedia, vol… [cited by applicant]
Anonymous, “svn_HEVC Software Revision 4998: /tags/HM-16.18+SCM-8.7”, URL: https://hevc.hhi.fraunhofer.de/svn/svn_HEVCSoftware/tags/HM-16.18+SCM-8.7/, Jan. 2018, 1 page. [cited by applicant]
Ha et al., “Deep Learning Based Single Image Super-resolution: A Survey”, International Journal of Automation and Computing, vol. 16, Issue No. 4, Jul. 19, 2019, 14 pages. [cited by applicant]
Wu et al., “Point Cloud Super Resolution with Adversarial Residual Graph Networks”, ARXIV Library, Cornell University, Aug. 6, 2019, 10 pages. [cited by applicant]
Cai et al., “[VPCC][New proposal] Adaptive occupancy map up-sampling”, International Organization for Standardization (ISO), Coding of Moving Pictures and Audio, ISO/IEC JTC1/SC29/WG11 MPEG2019/, Document: m46455, Marra… [cited by applicant]
Anonymous, Information Technology—Coded Representation of Immersive Media—Part 5: Visual Volumetric Video-based Coding (V3C) and Video-based Point Cloud Compression (V-PCC), International Organization for Standardizatio… [cited by applicant]
Anonymous, “Information technology—Coded Representation of Immersive Media—Part 5: Visual Volumetric Video-based Coding (V3C) and Video-based Point Cloud Compression (V-PCC)”, International Organization for Standardizat… [cited by applicant]
Anonymous, “Information Technology—Coded Representation of Immersive Media—Part 5: Video-based Point Cloud Compression”, International Organization for Standardization, MPEG Document N18030, ISO/IEC 23090-5:2018(E), Oct… [cited by applicant]
Hoppe et al., “Surface Reconstruction from Unorganized Points”, SIGGRAPH '92: Proceedings of the 19th Annual Conference on Computer Graphics and Interactive Techniques, Jul. 1992, 8 pages. [cited by applicant]