IP Library › Granted Patent US 12,666,077
Granted Patent B2
US 12,666,077 · App. 18/289,464 · Granted Jun 23, 2026

Methods and apparatuses for encoding/decoding a volumetric video, methods and apparatus for reconstructing a computer generated hologram

Inventors: Didier Doyen (Cesson-Sevigne, FR); Valter Drazic (Betton, FR); Vincent Brac De La Perriere (Rennes, FR); Guillaume Boisson (Pleumeleuc, FR)
Assignee: InterDigital CE Patent Holdings, SAS
H04N19/597G03H1/08H04N13/388H04N19/176H04N2013/0081
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,666,077
App. No.
18/289,464
Filed
Nov 3, 2023
Granted
Jun 23, 2026
Kind
B2
Art Unit
2484
USPC
375/240.12
Abstract

Methods and apparatuses for encoding/decoding data content representative of a volumetric video are provided, wherein, the encoding/decoding comprises encoding in/decoding from a bitstream, an indicator specifying whether data content has information representative of at least one set of depth layers, the information representative of a set of depth layers specifying a number of depth layers and a depth value for each of the depth layers for a layer-based representation of the volumetric video. Methods and apparatuses for reconstructing Computer Generated Holograms from a reconstructed layered-based representation of the volumetric video are also provided.

Claims (40)

1 . A method, comprising:

decoding, from a bitstream, data content representative of a volumetric video; and

decoding, from the bitstream, an indicator specifying whether the data content has information representative of at least two sets of depth layers, each depth layer in the at least two sets of depth layers associated with a group of pixels at an associated depth, the information representative of the at least two sets of depth layers specifying an indication of a number of sets of depth layers in the at least two sets of depth layers, and for each of the at least two sets of depth layers, a number of depth layers in a respective set of the at least two sets of depth layers and a depth value for each depth layer in the respective set of the at least two sets of depth layers for a layer-based representation of the volumetric video.

2 . The method of claim 1 , wherein responsive to the indicator specifying that the data content has information representative of the at least two sets of depth layers, the method further comprises:

decoding, from the bitstream, the information representative of the at least two sets of depth layers;

selecting a set of depth layers among the at least two sets of depth layers; and

reconstructing the layer-based representation of the volumetric video, the layer-based representation of the volumetric video comprising a number of depth layers and a depth value for each depth layer corresponding to a number of depth layers and a depth value for each depth layer of the selected set of depth layers.

3 . The method of claim 2 , wherein the data content representative of a volumetric video is a sequence of multiple plane images, and wherein the sequence of multiple plane images comprises at least one multiple plane image, the at least one multiple plane image comprising a plurality of layers.

4 . The method of claim 3 , wherein reconstructing the layer-based representation of the volumetric video further comprises:

decoding a subset of patches from an atlas of patches of the sequence of multiple plane images;

assigning each patch of the decoded subset of patches to a nearest depth layer in the selected set of depth layers; and

reconstructing each layer of the selected set of depth layers using an assigned patch.

5 . The method of claim 3 , wherein the indicator specifying whether data content has information representative of the at least two sets of depth layers and the information representative of the at least two sets of depth layers are signaled as a common atlas sequence parameter set.

6 . The method of claim 2 , wherein the data content representative of a volumetric video is a point cloud.

7 . The method of claim 6 , wherein reconstructing the layer-based representation of the volumetric video further comprises:

decoding a subset of samples of the point cloud;

assigning each sample of the decoded subset of samples to a nearest depth layer in the selected set of depth layers; and

reconstructing each layer of the selected set of depth layers using an assigned sample.

8 . The method of claim 7 , wherein the indicator specifying whether data content has information representative of the at least two sets of depth layers and information representative of the at least two sets of depth layers are signaled as general atlas frame parameter set of a raw byte sequence payload syntax.

9 . The method of claim 2 , further comprising reconstructing at least one computer generated hologram from the reconstructed layer-based representation of the volumetric video.

10 . A method, comprising:

encoding in a bitstream, data content representative of a volumetric video; and

encoding in the bitstream, an indicator specifying whether the data content has information representative of at least two sets of depth layers, each depth layer in the at least two sets of depth layers associated with a group of pixels at an associated depth, the information representative of the at least two sets of depth layers specifying an indication of a number of sets of depth layers in the at least two sets of depth layers, and for each of the at least two sets of depth layers, a number of depth layers in a respective set of the at least two sets of depth layers and a depth value for each depth layer in the respective set of the at least two sets of depth layers for a layer-based representation of the volumetric video.

11 . The method of claim 10 , wherein responsive to the indicator specifying that the data content has information representative of the at least two sets of depth layers, the method further comprising encoding the information representative of the at least two sets of depth layers.

12 . The method of claim 11 , wherein the data content representative of the volumetric video is a sequence of multiple plane images, and wherein the sequence of multiple plane images comprises at least one multiple plane image, the at least one multiple plane image comprising a plurality of layers.

13 . The method of claim 12 , wherein the indicator specifying whether the data content has information representative of the at least two sets of depth layers and the information representative of the at least two sets of depth layers are encoded in a common atlas sequence parameter set.

14 . The method of claim 11 , wherein the data content representative of the volumetric video is a point cloud.

15 . The method of claim 14 , wherein the indicator specifying whether the data content has information representative of the at least two sets of depth layers and information representative of the at least two sets of depth layers are signaled as general atlas frame parameter set of a raw byte sequence payload syntax.

16 . An apparatus comprising one or more processors configured for:

decoding, from a bitstream, a data content representative of a volumetric video;

decoding, from the bitstream, an indicator specifying whether the data content has information representative of at least two sets of depth layers, each depth layer in the at least two sets of depth layers associated with a group of pixels at an associated depth, the information representative of the at least two sets of depth layers specifying an indication of a number of sets of depth layers in the at least two sets of depth layers, and for each of the at least two sets of depth layers, a number of depth layers in a respective set of the at least two sets of depth layers and a depth value for each depth layer in the respective set of the at least two sets of depth layers for a layer-based representation of the volumetric video;

responsive to the indicator specifying that the data content has information representative of the at least two sets of depth layers, decoding, from the bitstream, the information representative of the at least two sets of depth layers;

selecting a set of depth layers among the at least two sets of depth layers;

reconstructing the layer-based representation of the volumetric video, the layer-based representation of the volumetric video comprising a number of depth layers and a depth value for each depth layer corresponding to a number of depth layers and a depth value for each depth layer of the selected set of depth layers; and

reconstructing at least one computer generated hologram from the reconstructed layer-based representation of the volumetric video.

17 . The apparatus of claim 16 , wherein selecting the set of depth layers comprises:

determining resources required to reconstruct the layer-based representation of the volumetric video and to reconstruct the at least one computer generated hologram using each of at least one set of depth layers; and

selecting a depth layer set with a largest number of depth layers for which the determined resources required to reconstruct the layer-based representation of the volumetric video and to generate the at least one computer generated hologram fits within a resources budget.

18 . The apparatus of claim 17 , wherein the determined resources comprise at least one of a number of processing cycles, an amount of decoding time, or an amount of memory.

19 . The apparatus of claim 18 , wherein the resources budget comprises at least one of a number of processing cycles available in the apparatus, an amount of decoding time available in the apparatus, or an amount of memory available in the apparatus.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 3, 2023
From: DOYEN, DIDIER; DRAZIC, VALTER; BRAC DE LA PERRIERE, VINCENT; BOISSON, GUILLAUME
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 065453/0310 →
Priority Claims (1)
EP 21305588 · May 6, 2021 · regional
Continuity (1)
Related Publication 20240244259A1 · Jul 18, 2024
References Cited (40)
US 6335765B1 · Daly · 2002 [cited by applicant]
US 8717405B2 · Li · 2014 [cited by applicant]
US 9020241B2 · Leichsenring · 2015 [cited by applicant]
US 9100642B2 · Ha · 2015 [cited by applicant]
US 11184599B2 · Harviainen · 2021 [cited by applicant]
US 20050286759A1 · Zitnick, III · 2005 [cited by applicant]
US 20100202540A1 · Fang · 2010 [cited by applicant]
US 20120127284A1 · Bar-Zeev · 2012 [cited by applicant]
US 20130069932A1 · Ha · 2013 [cited by applicant]
US 20130071012A1 · Leichsenring · 2013 [cited by applicant]
US 20160307372A1 · Pitts · 2016 [cited by applicant]
US 20160352791A1 · Adams · 2016 [cited by applicant]
US 20200219290A1 · Tourapis · 2020 [cited by examiner]
US 20200294271A1 · Lola · 2020 [cited by examiner]
US 20220159298A1 · Boyce · 2022 [cited by examiner]
US 20220191498A1 · Schwarz · 2022 [cited by examiner]
US 20220292763A1 · Ilola · 2022 [cited by examiner]
GB 2449631A · 2008 [cited by applicant]
GB 2452765A · 2009 [cited by applicant]
WO 2009109804A1 · 2009 [cited by applicant]
Petrovic, G., et al. “Framework for Layered 3D Video Streaming” Proc. of 27th Symposium on Information Theory in the Benelux, 08-09-06-2006, Noordwijk, The Netherlands 2006 (8 pages). [cited by applicant]
Shum, Heung-Yeung, et al. “Rendering with concentric mosaics.” In Proceedings of the 26th annual conference on Computer graphics and interactive techniques, 1999 (8 pages). [cited by applicant]
Tao, Michael W., et al. “Depth from combining defocus and correspondence using light-field cameras”, In: Proceedings of the IEEE International Conference on Computer Vision. 2013. p. 673-680 (8 pages). [cited by applicant]
Lloyd, Stuart P., “Least Square Quantization in PCM”, Institute of Electrical and Electronics Engineers (IEEE), IEEE Transactions on Information Theory, vol. 28, Issue 2, Mar. 1982, 9 pages. [cited by applicant]
“Iso/Iec Fdis 23090-5 Visual Volumetric Video-based Coding and Video-based Point Cloud Compression”, International Organization for Standardization, ISO/IEC JTC 1/SC 29/WG11, Document: N19579, Sep. 21, 2020, 353 pages. [cited by applicant]
“Infrastructure of audiovisual services—Coding of moving video—Advanced video coding for generic audiovisual services”, International Telecommunication Union, Telecommunication Standardization Sector of ITU, Series H: A… [cited by applicant]
Fleureau et al., “MIV CE1-Related—Activation of Transparency Attribute and MPI Profile in MIV”, International Organization for Standardization (ISO), Coding of Moving Pictures and Audio, ISO/IEC JTC 1/SC 29/WG 4, Docume… [cited by applicant]
Penner et al., “Soft 3D reconstruction for view synthesis”, Association for Computing Machinery (ACM), ACM Transactions on Graphics, vol. 36, Issue 6, Article No. 235, Dec. 2017, 11 pages. [cited by applicant]
Gilles et al., “Real-time layer-based computer-generated hologram calculation for the Fourier transform optical system”, Applied Optics, vol. 57, Issue No. 29, Oct. 10, 2018, 10 pages. [cited by applicant]
“Information Technology—Digital Compression and Coding of Continuous-Tone Still Images—Requirements and Guidelines, Technical Corrigendum 1: Patent Information Update”, International Telecommunication Union, Telecommuni… [cited by applicant]
Dore et al., “Activation of Transparency Attribute and MPI Profile in MIV”, International Organization for Standardization (ISO), ISO/IEC JTC 1/SC 29/WG 4, MPEG M55173, MPEG133, Online Meeting, Jan. 2021, 5 pages. [cited by applicant]
Vandame et al., “Pipeline for Real-Time Video View Synthesis”, Institute of Electrical and Electronics Engineers (IEEE), 2020 IEEE International Conference on Multimedia & Expo Workshops (ICMEW), Jul. 6, 2020, London, U… [cited by applicant]
“Infrastructure of audiovisual services—Coding of moving video—Reference software for ITU-T H.265 high efficiency video coding”, International Telecommunication Union, Telecommunication Standardization Sector of ITU, Se… [cited by applicant]
“Transmission of Non-Telephone Signals: Information Technology—Generic Coding of Moving Pictures and Associated Audio Information: Systems”, International Telecommunication Union, Telecommunication Standardization Secto… [cited by applicant]
“Infrastructure of audiovisual services—Coding of moving video—Information Technology—Generic coding of moving pictures and associated audio information: Video”, International Telecommunication Union, Telecommunication … [cited by applicant]
“Information technology—Coded Representation of Immersive Media—Part 5: Visual Volumetric Video-Based Coding (V3C) and Video-based Point Cloud Compression (V-PCC)”, International Organization for Standardization and the… [cited by applicant]
“Information technology—Coded Representation of Immersive Media—Part 12: Immersive Video, DIS Stage”, International Organization for Standardization and the International Electrotechnical Commission (ISO/IEC), ISO/IEC J… [cited by applicant]
“Infrastructure of audiovisual services—Coding of moving video—High Efficiency Video Coding”, International Telecommunication Union, Telecommunication Standardization Sector of ITU, Series H: Audiovisual and Multimedia … [cited by applicant]
MacQueen, J., “Some Methods for Classification and Analysis of Multivariate Observations”, University of California, Press, Proceedings of Fifth Berkeley Symposium on Mathematical Statistics and Probability, vol. 1: Sta… [cited by applicant]
“Committee Draft of MPEG Immersive Video”, International Organization for Standardization, ISO/IEC JTC 1/SC 29/WG 11, N19482, Jul. 3, 2020, 82 pages. [cited by applicant]