IP Library Granted Patent US 12,368,831
Granted Patent B2
US 12,368,831 · App. 18/012,083 · Granted Jul 22, 2025

Method and apparatus for encoding and decoding volumetric content in and from a data stream

Inventors: Julien Fleureau (Rennes, FR); Renaud Dore (Rennes, FR); Bertrand Chupeau (Rennes, FR); Franck Thudor (Rennes, FR)
Assignee: InterDigital CE Patent Holdings, SAS
H04N13/161H04N13/178H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,368,831
App. No.
18/012,083
Granted
Jul 22, 2025
Kind
B2
Abstract

Methods and apparatus for encoding and decoding a volumetric scene are disclosed. A set of attribute and geometry patches is obtained by projecting samples of the volumetric scene onto the patches according to projection parameters. If the geometry patch is comparable to a planar layer located at a constant depth according to the projection parameters, only the attribute patch is packed in an attribute atlas image and the depth value is encoded in metadata. Otherwise, both attribute and geometry patches are packed in an atlas. At the decoding, if metadata for an attribute patch indicates that its geometry may be determined from the projection parameters and a constant depth, the attributes are inverse projected on a planar layer. Otherwise, attributes are inverse projected according to the associated geometry patch.

Claims (42)

1. A method of synthetizing a viewport image, the method comprising:

obtaining, from a data stream, an attribute atlas image, a geometry atlas image, an atlas image packing patch pictures, a patch picture being a projection of a sample of a three-dimensional scene, and metadata comprising, for an attribute patch picture of the attribute atlas image:

projection parameters associated with the attribute patch picture, and

information indicating if the attribute patch picture is associated with a geometry patch picture of the geometry atlas image or if the attribute patch picture is associated with a depth value encoded in the metadata;

on a condition that an attribute patch picture is associated with a geometry patch picture, synthetizing the viewport image by inverse-projecting pixels of the attribute patch picture at a location determined by the geometry patch picture and projection parameters associated with the attribute patch picture; and

on a condition that an attribute patch picture is associated with a depth value, synthetizing the viewport image by inverse-projecting pixels of the attribute patch picture at a location determined by the depth value and projection parameters associated with the attribute patch picture.

2. The method of claim 1 , wherein pixels of the attribute atlas image encode two values for two different attributes, the two different attributes being inverse-projected together.

3. The method of claim 1 , wherein the data stream comprises two attribute atlases encoded according to a same packing layout, the metadata being generated for a pair of attribute patch pictures of each attribute atlases, the pair of attribute patch pictures being inverse-projected together.

4. A device for synthetizing the viewport image, the device comprising a processor configured for:

obtaining, from a data stream, an attribute atlas image, a geometry atlas image, an atlas image packing patch pictures, a patch picture being a projection of a sample of a three-dimensional scene, and metadata comprising, for an attribute patch picture of the attribute atlas image:

projection parameters associated with the attribute patch picture, and

information indicating if the attribute patch picture is associated with a geometry patch of the geometry atlas image or if the attribute patch picture is associated with a depth value encoded in the metadata;

on a condition that an attribute patch picture is associated with a geometry patch picture, synthetizing the viewport image by inverse-projecting pixels of the attribute patch picture at a location determined by the geometry patch picture and projection parameters associated with the attribute patch picture; and

on a condition that an attribute patch picture is associated with a depth value, synthetizing the viewport image by inverse-projecting pixels of the attribute patch picture at a location determined by the depth value and projection parameters associated with the attribute patch picture.

5. The device of claim 4 , wherein pixels of the attribute atlas image encode two values for two different attributes, the two different attributes being inverse-projected together.

6. The device of claim 4 , wherein the data stream comprises two attribute atlases encoded according to a same packing layout, the metadata being generated for a pair of attribute patch pictures of each attribute atlases, the pair of attribute patch pictures being inverse-projected together.

7. A method comprising:

obtaining a set of attribute patch pictures associated with a geometry patch picture, attribute and geometry patch pictures being obtained by projecting samples of a three-dimensional scene according to projection parameters;

for an attribute patch picture of the set of attribute patch pictures,

packing the attribute patch picture in an attribute atlas image; and

if the geometry patch picture associated with the attribute patch picture is comparable to a planar layer at a location determined by a depth value and the projection parameters, generating metadata comprising the projection parameters, the depth value and an information indicating that the attribute patch picture is associated with the depth value or,

in the other case, packing the geometry patch picture in a geometry atlas image and generating metadata comprising the projection parameters and information indicating that the attribute patch picture is associated with the geometry patch picture; and

encoding the attribute atlas image, the geometry atlas image and the metadata in a data stream.

8. The method of claim 7 , wherein pixels of the attribute atlas image encode two values for two different attributes.

9. The method of claim 7 , wherein the set of attribute patch pictures comprises pairs of attribute patch pictures for two different attributes, the pairs of attribute patch pictures being packed in two attribute atlases according to a same packing layout and the metadata being generated for a pair of attribute patch pictures.

10. A device comprising a processor configured for:

obtaining a set of attribute patch pictures associated with a geometry patch picture attribute and geometry patch pictures being obtained by projecting samples of a three-dimensional scene according to projection parameters;

for an attribute patch picture of the set of attribute patch pictures,

packing the attribute patch picture in an attribute atlas image; and

if the geometry patch picture associated with the attribute patch picture is comparable to a planar layer at a location determined by a depth value and the projection parameters, generating metadata comprising the projection parameters, the depth value and an information indicating that the attribute patch picture is associated with the depth value or,

in the other case, packing the geometry patch picture in a geometry atlas image and generating metadata comprising the projection parameters and information indicating that the attribute patch picture is associated with the geometry patch picture; and

encoding the attribute atlas image, the geometry atlas image and the metadata in a data stream.

11. The device of claim 10 , wherein pixels of the attribute atlas image encode two values for two different attributes.

12. The device of claim 10 , wherein the set of attribute patch pictures comprises pairs of attribute patch pictures for two different attributes, the pairs of attribute patch pictures being packed in two attribute atlases according to a same packing layout and the metadata being generated for a pair of attribute patch pictures.

13. A non-transitory computer readable medium containing an attribute atlas image, a geometry atlas image, an atlas image packing patch pictures, a patch picture being a projection of a sample of a three-dimensional scene, and metadata comprising, for an attribute patch picture of the attribute atlas image:

projection parameters associated with the attribute patch picture; and

information representative of an inverse-projecting mode indicating if the attribute patch picture is associated with a geometry patch of the geometry atlas image or if the attribute patch picture is associated with a depth value encoded in the metadata.

14. The non-transitory computer readable medium of claim 13 , wherein pixels of the attribute atlas image encode two values for two different attributes.

15. The non-transitory computer readable medium of claim 13 , comprising two attribute atlases encoded according to a same packing layout.

16. A non-transitory computer readable medium storing instructions that, when executed by one or more processors, perform the method of claim 1 .

17. The non-transitory computer readable medium of claim 16 , wherein pixels of the attribute atlas image encode two values for two different attributes, the two different attributes being inverse-projected together.

18. The non-transitory computer readable medium of claim 16 , wherein the data stream comprises two attribute atlases encoded according to a same packing layout, the metadata being generated for a pair of attribute patch pictures of each attribute atlases, the pair of attribute patch pictures being inverse-projected together.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2023
From: INTERDIGITAL VC HOLDINGS FRANCE, SAS
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 064460/0921 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 21, 2022
From: FLEUREAU, JULIEN; DORE, RENAUD; CHUPEAU, BERTRAND; THUDOR, FRANCK
To: INTERDIGITAL VC HOLDINGS FRANCE, SAS
Reel/Frame 062173/0589 →
Priority Claims (1)
EP 20305695 · Jun 24, 2020 · regional
Continuity (1)
Related Publication 20230239451A1 · Jul 27, 2023
References Cited (12)
US 20210074029A1 · Fleureau · 2021 [cited by examiner]
US 20220078486A1 · Hannuksela · 2022 [cited by examiner]
EP 3515068A1 · 2019 [cited by applicant]
WO WO2020055869A1 · 2020 [cited by applicant]
Tourapis et al., “H.264/14496-10 AVC Reference Software Manual”, Joint Video Team (JVT) of ISO/IEC MPEG & ITU-T VCEG (ISO/IEC JTC1/SC29/WG11 and ITU-T SG16 Q.6), Document No. JVT-AE010, 31st Meeting: London, Great Brita… [cited by applicant]
Salahieh et al., “Test Model for Immersive Video”, International Organisation for Standardisation, ISO/IEC JTC 1/SC 29/WG 11, Coding of Moving Pictures and Audio, Document: N18470, Geneva, Switzerland, Mar. 2019, 27 pag… [cited by applicant]
Fleureau et al., “Cardboard patches for MIV”, International Organisation for Standardisation, ISO/IEC JTC 1/SC 29/WG 11, Coding of Moving Pictures and Audio, Document: MPEG2020/M54417, online, Jul. 2020, 4 pages. [cited by applicant]
Anonymous, “AV1 Codec Library”, Alliance for Open Media, Url: https://aomedia.googlesource.com/aom, 13 pages. [cited by applicant]
Anonymous, Terminal Equipment and Protocols for Telematic Services—Information Technology—Digital Compression and Coding of Continuous-Tone Still Images—Requirements and Guidelines, International Telecommunication Union… [cited by applicant]
Fleureau et al., “Description of Technicolor Intel response to MPEG-I 3DoF+ Call for Proposal”, International Organisation for Standardisation, ISO/IEC JTC1/SC29/WG11, Coding of Moving Pictures and Audio, Document: MPEG… [cited by applicant]
Anonymous, “Series H: Audiovisual and Multimedia Systems—infrastructure of audiovisual services—Coding of moving video: High Efficiency Video Coding”, International Telecommunication Union, Recommendation ITU-T H.265, O… [cited by applicant]
ITU-T, “High Efficiency Video Coding”, Recommendation ITU-T H.265, Series H: Audiovisual and Multimedia Systems, Infrastructure of Audiovisual Services—Coding of Moving Video, Feb. 2018, 692 pages. [cited by applicant]