IP Library Granted Patent US 12,482,141
Granted Patent B2
US 12,482,141 · App. 17/765,558 · Granted Nov 25, 2025

Method and apparatus for encoding, transmitting and decoding volumetric video

Inventors: Bertrand Chupeau (Rennes, FR); Renaud Dore (Rennes, FR); Franck Thudor (Rennes, FR)
Assignee: InterDigital CE Patent Holdings, SAS
G06T9/001H04N19/17H04N19/182H04N19/186H04N19/30H04N19/597
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,482,141
App. No.
17/765,558
Granted
Nov 25, 2025
Kind
B2
Abstract

Methods, devices and streams are disclosed for encoding a depth atlas representative of the geometry of a volumetric video. Views to be encoded are analyzed to detect regions of the views with simple depth or color, that is regions for which the depth or color has a local variance lower than a given threshold. Resolution of such regions is reduced and the atlas comprises first regions in full resolution and downscaled second regions. Metadata indicating whether a patch is a downscaled region and, if so the downscaling factor, are associated with the atlas in the data stream. The decoder uses these metadata to compose the view from different patches.

Claims (27)

1 . A method for encoding a depth view of a Multiview plus Depth frame (MVD) in an atlas of patches, a patch being a rectangular region of the depth view, the method comprising:

dividing the depth view in first rectangular regions and second rectangular regions based on complexity of geometrical information within the rectangular regions, wherein the second rectangular regions are candidates for downscaling;

downscaling a resolution of the second rectangular regions in a vertical direction by a vertical factor and in a horizontal direction by a horizontal factor to obtain downscaled second rectangular regions; and

packing the first rectangular regions and the downscaled second rectangular regions in the atlas in association with metadata indicating for each patch of the atlas whether the patch is a first rectangular region or a downscaled second rectangular region and, if so, indicating the vertical factor and the horizontal factor used for the downscaling of the resolution.

2 . The method of claim 1 , wherein the vertical and horizontal factors are different for two second rectangular regions.

3 . The method of claim 1 , wherein every pixel of each second rectangular region has a valid depth value.

4 . The method of claim 1 , further comprising:

encoding in a data stream, information indicating that patches have different resolution, the atlas, and the metadata.

5 . A non-transitory computer readable medium storing instructions which, when executed by one or more processors, cause the one or more processors to perform the method of claim 1 .

6 . The method of claim 1 , wherein the first rectangular regions are regions comprising complex geometrical information and the second rectangular regions are regions having a planar surface.

7 . A device for encoding a depth view of a Multiview plus Depth frame (MVD) in an atlas of patches, a patch being a rectangular region of the depth view, the device comprising a processor configured to perform:

dividing the depth view in first rectangular regions and second rectangular regions based on complexity of geometrical information within the rectangular regions, wherein the second rectangular regions are candidates for downscaling;

downscaling a resolution of the second rectangular regions in a vertical direction by a vertical factor and in a horizontal direction by a horizontal factor to obtain downscaled second rectangular regions; and

packing the first rectangular regions and the downscaled second rectangular regions in the atlas in association with metadata indicating for each patch of the atlas whether the patch is a first rectangular region or a downscaled second rectangular region and, if so, indicating the vertical factor and the horizontal factor used for the downscaling of the resolution.

8 . The device of claim 7 , wherein the vertical and horizontal factors are different for two second rectangular regions.

9 . The device of claim 7 , wherein every pixel of each second rectangular region has a valid depth value.

10 . The device of claim 7 , wherein the processor is further configured to encode in a data stream, information indicating that patches have different resolution, the atlas, and the metadata.

11 . The device of claim 7 , wherein the first rectangular regions are regions comprising complex geometrical information and the second rectangular regions are regions having a planar surface.

12 . A method for decoding a depth view of a Multiview plus Depth frame (MVD) from an atlas packing patches, a patch being a rectangular region of the depth view, the method comprising:

obtaining metadata indicating for each patch of the atlas whether the patch is a first patch or a second patch being a downscaled rectangular region, and, if so, indicating a vertical factor and a horizontal factor, wherein the first patch is a rectangular region comprising complex geometrical information and the second patch is a rectangular region having a planar surface;

upscaling resolution of each second patch in a vertical direction by the vertical factor and in a horizontal direction by the horizontal factor; and

composing the depth view from each first patch and each upscaled second patch.

13 . A non-transitory computer readable medium storing instructions which, when executed by one or more processors, cause the one or more processors to perform the method of claim 12 .

14 . A device for decoding a depth view of a Multiview plus Depth frame (MVD) from an atlas packing patches, a patch being a rectangular region of the depth view, the device comprising a processor configured to perform:

obtaining metadata indicating for each patch of the atlas whether the patch is a first patch or a second patch being a downscaled rectangular region, and, if so, indicating a vertical factor and a horizontal factor, wherein the first patch is a rectangular region comprising complex geometrical information and the second patch is a rectangular region having a planar surface;

upscaling resolution of each second patch in a vertical direction by the vertical factor and in a horizontal direction by the horizontal factor; and

composing the depth view from each first patch and each upscaled second patch.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2023
From: INTERDIGITAL VC HOLDINGS FRANCE, SAS
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 064460/0921 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 31, 2022
From: CHUPEAU, BERTRAND; DORE, RENAUD; THUDOR, FRANCK
To: INTERDIGITAL VC HOLDINGS FRANCE, SAS
Reel/Frame 059457/0711 →
Priority Claims (1)
EP 19306264 · Oct 2, 2019 · regional
Continuity (1)
Related Publication 20220343549A1 · Oct 27, 2022
References Cited (21)
US 10424083B2 · Sinharoy · 2019 [cited by examiner]
US 20140267616A1 · Krig · 2014 [cited by applicant]
US 20190371051A1 · Dore et al. · 2019 [cited by applicant]
US 20190373287A1 · Lim et al. · 2019 [cited by applicant]
US 20210006834A1 · Salahieh · 2021 [cited by examiner]
US 20210067757A1 · Yun · 2021 [cited by examiner]
US 20220345756A1 · Kroon · 2022 [cited by examiner]
EP 3457688A1 · 2019 [cited by applicant]
EP 3474562A1 · 2019 [cited by applicant]
KR 1020180028299A · 2018 [cited by applicant]
KR 1020190105011A · 2019 [cited by applicant]
Salahieh et al., “Test Model for Immersive Video”, International Organization for Standardization, ISO/IEC JTC 1/SC 29/WG 11, Coding of Moving Pictures and Audio, Document: N18470, Geneva, Switzerland, March (Year: 2019… [cited by examiner]
Homayouni et al, Content adaptive depth map resampling scheme in multiview video plus depth, 2014 IEEE International Symposium on Circuits and Systems (ISCAS), pp. 538-541 (Year: 2014). [cited by examiner]
Boyce et al, “Working Draft 2 of Metadata of Immersive Video”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document MPEG2018/M18576, Gothenburg, Sweden, Jul. 2019, 38 pages. [cited by applicant]
Boyce et al, “Depth Coding for Immersive Video”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M49342, Gothenburg, Sweden, 6 pages. [cited by applicant]
Salahieh et al., “Test Model for Immersive Video”, International Organisation for Standardisation, ISO/IEC JTC 1/SC29/WG 11, Coding of Moving Pictures and Audio, Document: N18470, Geneva, Switzerland, Mar. 2019, 27 page… [cited by applicant]
Anonymous, “High Efficiency Video Coding”, ITU-T Telecommunication Standardization Sector of ITU, Series H: Audiovisual and Multimedia Systems, Infrastructure of audiovisual services—Coding of moving video, Recommendati… [cited by applicant]
Anonymous, “Information Technology—Coding of Audio-Visual Objects—Part 10: Advanced Video Coding”, International Standard, ISO/IEC 14496-10, Second Edition, Oct. 1, 2004, 280 pages. [cited by applicant]
Anonymous, “Terminal Equipment and Protocols for Telematic Services”, Information Technology—Digital Compression and Coding of Continuous-Tone Still images—Requirements and Guidelines, International Telecommunication Un… [cited by applicant]
ITU_T, “Advanced video coding for generic audiovisual services”, ITU-T H.264, International Telecommunication Union, ITU-T Telecommunication Standardization Sector of ITU, Series H: Audiovisual and Multimedia Systems, I… [cited by applicant]
ITU-T, “High Efficiency Video Coding”, Recommendation ITU-T H.265, Series H: Audiovisual and Multimedia Systems, Infrastructure of Audiovisual Services—Coding of Moving Video, Oct. 2014, 540 pages. [cited by applicant]