IP Library › Granted Patent US 12,739,355
Granted Patent B2
US 12,739,355 · App. 17/763,745 · Granted Sep 15, 2026

Method and apparatus for encoding, transmitting and decoding volumetric video

Inventors: Julien Fleureau (Rennes, FR); Franck Thudor (Rennes, FR); Renaud Dore (Rennes, FR)
Assignee: INTERDIGITAL CE PATENT HOLDINGS, SAS
H04N13/161H04N13/282H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,739,355
App. No.
17/763,745
Granted
Sep 15, 2026
Kind
B2
Abstract

Methods, devices and stream for encoding, decoding and transmitting a multi-views frame are disclosed. A non-pruned MVD frame is obtained and an acyclic graph representing pruning precedence relations between views is determined. The MVD is pruned by using these precedence relations. The pruned MVD and data representative of the graph are encoded in the data stream. At the decoding, the contribution of each view for a pixel of a viewport frame to generate is determined as a function of the decoded pruning graph.

Claims (37)

1 . A method for encoding views of a multi-view frame in a data stream, the method comprising:

obtaining an explicitly stored acyclic graph data structure linking views of the multi-view frame, links of the acyclic graph being representative of a pruning precedence relation, at least one basic view of the multi-view frame having no pruning precedence link;

pruning pixels of views of the multi-view frame in an order determined so that a given view is pruned after views linked to the given view by a pruning precedence link, wherein at least one pixel of the given view is pruned when the at least one pixel corresponds to information encoded in a pixel of a view transitively linked to the given view by a pruning precedence link such that the pruning is executable in a single pass along paths of the acyclic graph without requiring sequential multi-phase pruning; and

encoding the acyclic graph as a graph structure including identifiers for nodes and edges, the at least one basic view, and the pruned views of the multi-view frame in the data stream.

2 . The method of claim 1 , wherein the pruning pixels of views comprises replacing the value of the pixels by a determined value.

3 . The method of claim 1 , wherein the acyclic graph is signaled in the data stream as a list comprising, for each view of the multi-view frame, linked views.

4 . The method of claim 1 , wherein the acyclic graph encodes both basic and additional views in a single unified pruning hierarchy, such that pruning precedence applies uniformly to all views.

5 . A device for encoding views of a multi-view frame in a data stream, the device comprising a processor configured for:

obtaining an explicitly stored acyclic graph data structure linking views of the multi-view frame, links of the acyclic graph being representative of a pruning precedence relation, at least one basic view of the multi-view frame having no pruning precedence link;

pruning pixels of views of the multi-view frame in an order determined so that a given view is pruned after views linked to the given view by a pruning precedence link; wherein at least one pixel of the given view is pruned when the at least one pixel corresponds to information encoded in a pixel of a view transitively linked to the given view by a pruning precedence link such that the pruning is executable in a single pass along paths of the acyclic graph without requiring sequential multi-phase pruning; and

encoding the acyclic graph as a graph structure including identifiers for nodes and edges, the at least one basic view, and the pruned views of the multi-view frame in the data stream.

6 . The device of claim 5 , wherein the pruning pixels of views comprises replacing the value of the pixels by a determined value.

7 . The method of claim 5 , wherein the acyclic graph is signaled in the data stream as a list comprising, for each view of the multi-view frame, linked views.

8 . A method of decoding views of a multi-view frame from a data stream, the method comprising:

obtaining pruned views of the multi-view frame from the data stream and at least one basic view being unpruned;

obtaining an explicitly stored acyclic graph data structure from the data stream, the acyclic graph linking views of the multi-view frame, links of the acyclic graph being representative of a pruning precedence relation, the at least one basic view of the multi-view frame having no pruning precedence link; and

generating a viewport frame according to a viewing pose by determining the contribution of each view of the multi-view frame as a function of the pruning precedence relations of the acyclic graph, wherein the reconstruction of pruned pixels from the acyclic graph is performed in a single pass along graph paths without requiring sequential multi-phase pruning and without use of patch-group metadata, weights per patch, or per-visibility lists.

9 . The method of claim 8 , wherein a pruned pixel of a pruned view has a determined value.

10 . The method of claim 9 , wherein the determined value is not a color value or a depth value.

11 . The method of claim 10 , wherein the determined value comprises a unique marker value indicating pruning status that is neither a valid pixel color nor a depth, enabling efficient detection at the decoder.

12 . The method of claim 8 , wherein the acyclic graph is signaled in the data stream as a list comprising, for each view of the multi-view frame, linked views.

13 . The method of claim 8 , wherein reconstruction of pruned pixels from the acyclic graph is performed without use of patch-group metadata, weights per patch, or per-visibility lists.

14 . A device for decoding views of a multi-view frame from a data stream, the device comprising a processor configured for:

obtaining pruned views of the multi-view frame from the data stream and at least one basic view being unpruned;

obtaining an explicitly stored acyclic graph data structure from the data stream, the acyclic graph linking views of the multi-view frame, links of the acyclic graph being representative of a pruning precedence relation, the at least one basic view of the multi-view frame having no pruning precedence link; and

generating a viewport frame according to a viewing pose by determining the contribution of each view of the multi-view frame as a function of the pruning precedence relations of the acyclic graph, wherein the reconstruction of pruned pixels from the acyclic graph is performed in a single pass along graph paths without requiring sequential multi-phase pruning and without use of patch-group metadata, weights per patch, or per-visibility lists.

15 . The device of claim 14 , wherein a pruned pixel of a pruned view has a determined value.

16 . The device of claim 15 , wherein the determined value is not a color value or a depth value.

17 . The device of claim 14 , wherein the acyclic graph is signaled in the data stream as a list comprising, for each view of the multi-view frame, linked views.

18 . A non-transitory computer readable storage medium comprising instructions which, when executed by one or more processors, cause the one or more processors to perform the method of claim 1 .

19 . A non-transitory computer readable storage medium comprising instructions which, when executed by one or more processors, cause the one or more processors to perform the method of claim 8 .

20 . A non-transitory computer readable storage medium comprising instructions which, when executed by one or more processors, cause the one or more processors to:

store data representative of views of a multi-view frame comprising pruned views and at least one basic view being unpruned; and

store data representative of an explicitly stored acyclic graph data structure linking views of the multi-view frame, links of the graph being representative of a view pruning precedence relation, the at least one basic view of the multi-view frame having no pruning precedence link, wherein the acyclic graph encodes both basic and additional views in a single unified pruning hierarchy enabling pruning and reconstruction in a single pass without requiring sequential multi-phase pruning or reliance on patch-group metadata, weights per patch, or per-visibility lists.

21 . The non-transitory computer readable storage medium of claim 20 , wherein a pruned pixel of a pruned view has a determined value.

22 . The non-transitory computer readable storage medium of claim 21 , wherein the determined value is not a color value or a depth value.

23 . The non-transitory computer readable storage medium of claim 20 , wherein the acyclic graph is signaled in the data stream as a list comprising, for each view of the multi-view frame, linked views.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2023
From: INTERDIGITAL VC HOLDINGS FRANCE, SAS
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 064460/0921 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 25, 2022
From: FLEUREAU, JULIEN; THUDOR, FRANCK; DORE, RENAUD
To: INTERDIGITAL VC HOLDINGS FRANCE, SAS
Reel/Frame 059399/0082 →
Priority Claims (2)
EP 19306222 · Sep 30, 2019 · regional
EP 20305005 · Jan 7, 2020 · regional
Continuity (1)
Related Publication 20220368879A1 · Nov 17, 2022
References Cited (39)
US 20030085932A1 · Samra · 2003 [cited by applicant]
US 20130162625A1 · Schmit et al. · 2013 [cited by applicant]
US 20130170558A1 · Zhang · 2013 [cited by applicant]
US 20140111611A1 · Lecroart · 2014 [cited by applicant]
US 20140205015A1 · Rusert et al. · 2014 [cited by applicant]
US 20150009350A1 · Sarwari et al. · 2015 [cited by applicant]
US 20150181229A1 · Lin et al. · 2015 [cited by applicant]
US 20160088281A1 · Newton et al. · 2016 [cited by applicant]
US 20160360200A1 · Shimizu et al. · 2016 [cited by applicant]
US 20170064309A1 · Sethuraman et al. · 2017 [cited by applicant]
US 20180084260A1 · Chien et al. · 2018 [cited by applicant]
US 20180131951A1 · Hannuksela · 2018 [cited by applicant]
US 20180288415A1 · Li · 2018 [cited by examiner]
US 20210006834A1 · Salahieh · 2021 [cited by examiner]
CN 103024597A · 2013 [cited by applicant]
CN 103098468A · 2013 [cited by applicant]
CN 104412587A · 2015 [cited by applicant]
CN 105165008A · 2015 [cited by applicant]
CN 105830443A · 2016 [cited by applicant]
CN 106416250A · 2017 [cited by applicant]
CN 107925777A · 2018 [cited by applicant]
CN 108288270A · 2018 [cited by applicant]
CN 109691106A · 2019 [cited by applicant]
JP 2012505569A · 2012 [cited by applicant]
JP 2014520409A · 2014 [cited by applicant]
WO 2010041998A1 · 2010 [cited by applicant]
Shin et al. “[MPEG-I Visual] CE2-Related: Priority based View Pruning.”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M49150, Gothenburg, Sweden, Jul. 2019, 6 pages. (Year: … [cited by examiner]
Shin et al, “ETRI response to Immersive Video CE-2: Optimized pruning order”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M50030, Geneva, Switzerland, Oct. 2019, 6 pages. [cited by applicant]
Shin et al, “[MPEG-I Visual] CE2-related: Priority based View Pruning”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M49150, Gothenburg, Sweden, Jul. 2019, 6 pages. [cited by applicant]
Fleureau et al, “An Immersive Video Experience with Real-Time View Synthesis Leveraging the Upcoming MIV Distribution Standard”, Institute of Electrical and Electronics Engineers, 2020 IEEE International Conference on M… [cited by applicant]
Fleureau et al, “CE2.6.2 Graph-Based Pruning for Natural Contents”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M52414, Brussels, Belgium, Jan. 2020, 14 pages. [cited by applicant]
Fleureau et al, “CE2.6.2 Graph-Based Pruning for Natural Contents”, International Organization for Standardization, ISO/IEC JTC1/SC29/WG11, Document: MPEG2019/M52414, Brussels, Belgium, Jan. 2020, 24 pages. [cited by applicant]
Salahieh et al., “Test Model for Immersive Video”, International Organisation for Standardisation, ISO/IEC JTC 1/SC 29/WG 11, Coding of Moving Pictures and Audio, Document: N18470, Geneva, Switzerland, Mar. 2019, 27 pag… [cited by applicant]
Anonymous, “Information Technology—Coding of Audio-Visual Objects—Part 10: Advanced Video Coding”, International Standard, ISO/IEC 14496-10, Second Edition, Oct. 1, 2004, 280 pages. [cited by applicant]
Anonymous, “Series H: Audiovisual and Multimedia Systems—infrastructure of audiovisual services—Coding of moving video: High Efficiency Video Coding”, International Telecommunication Union, Recommendation ITU-T H.265, O… [cited by applicant]
Anonymous, “Terminal Equipment and Protocols for Telematic Services”, Information Technology—Digital Compression and Coding of Continuous-Tone Still images—Requirements and Guidelines, International Telecommunication Un… [cited by applicant]
ITU_T, “Advanced video coding for generic audiovisual services”, ITU-T H.264, International Telecommunication Union, ITU-T Telecommunication Standardization Sector of ITU, Series H: Audiovisual and Multimedia Systems, I… [cited by applicant]
He, et al., “JVET AHG Report: 360° video coding tools, software and test conditions (AHG6)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 JVET-00006 V2, Jul. 3-12, 2019, 3pp. [cited by applicant]
Zhao, “Research on 3D video visual quality and enhancement processing”, China Master's Theses Full-text Database (Information Technology Section), Jun. 30, 2014, 131 pages. [cited by applicant]