IP Library Granted Patent US 12,356,002
Granted Patent B2
US 12,356,002 · App. 18/392,158 · Granted Jul 8, 2025

Multi-view coding with effective handling of renderable portions

Inventors: Sebastian Bosse (Berlin, DE); Heiko Schwarz (Panketel, DE); Thomas Wiegand (Berlin, DE); Tobias Hinz (Berlin, DE)
Assignee: Dolby Video Compression, LLC
H04N19/52H04N19/463H04N19/553H04N19/597
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,356,002
App. No.
18/392,158
Granted
Jul 8, 2025
Kind
B2
Abstract

A proposed intermediate way of handling the renderable portion of the first view results in more efficient coding. Instead of omitting the coding of the renderable portion completely, even more efficient coding of multi-view signals entails merely suppressing the coding of the residual signal within the renderable portion, whereas the prediction parameter coding still takes place from the non-renderable portion of the multi-view signal across the renderable portion so that prediction parameters for the renderable portion may be exploited for predicting parameters for the non-renderable portion. The additional coding rate for transmitting the prediction parameters for the renderable portion may be kept low as this merely aims at forming a continuation of the parameter history across the renderable portion to serve as a basis for prediction parameters of other portions of the multi-view signal.

Claims (35)

1. A decoder for reconstructing a multi-view signal from a data stream, comprising:

a data stream extractor configured to extract, using a processor from the data stream, a rendering flag associated with a renderable portion of a first view of the multi-view signal, the rendering flag being indicative of whether the renderable portion is to be replaced with a rendered portion generated, by view synthesis, based on a second view of the multi-view signal; and

a reconstructor configured to reconstruct, using the processor, at least a portion of the first view based on a residual signal, wherein the decoder is configured to decode in a first coding mode or a second coding mode depending on the rendering flag, wherein the decoder is configured to, in the first coding mode, render, by view synthesis, the renderable portion of the first view from the second view to generate the rendered portion and replace the renderable portion of a reconstructed first view by the rendered portion, with not performing the replacement in the second coding mode.

2. The decoder of claim 1 , wherein the data stream extractor is configured to extract, using the processor, a prediction parameter from the data stream, the decoder further comprising:

a view predictor configured to perform, using the processor, a block-based prediction of a block of a non-renderable portion of the first view based on the prediction parameter and a reference signal derived from a previously-reconstructed portion of the multi-view signal to acquire a prediction signal,

wherein the reconstructor configured to reconstruct the at least a portion of the first view based on the prediction signal and the residual signal.

3. The decoder of claim 2 , wherein the view predictor is configured to, in performing the block-based prediction, predict the block within the non-renderable portion using, as the prediction parameter, a motion vector or a disparity vector.

4. The decoder of claim 2 , wherein the view predictor is configured to perform the block-based prediction of a block of the renderable portion of the first view using another prediction parameter, wherein the block of the renderable portion and the block of the non-renderable portion are within different pictures of the first view.

5. The decoder of claim 1 , further comprising a determinator configured to determine the renderable portion of the first view using a depth map of a scene to which the first and second views belong.

6. An encoder for encoding a multi-view signal into a data stream, comprising:

a determinator configured to determine, using a processor, a renderable portion of a first view of the multi-view signal which is renderable, by view synthesis, based on a second view of the multi-view signal;

obtaining, using the processor, a rendering flag associated with the renderable portion of the first view; and

a data stream generator configured to encode, using the processor, the rendering flag and a residual signal into the data stream,

wherein, based on a first coding mode associated with the rendering flag, the renderable portion of the first view is rendered, by view synthesis, from the second view to generate the rendered portion and replace the renderable portion of a reconstructed first view by the rendered portion, while based on a second coding mode, the replacement of the renderable portion is not performed.

7. The encoder of claim 6 , further comprising a view predictor configured to perform, using the processor, a block-based prediction of a block of a non-renderable portion of the first view based on a prediction parameter and a previously-encoded portion of the multi-view signal to acquire a prediction signal.

8. The encoder of claim 7 , wherein the view predictor is configured to, in performing the block-based prediction, predict the block within the non-renderable portion using, as the prediction parameter, a motion vector or a disparity vector.

9. The encoder of claim 7 , wherein the view predictor is configured to perform the block-based prediction of a block of the renderable portion of the first view using another prediction parameter, wherein the block of the renderable portion and the block of the non-renderable portion are within different pictures of the first view.

10. The encoder of claim 7 , further comprising:

a reconstructor configured to reconstruct the first view, including the renderable portion, from the prediction signal to acquire a reference signal, wherein the view predictor is configured to perform the block-based prediction based on the reference signal.

11. The encoder of claim 7 , wherein the determinator is configured to render the renderable portion to acquire the rendered signal, and the encoder further comprises:

a reconstructor configured to reconstruct the non-renderable portion based on the residual signal and the prediction signal to acquire a reconstructed signal, wherein the reconstructed signal and the rendered signal form a reference signal, which is used by the view predictor to perform the block-based prediction.

12. The encoder of claim 6 , wherein the data stream generator is configured to encode position information indicating a position of the renderable portion into the data stream.

13. The encoder of claim 6 , wherein the determinator configured to determine the renderable portion of the first view using a depth map of a scene to which the first and second views belong.

14. A method for decoding a multi-view signal from a data stream, the method comprising:

extracting, using a processor from the data stream, a rendering flag associated with a renderable portion of a first view of the multi-view signal, the rendering flag being indicative of whether the renderable portion is to be replaced with a rendered portion generated, by view synthesis, based on a second view of the multi-view signal; and

reconstructing, using the processor, at least a portion of the first view based on a residual signal, wherein the decoder is configured to decode in a first coding mode or a second coding mode depending on the rendering flag, wherein the decoder is configured to, in the first coding mode, render, by view synthesis, the renderable portion of the first view from the second view to generate the rendered portion and replace the renderable portion of a reconstructed first view by the rendered portion, with not performing the replacement in the second coding mode.

15. The method of claim 14 , further comprising:

extracting, using the processor, a prediction parameter from the data stream;

performing, using the processor, a block-based prediction of a block of a non-renderable portion of the first view based on the prediction parameter and a reference signal derived from a previously-reconstructed portion of the multi-view signal to acquire a prediction signal; and

reconstructing the at least a portion of the first view based on the prediction signal and the residual signal.

16. The method of claim 15 , wherein performing the block-based prediction comprises predicting the block within the non-renderable portion using, as the prediction parameter, a motion vector or a disparity vector.

17. The method of claim 15 , further comprising performing the block-based prediction of a block of the renderable portion of the first view using another prediction parameter, wherein the block of the renderable portion and the block of the non-renderable portion are within different pictures of the first view.

18. The method of claim 14 , further comprising:

determining the renderable portion of the first view using a depth map of a scene to which the first and second views belong.

19. A method for storing a data stream, comprising storing, on a digital storage medium, a data stream from which a multi-view signal is decoded by a method according to claim 14 .

Assignments (2)
CHANGE OF NAME Recorded Jan 30, 2026
From: GE VIDEO COMPRESSION, LLC
To: DOLBY VIDEO COMPRESSION, LLC
Reel/Frame 074536/0781 →
CHANGE OF NAME Recorded Nov 26, 2024
From: GE VIDEO COMPRESSION, LLC
To: DOLBY VIDEO COMPRESSION, LLC
Reel/Frame 069450/0772 →
Continuity (8)
Continuation 17863078 · Jul 12, 2022
Continuation 17111331 · Dec 3, 2020
Continuation 16548569 · Aug 22, 2019
Continuation 15687920 · Aug 28, 2017
Continuation 14271481 · May 7, 2014
Continuation PCTEP2012072182 · Nov 8, 2012
Provisional Application 61558647 · Nov 11, 2011
Related Publication 20240179340A1 · May 30, 2024
References Cited (59)
US 9774850B2 · Bosse et al. · 2017 [cited by applicant]
US 10264277B2 · Bosse et al. · 2019 [cited by applicant]
US 10440385B2 · Bosse et al. · 2019 [cited by applicant]
US 10880571B2 · Bosse et al. · 2020 [cited by applicant]
US 10887617B2 · Bosse et al. · 2021 [cited by applicant]
US 20030202592A1 · Sohn et al. · 2003 [cited by applicant]
US 20030234859A1 · Malzbender · 2003 [cited by applicant]
US 20060291561A1 · Seong · 2006 [cited by applicant]
US 20070064800A1 · Ha · 2007 [cited by applicant]
US 20070109300A1 · Li · 2007 [cited by applicant]
US 20070147502A1 · Nakamura · 2007 [cited by applicant]
US 20070183496A1 · Kadono · 2007 [cited by applicant]
US 20080174594A1 · Li · 2008 [cited by applicant]
US 20090015662A1 · Kim · 2009 [cited by applicant]
US 20090103616A1 · Ho · 2009 [cited by applicant]
US 20100020884A1 · Pandit · 2010 [cited by applicant]
US 20100284466A1 · Pandit · 2010 [cited by applicant]
US 20110038418A1 · Pandit · 2011 [cited by applicant]
US 20110116547A1 · Chen · 2011 [cited by applicant]
US 20110142138A1 · Tian · 2011 [cited by applicant]
US 20110222602A1 · Sung · 2011 [cited by applicant]
US 20110255592A1 · Sung · 2011 [cited by applicant]
US 20110261050A1 · Smolic · 2011 [cited by applicant]
US 20120027291A1 · Shimizu · 2012 [cited by applicant]
EP 1978750A2 · 2008 [cited by applicant]
KR 1020100014553A · 2010 [cited by applicant]
WO 2008007913A1 · 2008 [cited by applicant]
WO 2008133455A1 · 2008 [cited by applicant]
Office Action issued in corresponding European Patent Application No. 18185225.2 dated May 17, 2022. [cited by applicant]
Final Office Action issued in corresponding U.S. Appl. No. 17/129,457 dated Apr. 5, 2022. [cited by applicant]
Communication pursuant to Article 94(3) EPC issued in corresponding European Patent Application No. 18185225.2 dated Dec. 18, 2020. [cited by applicant]
Official Communication issued in International Patent Application No. PCT /EP2012/072182 mailed on Feb. 15, 2013. [cited by applicant]
Schwarz et al., “Description of 3D Video Technology Proposal by Fraunhofer HHI (HEVC compatible: configuration B)”, 98 MPEG Meeting, No. M22571, Geneva, Switzerland, Nov. 22, 2011, 46 pages. [cited by applicant]
Bartnik et al., “HEVC Extension for Multiview Video Coding and Multiview Video plus Depth Coding”, 43, VCEG Meeting, San Jose, CA, No. VCEG-AR13, Feb. 4, 2012, 42 pages. [cited by applicant]
Smolic et al., “An Overview of Available and Emerging 3D Video Formats and Depth Enhanced Stereo as Efficient Generic Solution”, Picture Coding Symposium 2009, May 6, 2009, 4 pages. [cited by applicant]
Martinian, Emin, et al., “View Synthesis for Multiview Video Compression,” Mitsubishi Electric Research Labs, Picture Coding Symposium, Apr. 24, 2004, 5 pages. [cited by applicant]
Survey of Algorithms used for Multi-view Coding (MVC), 1SO/IEC JTC1/SC29/WG11, MPEG2005/N6909, Jan. 2005, Hong Kong, China, 10 pages. [cited by applicant]
Schwarz, et al., “Description of 3D Video Technology Proposal by Fraunhofer HH1 (HEVC compatible; configuration BJ” ISO/IEC JTC1/SC29/WG11 MPEG2011/M22571, Nov. 2011, Geneva, Switzerland, 46 pages. [cited by applicant]
Office Action dated Aug. 25, 2016, issued in parallel U.S. Appl. No. 14/273,701, 45 pages. [cited by applicant]
Office Action issued on Sep. 22, 2016 in U.S. Appl. No. 14/271,481. [cited by applicant]
Notice of Allowance mailed May 26, 2017 in U.S. Appl. No. 14/271,481. [cited by applicant]
Office Action mailed Oct. 18, 2017 issued in U.S. Appl. No. 14/273,701. [cited by applicant]
Office Action mailed Apr. 13, 2018 in U.S. Appl. No. 14/273,701. [cited by applicant]
Konieczny et al., Depth-Based Inter-view Prediction of Motion Vectors for Improving Multiview video coding; 2010; Poznan University of Technology, pp. 1-4. [cited by applicant]
Notice of Reasons for Rejection Korean Patent Application No. 10-2018-7030649 dated Dec. 6, 2018. [cited by applicant]
Notice of Allowance dated Dec. 5, 2018 in U.S. Appl. No. 14/273,701. [cited by applicant]
Non-final Office Action U.S. Appl. No. 15/687,920 dated Jan. 25, 2019. [cited by applicant]
Extended European Search Report EP Application No. 18185225.2 dated Jan. 29, 2019. [cited by applicant]
Notice of Allowance U.S. Appl. No. 15/687,920 dated May 22, 2019. [cited by applicant]
Non-final Office Action U.S. Appl. No. 17/111,331 dated Aug. 19, 2021. [cited by applicant]
Notice of Allowance U.S. Appl. No. 17/111,331 dated Mar. 31, 2022. [cited by applicant]
Berg, 140 F.3d 1428, 46 USPQ2d 1226 (Fed. Cir. 1998). [cited by applicant]
Goodman, 11 F.3d 1046, 29 USPQ2d 2010 (Fed. Cir. 1993). [cited by applicant]
Langi, 759 F.2d 887, 225 USPQ 645 (Fed. Cir. 1985). [cited by applicant]
Van Ornum, 686 F.2d 937, 214 USPQ 761 (CCPA 1982). [cited by applicant]
Vogel, 422 F.2d 438, 164 USPQ 619 (CCPA 1970). [cited by applicant]
Thorington, 418 F.2d 528, 163 USPQ 644 (CCP A 1969). [cited by applicant]
Notice of Allowance issued in corresponding U.S. Appl. No. 17/129,457 dated Feb. 10, 2023. [cited by applicant]
Office Action issued in corresponding U.S. Appl. No. 17/129,457 dated Aug. 15, 2022. [cited by applicant]