IP Library Granted Patent US 12,457,311
Granted Patent B2
US 12,457,311 · App. 18/609,882 · Granted Oct 28, 2025

Efficient multi-view coding using depth-map estimate for a dependent view

Inventors: Heiko Schwarz (Berlin, DE); Thomas Wiegand (Berlin, DE)
Assignee: Dolby Video Compression, LLC
H04N13/128H04N13/161H04N19/194H04N19/46H04N19/513H04N19/52H04N19/597H04N19/895H04N2213/003
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,457,311
App. No.
18/609,882
Granted
Oct 28, 2025
Kind
B2
Abstract

The usual coding order according to which the reference view is coded prior to the dependent view, and within each view, a depth map is coded subsequent to the respective picture, may be maintained and does lead to a sacrifice of efficiency in performing inter-view redundancy removal by, for example, predicting motion data of the current picture of the dependent view from motion data of the current picture of the reference view. Rather, a depth map estimate of the current picture of the dependent view is obtained by warping the depth map of the current picture of the reference view into the dependent view, thereby enabling various methods of inter-view redundancy reduction more efficiently by bridging the gap between the views. According to another aspect, the following discovery is exploited: the overhead associated with an enlarged list of motion predictor candidates for a block of a picture of a dependent view is comparatively low compared to a gain in motion vector prediction quality resulting from an adding of a motion vector candidate which is determined from an, in disparity-compensated sense, co-located block of a reference view.

Claims (39)

1. An apparatus for reconstructing a multi-view signal coded into a data stream, comprising:

a depth estimator configured to obtain, using a processor, a depth map estimate of a current picture in a dependent view of the multi-view signal; and

a dependent-view reconstructor configured to, using the processor:

determine a disparity vector for at least one block of blocks of the current picture in the dependent view of the multi-view signal based on the depth map estimate, the disparity vector representing a disparity between the current picture of the dependent view and a current picture of a reference view of the multi-view signal at the at least one block of the current picture of the dependent view;

determine, for the at least one block of the current picture of the dependent view, a list of motion vector predictor candidates; and

identify, for the at least one block of the current picture of the dependent view, a motion vector predictor candidate from the list of motion vector predictor candidates; and

reconstruct the at least one block of the current picture of the dependent view by performing a motion-compensated prediction of the at least one block of the current picture of the dependent view using a motion vector which depends on the identified motion vector predictor candidate.

2. The apparatus according to claim 1 , wherein the dependent-view reconstructor is configured to extract, for the at least one block of the current picture of the dependent view, from the data stream index information specifying the motion vector predictor candidate from the list of motion vector predictor candidates.

3. The apparatus according to claim 1 , further comprising:

a reference-view reconstructor configured to reconstruct a depth map of the current picture of the reference view of the multi-view signal from a reference view depth-map portion of the multi-view data stream, wherein:

the depth estimator is configured to estimate the depth map of the current picture of the dependent view by warping the depth map of the current picture of the reference view into the dependent view, and

the dependent-view reconstructor is configured to, in determining the disparity vector for the at least one block, subject the depth data estimate at the at least one block to depth-to-disparity conversion to acquire the disparity vector.

4. The apparatus according to claim 1 , wherein the dependent-view reconstructor is configured to extract, for the at least one block of the current picture of the dependent view, a motion vector difference in relation to the identified motion vector predictor candidate and to perform the reconstruction of the at least one block of the current picture such that the used motion vector further depends on a sum of the motion vector difference and the identified motion vector predictor candidate.

5. The apparatus according to claim 1 , wherein the dependent-view reconstructor is configured to extract for the at least one block of the current picture of the dependent view, a reference picture index specifying a reference picture of a list of reference pictures comprising the current picture of the reference view and already decoded pictures of the dependent view,

wherein the dependent-view reconstructor is configured to, if the reference picture is one of the already decoded pictures of the dependent view, perform the motion-compensated prediction using the one already decoded picture of the dependent view as a reference, and if the reference picture is the current picture of the reference view, add the disparity vector or a modified disparity vector derived from the disparity vector to a list of disparity vector prediction candidates, extract index information specifying one disparity vector predictor candidate of the list of disparity vector predictor candidates from the multi-view data stream and reconstruct the at least one block of the current picture of the dependent view by performing a disparity-compensated prediction of the at least one block of the current picture of the dependent view using a disparity vector which depends on the specified disparity vector candidate using the current picture of the reference view as a reference.

6. The apparatus according to claim 1 , wherein the dependent-view reconstructor is further configured to, in deriving the list of motion vector predictor candidates, spatially and/or temporarily predict a further motion vector from spatially and/or temporarily neighboring blocks of the dependent view and add the further motion vector or a version derived from the further motion vector to the list of motion vector predictor candidates.

7. The apparatus according to claim 1 , wherein the dependent-view reconstructor is configured to perform the derivation of the list of motion vector predictor candidates via a list of motion/disparity vector predictor candidates being a list of motion/disparity parameter candidates each comprising a number of hypotheses, and, per hypothesis, a motion/disparity motion vector and a reference index specifying a reference picture out of a list of reference pictures comprising the current picture of a reference view and previously decoded pictures of the dependent view, wherein the dependent-view reconstructor is configured to add motion/disparity parameters to the list of motion/disparity parameter candidates which depend on motion/disparity parameters associated with the determined block of the current picture of the reference view, and to reconstruct the at least one block of the current picture of the dependent view by performing motion/disparity-compensated prediction on the at least one block of the current picture of the dependent view using motion/disparity parameters which depend on a motion/disparity parameter candidate specified by the index information.

8. An apparatus for encoding a multi-view signal into a data stream, comprising:

a depth estimator configured to obtain, using a processor, a depth map estimate of a current picture in a dependent view of the multi-view signal; and

a dependent-view encoder configured to, using the processor:

determine a disparity vector for at least one block of blocks of the current picture in the dependent view of the multi-view signal based on the depth map estimate, the disparity vector representing a disparity between the current picture of the dependent view and a current picture of a reference view of the multi-view signal at the at least one block of the current picture of the dependent view;

determine, for the at least one block of the current picture of the dependent view, a list of motion vector predictor candidates; and

encode the at least one block of the current picture of the dependent view by performing a motion-compensated prediction of the at least one block of the current picture of the dependent view using a motion vector which depends on a certain motion vector predictor candidate from the list of motion vector predictor candidates.

9. The apparatus according to claim 8 , wherein the dependent-view encoder is configured to insert, into the data stream, index information specifying the certain motion vector predictor candidate from the list of motion vector predictor candidates for the at least one block of the current picture of the dependent view.

10. The apparatus according to claim 8 , further comprising:

a reference-view encoder configured to encode a depth map of the current picture of the reference view of the multi-view signal into a reference view depth-map portion of the multi-view data stream, wherein:

the depth estimator is configured to estimate the depth map of the current picture of the dependent view by warping the depth map of the current picture of the reference view into the dependent view, and

the dependent-view encoder is configured to, in determining the disparity vector for the at least one block, subject the depth data estimate at the at least one block to depth-to-disparity conversion to acquire the disparity vector.

11. A method for encoding a multi-view signal into a data stream, comprising:

obtaining a depth map estimate of a current picture in a dependent view of the multi-view signal;

determining a disparity vector for at least one block of blocks of the current picture in the dependent view of the multi-view signal based on the depth map estimate, the disparity vector representing a disparity between the current picture of the dependent view and a current picture of a reference view of the multi-view signal at the at least one block of the current picture of the dependent view;

determining, for the at least one block of the current picture of the dependent view, a list of motion vector predictor candidates; and

encoding the at least one block of the current picture of the dependent view by performing a motion-compensated prediction of the at least one block of the current picture of the dependent view using a motion vector which depends on a certain motion vector predictor candidate from the list of motion vector predictor candidates.

12. The method according to claim 11 , further comprising:

inserting, into the data stream, index information specifying the certain motion vector predictor candidate from the list of motion vector predictor candidates for the at least one block of the current picture of the dependent view.

13. A method for storing a data stream, comprising storing, on a digital storage medium, a data stream into which a multi-view signal is encoded by a method according to claim 11 .

14. The method according to claim 13 , further comprising:

encoding a depth map of the current picture of the reference view of the multi-view signal into a reference view depth-map portion of the multi-view data stream, wherein the depth map of the current picture of the dependent view is estimated by warping the depth map of the current picture of the reference view into the dependent view; and

in determining the disparity vector for the at least one block, subjecting the depth data estimate at the at least one block to depth-to-disparity conversion to acquire the disparity vector.

Assignments (2)
CHANGE OF NAME Recorded Jan 30, 2026
From: GE VIDEO COMPRESSION, LLC
To: DOLBY VIDEO COMPRESSION, LLC
Reel/Frame 074536/0781 →
CHANGE OF NAME Recorded Nov 26, 2024
From: GE VIDEO COMPRESSION, LLC
To: DOLBY VIDEO COMPRESSION, LLC
Reel/Frame 069450/0772 →
Continuity (6)
Continuation 17576873 · Jan 14, 2022
Continuation 16871919 · May 11, 2020
Continuation 14273730 · May 9, 2014
Continuation PCTEP2012072300 · Nov 9, 2012
Provisional Application 61558660 · Nov 11, 2011
Related Publication 20240348764A1 · Oct 17, 2024
References Cited (132)
US 7876833B2 · Segall · 2011 [cited by applicant]
US 7885471B2 · Segall · 2011 [cited by applicant]
US 8204129B2 · He · 2012 [cited by applicant]
US 8369406B2 · Kim · 2013 [cited by applicant]
US 8594193B2 · He et al. · 2013 [cited by applicant]
US 10887575B2 · Schwarz et al. · 2021 [cited by applicant]
US 11523098B2 · Schwarz et al. · 2022 [cited by applicant]
US 20030138150A1 · Srinivasan · 2003 [cited by applicant]
US 20050157797A1 · Gaedke · 2005 [cited by applicant]
US 20070019726A1 · Cha · 2007 [cited by applicant]
US 20070147502A1 · Nakamura · 2007 [cited by applicant]
US 20070223582A1 · Borer · 2007 [cited by applicant]
US 20070230567A1 · Wang · 2007 [cited by applicant]
US 20080069247A1 · He · 2008 [cited by applicant]
US 20080089417A1 · Bao · 2008 [cited by applicant]
US 20080165848A1 · Ye · 2008 [cited by applicant]
US 20080170618A1 · Choi · 2008 [cited by applicant]
US 20080198924A1 · Ho et al. · 2008 [cited by applicant]
US 20080205508A1 · Ziauddin · 2008 [cited by applicant]
US 20080211901A1 · Civanlar · 2008 [cited by applicant]
US 20080267291A1 · Vieron · 2008 [cited by applicant]
US 20090010323A1 · Su et al. · 2009 [cited by applicant]
US 20090015662A1 · Kim · 2009 [cited by applicant]
US 20090074061A1 · Yin · 2009 [cited by applicant]
US 20090096643A1 · Chang · 2009 [cited by applicant]
US 20090103616A1 · Ho · 2009 [cited by applicant]
US 20090175349A1 · Ye · 2009 [cited by applicant]
US 20090190662A1 · Park et al. · 2009 [cited by applicant]
US 20090290643A1 · Yang · 2009 [cited by applicant]
US 20100034260A1 · Shimizu et al. · 2010 [cited by applicant]
US 20100111183A1 · Jeon et al. · 2010 [cited by applicant]
US 20100158129A1 · Lai et al. · 2010 [cited by applicant]
US 20100189177A1 · Shimizu et al. · 2010 [cited by applicant]
US 20100220791A1 · Lin · 2010 [cited by applicant]
US 20100220795A1 · Mn · 2010 [cited by applicant]
US 20100260260A1 · Wiegand · 2010 [cited by applicant]
US 20100260268A1 · Cowan · 2010 [cited by applicant]
US 20100284466A1 · Pandit · 2010 [cited by applicant]
US 20100309292A1 · Ho et al. · 2010 [cited by applicant]
US 20100316139A1 · Le Leannec · 2010 [cited by applicant]
US 20100329342A1 · Joshi · 2010 [cited by applicant]
US 20110038418A1 · Pandit · 2011 [cited by applicant]
US 20110044550A1 · Tian · 2011 [cited by applicant]
US 20110090959A1 · Wiegand · 2011 [cited by applicant]
US 20110142138A1 · Tian · 2011 [cited by applicant]
US 20110222602A1 · Sung et al. · 2011 [cited by applicant]
US 20110286520A1 · Xu · 2011 [cited by applicant]
US 20110317930A1 · Kim · 2011 [cited by applicant]
US 20120023250A1 · Chen · 2012 [cited by applicant]
US 20120075436A1 · Chen · 2012 [cited by applicant]
US 20120082222A1 · Wang · 2012 [cited by applicant]
US 20120114036A1 · Po · 2012 [cited by applicant]
US 20120229602A1 · Chen et al. · 2012 [cited by applicant]
US 20120236115A1 · Zhang · 2012 [cited by applicant]
US 20120269270A1 · Chen · 2012 [cited by applicant]
US 20130229485A1 · Rusanovskyy · 2013 [cited by applicant]
US 20130293676A1 · Sugio et al. · 2013 [cited by applicant]
US 20130322536A1 · Yang · 2013 [cited by applicant]
US 20130335522A1 · Zhang · 2013 [cited by applicant]
CN 101600108A · 2006 [cited by applicant]
CN 101170702A · 2008 [cited by applicant]
CN 101248669A · 2008 [cited by applicant]
CN 101491096A · 2009 [cited by applicant]
CN 101669367A · 2010 [cited by applicant]
CN 101690220A · 2010 [cited by applicant]
CN 101690220B · 2010 [cited by applicant]
CN 101867813A · 2010 [cited by applicant]
CN 101911700A · 2010 [cited by applicant]
CN 101917619A · 2010 [cited by applicant]
CN 101986716A · 2011 [cited by applicant]
CN 102077246A · 2011 [cited by applicant]
EP 2348733A2 · 2011 [cited by applicant]
JP 200822549A · 2008 [cited by applicant]
JP 2009213161A · 2009 [cited by applicant]
JP 2009543508A · 2009 [cited by applicant]
JP 2010525724A · 2010 [cited by applicant]
JP 2010537484A · 2010 [cited by applicant]
JP 20100537484A · 2010 [cited by applicant]
JP 2011509631A · 2011 [cited by applicant]
KR 1020090046826A · 2009 [cited by applicant]
KR 102029401B1 · 2019 [cited by applicant]
KR 102090106B1 · 2020 [cited by applicant]
WO 2007037645A1 · 2007 [cited by applicant]
WO 2007081756A2 · 2007 [cited by applicant]
WO 2008007913A1 · 2008 [cited by applicant]
WO 20008007913A1 · 2008 [cited by applicant]
WO 2009005626A2 · 2009 [cited by applicant]
WO 2010043773A1 · 2010 [cited by applicant]
WO 2013068547A2 · 2013 [cited by applicant]
WO 2013072484A1 · 2013 [cited by applicant]
Notice of Allowance mailed Jan. 14, 2020 in U.S. Appl. No. 14/277,850. [cited by applicant]
Notice of Allowance mailed Feb. 11, 2020 in U.S. Appl. No. 14/273,730. [cited by applicant]
Notice of Allowance issued Sep. 20, 2021 in U.S. Appl. No. 16/871,919. [cited by applicant]
Tanimoto et al., “Multi-view Depth Map of Rena and Akko & Kayo”, ISO/IEC JTC1/SC29/WG11 M14888, Oct. 2007. 5 pages, Shenzhen, Chin. [cited by applicant]
Notice of Allowance mailed Sep. 3, 2020 in U.S. Appl. No. 16/859,433. [cited by applicant]
Notice of Allowance issued in corresponding U.S. Appl. No. 16/592,433 mailed Sep. 3, 2020. [cited by applicant]
Office Action issued in corresponding U.S. Appl. No. 16/871,866 dated Oct. 9, 2020. [cited by applicant]
Notice of Allowance issued in corresponding U.S. Appl. No. 16/848,297 dated Jul. 22, 2021. [cited by applicant]
Non-final Office Action issued in U.S. Appl. No. 16/871,919 dated Jun. 10, 2021. [cited by applicant]
Haitao Yang et al., Regional Disparity Based Motion and Disparity Prediction for MVC, Joint Video Team (JVT) of ISO/IEC MPEG & ITU-T VCEG{ISO/IEC JTC1/SC29/WG11 and ITU-T SG16 Q.6), 22nd Meeting: Marrakech, Morocco, Jan… [cited by applicant]
Bartnik et al., “HEVC Extension for Multiview Video Coding and Multiview Video plus Depth Coding”; ITU-Telecommunications Standardization Sector Study Group 16 Question 6, Document VCEG-AR13, Feb. 3-10, 2012, 42 pages. [cited by applicant]
Ekmekcioglu et al .. “Content Adaptive Enhancement of Multi-View Depth Maps for Free Viewpoint Video”, IEEE Journal of Selected Topics in Signal Processing, vol. 5, No. 2, Apr. 1, 2011, pp. 352-361. [cited by applicant]
ITU-T and ISO/IEC JTC 1, “Advanced Video Coding for Generic Audiovisual Services”, Recommendation ITU-T H.264 and ISO/IEC 14496-10 (MPEG-4 AC). 2010. [cited by applicant]
Merkle et al., “3D Video Coding: An Overview of Present and Upcoming Standards”, Visual Communications and Image Processing, Jul. 11, 2010, 7 pages, Huang Shan, An Hui, China. [cited by applicant]
Na, S. et al., “Multi-View Depth Video Coding using Depth View Synthesis”, IEEE International Symposium on Circuits and Systems, May 18, 2008, pp. 1400-1403. [cited by applicant]
Office Action mailed Sep. 30, 2016 in U.S. Appl. No. 14/273,730. [cited by applicant]
Schwarz et al., “Description of 3D Video Coding Technology Proposal by Fraunhofler HHI (HEVC compatible, Configuration A)”, ISO/IEC JTC1/SC29/WG11, MPEG2011/M22570, Nov. 2011, 46 pages, Geneva, Switzerland. [cited by applicant]
Schwarz, H. et al., “Overview of the Scalable Video Coding Extension of the H.264/AVC Standard”, IEEE Transactions on Circuits and Systems for Video Technology, vol. 17, No. 9, pp. 1103-1120, Sep. 2007. [cited by applicant]
Seo et al., “Motion Information Sharing Mode for Depth Video Coding”, IEEE, 3DTV-Conference: The True Vision-Capture, Transmission and Display of 3D Video, Jun. 7, 2010, pp. 1-4. [cited by applicant]
Shimizu, S. et al., View Scalable Multiview Video Coding Using 3-D Warping With Depth Map, IEEE Transactions on Circuits and Systems for Video Technology, vol. 17, No. 11, Nov. 2007, pp. 1485-1495. [cited by applicant]
Stefanoski, N. et al., “Description of 3D Video Coding Technology Proposal by Disney Research Zurich and Fraunhofer HHI”, Motion Picture Expert Group meeting, No. M22668, Nov. 22, 2011, 34 pages. [cited by applicant]
Tanimoto et al., View Synthesis Algorithm in View Synthesis Reference Software 2.0 {VSRS2.0), ISO/IEC JTC1/SC29/WG11, MPEG 2008/M16090, Feb. 2009, 5 pages, Lausanne, Switzerland. [cited by applicant]
Vetro et al., “Overview of the Stereo and Multiview Video Coding Extensions of the H.264/MPEG-4 AVG Standard”, Proceedings of the IEEE, vol. 99, No. 4, Apr. 11, 2011, pp. 626-642. [cited by applicant]
Vetro et al., “Joint Multiview Video Model {JMVM) 6.0”, Joint Video Team {JVT) of ISO/IEC MPEG & ITU-T VCEG ISO/IEC JTC1/SC29/WG11 and ITU-T SG16 0.6, 25th Meeting, Document JVT-Y207, Oct. 2007, 10 pages, Shenzhen, Chin… [cited by applicant]
Zheng et al., “Affine Model for Disparity Estimation in Multiview Sequence Coding”, International Conference on Information Theory and Information Security {ICITIS), Dec. 17, 2010, pp. 736-740, Nanning, China. [cited by applicant]
Office Action mailed Dec. 27, 2016 in U.S. Appl. No. 14/277,850. [cited by applicant]
Office Action issued Mar. 15, 2017 in U.S. Appl. No. 14/272,671. [cited by applicant]
Office Action issued Oct. 20, 2017 in U.S. Appl. No. 14/273,730. [cited by applicant]
Office Action issued Oct. 3, 2017 in U.S. Appl. No. 14/277,850. [cited by applicant]
Non-final Office Action U.S. Appl. No. 14/277,850 dated May 10, 2018. [cited by applicant]
Konieczny et al., “Depth-Based Interview Prediction of Motion Vectors for Improved Multiview Video Coding”, 2010, EEE. pp. 1-4. [cited by applicant]
Notice of Allowance U.S. Appl. No. 14/272,671 dated Jun. 4, 2018. [cited by applicant]
Non-final Office Action U.S. Appl. No. 14/273,730 dated Jun. 1, 2018. [cited by applicant]
Final Office Action U.S. Appl. No. 14/277,850 dated Dec. 10, 2018. [cited by applicant]
Final Office Action U.S. Appl. No. 14/273,730 dated Jan. 28, 2019. [cited by applicant]
Non-final Office Action U.S. Appl. No. 16/120,731 dated Mar. 21, 2019. [cited by applicant]
Non-final Office Action U.S. Appl. No. 14/277,850 dated May 30, 2019. [cited by applicant]
Notice of Allowance U.S. Appl. No. 16/120,731 dated Jul. 3, 2019. [cited by applicant]
Non-final Office Action U.S. Appl. No. 14/273,730 dated Jun. 14, 2019. [cited by applicant]
Non-Final Office Action issued in corresponding U.S. Appl. No. 17/129,450 dated Mar. 3, 2022. [cited by applicant]
Notice of Allowance issued in corresponding U.S. Appl. No. 17/129,450 dated Jul. 25, 2022. [cited by applicant]
Office Action issued in corresponding U.S. Appl. No. 17/531,478 dated Feb. 2, 2023. [cited by applicant]