IP Library › Granted Patent US 10,158,884
Granted Patent B2
US 10,158,884 · App. 14/662,071 · Granted Dec 18, 2018

Simplified merge list construction process for 3D-HEVC

Inventors: Li Zhang (San Diego, CA); Ying Chen (San Diego, CA)
Assignee: QUALCOMM Incorporated
H04N19/597H04N19/30H04N19/51H04N19/70H04N19/52
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,158,884
App. No.
14/662,071
Granted
Dec 18, 2018
Kind
B2
Abstract

A device for encoding video data includes a memory configured to store video data and a video encoder comprising one or more processors configured to, for a current layer being encoded, determine that the current layer has no direct reference layers, based on determining that the current layer has no direct reference layers, set at least one of a first syntax element, a second syntax element, a third syntax element, or a fourth syntax element to a disabling value indicating that a coding tool corresponding to the syntax element is disabled for the current layer.

Claims (45)

1. A method of encoding three-dimensional (3D) video data, the method comprising:

for a current layer being encoded, determining that the current layer has no direct reference layers that are used to predict the current layer;

in response to determining that the current layer has no direct reference layers, encoding the current layer in accordance with a constraint that requires setting a first syntax element, a second syntax element, a third syntax element, and a fourth syntax element to a disabling value indicating that a coding tool corresponding to the syntax element is disabled;

in response to the constraint, setting the first syntax element to a disabling value, wherein the disabling value for the first syntax element indicates that inter-view motion parameter prediction is disabled for the current layer;

in response to the constraint, setting the second syntax element to a disabling value, wherein the disabling value for the second syntax element indicates that view synthesis prediction merge candidates are disabled for the current layer;

in response to the constraint, setting the third syntax element to a disabling value, wherein the disabling value for the third syntax element indicates that accessing depth view components is disabled for the derivation process for a disparity vector for the current layer;

in response to the constraint, setting the fourth syntax element to a disabling value, wherein the disabling value for the fourth syntax element indicates that inter-view residual prediction is disabled for the current layer; and

generating an encoded bitstream of 3D video data comprising the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element.

2. The method of claim 1 , wherein the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element are view-level syntax elements.

3. The method of claim 1 , wherein the coding tool corresponding to the syntax element comprises a block-level coding tool.

4. A device for encoding three-dimensional (3D) video data, the device comprising:

a memory configured to store the 3D video data; and

a video encoder comprising one or more processors configured to:

for a current layer being encoded, determine that the current layer has no direct reference layers that are used to predict the current layer;

in response to determining that the current layer has no direct reference layers, encode the current layer in accordance with a constraint that requires setting a first syntax element, a second syntax element, a third syntax element, and a fourth syntax element to a disabling value indicating that a coding tool corresponding to the syntax element is disabled;

in response to the constraint, set the first syntax element to a disabling value, wherein the disabling value for the first syntax element indicates that inter-view motion parameter prediction is disabled for the current layer;

in response to the constraint, set the second syntax element to a disabling value, wherein the disabling value for the second syntax element indicates that view synthesis prediction merge candidates are disabled for the current layer;

in response to the constraint, set the third syntax element to a disabling value, wherein the disabling value for the third syntax element indicates that accessing depth view components is disabled for the derivation process for a disparity vector for the current layer;

in response to the constraint, set the fourth syntax element to a disabling value, wherein the disabling value for the fourth syntax element indicates that inter-view residual prediction is disabled for the current layer; and

generate an encoded bitstream of the 3D video data comprising the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element.

5. The device of claim 4 , wherein the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element are view-level syntax elements.

6. The device of claim 4 , wherein the coding tool corresponding to the syntax element comprises a block-level coding tool.

7. An apparatus for encoding three-dimensional (3D) video data, the apparatus comprising:

means for determining that a current layer being encoded has no direct reference layers that are used to predict the current layer;

means for encoding the current layer in accordance with a constraint that requires setting a first syntax element, a second syntax element, a third syntax element, and a fourth syntax element to a disabling value indicating that a coding tool corresponding to the syntax element is disabled in response to determining that the current layer has no direct reference layers;

means for setting the first syntax element to a disabling value in response to the constraint, wherein the disabling value for the first syntax element indicates that inter-view motion parameter prediction is disabled for the current layer;

means for setting the second syntax element to a disabling value in response to the constraint, wherein the disabling value for the second syntax element indicates that view synthesis prediction merge candidates are disabled for the current layer;

means for setting the third syntax element to a disabling value in response to the constraint, wherein the disabling value for the third syntax element indicates that accessing depth view components is disabled for the derivation process for a disparity vector for the current layer;

means for setting the fourth syntax element to a disabling value in response to the constraint, wherein the disabling value for the fourth syntax element indicates that inter-view residual prediction is disabled for the current layer; and

means for generating an encoded bitstream of 3D video data comprising the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element.

8. The apparatus of claim 7 , wherein the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element are view-level syntax elements.

9. The apparatus of claim 7 , wherein the coding tool corresponding to the syntax element comprises a block-level coding tool.

10. The apparatus of claim 7 , wherein the means for setting the at least one of the first syntax element, the second syntax element, the third syntax element, or the fourth syntax element to the disabling value indicating that the coding tool corresponding to the syntax element is disabled in response to determining that the current layer has no direct reference layers.

11. A non-transitory computer-readable storage medium storing instructions that when executed by one or more processors cause the one or more processors to:

for a current layer being encoded, determine that the current layer has no direct reference layers that are used to predict the current layer;

in response to determining that the current layer has no direct reference layers, encode the current layer in accordance with a constraint that requires setting a first syntax element, a second syntax element, a third syntax element, and a fourth syntax element to a disabling value indicating that a coding tool corresponding to the syntax element is disabled;

in response to the constraint, set the first syntax element to a disabling value, wherein the disabling value for the first syntax element indicates that inter-view motion parameter prediction is disabled for the current layer;

in response to the constraint, set the second syntax element to a disabling value, wherein the disabling value for the second syntax element indicates that view synthesis prediction merge candidates are disabled for the current layer;

in response to the constraint, set the third syntax element to a disabling value, wherein the disabling value for the third syntax element indicates that accessing depth view components is disabled for the derivation process for a disparity vector for the current layer;

in response to the constraint, set the fourth syntax element to a disabling value, wherein the disabling value for the fourth syntax element indicates that inter-view residual prediction is disabled for the current layer; and

generate an encoded bitstream of the 3D video data comprising the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element.

12. The non-transitory computer-readable storage medium of claim 11 , wherein the first syntax element, the second syntax element, the third syntax element, and the fourth syntax element are view-level syntax elements.

13. The non-transitory computer-readable storage medium of claim 11 , wherein the coding tool corresponding to the syntax element comprises a block-level coding tool.

14. The device of claim 4 , wherein the device comprises a wireless communication device, further comprising a transmitter configured to transmit the encoded bitstream of the 3D video data.

15. The device of claim 14 , wherein the wireless communication device comprises a telephone handset and wherein the transmitter is configured to modulate, according to a wireless communication standard, a signal comprising the encoded bitstream of the 3D video data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 30, 2015
From: ZHANG, LI; CHEN, YING
To: QUALCOMM INCORPORATED
Reel/Frame 035288/0832 →
Continuity (2)
Provisional Application 61955720 · Mar 19, 2014
Related Publication 20150271524A1 · Sep 24, 2015
Cited By (10)
US 12,200,221 US 12,348,706 US 12,382,085 US 12,395,661 US 12,407,812 US 12,483,691 US 12,615,394 US 12,621,481 US 12,688,645 US 12,720,068