IP Library Granted Patent US 12,335,503
Granted Patent B2
US 12,335,503 · App. 18/404,209 · Granted Jun 17, 2025

Derivation on sublayer-wise output layer set

Inventors: Byeongdoo Choi (Palo Alto, CA); Shan Liu (San Jose, CA); Stephan Wenger (Hillsborough, CA)
Assignee: TENCENT AMERICA LLC
H04N19/44H04N19/503
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,335,503
App. No.
18/404,209
Granted
Jun 17, 2025
Kind
B2
Abstract

A method and system for decoding a coded video sequence may include obtaining the coded video sequence, and decoding the coded video sequence. A a value of a temporal sublayer identifier of a video coding layer (VCL) network abstraction layer (NAL) unit in the coded video sequence is constrained to be less than or equal to a value of vps_max_sublayers_minus1, that specifies a maximum number of temporal sublayers that may be present in a layer in each coded video sequence referring to the video parameter set (VPS), in the VPS referred to by the VCL NAL unit.

Claims (34)

1. A method for decoding a coded video sequence by at least one processor, the method comprising:

obtaining the coded video sequence; and

decoding the coded video sequence,

wherein a value of a temporal sublayer identifier of a video coding layer (VCL) network abstraction layer (NAL) unit in the coded video sequence is constrained to be less than or equal to a value of vps_max_sublayers_minus1, that specifies a maximum number of temporal sublayers in each coded video sequence referring to the video parameter set (VPS), in the VPS referred to by the VCL NAL unit,

wherein a value of max_tid_il_ref_pics_plus1[i] being zero specifies that inter-layer prediction is not used by non-intra random access point (IRAP) pictures of the i-th layer.

2. The method of claim 1 , wherein a value of max_tid_il_ref_pics_plus1[i] being greater than zero specifies that, for decoding pictures of the i-th layer, a picture with a temporal sublayer identifier greater than max_tid_il_ref_pics_plus1[i]−1 is not used as an inter-layer reference picture (ILRP).

3. The method of claim 1 , wherein when not present, a value of max_tid_il_ref_pics_plus1[i] is inferred to be equal to vps_max_sublayers_minus1+1.

4. The method of claim 1 , wherein max_tid_il_ref_pics_plus1[i] is less than or equal to vps_max_sublayers_minus1+1.

5. The method of claim 1 , wherein a sublayer-wise output layer set is not derived for independent layers.

6. The method of claim 1 ,

a first variable NumSubLayersInLayerInOLS[i][j] specifies the number of sublayers in a j-th layer output layer in an i-th OLS, wherein a second variable OutputLayerIdInOls[i][j] specifies a nuh_layer_id value of the j-th output layer in the i-th OLS, wherein a third variable LayerUsedAsOutputLayerFlag[k] specifies whether a k-th layer is used as an output layer in at least one OLS, and wherein the first variable, the second variable, and the third variable are derived from a value of variable max_tid_il_ref_pics_plus1[i].

7. The method of claim 1 , wherein max_tid_il_ref_pics_plus1 and layerIncludedInOlsFlag are constrained not be derived for independent layers.

8. A method for encoding a video sequence by at least one processor, the method comprising:

determining a value of a temporal sublayer identifier max_tid_il_ref_pics_plus1[i] of a video coding layer (VCL) network abstraction layer (NAL) unit in a video sequence, wherein max_tid_il_ref_pics_plus1[i] is constrained to be less than or equal to a value of vps_max_sublayers_minus1 that specifies a maximum number of temporal sublayers in each coded video sequence referring to a video parameter set (VPS), in the VPS referred to by the VCL NAL unit;

setting the value of max_tid_il_ref_pics_plus1[i] in the video sequence, wherein the value of max_tid_il_ref_pics_plus1[i] being zero specifying that inter-layer prediction is not used by non-intra random access point (IRAP) pictures of an i-th layer; and

encoding the video sequence comprising the value of max_tid_il_ref_pics_plus1[i].

9. The method of claim 8 , wherein a sublayer-wise output layer set is not derived for independent layers.

10. The method of claim 8 , wherein the value of max_tid_il_ref_pics_plus1[i] being greater than zero specifies that, for decoding pictures of an i-th layer, a picture with a temporal sublayer identifier greater than max_tid_il_ref_pics_plus1[i]−1 is not used as an inter-layer reference picture (ILRP).

11. The method of claim 8 , wherein when not present, the value of max_tid_il_ref_pics_plus1[i] is inferred to be equal to vps_max_sublayers_minus1+1.

12. The method of claim 8 , wherein max_tid_il_ref_pics_plus1[i] is less than or equal to vps_max_sublayers_minus1+1.

13. The method of claim 8 , wherein

a first variable NumSubLayersInLayerInOLS[i][j] specifies the number of sublayers in a j-th layer output layer in an i-th OLS, wherein a second variable OutputLayerIdInOls[i][j] specifies a nuh_layer_id value of the j-th output layer in the i-th OLS, wherein a third variable LayerUsedAsOutputLayerFlag[k] specifies whether a k-th layer is used as an output layer in at least one OLS, and wherein the first variable, the second variable, and the third variable are derived from a value of variable max_tid_il_ref_pics_plus1[i].

14. The method of claim 8 , wherein max_tid_il_ref_pics_plus1 and layerIncludedInOlsFlag are not derived for the independent layers.

15. A method of processing visual media by at least one processor, the method comprising:

performing a conversion between a visual media file and a bitstream of a visual media data; and

processing at least one of the visual media file and the bitstream of the visual media data based on the conversion,

wherein a value of a temporal sublayer identifier max_tid_il_ref_pics_plus1[i] of a video coding layer (VCL) network abstraction layer (NAL) unit in a video sequence of the bitstream is constrained to be less than or equal to a value of vps_max_sublayers_minus1 that specifies a maximum number of temporal sublayers in each coded video sequence referring to a video parameter set (VPS), in the VPS referred to by the VCL NAL unit, and

wherein the value of max_tid_il_ref_pics_plus1[i] being zero specifies that inter-layer prediction is not used by non-intra random access point (IRAP) pictures of an i-th layer.

16. The method of claim 15 , wherein a sublayer-wise output layer set is not derived for independent layers.

17. The method of claim 15 , wherein the value of max_tid_il_ref_pics_plus1[i] being greater than zero specifies that, for decoding pictures of an i-th layer, a picture with a temporal sublayer identifier greater than max_tid_il_ref_pics_plus1[i]−1 is not used as an inter-layer reference picture (ILRP).

18. The method of claim 15 , wherein when not present, the value of max_tid_il_ref_pics_plus1[i] is inferred to be equal to vps_max_sublayers_minus1+1.

19. The method of claim 15 , wherein max_tid_il_ref_pics_plus1[i] is less than or equal to vps_max_sublayers_minus1+1.

20. The method of claim 15 , wherein

a first variable NumSubLayersInLayerInOLS[i][j] specifies the number of sublayers in a j-th layer output layer in an i-th OLS, wherein a second variable OutputLayerIdInOls[i][j] specifies a nuh_layer_id value of the j-th output layer in the i-th OLS, wherein a third variable LayerUsedAsOutputLayerFlag[k] specifies whether a k-th layer is used as an output layer in at least one OLS, and wherein the first variable, the second variable, and the third variable are derived from a value of variable max_tid_il_ref_pics_plus1[i].

Continuity (5)
Continuation 18524806 · Nov 30, 2023
Continuation 17582570 · Jan 24, 2022
Continuation 17097636 · Nov 13, 2020
Provisional Application 63000980 · Mar 27, 2020
Related Publication 20240187621A1 · Jun 6, 2024
References Cited (15)
US 20150195561A1 · Wang · 2015 [cited by examiner]
US 20160191926A1 · Deshpande et al. · 2016 [cited by applicant]
US 20160366428A1 · Deshpande · 2016 [cited by applicant]
US 20180376154A1 · Deshpande · 2018 [cited by applicant]
US 20190058895A1 · Deshpande · 2019 [cited by examiner]
US 20200077105A1 · Hannuksela · 2020 [cited by applicant]
WO 2015194183A1 · 2015 [cited by applicant]
Benjamin Bross, “Versatile Video Coding (Draft 8)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-Q2001-vE, Jan. 7-17, 2020, pp. 1-511, 17th Meeting, Brussels, BE. [cited by applicant]
Choi et al., “AHG8/AHG9: On derivation of sublayer No. in output layer set”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R0119, 18th Meeting by teleconference, Apr. 15-24, 202… [cited by applicant]
Extended European Search Report issued Jun. 20, 2022 in European Application No. 21772938.3. [cited by applicant]
International Search Report dated Apr. 23, 2021 in Application No. PCT/US21/16019. [cited by applicant]
International Search Report dated Feb. 1, 2021 in Application No. PCT/US21/16019. [cited by applicant]
Sullivan et al., “Meeting Report of the 18th Meeting of the Joint Video Experts Team (JVET)”, by teleconference, Apr. 15-24, 2020, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-… [cited by applicant]
Vietnamese Office Communication dated Sep. 11, 2023. [cited by applicant]
Written Opinion of the International Search Authority dated Apr. 23, 2021 in Application No. PCT/US21/16019. [cited by applicant]