IP Library › Granted Patent US 12,231,659
Granted Patent B2
US 12,231,659 · App. 17/990,247 · Granted Feb 18, 2025

Number restriction for sublayers

Inventors: Ye-kui Wang (San Diego, CA); Zhipin Deng (Beijing, CN)
Assignees: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.; BYTEDANCE INC.
H04N19/196H04N19/117H04N19/167H04N19/172H04N19/70H04N19/82
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,231,659
App. No.
17/990,247
Granted
Feb 18, 2025
Kind
B2
Abstract

Methods, apparatus, systems for performing video processing such as video encoding, video decoding, or video transcoding are described. One example method of video processing includes performing a conversion between a video comprising one or more video layers and a bitstream of the video according to a rule. The rule specifies that responsive to a value of a first field in a sequence parameter set (SPS) that is referred to by a video layer, a second field in a video parameter set referred to by the SPS that is indicative of a maximum number of sublayers allowed in the bitstream is construed to be equal to a third field in the SPS. The third field is indicative of a maximum number of sublayers allowed in the video layer.

Claims (49)

1. A method of video processing, comprising:

performing a conversion between a video and a bitstream of the video according to a first rule,

determining whether a value of a first field of a sequence parameter set is set equal to 0 or is greater than 0;

wherein the first rule specifies that responsive to the value of the first field of the sequence parameter set that is indicative of an identifier of a video parameter set being equal to 0, a second field of the video parameter set that is indicative of a first maximum number of sublayers allowed in a layer specified by the video parameter set is construed to be equal to a third field in the sequence parameter set, and the value of the third field is in a range of 0 to 6, inclusive,

wherein the third field is indicative of a second maximum number of sublayers allowed in each coded layer video sequence referring to the sequence parameter set,

wherein the first rule further specifies that responsive to the value of the first field being greater than 0, the first field specifies the identifier of the video parameter set referred to by the sequence parameter set, and the value of the third field is in a range of 0 to a value of the second field, inclusive, and

wherein the conversion is performed according to a second rule, and wherein the second rule specifies that, responsive to the value of the first field being equal to 0, a total number of output layer sets specified by the video parameter set and a number of layers in a first output layer set is equal to 1, and a layer identifier of the video parameter set is equal to a network abstraction layer (NAL) unit header identifier (nuh_layer_id) of a layer in a coded video sequence.

2. The method of claim 1 , wherein the value of the second field is equal to the first maximum number minus 1.

3. The method of claim 1 , wherein the value of the third field is equal to the second maximum number minus 1.

4. The method of claim 1 , wherein the conversion is performed according to a third rule, wherein the third rule specifies that pictures within a layer are permitted to reference picture parameter sets that have different values of syntax elements indicative of whether picture partitioning is enabled for a corresponding picture.

5. The method of claim 4 , wherein the syntax elements equal to a first value specifies that no picture partitioning is applied to each picture referring to a picture parameter set; and the syntax elements equal to a second value specifies that each picture referring to a corresponding picture parameter set is allowed to be partitioned into more than one tile or slice.

6. The method of claim 1 , wherein the conversion is performed according to a fourth rule, and wherein the fourth rule specifies that a responsive to a syntax element indicative of whether loop filtering across tiles of a picture is enabled being absent from a picture parameter set referred to by the picture, the syntax element is inferred to have a particular value.

7. The method of claim 6 , wherein the particular value is 0.

8. The method of claim 7 , wherein the syntax element equal to 0 specifies that in-loop filtering operations across the tiles of the picture are disabled for pictures referring to the picture parameter set.

9. The method of claim 1 , wherein the conversion includes encoding the video into the bitstream.

10. The method of claim 1 , wherein the conversion includes decoding the video from the bitstream.

11. An apparatus for processing video data comprising a processor and a non-transitory memory with instructions thereon, wherein the instructions upon execution by the processor, cause the processor to:

perform a conversion between a video and a bitstream of the video according to a first rule,

wherein the first rule specifies that responsive to a value of a first field of a sequence parameter set that is indicative of an identifier of a video parameter set being equal to 0, a second field of the video parameter set that is indicative of a first maximum number of sublayers allowed in a layer specified by the video parameter set is construed to be equal to a third field in the sequence parameter set, and the value of the third field is in a range of 0 to 6, inclusive,

wherein the third field is indicative of a second maximum number of sublayers allowed in each coded layer video sequence referring to the sequence parameter set, and

wherein the first rule further specifies that responsive to the value of the first field being greater than 0, the first field specifies the identifier of the video parameter set referred to by the sequence parameter set, and the value of the third field is in a range of 0 to a value of the second field, inclusive, and

wherein the conversion is performed according to a second rule, and wherein the second rule specifies that, responsive to the value of the first field being equal to 0, a total number of output layer sets specified by the video parameter set and a number of layers in a first output layer set is equal to 1, and a layer identifier of the video parameter set is equal to a network abstraction layer (NAL) unit header identifier (nuh_layer_id) of a layer in a coded video sequence.

12. The apparatus of claim 11 , wherein the conversion is performed according to, a third rule, or a fourth rule,

wherein the third rule specifies that pictures within a layer are permitted to reference picture parameter sets that have different values of syntax elements indicative of whether picture partitioning is enabled for a corresponding picture, or

wherein the fourth rule specifies that a responsive to a syntax element indicative of whether loop filtering across tiles of a picture is enabled being absent from a picture parameter set referred to by the picture, the syntax element is inferred to have a particular value.

13. The apparatus of claim 12 , wherein the syntax elements equal to a first value specifies that no picture partitioning is applied to each picture referring to a picture parameter set; and the syntax elements equal to a second value specifies that each picture referring to a corresponding picture parameter set is allowed to be partitioned into more than one tile or slice.

14. The apparatus of claim 12 , wherein the particular value is 0.

15. The apparatus of claim 11 , wherein the value of the second field is equal to the first maximum number minus 1.

16. The apparatus of claim 11 , wherein the value of the third field is equal to the second maximum number minus 1.

17. A non-transitory computer-readable storage medium storing instructions that cause a processor to:

perform a conversion between a video and a bitstream of the video according to a first rule,

wherein the first rule specifies that responsive to a value of a first field of a sequence parameter set that is indicative of an identifier of a video parameter set being equal to 0, a second field of the video parameter set that is indicative of a first maximum number of sublayers allowed in a layer specified by the video parameter set is construed to be equal to a third field in the sequence parameter set, and the value of the third field is in a range of 0 to 6, inclusive,

wherein the third field is indicative of a second maximum number of sublayers allowed in each coded layer video sequence referring to the sequence parameter set, and

wherein the first rule further specifies that responsive to the value of the first field being greater than 0, the first field specifies refers to the identifier of the video parameter set referred to by the sequence parameter set, and the value of the third field is in a range of 0 to a value of the second field, inclusive, and

wherein the conversion is performed according to a second rule, and wherein the second rule specifies that, responsive to the value of the first field being equal to 0, a total number of output layer sets specified by the video parameter set and a number of layers in a first output layer set is equal to 1, and a layer identifier of the video parameter set is equal to a network abstraction layer (NAL) unit header identifier (nuh_layer_id) of a layer in a coded video sequence.

18. The non-transitory computer-readable storage medium of claim 17 , wherein the conversion is performed according to a second rule, a third rule or a fourth rule,

wherein the third rule specifies that pictures within a layer are permitted to reference picture parameter sets that have different values of syntax elements indicative of whether picture partitioning is enabled for a corresponding picture, or

wherein the fourth rule specifies that a responsive to a syntax element indicative of whether loop filtering across tiles of a picture is enabled being absent from a picture parameter set referred to by the picture, the syntax element is inferred to have a particular value.

19. A method for storing a bitstream, comprising:

determining whether a value of a first field of a sequence parameter set is set equal to 0 or is greater than 0;

generating the bitstream of a video according to a first rule; and

storing the bitstream in a non-transitory computer-readable recording medium,

wherein the first rule specifies that responsive to the value of the first field of the sequence parameter set that is indicative of an identifier of a video parameter set being equal to 0, a second field of the video parameter set that is indicative of a first maximum number of sublayers allowed in a layer specified by the video parameter set is construed to be equal to a third field in the sequence parameter set, and the value of the third field is in a range of 0 to 6, inclusive,

wherein the third field is indicative of a second maximum number of sublayers allowed in each coded layer video sequence referring to the sequence parameter set, and

wherein the first rule further specifies that responsive to the value of the first field being greater than 0, the first field specifies the identifier of the video parameter set referred to by the sequence parameter set, and the value of the third field is in a range of 0 to a value of the second field, inclusive, and

wherein the generating is performed according to a second rule, and wherein the second rule specifies that, responsive to the value of the first field being equal to 0, a total number of output layer sets specified by the video parameter set and a number of layers in a first output layer set is equal to 1, and a layer identifier of the video parameter set is equal to a network abstraction layer (NAL) unit header identifier (nuh_layer_id) of a layer in a coded video sequence.

20. The method of claim 19 , wherein the generating is performed according to a third rule or a fourth rule,

wherein the third rule specifies that pictures within a layer are permitted to reference picture parameter sets that have different values of syntax elements indicative of whether picture partitioning is enabled for a corresponding picture, or

wherein the fourth rule specifies that a responsive to a syntax element indicative of whether loop filtering across tiles of a picture is enabled being absent from a picture parameter set referred to by the picture, the syntax element is inferred to have a particular value.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 20, 2023
From: WANG, YE-KUI
To: BYTEDANCE INC.
Reel/Frame 062740/0017 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 20, 2023
From: DENG, ZHIPIN
To: BEIJING ZITIAO NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 062740/0035 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 20, 2023
From: BEIJING ZITIAO NETWORK TECHNOLOGY CO., LTD.
To: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 062740/0051 →
Priority Claims (1)
WO PCT/CN2020/091758 · May 22, 2020 · international
Continuity (2)
Continuation PCTCN2021095124 · May 21, 2021
Related Publication 20230078115A1 · Mar 16, 2023
References Cited (27)
US 11425422B2 · He · 2022 [cited by examiner]
US 11601655B2 · Chen · 2023 [cited by examiner]
US 11611778B2 · Samuelsson · 2023 [cited by examiner]
CN 105519119A · 2016 [cited by applicant]
JP 7513756B2 · 2024 [cited by applicant]
WO 2015102042A1 · 2015 [cited by applicant]
WO 2015138979A2 · 2015 [cited by applicant]
WO 2021195514A1 · 2021 [cited by applicant]
WO 2021235412A1 · 2021 [cited by applicant]
Document: JVET-R2001-vA, Bross, B., et al., “Versatile Video Coding (Draft 9),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr. 15-24, 2020, 524 pa… [cited by applicant]
“Series H: Audiovisual and Multimedia Systems Infrastructure of audiovisual services—Coding of moving video High efficiency video coding,” ITU-T and ISO/IEC, Rec. ITU-T H.265 | ISO/IEC 23008-2 (in force edition), Feb. 2… [cited by applicant]
Document: JVET-G1001-v1, Chen, J., et al., “Algorithm Description of Joint Exploration Test Model 7 (JEM 7),” Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 7th Meeting: Torino, IT… [cited by applicant]
Document: JVET-Q2002-v3, Chen, J., et al., “Algorithm description for Versatile Video Coding and Test Model 8 (VTM 8),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 17th Meeting: Bru… [cited by applicant]
VTM software, Retrieved from the internet: https://vcgit.hhi.fraunhofer.de/jvet/VVCSoftware_VTM.git, Feb. 17, 2023, 5 pages. [cited by applicant]
Document: JCTVC-AC1005-v2, Boyce, J., et al., “HEVC Additional Supplemental Enhancement Information (Draft 4),” Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 29th Me… [cited by applicant]
Document: JVET-R0278, Seregin, V., et al., “AHG8: On SPS sharing and slice type constraint,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr. 15-24,… [cited by applicant]
Document: JVET-Q0402-v2, Skupin, R., et al., “AHG12: On subpicture and scalability,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 17th Meeting: Brussels, BE, Jan. 7-17, 2020, 7 pages. [cited by applicant]
Document: JVET-S0049-v1, Wang, Y., et al., “AHG9/AHG8/AHG12: On parameter sets and picture header,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 19th Meeting: by teleconference, Jun.… [cited by applicant]
Document: JVET-S0200, Jhu, H., et al., “AHG9: On ph_inter_slice_allowed_flag in GDR picture,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 19th Meeting: by teleconference, Jun. 22-Ju… [cited by applicant]
Document: JVET-R0125, Choi, B., et al., “AHG8/AHG9: On signaling max No. of sublayers,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr. 15-24, 2020… [cited by applicant]
Document: JVET-R0199-v1, Kim, D., et al., “AHG9: On vps_max_sublayers_minus1,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr. 15-24, 2020, 3 pages. [cited by applicant]
Document: JVET-R0158-v1, Wang, B., et al., “Semantic bug fixes for syntax elements in VPS and SPS,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr.… [cited by applicant]
Document: JVET-R0222, Luo, J., et al., “AHG9: Sps sublayer syntax cleanup,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/JTC 1/SC 29/WG 11 18th Meeting: by teleconference, Apr. 15-24, 2020, 3 pages. [cited by applicant]
Foreign Communication From A Related Counterpart Application, International Application No. PCT-CN2021-095115, International Search Report dated Aug. 20, 2021, 11 pages. [cited by applicant]
Foreign Communication From A Related Counterpart Application, International Application No. PCT-CN2021-095124, International Search Report dated Aug. 20, 2021, 10 pages. [cited by applicant]
Document: JVET-R0339-vB, Sullivan, G., et al., “Agenda and report of the Ccategory 1 AHG pre-meeting for the 18th JVET meeting,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 18th Mee… [cited by applicant]
Extended European Search Report from European Application No. 21807970.5 dated Nov. 29, 2023, 11 pages. [cited by applicant]