IP Library Granted Patent US 12,309,356
Granted Patent B2
US 12,309,356 · App. 17/678,693 · Granted May 20, 2025

Method for output layer set for multilayered video stream

Inventors: Byeongdoo Choi (Palo Alto, CA); Shan Liu (San Jose, CA); Stephan Wenger (Hillsborough, CA)
Assignee: TENCENT AMERICA LLC
H04N19/105H04L65/75H04L65/752H04N19/119H04N19/187H04N19/46H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,309,356
App. No.
17/678,693
Granted
May 20, 2025
Kind
B2
Abstract

Systems and methods for coding and decoding are provided. A method includes encoding a video stream, the coded video stream including a video parameter set (VPS) and video data partitioned into a plurality of layers; and sending the coded video stream to a decoder, wherein the VPS includes, (1) at least one first syntax element that specifies at least one first layer, from among the plurality of layers, to be outputted in an output layer set, and (2) at least one second syntax element that indicates profile-tier-level information of the output layer set.

Claims (45)

1. A method of video encoding performed by at least one processor, the method comprising:

obtaining a video stream including a video parameter set (VPS) and video data partitioned into a plurality of layers;

when a maximum allowed number of layers for each coded video sequence referring to the VPS is greater than 1:

determining a first value for at least one first syntax element that specifies output layer set (OLS) information for the plurality of layers associated with the video data;

determining a second value for at least one second syntax element that indicates a number of profile-tier-level information sets in the VPS; and

setting respective syntax elements in the VPS using the first value and the second value;

when the maximum allowed number of layers for each coded video sequence referring to the VPS is not greater than 1, forgoing setting the respective syntax elements in the VPS using the first value and the second value; and

encoding the video stream, wherein the encoded video stream includes the VPS.

2. The method of claim 1 , further comprising determining a third value for a set of one or more syntax elements specifying profile-tier-level information of an OLS in the plurality of layers.

3. The method of claim 1 , further comprising:

determining a fourth value for a fourth syntax element that specifies a mode of output layer signaling for an OLS; and

setting a syntax element in the VPS using the fourth value, wherein the VPS of the encoded video stream comprises the fourth value.

4. The method of claim 3 , wherein the at least one first syntax element is signalled within the VPS based on the mode specified by the fourth syntax element.

5. The method of claim 4 , wherein the at least one first syntax element includes a flag specifying whether one of the plurality of layers is to be output.

6. The method of claim 1 , further comprising:

determining a fourth value for a fourth syntax element that specifies a mode of output layer set signaling for a plurality of output layer sets; and

setting a syntax element in the VPS using the fourth value, wherein the VPS of the encoded video stream comprises the fourth value.

7. The method of claim 1 , wherein the VPS further includes:

an ols_mode_idc parameter specifying whether, for each output layer set, all layers or only a highest layer in the output layer set is an output layer;

a first flag that specifies whether layers specified by the VPS may use inter-layer prediction; and

a second flag that specifies whether each output layer set contains only one layer or at least one output layer set contains more than one layer.

8. A method of video decoding performed by at least one processor, the method comprising:

receiving an encoded video stream including a video parameter set (VPS) and video data partitioned into a plurality of layers;

when a maximum allowed number of layers for each coded video sequence referring to the VPS is greater than 1:

identifying a first parameter from the VPS that specifies output layer set (OLS) information for the plurality of layers associated with the video data;

identifying a second parameter from the VPS that indicates a number of profile-tier-level information sets in the VPS;

decoding a portion of the video data of the encoded video stream based on the first parameter and the second parameter; and

when the maximum allowed number of layers for each coded video sequence referring to the VPS is not greater than 1, decoding the portion of the video data of the encoded video stream without identifying the first parameter or the second parameter.

9. The method of claim 8 , wherein the second parameter is identified based on a set of syntax elements specifying profile-tier-level information.

10. The method of claim 8 , further comprising identifying a fourth parameter from the VPS that specifies a mode of output layer signaling for an OLS, wherein the portion of the video data is decoded based on the fourth parameter.

11. The method of claim 8 , wherein the first parameter is identified based on a flag specifying whether one of the plurality of layers is to be output.

12. The method of claim 8 , further comprising identifying a fourth parameter from the VPS that specifies a mode of output layer set signaling for a plurality of output layer sets, wherein the portion of the video data is decoded based on the fourth parameter.

13. The method of claim 8 , further comprising identifying an ols_mode_idc parameter specifying whether, for each output layer set, all layers or only a highest layer in the output layer set is an output layer.

14. The method of claim 8 , further comprising identifying a flag that specifies whether layers specified by the VPS may use inter-layer prediction.

15. The method of claim 8 , further comprising identifying a flag that specifies whether each output layer set contains only one layer or at least one output layer set contains more than one layer.

16. A method of processing visual media data, comprising:

obtaining a visual media file; and

performing a conversion between the visual media file and a bitstream of a visual media data, wherein:

the bitstream comprises video data partitioned into a plurality of layers and a video parameter set (VPS),

when a maximum allowed number of layers for each coded video sequence referring to the VPS is greater than 1, the VPS includes:

a first syntax element specifying output layer set (OLS) information for the plurality of layers associated with the video data, and

a second syntax element indicating a number of profile-tier-level information sets in the VPS; and

when the maximum allowed number of layers for each coded video sequence referring to the VPS is not greater than 1, the VPS does not include the first syntax element or the second syntax element.

17. The method of claim 16 , wherein a fourth syntax element of the bitstream specifies a mode of output layer signaling for the output layer set.

18. The method of claim 16 , wherein the first syntax element specifies a number of OLSs in a coded video sequence referring to the VPS.

Continuity (3)
Continuation 16987911 · Aug 7, 2020
Provisional Application 63001018 · Mar 27, 2020
Related Publication 20220182679A1 · Jun 9, 2022
References Cited (33)
US 9485508B2 · Wang · 2016 [cited by examiner]
US 11297350B1 · Choi et al. · 2022 [cited by applicant]
US 20140301469A1 · Wang · 2014 [cited by examiner]
US 20150103888A1 · Chen · 2015 [cited by examiner]
US 20150373361A1 · Wang · 2015 [cited by examiner]
US 20160316210A1 · Lee · 2016 [cited by examiner]
US 20170019673A1 · Tsukuba · 2017 [cited by examiner]
US 20170180744A1 · Deshpande · 2017 [cited by examiner]
US 20170347026A1 · Hannuksela · 2017 [cited by applicant]
US 20190158880A1 · Deshpande · 2019 [cited by applicant]
US 20210092406A1 · Seregin · 2021 [cited by examiner]
US 20210235124A1 · Seregin · 2021 [cited by examiner]
US 20210274204A1 · He · 2021 [cited by examiner]
US 20210352328A1 · Deshpande · 2021 [cited by examiner]
RU 2646381C2 · 2016 [cited by applicant]
RU 2610670C1 · 2017 [cited by applicant]
WO WO2021061394A1 · 2021 [cited by examiner]
International Search Report dated Feb. 8, 2021 from the International Searching Authority in International Application No. PCT/US2020/059697. [cited by applicant]
Written Opinion dated Feb. 8, 2021 from the International Searching Authority in International Application No. PCT/US2020/059697. [cited by applicant]
Benjamin Bross, et al, “Versatile Video Coding (Draft 8)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Jan. 7-17, 2020, 513 pages, Brussels, BE. [cited by applicant]
Extended European Search Report dated Jun. 7, 2022 in European Application No. 20926365.6. [cited by applicant]
Xuewei Meng et al., “CE5-related: On CC-ALF slice and picture header syntax”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 2020, Document: JVET-Q0326-v1, pp. 1-4 (4 pages total). [cited by applicant]
Office Action dated Oct. 31, 2022 issued Indian Application No. 202147051890. [cited by applicant]
Office Action dated Oct. 31, 2022 issued in Japanese Application No. 2021-562370. [cited by applicant]
Byeongdoo Choi et al., “AHG8: Output layer set and PTL signaling”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-P0225-v3, Oct. 1-11, 2019, 16th Meeting, Geneva, CH (4 pages tot… [cited by applicant]
Tencent America LLC, Canadian Office Action, CA Patent Application No. 3,137,047, Sep. 16, 2024, 6 pgs. [cited by applicant]
Tencent America LLC, Vietnamese Office Action, VN Patent Application No. 1-2021-07253 Sep. 9, 2024, 4 pgs. [cited by applicant]
Tencent America LLC, European Office Action, EP Patent Application No. 20926365.6, May 6, 2024, 6 pgs. [cited by applicant]
Tencent America LLC, Korean Office Action, KR Patent Application No. 10-2021-7036630, Sep. 5, 2024, 9 pgs. [cited by applicant]
Adarsh K. Ramasubramonian, “MV-HEVC/SHVC HLS: VPS Extension Syntax Elements with UE (V) Coded Syntax Elements”, Joint Collaborative Team on Video Coding JCT-VC of ITU-T SG 16 WP 3 and ISO/IEC JTC I/SC 29/WG 11, 16th Mee… [cited by applicant]
Sachin Deshpande et al., “AHG9: On PTL and HRD Parameters Signalling in VPS”, Document: JVET-Q0786-v2, Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 17th Meeting: Brussels, BE, Jan. 7-17, … [cited by applicant]
Tencent Technology, Canadian Office Action, CA Patent Application No. 3137047, Oct. 5, 2023, 6 pgs. [cited by applicant]
Tencent Technology, Indonesian Office Action, ID Patent Application No. P00202108588, Sep. 29, 2023, 4 pgs. [cited by applicant]