IP Library Granted Patent US 12,477,154
Granted Patent B2
US 12,477,154 · App. 17/853,153 · Granted Nov 18, 2025

Encoder, decoder and corresponding methods and apparatus

Inventors: Xiang Ma (Moscow, RU); Haitao Yang (Shenzhen, CN)
Assignee: HUAWEI TECHNOLOGIES CO., LTD.
H04N19/70H04N19/105H04N19/172H04N19/186H04N19/30H04N19/50
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,477,154
App. No.
17/853,153
Granted
Nov 18, 2025
Kind
B2
Abstract

A method for decoding a coded video bitstream is provided. The method includes: obtaining a reference layer syntax element by parsing the coded video bitstream, wherein a value of the reference layer syntax element specifying whether a layer with index k is a direct reference layer for a layer with index i; determining whether the layer with index j is a reference layer for the layer with index i based on the value of the reference layer syntax element; and when a condition is satisfied, predicting a picture of the layer with index i based on the layer with index j, wherein the value of a chroma format related syntax element applied to the layer with index i is the same as the value of a chroma format related syntax element applied to the layer with index j.

Claims (29)

1 . A method for decoding a coded video bitstream, comprising:

obtaining a reference layer syntax element by parsing the coded video bitstream, wherein a value of the reference layer syntax element specifying whether a layer with index k is a direct reference layer for a layer with index i, wherein i and k are integers larger than or equal to 0;

determining whether a layer with index j is a reference layer for the layer with index i based on the value of the reference layer syntax element, wherein the layer with index j is determined to be a reference layer for the layer with index i if the layer with index i is a reference layer for the layer with index k and the value of the reference layer syntax element specifies that the layer with index k is a direct reference layer for the layer with index i, wherein j is an integer larger than or equal to 0; and

when the layer with index j is the reference layer for the layer with index i, predicting a picture of the layer with index i based on the layer with index j, wherein a value of a chroma format related syntax element applied to the layer with index i is the same as a value of a chroma format related syntax element applied to the layer with index j.

2 . The method of claim 1 , further comprising:

when the layer with index j is the reference layer for the layer with index i, and the value of the chroma format related syntax element applied to the layer with index i is not the same as the value of the chroma format related syntax element applied to the layer with index j, stopping decoding the coded video bitstream.

3 . The method of claim 1 , further comprising:

when the layer with index j is not the reference layer for the layer with index i, predicting the picture of the layer with index i without using the layer with index j.

4 . A method for decoding a coded video bitstream, comprising:

obtaining a reference layer syntax element by parsing the coded video bitstream, wherein a value of the reference layer syntax element specifying whether a layer with index j is a direct reference layer for a layer with index i, wherein i and j are integers larger than or equal to 0; and

when a condition is satisfied, predicting a picture of the layer with index i based on the layer with index j, wherein the layer with index j is a direct reference layer for the layer with index i, wherein a value of a chroma format related syntax element applied to the layer with index i is the same as a value of a chroma format related syntax element applied to the layer with index j, and wherein the condition comprises the value of the reference layer syntax element specifies the layer with index j is a direct reference layer for the layer with index i.

5 . The method of claim 4 , wherein the reference layer syntax element is a video parameter set (VPS) level syntax element that is applied to the layer with index j and the layer with index i.

6 . The method of claim 4 , wherein the chroma format related syntax element is a sequence parameter set (SPS) level syntax element that is applied to the layer with index j or the layer with index i.

7 . The method of claim 4 , further comprising:

obtaining the chroma format related syntax element applied to the layer with index i and the chroma format related syntax element applied to the layer with index j by parsing the coded video bitstream; and wherein the condition further comprises the value of the chroma format related syntax element applied to the layer with index i is the same as the value of the chroma format related syntax element applied to the layer with index j.

8 . The method of claim 4 , further comprising:

when the condition is satisfied, determining, without obtaining the chroma format related syntax element applied to the layer with index i by parsing the coded video bitstream, that the value of the chroma format related syntax element applied to the layer with index i is the value of the chroma format related syntax element applied to the layer with index j.

9 . A method for encoding a video, comprising:

determining whether a layer with index j is a direct reference layer for a layer with index i, wherein i and j are integers larger than or equal to 0;

when the layer with index j is the direct reference layer for the layer with index i, encoding a reference layer syntax element with a value specifying the layer with index j is a direct reference layer for the layer with index i into a video bitstream, and encoding a chroma format related syntax element applied to the layer with index i and a chroma format related syntax element applied to the layer with index j into the video bitstream, wherein the value of the chroma format related syntax element applied to the layer with index i is the same as the value of the chroma format related syntax element applied to the layer with index j; and

when the layer with index i is the direct reference layer for the layer with index i, predicting a picture of the layer with index i based on the layer with index j.

10 . The method of claim 9 , further comprising:

when the layer with index j is not a reference layer for the layer with index i, predicting a picture of the layer with index i without using the layer with index j.

11 . A decoder, comprising:

one or more processors; and

a non-transitory computer-readable storage medium coupled to the one or more processors and storing programming instructions, which when executed by the one or more processors, cause the decoder to perform the method according to claim 2 .

12 . An encoder, comprising:

one or more processors; and

a non-transitory computer-readable storage medium coupled to the processors and storing programming instructions, which when executed by the one or more processors, cause the encoder to perform the method according to claim 9 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 13, 2023
From: MA, XIANG; YANG, HAITAO
To: HUAWEI TECHNOLOGIES CO., LTD.
Reel/Frame 062674/0612 →
Priority Claims (2)
WO PCT/CN2019/130804 · Dec 31, 2019 · international
WO PCT/CN2020/070153 · Jan 2, 2020 · international
Continuity (2)
Continuation PCTCN2020142501 · Dec 31, 2020
Related Publication 20220345748A1 · Oct 27, 2022
References Cited (25)
US 20130279576A1 · Chen et al. · 2013 [cited by applicant]
US 20140301484A1 · Wang · 2014 [cited by examiner]
US 20150103886A1 · He et al. · 2015 [cited by applicant]
US 20150195554A1 · Misra et al. · 2015 [cited by applicant]
US 20160241869A1 · Choi et al. · 2016 [cited by applicant]
US 20170019666A1 · Deshpande · 2017 [cited by applicant]
US 20170041641A1 · Deshpande et al. · 2017 [cited by applicant]
US 20180115787A1 · Koo et al. · 2018 [cited by applicant]
US 20190261011A1 · Hannuksela · 2019 [cited by applicant]
CN 105122816A · 2015 [cited by applicant]
CN 105191315A · 2015 [cited by applicant]
CN 105393542A · 2016 [cited by applicant]
CN 106416250A · 2017 [cited by applicant]
JP 2017500764A · 2017 [cited by applicant]
RU 2639958C2 · 2017 [cited by applicant]
WO 2015054634A2 · 2015 [cited by applicant]
Takeshi Chujoh et al., On improvement of collocated_ref_idx [online], JVET-Q0130-v1, Internet URL: https://jvet-experts.org/doc_end_user/documents/17_Brussels/wg11/JVET-Q0130-v1.zip>, Dec. 30, 2019. [cited by applicant]
Byeongdoo Choi et al., Video parameter set design [online], JCTVC-L0132, Internet URL: http://phenix.it-sudparis.eu/jct/doc_end_user/documents/12_Geneva/wg11/JCTVC-L0132-v2.zip>, Jan. 14, 2013. [cited by applicant]
Jill M. Boyce, et.al., Overview of SHVC: Scalable Extensions of the High Efficiency Video Coding Standard [online], Published in: IEEE Transactions on Circuits and Systems for Video Technology (vol. 26, Issue: 1, Jan. 2… [cited by applicant]
Document: JVET-Q0172-v1, Tzu-Der Chuang et al, AHG9: Chroma format and bitdepth constraints for multi-layer structures, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 17th Meeting: Br… [cited by applicant]
ITU-T H.264(Jun. 2019), Series H: Audiovisual and Multimedia Systems, Infrastructure of audiovisual services-Coding of moving video, Advanced video coding for generic audiovisual services, total 836 pages. [cited by applicant]
Document: JVET-P2001-vE, Benjamin Bross et al, Versatile Video Coding (Draft 7), Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 16th Meeting: Geneva, CH, Oct. 1-11, 2019, 491 pages. [cited by applicant]
ITU-T H.265(Jun. 2019), Series H: Audiovisual and Multimedia Systems, Infrastructure of audiovisual services-Coding of moving video, High efficiency video coding, total 696 pages. [cited by applicant]
Biatek T et al: “MIP with fixed-length mode coding and memory reduction”, 17. JVET Meeting; Jan. 7, 2020-Jan. 17, 2020; Brussels; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16 ), No. JVET-Q… [cited by applicant]
T-D Chuang et al: “AHG9: Chroma format and bitdepth constraint for multi-layer structure”, 17. JVET Meeting; Jan. 7, 2020-Jan. 17, 2020; Brussels; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-TSG.… [cited by applicant]