IP Library Granted Patent US 12,192,511
Granted Patent B2
US 12,192,511 · App. 17/956,711 · Granted Jan 7, 2025

Methods and apparatuses for signaling of syntax elements in video coding

Inventors: Yi-Wen Chen (San Diego, CA); Xiaoyu Xiu (San Diego, CA); Tsung-Chuan Ma (San Diego, CA); Hong-Jheng Jhu (San Diego, CA); Wei Chen (San Diego, CA); Xianglin Wang (San Diego, CA); Bing Yu (Beijing, CN)
Assignee: BEIJING DAJIA INTERNET INFORMATION TECHNOLOGY CO., LTD.
H04N19/521H04N19/105H04N19/159H04N19/174H04N19/573H04N19/577H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,192,511
App. No.
17/956,711
Granted
Jan 7, 2025
Kind
B2
Abstract

Methods and apparatuses for video coding are provided. The method includes that a decoder determines whether one or more reference picture lists are signaled in a picture header (PH) associated with a picture and whether the one or more reference picture lists indicate that one or more slices associated with the picture are bi-predictive. The method further includes that the decoder adds one or more constraints to one or more syntax elements in the PH in response to determining that the one or more reference picture lists are signaled in the PH and the one or more reference picture lists indicate that the one or more slices are not bi-predictive.

Claims (50)

1. A method for video decoding, comprising:

determining, by a decoder, whether information of one or more reference picture lists is signaled in a picture header (PH) associated with a picture, and determining from the information of the one or more reference picture lists whether one or more slices associated with the picture are bi-predictive in response to determining that the information of the one or more reference picture lists is signaled in the PH; and

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are not bi-predictive, not parsing, by the decoder, one or more syntax elements in the PH.

2. The method according to claim 1 , wherein the one or more syntax elements comprise one or more flags applicable for the one or more slices.

3. The method according to claim 1 , further comprising:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether a corresponding motion vector difference (MVD) coding syntax structure is not parsed and whether two variables are set to be zero for the one or more slices associated with the PH, wherein the two variables respectively specify a difference between a list vector component and a prediction corresponding to the list vector component;

in response to determining that the one or more syntax elements equal to 0, parsing the MVD coding syntax structure for the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, not parsing the MVD coding syntax structure to decode the one or more slices.

4. The method according to claim 1 , further comprising:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether bi-directional optical flow (BDOF) inter prediction based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enabling the BDOF inter prediction based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disabling the BDOF inter prediction based inter bi-prediction to decode the one or more slices.

5. The method according to claim 1 , further comprising:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether decoder motion vector refinement (DMVR) based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enabling the DMVR based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disabling the DMVR based inter bi-prediction to decode the one or more slices.

6. An apparatus for video decoding, comprising:

one or more processors; and

a memory configured to store instructions executable by the one or more processors; wherein the one or more processors, upon execution of the instructions, are configured to:

determine whether information of one or more reference picture lists is signaled in a picture header (PH) associated with a picture, and determine from the information of the one or more reference picture lists whether one or more slices associated with the picture are bi-predictive in response to determining that the information of the one or more reference picture lists is signaled in the PH; and

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are not bi-predictive, not parse one or more syntax elements in the PH.

7. The apparatus according to claim 6 , wherein the one or more syntax elements comprise one or more flags applicable for the one or more slices.

8. The apparatus according to claim 6 , wherein the one or more processors are further configured to:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtain the one or more syntax elements to specify whether a corresponding motion vector difference (MVD) coding syntax structure is not parsed and whether two variables are set to be zero for the one or more slices associated with the PH, wherein the two variables respectively specify a difference between a list vector component and a prediction corresponding to the list vector component;

in response to determining that the one or more syntax elements equal to 0, parse the MVD coding syntax structure for the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, not parse the MVD coding syntax structure to decode the one or more slices.

9. The apparatus according to claim 6 , wherein the one or more processors are further configured to:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtain the one or more syntax elements to specify whether bi-directional optical flow (BDOF) inter prediction based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enable the BDOF inter prediction based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disable the BDOF inter prediction based inter bi-prediction to decode the one or more slices.

10. The apparatus according to claim 6 , wherein the one or more processors are further configured to:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtain the one or more syntax elements to specify whether decoder motion vector refinement (DMVR) based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enable the DMVR based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disable the DMVR based inter bi-prediction to decode the one or more slices.

11. A non-transitory computer-readable storage medium for video decoding storing a bitstream to be decoded by a method for video decoding comprising:

determining whether information of one or more reference picture lists is signaled in a picture header (PH) associated with a picture, and determining from the information of the one or more reference picture lists whether one or more slices associated with the picture are bi-predictive in response to determining that the information of the one or more reference picture lists is signaled in the PH; and

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are not bi-predictive, not parsing one or more syntax elements in the PH.

12. The non-transitory computer-readable storage medium according to claim 11 , wherein the one or more syntax elements comprise one or more flags applicable for the one or more slices.

13. The non-transitory computer-readable storage medium according to claim 11 , wherein the method further comprises:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether a corresponding motion vector difference (MVD) coding syntax structure is not parsed and whether two variables are set to be zero for the one or more slices associated with the PH, wherein the two variables respectively specify a difference between a list vector component and a prediction corresponding to the list vector component;

in response to determining that the one or more syntax elements equal to 0, parsing the MVD coding syntax structure for the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, not parsing the MVD coding syntax structure to decode the one or more slices.

14. The non-transitory computer-readable storage medium according to claim 11 , wherein the method further comprises:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether bi-directional optical flow (BDOF) inter prediction based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enabling the BDOF inter prediction based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disabling the BDOF inter prediction based inter bi-prediction to decode the one or more slices.

15. The non-transitory computer-readable storage medium according to claim 11 , wherein the method further comprises:

in response to determining that the information of the one or more reference picture lists is signaled in the PH and determining from the information of the one or more reference picture lists that the one or more slices are bi-predictive, obtaining the one or more syntax elements to specify whether decoder motion vector refinement (DMVR) based inter bi-prediction is disabled for the one or more slices associated with the PH;

in response to determining that the one or more syntax elements equal to 0, enabling the DMVR based inter bi-prediction to decode the one or more slices; and

in response to determining that the one or more syntax elements equal to 1, disabling the DMVR based inter bi-prediction to decode the one or more slices.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 4, 2022
From: CHEN, YI-WEN; XIU, XIAOYU; MA, TSUNG-CHUAN; JHU, HONG-JHENG; CHEN, WEI; WANG, XIANGLIN; YU, BING
To: BEIJING DAJIA INTERNET INFORMATION TECHNOLOGY CO., LTD.
Reel/Frame 061307/0819 →
Continuity (3)
Continuation PCTUS2021023727 · Mar 23, 2021
Provisional Application 63003226 · Mar 31, 2020
Related Publication 20230031699A1 · Feb 2, 2023
References Cited (46)
US 8254455B2 · Wu · 2012 [cited by examiner]
US 10931945B2 · Francois · 2021 [cited by examiner]
US 11496771B2 · Seregin · 2022 [cited by examiner]
US 11533472B2 · Chen · 2022 [cited by examiner]
US 20130268621A1 · Mese · 2013 [cited by examiner]
US 20130272375A1 · Yu · 2013 [cited by examiner]
US 20130336407A1 · Chen · 2013 [cited by examiner]
US 20140086324A1 · Ramasubramonian · 2014 [cited by examiner]
US 20160330255A1 · Denoual · 2016 [cited by examiner]
US 20170302951A1 · Laxman et al. · 2017 [cited by applicant]
US 20210195179A1 · Coban · 2021 [cited by examiner]
US 20210266600A1 · Seregin · 2021 [cited by examiner]
US 20210274215A1 · Kang · 2021 [cited by examiner]
US 20210314624A1 · Coban · 2021 [cited by examiner]
US 20210368208A1 · Samuelsson · 2021 [cited by examiner]
US 20220053207A1 · Deshpande · 2022 [cited by applicant]
US 20230026475A1 · Deshpande · 2023 [cited by examiner]
US 20230040224A1 · Chen · 2023 [cited by examiner]
US 20230115242A1 · Laroche · 2023 [cited by examiner]
US 20230353749A1 · Nam · 2023 [cited by examiner]
US 20240137546A1 · Kuo · 2024 [cited by examiner]
CN 110868613A · 2020 [cited by applicant]
CN 110944172A · 2020 [cited by applicant]
WO 2014104242A1 · 2014 [cited by applicant]
WO 2019147628A1 · 2019 [cited by applicant]
WO 2019192301A1 · 2019 [cited by applicant]
WO 2019194435A1 · 2019 [cited by applicant]
WO 2021201759A1 · 2021 [cited by applicant]
International Search Report of International Application No. PCT Application No. PCT/US2021/023727 dated Jul. 7, 2021, (3p). [cited by applicant]
Bross, Benjamin et al., “Versatile Video Coding (Draft 8)”, JVET-Q2001-vE, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 17th Meeting: Brussels, (512p). [cited by applicant]
Wan, Wade et al., “AHG8: RPR Scaling Window Issues”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-Q0487-v2, 17th Meeting: Brussels, BE Jan. 7-17, 2020, (6p). [cited by applicant]
Hendry et al., “[AHG9]: on signalling of TMVP enabled flag and collocated reference picture”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-Q0207, 17th Meeting: Brussels, BE Jan… [cited by applicant]
Sun, Yucheng et al., Non-CE4: Constraints on block size for ATMVP, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 17th Meeting: Brussels, BE, Jan. 7-17, 2020, (2p). [cited by applicant]
Benjamin Bross et al., “Versatile Video Coding (Draft 8)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11,JVET-Q2001-vED, 17th Meeting: Brussels, BE, Jan. 7-17, 2020,(512p). [cited by applicant]
MediaTek Inc, Shih-Ta Hsiang et al., “AHG9: Overhead reduction for picture header and slice header” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R0052-v1, 18th Meeting: by tele… [cited by applicant]
Bytedance Inc, Li Zhang, et al., “AHG9: On allowed slice types in a picture”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R0061-v1, 18th Meeting: by teleconference, Apr. 15-24… [cited by applicant]
Sharp Corporation., Takeshi Chujoh et al., Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R0137-v14. 18th Meeting: by teleconference, Apr. 15-24, 2020, (8p). [cited by applicant]
Ericsson, Martin Pettersson et al., “AHG9: On B-slice signaling in the PH and derivation of slice_type”, oint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11,JVET-R0250, 18th Meeting: by tele… [cited by applicant]
Kwai Inc., Yi-Wen Chen et al., “AHG9: On syntax signaling conditions in picture header”, oint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11,JVET-R0324, 18th Meeting: by teleconference, Apr.… [cited by applicant]
S-T Hsiang et al: “AHG9: Overhead reduction for picture header”, 17. JVET Meeting; Jan. 7, 2020-Jan. 17, 2020; Brussels; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16) No. JVET-Q0176 ;m5176… [cited by applicant]
Y-W Chen(Kwai)et al: “AHG9:On syntax signalling conditions in picture header”, 130. MPEG Meeting; Apr. 20, 2020-Apr. 24, 2020; Alpbach; (Motion Picture Expert Group or ISO/IEC JTC1/SC29/WG11) No. m53308; JVET-R0324 Apr.… [cited by applicant]
Y-K Wang(Bytedance) et al: “AHG9:A summary of proposals on PH and SH syntax”, 130.MPEG Meeting; Apr. 20, 2020-Apr. 24, 2020; Alpbach; (Motion Picture Expert Group or ISO JTC1/SC29/WG11) No. m53763 ; JVET-R0410 Apr. 15, … [cited by applicant]
Benjamin Bross et al: “Versatile Video Coding(Draft 7)”, 16. JVET Meeting; Oct. 1, 2019-Oct. 11, 2019; Geneva; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16) No. JVET-P2001-vE; JVET-P2001 ;… [cited by applicant]
Esenlik(Huawei) S et al: “AHG9: Slice Level control of coding tools”, 17. JVET Meeting; Jan. 7, 2020-Jan. 17, 2020; Brussels; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16) No. JVET-Q0360;m… [cited by applicant]
Office Action issued to Korean Application No. 2023-052605144 dated Jun. 9, 2023 with English translation, (9p). [cited by applicant]
Bross, Benjamin, et al. “Versatile Video Coding (Draft 8)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG11, JVET-Q2001-v14, 17th Meeting Brussels, Feb. 2020, (80p). [cited by applicant]