IP Library Granted Patent US 12,388,985
Granted Patent B2
US 12,388,985 · App. 18/733,595 · Granted Aug 12, 2025

Methods and apparatuses for video coding with triangle prediction

Inventors: Xianglin Wang (San Diego, CA); Yi-Wen Chen (San Diego, CA); Xiaoyu Xiu (San Diego, CA); Tsung-Chuan Ma (San Diego, CA); Hong-Jheng Jhu (San Diego, CA); Shuiming Ye (San Diego, CA)
Assignee: BEIJING DAJIA INTERNET INFORMATION TECHNOLOGY CO., LTD.
H04N19/105H04N19/119H04N19/13H04N19/139H04N19/46H04N19/52H04N19/573
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,388,985
App. No.
18/733,595
Granted
Aug 12, 2025
Kind
B2
Abstract

Methods, apparatuses, and non-transitory computer-readable storage mediums are provided for video decoding. The method includes: constructing, by a decoder, a first merge list comprising a plurality of candidates, based on a merge list construction process for regular merge prediction, wherein each one of the plurality of candidates is a motion vector comprising a List 0 motion vector, or a List 1 motion vector, or both; receiving, by the decoder, a first/second index value to indicate a first/second candidate that is chosen from the first merge list; receiving, by the decoder, a first/second binary flag that is coded using a first/second context-adaptive binary arithmetic coding (CABAC) context modeling method to indicate whether a List 0 motion vector of the first/second candidate or a List 1 motion vector of the first/second candidate is selected for a first PU of the geometric prediction.

Claims (40)

1. A method for video decoding with geometric prediction, comprising:

constructing, by a decoder, a first merge list comprising a plurality of candidates, based on a merge list construction process for regular merge prediction, wherein each one of the plurality of candidates is a motion vector comprising a List 0 motion vector, or a List 1 motion vector, or both;

receiving, by the decoder, a first index value to indicate a first candidate that is chosen from the first merge list;

receiving, by the decoder, a second index value to indicate a second candidate that is chosen from the first merge list;

receiving, by the decoder, a first binary flag that is coded using a first context-adaptive binary arithmetic coding (CABAC) context modeling method to indicate whether a List 0 motion vector of the first candidate or a List 1 motion vector of the first candidate is selected for a first PU of the geometric prediction; and

receiving, by the decoder, a second binary flag that is coded using a second CABAC context modeling method to indicate whether a List 0 motion vector of the second candidate or a List 1 motion vector of the second candidate is selected for a second PU of the geometric prediction,

wherein the first CABAC context modeling method or the second CABAC context modeling method comprises:

receiving each binary flag as a CABAC context bin, wherein a CABAC probability under each context model is initialized differently in response to whether a current picture uses a backward prediction.

2. The method for video decoding with geometric prediction of claim 1 , wherein at least two context models are available to decode the second binary flag, with a context model selected based on a value of the first binary flag.

3. The method for video decoding with geometric prediction of claim 1 , wherein the CABAC probability under each context model is initialized differently further comprises:

a higher CABAC initial probability is used for each binary flag to indicate selecting a List 0 motion vector in case the current picture does not use the backward prediction than in case the current picture uses the backward prediction.

4. The method for video decoding with geometric prediction of claim 1 , further comprising:

decoding, by the decoder, the first binary flag as a CABAC bypass bin, and decoding the second binary flag as a CABAC context coded bin.

5. An apparatus for video decoding with geometric prediction, comprising:

one or more processors; and

a memory configured to store instructions executable by the one or more processors;

wherein the one or more processors, upon execution of the instructions, are configured to:

construct a first merge list comprising a plurality of candidates, based on a merge list construction process for regular merge prediction, wherein each one of the plurality of candidates is a motion vector comprising a List 0 motion vector, or a List 1 motion vector, or both;

receive a first index value to indicate a first candidate that is chosen from the first merge list;

receive a second index value to indicate a second candidate that is chosen from the first merge list;

receive a first binary flag that is coded using a first context-adaptive binary arithmetic coding (CABAC) context modeling method to indicate whether a List 0 motion vector of the first candidate or a List 1 motion vector of the first candidate is selected for a first PU of the geometric prediction; and

receive a second binary flag that is coded using a second CABAC context modeling method to indicate whether a List 0 motion vector of the second candidate or a List 1 motion vector of the second candidate is selected for a second PU of the geometric prediction,

wherein the first CABAC context modeling method or the second CABAC context modeling method comprises:

receiving each binary flag as a CABAC context bin, wherein a CABAC probability under each context model is initialized differently in response to whether a current picture uses a backward prediction.

6. The apparatus for video decoding with geometric prediction of claim 5 , wherein at least two context models are available to decode the second binary flag, with a context model selected based on a value of the first binary flag.

7. The apparatus for video decoding with geometric prediction of claim 5 , wherein a higher CABAC initial probability is used for each binary flag to indicate selecting a List 0 motion vector in case the current picture does not use the backward prediction than in case the current picture uses the backward prediction.

8. The apparatus for video decoding with geometric prediction of claim 5 , wherein the one or more processors are further configured to:

decode the first binary flag as a CABAC bypass bin, and decode the second binary flag as a CABAC context coded bin.

9. A non-transitory computer-readable storage medium for video decoding stored computer-executable instructions and a bitstream that, when executed by one or more processors, cause the one or more processors to perform the following method with the bitstream:

constructing a first merge list comprising a plurality of candidates, based on a merge list construction process for regular merge prediction, wherein each one of the plurality of candidates is a motion vector comprising a List 0 motion vector, or a List 1 motion vector, or both;

receiving a first index value to indicate a first candidate that is chosen from the first merge list;

receiving a second index value to indicate a second candidate that is chosen from the first merge list;

receiving a first binary flag that is coded using a first context-adaptive binary arithmetic coding (CABAC) context modeling method to indicate whether a List 0 motion vector of the first candidate or a List 1 motion vector of the first candidate is selected for a first PU of the geometric prediction; and

receiving a second binary flag that is coded using a second CABAC context modeling method to indicate whether a List 0 motion vector of the second candidate or a List 1 motion vector of the second candidate is selected for a second PU of the geometric prediction,

wherein the first CABAC context modeling method or the second CABAC context modeling method comprises:

receiving each binary flag as a CABAC context bin, wherein a CABAC probability under each context model is initialized differently in response to whether a current picture uses a backward prediction.

10. The non-transitory computer-readable storage medium of claim 9 , wherein at least two context models are available to decode the second binary flag, with a context model selected based on a value of the first binary flag.

11. The non-transitory computer-readable storage medium of claim 9 , wherein a higher CABAC initial probability is used for each binary flag to indicate selecting a List 0 motion vector in case the current picture does not use the backward prediction than in case the current picture uses the backward prediction.

12. The non-transitory computer-readable storage medium of claim 9 , wherein the computer-executable instructions cause the one or more processors to further perform:

decoding the first binary flag as a CABAC bypass bin, and decoding the second binary flag as a CABAC context coded bin.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 14, 2024
From: WANG, XIANGLIN; CHEN, YI-WEN; XIU, XIAOYU; MA, TSUNG-CHUAN; JHU, HONG-JHENG; YE, SHUIMING
To: BEIJING DAJIA INTERNET INFORMATION TECHNOLOGY CO., LTD.
Reel/Frame 067737/0354 →
Continuity (4)
Continuation 17522785 · Nov 9, 2021
Continuation PCTUS2020032405 · May 11, 2020
Provisional Application 62846560 · May 10, 2019
Related Publication 20240323354A1 · Sep 26, 2024
References Cited (14)
US 10248966B2 · Helle et al. · 2019 [cited by applicant]
US 20130202038A1 · Seregin · 2013 [cited by applicant]
US 20140161179A1 · Seregin · 2014 [cited by applicant]
US 20140192883A1 · Seregin · 2014 [cited by examiner]
US 20170310988A1 · Lin et al. · 2017 [cited by applicant]
US 20210152852A1 · Andersson · 2021 [cited by examiner]
US 20210203931A1 · Yang · 2021 [cited by examiner]
US 20210227247A1 · Blaeser · 2021 [cited by examiner]
RWTH Aachen University, Blaser et al., “CE10-related: Bi-directional motion vector storage for triangular prediction,” Joint Video Experts Team (JVET), of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-M0581, 13th… [cited by applicant]
Fraunhofer Hhi, Winken et al., “CE4 Restricted merge (Test 4.2.2),” Joint Video Experts Team (JVET), of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-K0279-v1, 11th Meeting: Ljubljana, SI, Jul. 10-18, 2018, (4p). [cited by applicant]
International Search Report of PCT Application No. PCT/US2020/032405 dated Aug. 24, 2020, (11p). [cited by applicant]
Wang, Xianglin, et al., “CE4-related: An improved method for triangle merge list construction,” Joint Video Experts Team (JVET), of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-N0340, 14th Meeting: Geneva, CH, M… [cited by applicant]
Meng, Xuewei, et al., “CE4-related: Further simplification of triangle prediction merging candidate list derivation,” Joint Video Experts Team (JVET), of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-N0487, 14th … [cited by applicant]
Chuang, Tzu-Der, et al., “CE4-4.1: Simplification of triangle merging candidate list derivation,” Joint Video Experts Team (JVET), of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-N0083, 14th Meeting: Geneva, CH,… [cited by applicant]