IP Library › Granted Patent US 12,192,472
Granted Patent B2
US 12,192,472 · App. 17/674,539 · Granted Jan 7, 2025

Support of multiple chroma format in VVC

Inventors: Xin Zhao (San Diego, CA); Xiaozhong Xu (State College, PA); Xiang Li (Los Gatos, CA); Shan Liu (San Jose, CA)
Assignee: TENCENT AMERICA LLC
H04N19/137H04N19/119H04N19/159H04N19/176H04N19/186H04N19/80H04N19/96
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,192,472
App. No.
17/674,539
Granted
Jan 7, 2025
Kind
B2
Abstract

A method and apparatus for encoding or decoding a video sequence includes encoding or decoding the video sequence using a 4:4:4 chroma format, or encoding or decoding the video sequence using a 4:2:2 chroma format, wherein when encoding or decoding the video sequence using the 4:4:4 chroma format, copying an affine motion vector of one 4×4 luma block using an operation other than an averaging operation and associating the affine motion vector to a co-located 4×4 chroma block, and when encoding or decoding the video sequence using the 4:2:2 chroma format, associating each 4×4 chroma block with two 4×4 co-located luma blocks such that an affine motion vector of one 4×4 chroma block is an average of the motion vectors of the two co-located luma blocks.

Claims (39)

1. A method for decoding a video sequence, the method comprising:

decoding the video sequence using one of a 4:4:4 chroma format and a 4:2:2 chroma format;

when the video sequence is decoded using the 4:4:4 chroma format, copying an affine motion vector of one 4×4 luma block using an operation other than an averaging operation and associating the affine motion vector to a co-located 4×4 chroma block; and

when the video sequence is decoded using the 4:2:2 chroma format, associating each 4×4 chroma block with two 4×4 co-located luma blocks such that an affine motion vector of one 4×4 chroma block is an average of the motion vectors of the two co-located luma blocks, wherein a maximum vertical transform size is the same among different color components, and wherein a maximum horizontal transform size for color components is half of a maximum horizontal transform size for luma components.

2. The method of claim 1 , further comprising:

regardless of the chroma format, dividing a current 4×4 chroma block into four 2×2 sub-blocks;

deriving a first affine motion vector of a co-located luma block for a top-left 2×2 chroma sub-block;

deriving a second affine motion vector of the co-located luma block for a bottom-right 2×2 chroma block; and

deriving an affine motion vector of the current 4×4 chroma block using the average of the first affine motion vector and the second affine motion vector.

3. The method of claim 1 , further comprising aligning an interpolation filter used for motion compensation between luma and chroma components.

4. The method of claim 3 , wherein when the video sequence is input using a 4:2:0 chroma format, applying an 8-tap interpolation filter for luma components and chroma components.

5. The method of claim 1 , further comprising decoding as three separate trees, components Y, Cb, and Cr, wherein each tree of the three separate trees decodes one component of the components Y, Cb, and Cr.

6. The method of claim 5 , wherein the decoding as three separate trees is performed for an I slice or an I tile group.

7. The method of claim 1 , wherein at least one of a Position-Dependent Predictor combination (PDPC), a Multiple Transform Selection (MTS), a Non-Separable Secondary Transform (NSST), an Intra-Sub Partitioning (ISP), and a Multiple reference line (MRL) intra prediction is applied to both a luma component and a chroma component.

8. The method of claim 7 , wherein:

when the Multiple reference line (MRL) intra prediction is applied to both the luma component and the chroma component, and when decoding the video sequence is performed using the 4:4:4 chroma format, the method further comprises selecting an Nth reference for intra prediction, and using a same reference line without explicit signaling for chroma components;

when the Intra-Sub Partitioning (ISP) is applied to both the luma component and the chroma component, the method further comprises applying the Intra-Sub Partitioning (ISP) at a block level for a current block for components Y, Cb, and Cr; and

when different trees are used for different color components, the method further comprises implicitly deriving coding parameters for U and V components from collocated Y components without signaling.

9. A method for encoding a video sequence, the method comprising:

encoding the video sequence using one of a 4:4:4 chroma format and a 4:2:2 chroma format;

when the video sequence is being encoded using the 4:4:4 chroma format, copying an affine motion vector of one 4×4 luma block using an operation other than an averaging operation and associating the affine motion vector to a co-located 4×4 chroma block, and

when the video sequence is being encoded using the 4:2:2 chroma format; associating each 4×4 chroma block with two 4×4 co-located luma blocks such that an affine motion vector of one 4×4 chroma block is an average of the motion vectors of the two co-located luma blocks, wherein a maximum vertical transform size is the same among different color components, and wherein a maximum horizontal transform size for color components is half of a maximum horizontal transform size for luma components.

10. The method of claim 9 , further comprising:

regardless of the chroma format, dividing a current 4×4 chroma block into four 2×2 sub-blocks;

deriving a first affine motion vector of a co-located luma block for a top-left 2×2 chroma sub-block;

deriving a second affine motion vector of the co-located luma block for a bottom-right 2×2 chroma block; and

deriving an affine motion vector of the current 4×4 chroma block using the average of the first affine motion vector and the second affine motion vector.

11. The method of claim 9 , further comprising aligning an interpolation filter used for motion compensation between luma and chroma components.

12. A method of processing visual media data, comprising:

obtaining a visual media file; and

performing a conversion between the visual media file and a bitstream of a visual media data, wherein the bitstream comprises an encoded video sequence with a 4:4:4 chroma format and a 4:2:2 chroma format, and wherein performing the conversion comprises:

using, in the 4:4:4 chroma format, an affine motion vector copied from a 4×4 luma block via an operation other than an averaging operation for a co-located 4×4 chroma block, and

using, in the 4:2:2 chroma format, an affine motion vector for a 4×4 chroma block that is an average of motion vectors of two co-located luma blocks, wherein a maximum vertical transform size is the same among different color components, and wherein a maximum horizontal transform size for color components is half of a maximum horizontal transform size for luma components.

13. The method of claim 12 , wherein performing the conversion further includes:

dividing a current 4×4 chroma block into four 2×2 sub-blocks;

deriving a first affine motion vector of a co-located luma block for a top-left 2×2 chroma sub-block;

deriving a second affine motion vector of the co-located luma block for a bottom-right 2×2 chroma block; and

deriving an affine motion vector of the current 4×4 chroma block using the average of the first affine motion vector and the second affine motion vector.

14. The method of claim 12 , wherein performing the conversion further includes aligning an interpolation filter used for motion compensation between luma and chroma components.

Continuity (3)
Continuation 16815729 · Mar 11, 2020
Provisional Application 62817517 · Mar 12, 2019
Related Publication 20220224909A1 · Jul 14, 2022
References Cited (41)
US 11290722B2 · Zhao et al. · 2022 [cited by applicant]
US 20150030067A1 · Zhao · 2015 [cited by examiner]
US 20170094288A1 · Hannuksella · 2017 [cited by applicant]
US 20180270500A1 · Li et al. · 2018 [cited by applicant]
US 20180302631A1 · Chiang et al. · 2018 [cited by applicant]
US 20180332284A1 · Liu · 2018 [cited by examiner]
US 20180376148A1 · Zhang et al. · 2018 [cited by applicant]
US 20190068969A1 · Rusanovskyy et al. · 2019 [cited by applicant]
US 20190068989A1 · Lee · 2019 [cited by applicant]
US 20190166382A1 · He · 2019 [cited by examiner]
US 20200059659A1 · Chen · 2020 [cited by examiner]
US 20200275118A1 · Wang · 2020 [cited by examiner]
US 20210203947A1 · He · 2021 [cited by examiner]
US 20220224909A1 · Zhao et al. · 2022 [cited by applicant]
EP 3939258A1 · 2020 [cited by applicant]
EP 3928510A1 · 2021 [cited by applicant]
JP 2015537448A · 2015 [cited by applicant]
WO WO2017205648A1 · 2017 [cited by applicant]
WO 2018033661A1 · 2018 [cited by applicant]
WO WO2020172292A1 · 2020 [cited by applicant]
WO WO2020185876A1 · 2020 [cited by applicant]
Tamse and Park (Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019 “CE2-related: MV Derivation for Affine Chroma” (Year: 2019). [cited by examiner]
Tamse and Park (Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019 (Year: 2019). [cited by examiner]
Notice of Reason for Refusal dated Jul. 12, 2022 from the Japanese Patent Office in Japanese Application No. 2021-532208. [cited by applicant]
Supervised by Ei Okubo, “Impress Standard Textbook Series H.265 / HEVC Textbook”, First Edition, Impress Japan Co., Ltd., 2013, pp. 41-43 (11 pages total). [cited by applicant]
Chen et al., “Algorithm description for Versatile Video Coding and Test Model 4 (VTM 4)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 13th Meeting: Marrakech, MA, JVET-M1002-v2, Ja… [cited by applicant]
Chen et al., “Algorithm description for Versatile Video Coding and Test Model 8 (VTM 8)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 17th Meeting: Brussels, BE, JVET-Q2002-v2, Jan… [cited by applicant]
Extended European Search Report dated Nov. 9, 2022 from the European Patent Office in EP Application No. 20769081.9. [cited by applicant]
Communication dated Nov. 29, 2022 from the European Patent Office in EP Application No. 20769081.9. [cited by applicant]
Office Action issued Mar. 8, 2022 in Indian Application No. 202137041474. [cited by applicant]
Written Opinion and International Search Report dated Jun. 15, 2020, from the International Searching Authority in International Application No. PCT/US2020/22066. [cited by applicant]
Benjamin Bross et al., “Versatile Video Coding (Draft 4)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, Document: JVET-M1001-v6, 295 pgs. [cited by applicant]
Tencent Technology, Canadian Office Action, CA Patent Application No. 3,132,485, Aug. 23, 2023, 5 pgs. [cited by applicant]
Benjamin Bross et al., “Versatile Video Coding (Draft 3)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document: JVET-L1001-v9, 12th Meeting: Macao, CN, Oct. 3-12, 2018, 235 pgs. [cited by applicant]
Tencent Technology, Korean Office Action, KR Patent Application No. 10-2021-7015100, Mar. 11, 2024, 12 pgs. [cited by applicant]
Tencent Technology, Australian Office Action, AU Patent Application No. 2023214361, Nov. 16, 2023, 3 pgs. [cited by applicant]
Tencent Technology, Japanese Office Action, JP Patent Application No. 2023036984, Jan. 9, 2024, 14 pgs. [cited by applicant]
Jianle Chen et al., “Algorithm Description for Versatile Video Coding and Test Model 4”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, Do… [cited by applicant]
Tencent Technology, Australian Office Action, AU Patent Application No. 2023214363, Oct. 30, 2023, 3 pgs. [cited by applicant]
Tencent Technology, Canadian Office Action, CA Patent Application No. 3,132,485, May 17, 2024, 3 pgs. [cited by applicant]
Tencent Technology, Vietnamese Office Action, VN Patent Application No. 1-2021-05806, May 30, 2024, 3 pgs. [cited by applicant]