IP Library Granted Patent US 12,200,218
Granted Patent B2
US 12,200,218 · App. 18/474,512 · Granted Jan 14, 2025

Single-line cross component linear model prediction mode

Inventors: Kai Zhang (San Diego, CA); Li Zhang (San Diego, CA); Hongbin Liu (Beijing, CN); Yue Wang (Beijing, CN)
Assignees: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.; BYTEDANCE INC.
H04N19/132H04N19/105H04N19/117H04N19/176H04N19/186H04N19/30H04N19/80
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,200,218
App. No.
18/474,512
Granted
Jan 14, 2025
Kind
B2
Abstract

Devices, systems, and methods for digital video coding that include cross-component prediction are described. In a representative aspect, a method for video coding includes receiving a bitstream representation of a current block of video data comprising a luma component and a chroma component, determining parameters of a linear model based on a first set of samples that are generated by down-sampling a second set of samples of the luma component, and processing, based on the parameters of the linear model, the bitstream representation to generate the current block.

Claims (49)

1. A method of processing video data, comprising:

determining, for a conversion between a current video block of a video that is a chroma block and a bitstream of the video, a corresponding luma block for the current video block, wherein a color format for the current video block and the corresponding luma block is 4:2:0, and a cross component linear model is applied for the current video block;

generating down-sampled inside luma samples of the corresponding luma block;

generating down-sampled above neighboring luma samples of the corresponding luma block, wherein different down-sampled filtering schemes are used to generate the down-sampled above neighboring luma samples based on a position of the corresponding luma block, and wherein in response to the position of the corresponding luma block meeting a first position rule, a first down-sampled filtering scheme is used to generate the down-sampled above neighboring luma samples, and wherein the first position rule is that a top boundary of the corresponding luma block is overlapping with a top boundary of a current luma coding tree block (CTB) including the corresponding luma block, and wherein only one above luma sample row and a horizontal 3-tap filter are used to generate the down-sampled above neighboring luma samples in the first down-sampled filtering scheme;

deriving parameters of the cross component linear model at least based on the down-sampled above neighboring luma samples;

generating predicted chroma samples of the current video block based on the parameters of the cross component linear model and the down-sampled inside luma samples; and

performing the conversion based on the predicted chroma samples.

2. The method of claim 1 , wherein the only one above luma sample row is adjacent to the corresponding luma block.

3. The method of claim 2 , wherein the only one above luma sample row comprises above adjacent luma samples and above-right adjacent luma samples.

4. The method of claim 2 , wherein the first down-sampled filtering scheme calculates down-sampled above neighboring luma sample d[i] from luma sample a[i] of the only one above luma sample row as:

d[i ]=( a[ 2 i− 1]+2* a[ 2 i]+a[ 2 i+ 1]+2)>>2.

5. The method of claim 4 , wherein in a case that the luma sample a [i] is unavailable, apply a padding process.

6. The method of claim 5 , wherein in a case that a [ 2 i - 1 ] is unavailable, down-sampled above neighboring luma samples d[i] are derived as:

d[i ]=(3* a[ 2 i]+a[ 2 i+ 1]+2)>>2.

7. The method of claim 1 , wherein in response to the position of the corresponding luma block not meeting the first position rule, a second down-sampled filtering scheme is used to generate the down-sampled above neighboring luma samples.

8. The method of claim 7 , wherein multiple above luma sample rows and a filter with taps more than 3 are used to generate the down-sampled above neighboring luma samples in the second down-sampled filtering scheme.

9. The method of claim 8 , wherein at least one of multiple above luma sample rows is not adjacent to a block corresponding to the at least one of multiple above luma sample rows.

10. The method of claim 7 , further comprising:

applying a third down-sampled filtering scheme to generate the down-sampled inside luma samples of the corresponding luma block, wherein the third down-sampled filtering scheme uses same filtering taps and filtering coefficients with the second down-sampled filtering scheme.

11. The method of claim 7 , further comprising:

applying a fourth down-sampled filtering scheme to generate down-sampled left neighboring luma samples of the corresponding luma block, wherein the fourth down-sampled filtering scheme uses same filtering taps and filtering coefficients with the second down-sampled filtering scheme;

and the parameters of the cross component linear model are further derived based on the down-sampled left neighboring luma samples.

12. The method of claim 1 , wherein the first down-sampled filtering scheme corresponds to down-sampling above neighboring luma samples to lower left positions.

13. The method of claim 1 , wherein the conversion includes encoding the current video block into the bitstream.

14. The method of claim 1 , wherein the conversion includes decoding the current video block from the bitstream.

15. An apparatus for processing video data comprising a processor and a non-transitory memory with instructions thereon, wherein the instructions upon execution by the processor, cause the processor to:

determine, for a conversion between a current video block of a video that is a chroma block and a bitstream of the video, a corresponding luma block for the current video block, wherein a color format for the current video block and the corresponding luma block is 4:2:0, and a cross component linear model is applied for the current video block;

generate down-sampled inside luma samples of the corresponding luma block;

generate down-sampled above neighboring luma samples of the corresponding luma block, wherein different down-sampled filtering schemes are used to generate the down-sampled above neighboring luma samples based on a position of the corresponding luma block, and wherein in response to the position of the corresponding luma block meeting a first position rule, a first down-sampled filtering scheme is used to generate the down-sampled above neighboring luma samples, and wherein the first position rule is that a top boundary of the corresponding luma block is overlapping with a top boundary of a current luma coding tree block (CTB) including the corresponding luma block, and wherein only one above luma sample row and a horizontal 3-tap filter are used to generate the down-sampled above neighboring luma samples in the first down-sampled filtering scheme;

derive parameters of the cross component linear model at least based on the down-sampled above neighboring luma samples;

generate predicted chroma samples of the current video block based on the parameters of the cross component linear model and the down-sampled inside luma samples; and

perform the conversion based on the predicted chroma samples.

16. The apparatus of claim 15 , wherein the only one above luma sample row is adjacent to the corresponding luma block.

17. The apparatus of claim 16 , wherein the only one above luma sample row comprises above neighboring luma samples and above-right neighboring luma samples.

18. A non-transitory computer-readable storage medium storing instructions that cause a processor to:

determine, for a conversion between a current video block of a video that is a chroma block and a bitstream of the video, a corresponding luma block for the current video block, wherein a color format for the current video block and the corresponding luma block is 4:2:0, and a cross component linear model is applied for the current video block;

generate down-sampled inside luma samples of the corresponding luma block;

generate down-sampled above neighboring luma samples of the corresponding luma block, wherein different down-sampled filtering schemes are used to generate the down-sampled above neighboring luma samples based on a position of the corresponding luma block, and wherein in response to the position of the corresponding luma block meeting a first position rule, a first down-sampled filtering scheme is used to generate the down-sampled above neighboring luma samples, and wherein the first position rule is that a top boundary of the corresponding luma block is overlapping with a top boundary of a current luma coding tree block (CTB) including the corresponding luma block, and wherein only one above luma sample row and a horizontal 3-tap filter are used to generate the down-sampled above neighboring luma samples in the first down-sampled filtering scheme;

derive parameters of the cross component linear model at least based on the down-sampled above neighboring luma samples;

generate predicted chroma samples of the current video block based on the parameters of the cross component linear model and the down-sampled inside luma samples; and

perform the conversion based on the predicted chroma samples.

19. The non-transitory computer-readable storage medium of claim 18 , wherein the only one above luma sample row is adjacent to the corresponding luma block.

20. A non-transitory computer-readable recording medium storing a bitstream which is generated by a method performed by a video processing apparatus, wherein the method comprises:

determining, for a current video block of a video that is a chroma block, a corresponding luma block for the current video block, wherein a color format for the current video block and the corresponding luma block is 4:2:0, and a cross component linear model is applied for the current video block;

generating down-sampled inside luma samples of the corresponding luma block;

generating down-sampled above neighboring luma samples of the corresponding luma block, wherein different down-sampled filtering schemes are used to generate the down-sampled above neighboring luma samples based on a position of the corresponding luma block, and wherein in response to the position of the corresponding luma block meeting a first position rule, a first down-sampled filtering scheme is used to generate the down-sampled above neighboring luma samples, wherein the first position rule is that a top boundary of the corresponding luma block is overlapping with a top boundary of a current luma coding tree block (CTB) including the corresponding luma block, and wherein only one above luma sample row and a horizontal 3-tap filter are used to generate the down-sampled above neighboring luma samples in the first down-sampled filtering scheme;

deriving parameters of the cross component linear model at least based on the down-sampled above neighboring luma samples;

generating predicted chroma samples of the current video block based on the parameters of the cross component linear model and the down-sampled inside luma samples; and

generating the bitstream based on the predicted chroma samples.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 2, 2023
From: LIU, HONGBIN; WANG, YUE
To: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 065094/0438 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 2, 2023
From: ZHANG, KAI; ZHANG, LI
To: BYTEDANCE INC.
Reel/Frame 065094/0532 →
Priority Claims (3)
WO PCT/CN2018/105182 · Sep 12, 2018 · international
WO PCT/CN2018/108681 · Sep 29, 2018 · international
WO PCT/CN2019/088005 · May 22, 2019 · international
Continuity (4)
Continuation 17512488 · Oct 27, 2021
Continuation 17115388 · Dec 8, 2020
Continuation PCTIB2019057699 · Sep 12, 2019
Related Publication 20240015298A1 · Jan 11, 2024
References Cited (125)
US 6259738B1 · Yamaguchi · 2001 [cited by applicant]
US 9565428B2 · Guo · 2017 [cited by applicant]
US 10045023B2 · Pettersson · 2018 [cited by applicant]
US 10200700B2 · Zhang · 2019 [cited by applicant]
US 10368094B2 · Budagavi · 2019 [cited by applicant]
US 10390015B2 · Hu · 2019 [cited by applicant]
US 10419757B2 · Chen · 2019 [cited by applicant]
US 10469847B2 · Xiu · 2019 [cited by applicant]
US 10609402B2 · Zhao · 2020 [cited by applicant]
US 10939128B2 · Zhang · 2021 [cited by applicant]
US 10979717B2 · Zhang · 2021 [cited by applicant]
US 11172202B2 · Zhang · 2021 [cited by examiner]
US 11218702B2 · Zhang · 2022 [cited by examiner]
US 11463713B2 · Deng · 2022 [cited by applicant]
US 11616965B2 · Deng · 2023 [cited by applicant]
US 11638039B2 · Zhao · 2023 [cited by applicant]
US 11677956B2 · Zhang · 2023 [cited by examiner]
US 11812026B2 · Zhang · 2023 [cited by examiner]
US 20120287995A1 · Budagavi · 2012 [cited by applicant]
US 20120328013A1 · Budagavi · 2012 [cited by applicant]
US 20130128966A1 · Gao · 2013 [cited by applicant]
US 20130136174A1 · Xu · 2013 [cited by applicant]
US 20130188703A1 · Liu · 2013 [cited by applicant]
US 20130188705A1 · Liu · 2013 [cited by applicant]
US 20140233650A1 · Zhang · 2014 [cited by applicant]
US 20150036745A1 · Hsu · 2015 [cited by applicant]
US 20150098510A1 · Ye · 2015 [cited by applicant]
US 20150365684A1 · Chen · 2015 [cited by applicant]
US 20160198190A1 · Budagavi · 2016 [cited by applicant]
US 20160219283A1 · Chen · 2016 [cited by applicant]
US 20160277762A1 · Zhang · 2016 [cited by applicant]
US 20170150186A1 · Zhang · 2017 [cited by applicant]
US 20170244975A1 · Huang · 2017 [cited by applicant]
US 20170359595A1 · Zhang · 2017 [cited by applicant]
US 20170359597A1 · Budagavi · 2017 [cited by applicant]
US 20170366818A1 · Zhang · 2017 [cited by applicant]
US 20180048889A1 · Zhang · 2018 [cited by applicant]
US 20180063527A1 · Chen · 2018 [cited by applicant]
US 20180063531A1 · Hu · 2018 [cited by applicant]
US 20180063553A1 · Zhang · 2018 [cited by applicant]
US 20180077426A1 · Zhang · 2018 [cited by applicant]
US 20180131962A1 · Chen · 2018 [cited by applicant]
US 20180176594A1 · Zhang · 2018 [cited by applicant]
US 20180205946A1 · Zhang · 2018 [cited by applicant]
US 20180316918A1 · Drugeon · 2018 [cited by applicant]
US 20190014316A1 · Panusopone · 2019 [cited by applicant]
US 20190045184A1 · Zhang · 2019 [cited by applicant]
US 20190110045A1 · Zhao · 2019 [cited by applicant]
US 20190166382A1 · He · 2019 [cited by applicant]
US 20190306516A1 · Misra · 2019 [cited by applicant]
US 20190342546A1 · Lin · 2019 [cited by applicant]
US 20200068203A1 · Zhang · 2020 [cited by applicant]
US 20200120359A1 · Hanhart · 2020 [cited by applicant]
US 20200128272A1 · Jangwon · 2020 [cited by applicant]
US 20200177878A1 · Choi · 2020 [cited by applicant]
US 20200195930A1 · Choi · 2020 [cited by applicant]
US 20200195970A1 · Ikai · 2020 [cited by applicant]
US 20200195976A1 · Zhao · 2020 [cited by applicant]
US 20200252619A1 · Zhang · 2020 [cited by applicant]
US 20200288135A1 · Laroche · 2020 [cited by applicant]
US 20200296380A1 · Aono · 2020 [cited by applicant]
US 20200359051A1 · Zhang · 2020 [cited by applicant]
US 20200366896A1 · Zhang · 2020 [cited by applicant]
US 20200366910A1 · Zhang · 2020 [cited by applicant]
US 20200366933A1 · Zhang · 2020 [cited by applicant]
US 20200382769A1 · Zhang · 2020 [cited by applicant]
US 20200382800A1 · Zhang · 2020 [cited by applicant]
US 20200389650A1 · Laroche · 2020 [cited by applicant]
US 20200413049A1 · Biatek · 2020 [cited by applicant]
US 20210092395A1 · Zhang · 2021 [cited by applicant]
US 20210092396A1 · Zhang · 2021 [cited by applicant]
CN 104380741A · 2015 [cited by applicant]
CN 104718759A · 2015 [cited by applicant]
CN 106664425A · 2017 [cited by applicant]
CN 107079166A · 2017 [cited by applicant]
CN 107409209A · 2017 [cited by applicant]
CN 108293130A · 2018 [cited by applicant]
CN 110896478A · 2020 [cited by applicant]
EP 3799428A1 · 2021 [cited by applicant]
GB 2495942A · 2013 [cited by applicant]
GB 201716538 · 2017 [cited by applicant]
WO 2017139937A1 · 2017 [cited by applicant]
WO 2017214420A1 · 2017 [cited by applicant]
WO 2018053293A1 · 2018 [cited by applicant]
WO 2018064948A1 · 2018 [cited by applicant]
WO 2018140587A1 · 2018 [cited by applicant]
Enhanced cross component linear model intra prediction; Zhang—Oct. 2016. (Year: 2017). [cited by examiner]
Reduced No. of reference samples for CCLM parameter; Choi—Oct. 2017. (Year: 2017). [cited by examiner]
CCLM prediction with single-line neighboring luma samples; Zhang—Oct. 2018. (Year: 2018). [cited by examiner]
Enhanced cross component linear model for chroma intra prediction; Zhang—Aug. 2018. (Year: 2018). [cited by examiner]
Ma et al. “CE3: Multi-directional LM (MDLM) (Test 5.4.1 and 5.4.2)” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/VVG 11, 12th Meeting: Macao, CN, Oct. 3-12, 2018, document JVET-L0338, 2018. [cited by applicant]
Wang et al. “CE3-1.5: CCLM Derived with Four Neighbouring Samples,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/VVG 11, 14th Meeting: Geneva, CH, Mar. 1-27, 2019, document JVET-N0271 2019. [cited by applicant]
Flynn et al. Overview of the Range Extensions for the HEVC Standard: Tools, Profiles, and Performance, IEEE Transactions on Circuits and Systems for Video Technology, Institute of Electrical and Electronic Engineers, US… [cited by applicant]
Chen et al. “Description of SOR, HOR and 360 degree Video Coding Technology Proposal by Huawei, GoPro, HiSilicon, and Samsung,” Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/VVG 1, 10th… [cited by applicant]
Van Der Auwera et al. “Description of Core Experiment 3: Intra Prediction and Mode Coding,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/VVG 11, 10th Meeting, San Diego, USA, Apr. 10-20, 2… [cited by applicant]
Zhang et al. “New Modes for Chroma Intra Prediction,” Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG16 WP3 and ISO/IEC JTC1/SC29/VVG11, 7th Meeting, Geneva, CH, Nov. 21-30, 2011, document JCTVC-G358, 2011. [cited by applicant]
Zhang et al. “CE3-Related: CCLM Prediction with Single-Line Neighbouring Luma Samples,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/VVG 11, 12th Meeting, Macao, CN, Oct. 3-12, 2018, docum… [cited by applicant]
Misra et al. “Description of SOR and HOR Video Coding Technology Proposal by Sharp and Foxconn,” Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 10th Meeting, San Diego, US, 10-20,… [cited by applicant]
Ramasubramonian et al. CE3: Combined Test of CE3.2.1 and CE3.2.2 (Test 3.2.2.1), Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, document … [cited by applicant]
Chen et al. “Chroma Intra Prediction by Reconstructed Luma Samples,” Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP3 and ISO/IEC JTC1/SC29/WG11, 3rd Meeting, Guangzhou, CN, Oct. 7-15, 2010, JCTVC-C2… [cited by applicant]
Chen et al. “CE6.a: Chroma Intra Prediction by Reconstructed Luma Samples,” Joint Collaborative Team on Video Coding (JCT-VG) of ITU-T SG 16 WP 3 and ISO/IEC JTC1/SC29/WG11, 4th Meeting: Daegu, KR, Jan. 20-28, 2011, doc… [cited by applicant]
Choi et al. “CE3-Related: Reduced No. of Reference Samples for CCLM Parameter Calculation,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN, Oct. 3-12, 2018, do… [cited by applicant]
Zhang et al. “Enhanced Cross-Component Linear Model Intra Prediction,” Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 4th Meeting: Chengdu, CN, Oct. 15-21, 2016, JVET-D0110, 2018. [cited by applicant]
Zhang et al. “Enhanced Cross-Component Linear Model for Chroma Intra-Prediction in Video Coding,” IEEE Transactions on Image Processing, Aug. 2018, 27(8): 3983-3997. [cited by applicant]
Van Der Auwera et al. “Description of Core Experiment 3: Intra Prediction and Mode Coding,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 11th Meeting: Ljubljana, SI, Jul. 10-18, 2018… [cited by applicant]
Van Der Auwera et al. “Extension of Simplified PDPC to Diagonal Intra Modes,” Joint Video Experts Team (JVET) pf ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 10th Meeting: San Diego, USA, Apr. 10-20, 2018, document UV… [cited by applicant]
Chen et al. “CE6.a.4: Chroma Intra Prediction by Reconstructed Luma Samples,” Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG16 WP3 and ISO/IEC JTC1/SC29/WG11 5th Meeting: Geneva, Mar. 16-23, 2011, documen… [cited by applicant]
Document: JVET-K0204-v1, Laroche, G., “Non-CE3: On cross-component linear model simplification,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 11th Meeting: Ljubljana, SI, Jul. 10-18,… [cited by applicant]
Bross et al. “CE3: Multiple Reference Line Intra Prediction (Test 1.1.1, 1.1.2, 1.1.3 and 1.1.4),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN 3-12, Oct. 20… [cited by applicant]
Notice of Allowance from U.S. Appl. No. 17/646,196 dated Jan. 25, 2023. [cited by applicant]
Examination Report from Patent Application No. GB2102253.8 dated Apr. 20, 2022 (3 pages). [cited by applicant]
Intemational Search Report and Written Opinion from International Patent Application No. PCT/IB2019/056967 dated Feb. 4, 2020 (24 pages). [cited by applicant]
Intemational Search Report and Written Opinion from International Patent Application No. PCT/CN2020/071659 dated Mar. 27, 2020 (12 pages). [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/115,290 dated Feb. 4, 2021. [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/115,290 dated Jun. 4, 2021. [cited by applicant]
International Search Report and Written Opinion from International Patent Application No. PCT/IB2019/057698 dated Feb. 4, 2020 (22 pages). [cited by applicant]
International Search Report and Written Opinion from International Patent Application No. PCT/IB2019/057699 dated Dec. 11, 2019 (17 pages). [cited by applicant]
International Search Report and Written Opinion from International Patent Application No. PCT/IB2019/057700 dated Dec. 11, 2019 (17 pages). [cited by applicant]
Intemational Search Report and Written Opinion from International Patent Application No. PCT/CN2020/091830 dated Aug. 24, 2020 (10 pages). [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/115,388 dated Feb. 9, 2021. [cited by applicant]
Notice of Allowance from U.S. Appl. No. 17/115,388 dated Sep. 28, 2021. [cited by applicant]
Notice of Allowance from U.S. Appl. No. 17/115,290 dated Oct. 14, 2021. [cited by applicant]
Notice of Allowance from U.S. Appl. No. 17/512,488 dated Jul. 7, 2023, 13 pages. [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/512,488 dated Feb. 15, 2023, 22 pages. [cited by applicant]
Document: JVET-M0097-v1, Ramasubramonian, A., et al., “CE3: On MMLM (Test 2.1),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, 4 pages. [cited by applicant]