IP Library › Granted Patent US 12,388,991
Granted Patent B2
US 12,388,991 · App. 18/367,982 · Granted Aug 12, 2025

Feature based cross-component sample offset optimization and signaling improvement

Inventors: Samruddhi Yashwant Kahu (Laguna Hills, CA); Xin Zhao (San Jose, CA); Shan Liu (San Jose, CA)
Assignee: Tencent America LLC
H04N19/117H04N19/11H04N19/124H04N19/13H04N19/132H04N19/14H04N19/172H04N19/176H04N19/182H04N19/70H04N19/80
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,388,991
App. No.
18/367,982
Granted
Aug 12, 2025
Kind
B2
Abstract

In some examples, an apparatus for video decoding includes receiving circuitry and processing circuitry. The processing circuitry receives, from a coded video bitstream, coded information associated with one or more blocks in a current picture, the coded information indicates that a cross-component sample offset (CCSO) filter is used for decoding the one or more blocks. The processing circuitry determines one or more features associated with the one or more blocks, the one or more features is used as a context for signaling a value associated with the CCSO filter. The processing circuitry determines at least a signaled value associated with the CCSO filter based on a context model, the context model using the one or more features as a context, and reconstructs samples in the one or more blocks based on at least the signaled value associated with the CCSO filter.

Claims (57)

1. A method of video processing in a decoder, comprising:

receiving, from a coded video bitstream, coded information associated with one or more blocks in a current picture, the coded information indicating that a cross-component sample offset (CCSO) filter is used for decoding the one or more blocks, the CCSO filter being associated with a plurality of CCSO filter parameters;

determining one or more features associated with the one or more blocks, the one or more features being used as a context for signaling a value associated with the CCSO filter;

determining, based on the one or more features, a context model that is used for signaling the value of at least a parameter of the CCSO filter in the plurality of CCSO filter parameters, the context model including different probability values associated with candidate values of the parameter of the CCSO filter with the one or more features being used as the context;

determining at least a signaled value associated with the CCSO filter based on the context model; and

reconstructing samples in the one or more blocks based on at least the signaled value associated with the CCSO filter.

2. The method of claim 1 , wherein the context for signaling the value associated with the CCSO filter comprises at least one of whether a specific CCSO filter parameter is signaled, and/or the context used for entropy coding the specific CCSO filter parameter.

3. The method of claim 1 , wherein the one or more features is indicative of texture information in the one or more blocks, the one or more features comprises at least one of

an edge direction for one or more edges in the one or more blocks;

a range of pixel values in the one or more blocks;

a histogram of pixel values in the one or more blocks; and

gradient information of pixels in the one or more blocks.

4. The method of claim 1 , further comprising:

determining the context model based on the one or more features, the context model comprising an assignment of at least a first probability value to a first candidate value of a filter parameter for the CCSO filter and a second probability value to a second candidate value of the filter parameter for the CCSO filter, the first probability value being higher than the second probability value when the first candidate value is associated with the one or more features.

5. The method of claim 4 , further comprising:

determining, based on the context model, a specific filter parameter for the CCSO filter is not signaled.

6. The method of claim 4 , wherein the one or more features comprise an edge direction, the signaled value is indicative of a filter shape, and the determining the context model further comprises at least one of:

when the edge direction is vertical, determining the context model that comprises a highest probability assigned to a vertical filter shape;

when the edge direction is horizontal, determining the context model that comprises a highest probability assigned to a horizontal filter shape; or

when the edge direction is diagonal, determining the context model that comprises a highest probability assigned to a diagonal filter shape.

7. The method of claim 4 , wherein the one or more features comprise a range of pixel values in the one or more blocks, and the context model comprises an assignment of probability values to candidate band numbers.

8. The method of claim 4 , wherein the one or more features comprise a range of pixel values in the one or more blocks, and the context model comprises an assignment of probability values to candidate quantization steps.

9. The method of claim 1 , wherein the determining at least the signaled value associated with the CCSO filter based the context model further comprises:

determining the signaled value indicative of an on/off decision for the CCSO filter based on the context model that uses the one or more features as the context.

10. The method of claim 9 , wherein the on/off decision is of at least one of a frame level, a tile level, and a block level.

11. The method of claim 9 , wherein the one or more features is indicative of a flatness, and the context model comprises an assignment of a higher probability value to a first candidate value of an off decision for the CCSO filter than to a second candidate value of an on decision for the CCSO filter.

12. The method of claim 9 , wherein the one or more blocks comprises a current block and a neighboring block of the current block, and the context model uses a flatness of the neighboring block as context.

13. The method of claim 9 , wherein the one or more features comprise at least one of a first coded syntax value indicative of whether a skip mode being applied, a second coded syntax value indicative of whether intra or inter prediction mode being applied, a third coded syntax value indicative of whether residuals being zero or not, a fourth coded syntax value indicative of an intra prediction mode direction, and a fifth coded syntax value indicative of a coding block partitioning size.

14. A method of video processing in an encoder, comprising:

determining one or more features associated with one or more blocks in a current picture;

determining to use a cross-component sample offset (CCSO) filter having a plurality of CCSO filter parameters for coding the one or more blocks;

determining at least a value of a parameter of the CCSO filter by skipping a check of at least a candidate value of the parameter based on the one or more features;

determining, based on the one or more features, a context model for signaling the value of the parameter, the context model including different probability values associated with candidate values of the parameter with the one or more features being used as a context; and

generating coded information of the one or more blocks based on the context model, the coded information indicating the value of the parameter associated with the CCSO filter according to the context model.

15. The method of claim 14 , wherein the context for signaling the value associated with the CCSO filter comprises at least one of whether a specific CCSO filter parameter is signaled, and/or the context used for entropy coding the specific CCSO filter parameter.

16. The method of claim 14 , wherein the one or more features is indicative of texture information in the one or more blocks, the one or more features comprises at least one of:

an edge direction for one or more edges in the one or more blocks;

a range of pixel values in the one or more blocks;

a histogram of pixel values in the one or more blocks; and

gradient information of pixels in the one or more blocks.

17. The method of claim 14 , wherein the determining the context model comprises:

determining the context model based on the one or more features, the context model comprising an assignment of at least a first probability value to a first candidate value of a filter parameter for the CCSO filter and a second probability value to a second candidate value of the filter parameter for the CCSO filter, the first probability value being higher than the second probability value when the first candidate value is associated with the one or more features.

18. The method of claim 17 , further comprising:

determining, based on the context model, not to signal a specific filter parameter for the CCSO filter.

19. The method of claim 17 , wherein the one or more features comprise an edge direction, the value is indicative of a filter shape, and the determining the context model further comprises at least one of:

when the edge direction is vertical, determining the context model that comprises a highest probability assigned to a vertical filter shape;

when the edge direction is horizontal, determining the context model that comprises a highest probability assigned to a horizontal filter shape; or

when the edge direction is diagonal, determining the context model that comprises a highest probability assigned to a diagonal filter shape.

20. A method of processing visual media data, the method comprising:

processing a bitstream that includes the visual media data according to a format rule, wherein

the bitstream carries a plurality of pictures; and

the format rule specifies that:

coded information associated with one or more blocks in a current picture indicates that a cross-component sample offset (CCSO) filter is used for coding the one or more blocks, the CCSO filter being associated with a plurality of CCSO filter parameters;

one or more features associated with the one or more blocks are determined, the one or more features being used as a context for signaling a value associated with the CCSO filter;

a context model for signaling the value of at least a parameter of the CCSO filter in the plurality of CCSO filter parameters is determined based on the one or more features, the context model including different probability values associated with candidate values of the parameter of the CCSO filter with the one or more features being used as the context;

at least a signaled value associated with the CCSO filter is determined based on the context model; and

samples in the one or more blocks are reconstructed based on at least the signaled value associated with the CCSO filter.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2023
From: KAHU, SAMRUDDHI YASHWANT; ZHAO, XIN; LIU, SHAN
To: TENCENT AMERICA LLC
Reel/Frame 064895/0906 →
Continuity (2)
Provisional Application 63448617 · Feb 27, 2023
Related Publication 20240291976A1 · Aug 29, 2024
References Cited (49)
US 11595644B2 · Du et al. · 2023 [cited by applicant]
US 11683530B2 · Du et al. · 2023 [cited by applicant]
US 20110200101A1 · Zan · 2011 [cited by examiner]
US 20130336383A1 · Xu et al. · 2013 [cited by applicant]
US 20170127078A1 · Kobayashi · 2017 [cited by applicant]
US 20180352225A1 · Guo · 2018 [cited by examiner]
US 20190045186A1 · Zhang et al. · 2019 [cited by applicant]
US 20190052877A1 · Zhang et al. · 2019 [cited by applicant]
US 20190182482A1 · Vanam et al. · 2019 [cited by applicant]
US 20210051320A1 · Tourapis et al. · 2021 [cited by applicant]
US 20220101095A1 · Li et al. · 2022 [cited by applicant]
US 20220109832A1 · Du · 2022 [cited by examiner]
US 20220109848A1 · Wang · 2022 [cited by applicant]
US 20220109853A1 · Zhang · 2022 [cited by examiner]
US 20220248007A1 · Misra et al. · 2022 [cited by applicant]
US 20220272335A1 · Liu et al. · 2022 [cited by applicant]
US 20220272347A1 · Zhu et al. · 2022 [cited by applicant]
US 20220272348A1 · Zhang et al. · 2022 [cited by applicant]
US 20220279176A1 · Sarwer et al. · 2022 [cited by applicant]
US 20220286674A1 · Wang et al. · 2022 [cited by applicant]
US 20220295054A1 · Zhao et al. · 2022 [cited by applicant]
US 20220303586A1 · Du et al. · 2022 [cited by applicant]
US 20220321919A1 · Deshpande · 2022 [cited by applicant]
US 20220337853A1 · Li et al. · 2022 [cited by applicant]
CN 104702963B · 2017 [cited by applicant]
WO 2015163046A1 · 2015 [cited by applicant]
WO 2020259538A1 · 2020 [cited by applicant]
WO 2022040428A1 · 2022 [cited by applicant]
Chen et al., An Overview of Core Coding Tools in the AV1 Video Codec, 2018 Picture Coding Symposium (PCS), San Francisco, CA, USA, 2018, pp. 41-45. [cited by applicant]
Rivaz et al., AV1 Bitstream & Decoding Process Specification the Alliance for Open Media 681, Jan. 8, 2019, pp. 1-681. [cited by applicant]
D. Mukherjee, S. Li, Y. Chen, A. Anis, S. Parker, and J. Bankoski. “A switchable loop-restoration with side-information framework for the emerging AV1 video codec.” In 2017 IEEE International Conference on Image Process… [cited by applicant]
S. Midtskogen, A. Fuldseth, G. Bj, and T. Davies. “Integrating Thor tools into the emerging AV1 codec.” In 2017 IEEE International Conference on Image Processing (ICIP), pp. 930-933. IEEE, 2017. [cited by applicant]
S. Midtskogen, and J.-M. Valin. “The AV1 constrained directional enhancement filter (CDEF).” In 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 1193-1197. IEEE, 2018. [cited by applicant]
Daede, Thomas J., Nathan E. Egge, Jean-Marc Valin, Guillaume Martres, and Timothy B. Terriberry. “Daala: A perceptually-driven next generation video codec.” arXiv preprint arXiv:1603.03129 (2016). [cited by applicant]
X. Zhao, Y. Du, Y. Zheng, et al., “Improved CCSO with luma extension and band feature.” CWG-B099, Jan. 2022, pp. 1-5. [cited by applicant]
Y. Du, X. Zhao and S. Liu, “Cross-component sample offset for image and video coding.” In 2021 IEEE International Conference on Visual Communication and Image Processing (VCIP), IEEE, 2021, pp. 1-5. [cited by applicant]
C. Tsai, C. Fu, C. Chen, Y. Huang, S. Lei, “TE10 Subtest2: Coding Unit Synchronous Picture Quadtree-based Adaptive Loop Filter”, ITU-T SG16 WP3 and ISO/IEC JTC1/SC29/WG11 JCTVC-C143, Oct. 2010, pp. 1-12. [cited by applicant]
J. Taquet, P. Onno, C. Gisquet, G. Laroche, “CE5: Results of tests CE5-3.1, CE5-3.2, CE5-3.3 and CE5-3.4 on Non-Linear Adaptive Loop Filter.”, ISO/IEC JTC1/SC29/WVG11 JVET-N0242-v2, Mar. 2019, pp. 1-10. [cited by applicant]
K. Misra, F. Bossen, A. Segall, “Cross-Component Adaptive Loop Filter for chroma”, ISO/IEC JTC1/SC29/WG11 JVET O-0636-r1, Jul. 2019, pp. 1-9. [cited by applicant]
K. Misra, F. Bossen, A. Segall, ect, “CE5-related: On the design of CC-ALF”, ISO/IEC JTC1/SC29/WG11 JVET-P1008-v2, Oct. 2019, pp. 1-6. [cited by applicant]
B. Bross, J. Chen, S. Liu, Y. K. Wang, “Versatile Video Coding (Draft 8)”, ISO/IEC JTC1/SC29/WG11 JVET-Q2001, Jan. 2020, pp. 1-510. [cited by applicant]
International Search Report mailed May 3, 2022 for International Application No. PCT/US2022/014255. [cited by applicant]
Written Opinion mailed May 3, 2022 for International Application No. PCT/US2022/014255. [cited by applicant]
Bross, B. et al.; “Versatile Video Coding (Draft 6)”; Document: JVET-O2001-vE; Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG11, 15th Meeting: Gothenburg, SE, Jul. 3-12, 2019; 455 pages. [cited by applicant]
Bross, B et al.; “Versatile Video Coding (Draft 7)”; Document: JVET-P2001-vE; Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 16th Meeting: Geneva, CH, Oct. 1-11, 2019; 485 pages. [cited by applicant]
Office Action issued on European Application 22760638.1 on Jun. 9, 2023, 16p. [cited by applicant]
Bross, Benjamin et al., “Working Draft 5 of Versatile Video Coding”, JVET, International Organization for Standardization, ISO/IEC JTC 1/SC29/WG11 N18370, Coding of Moving Pictures and Audio, Mar. 2019, 406p, CH. [cited by applicant]
Bhat, Madhukar et al., “AHG10: Adaptive Coding Sub-set for encoder optimization”, Input Document to JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 15th Meeting, Jul. 3-12, 2019, 7p, SE. [cited by applicant]
International Search Report with Written Opinion mailed Dec. 12, 2023 for International Application No. PCT/US2023/074464. [cited by applicant]