IP Library › Granted Patent US 12,423,878
Granted Patent B2
US 12,423,878 · App. 17/826,806 · Granted Sep 23, 2025

Content-adaptive online training method and apparatus for deblocking in block-wise image compression

Inventors: Ding Ding (Palo Alto, CA); Wei Jiang (Sunnyvale, CA); Wei Wang (San Jose, CA); Shan Liu (San Jose, CA)
Assignee: Tencent America LLC
G06T9/002G06N3/08H04N19/44H04N19/86
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,423,878
App. No.
17/826,806
Granted
Sep 23, 2025
Kind
B2
Abstract

Aspects of the disclosure provide a method, an apparatus, and non-transitory computer-readable storage medium for video decoding. The apparatus includes processing circuitry that reconstructs blocks of an image that is to be reconstructed from a coded video bitstream. The processing circuitry decodes first deblocking information in the coded video bitstream including a first deblocking parameter of a deep neural network (DNN) in a video decoder. The first deblocking parameter of the DNN is an updated parameter that has been previously determined by a content adaptive training process. The processing circuitry determines the DNN for a first boundary region comprising a subset of samples in the reconstructed blocks based on the first deblocking parameter included in the first deblocking information. The processing circuitry deblocks the first boundary region comprising the subset of samples in the reconstructed blocks based on the determined DNN corresponding to the first deblocking parameter.

Claims (68)

1. A method for video decoding in a video decoder, comprising:

reconstructing blocks of an image that is to be reconstructed from a coded video bitstream;

decoding first deblocking information in the coded video bitstream including a first deblocking parameter of a deep neural network (DNN) in the video decoder, wherein the first deblocking parameter of the DNN is an updated parameter that has been previously determined by a content adaptive training process;

determining the DNN in the video decoder for a first boundary region comprising a subset of samples in the reconstructed blocks based on the first deblocking parameter included in the first deblocking information; and

deblocking the first boundary region comprising the subset of samples in the reconstructed blocks based on the determined DNN corresponding to the first deblocking parameter.

2. The method of claim 1 , wherein

the reconstructed blocks include first neighboring reconstructed blocks that have a first shared boundary and include the first boundary region of samples on both sides of the first shared boundary;

the first neighboring reconstructed blocks further include non-boundary regions that are outside the first boundary region; and

the first boundary region in the first neighboring reconstructed blocks is replaced with the deblocked first boundary region.

3. The method of claim 2 , wherein

the reconstructed blocks include second neighboring reconstructed blocks that have a second shared boundary and include a second boundary region of samples on both sides of the second shared boundary; and

the method further includes:

decoding second deblocking information in the coded video bitstream corresponding to the second boundary region, the second deblocking information indicating a second deblocking parameter that has been previously determined by a content adaptive training process, the second boundary region being different from the first boundary region;

updating the DNN based on the first deblocking parameter and the second deblocking parameter, the updated DNN corresponding to the second boundary region and being configured with the first deblocking parameter and the second deblocking parameter; and

deblocking the second boundary region based on the updated DNN corresponding to the second boundary region.

4. The method of claim 2 , wherein

the reconstructed blocks include second neighboring reconstructed blocks of the reconstructed blocks that have a second shared boundary and include a second boundary region having samples on both sides of the second shared boundary; and

the method further includes deblocking the second boundary region based on the determined DNN corresponding to the first boundary region.

5. The method of claim 2 , wherein

the first boundary region further includes samples on both sides of a third shared boundary between third two neighboring reconstructed blocks included in the reconstructed blocks, and

the first two neighboring reconstructed blocks are different from the third two neighboring reconstructed blocks.

6. The method of claim 1 , wherein the first deblocking parameter is a bias term or a weight coefficient in the DNN.

7. The method of claim 1 , wherein

the DNN is configured with initial parameters, and

the determining the DNN includes updating one of the initial parameters based on the first deblocking parameter.

8. The method of claim 7 , wherein

the first deblocking information indicates a difference between the first deblocking parameter and the one of the initial parameters, and

the method further includes determining the first deblocking parameter according to a sum of the difference and the one of the initial parameters.

9. The method of claim 1 , wherein a number of layers of the DNN is dependent on a size of the first boundary region.

10. An apparatus for video decoding, comprising:

processing circuitry configured to:

reconstruct blocks of an image that is to be reconstructed from a coded video bitstream;

decode first deblocking information in the coded video bitstream including a first deblocking parameter of a deep neural network (DNN) in the video decoder, wherein the first deblocking parameter of the DNN is an updated parameter that has been previously determined by a content adaptive training process;

determine the DNN in the video decoder for a first boundary region comprising a subset of samples in the reconstructed blocks based on the first deblocking parameter included in the first deblocking information; and

deblock the first boundary region comprising the subset of samples in the reconstructed blocks based on the determined DNN corresponding to the first deblocking parameter.

11. The apparatus of claim 10 , wherein

the reconstructed blocks include first neighboring reconstructed blocks that have a first shared boundary and include the first boundary region of samples on both sides of the first shared boundary;

the first neighboring reconstructed blocks further include non-boundary regions that are outside the first boundary region; and

the first boundary region in the first neighboring reconstructed blocks is replaced with the deblocked first boundary region.

12. The apparatus of claim 11 , wherein

the reconstructed blocks include second neighboring reconstructed blocks that have a second shared boundary and include a second boundary region of samples on both sides of the second shared boundary; and

the processing circuitry is configured to:

decode second deblocking information in the coded video bitstream corresponding to the second boundary region, the second deblocking information indicating a second deblocking parameter that has been previously determined by a content adaptive training process, the second boundary region being different from the first boundary region;

update the DNN based on the first deblocking parameter and the second deblocking parameter, the updated DNN corresponding to the second boundary region and being configured with the first deblocking parameter and the second deblocking parameter; and

deblock the second boundary region based on the updated DNN corresponding to the second boundary region.

13. The apparatus of claim 11 , wherein

the reconstructed blocks include second neighboring reconstructed blocks of the reconstructed blocks that have a second shared boundary and include a second boundary region having samples on both sides of the second shared boundary; and

the processing circuitry is configured to deblock the second boundary region based on the determined DNN corresponding to the first boundary region.

14. The apparatus of claim 11 , wherein

the first boundary region further includes samples on both sides of a third shared boundary between third two neighboring reconstructed blocks included in the reconstructed blocks, and

the first two neighboring reconstructed blocks are different from the third two neighboring reconstructed blocks.

15. The apparatus of claim 10 , wherein the first deblocking parameter is a bias term or a weight coefficient in the DNN.

16. The apparatus of claim 10 , wherein

the DNN is configured with initial parameters, and

the processing circuitry is configured to update one of the initial parameters based on the first deblocking parameter.

17. The apparatus of claim 16 , wherein

the first deblocking information indicates a difference between the first deblocking parameter and the one of the initial parameters, and

the processing circuitry is configured to determine the first deblocking parameter according to a sum of the difference and the one of the initial parameters.

18. The apparatus of claim 10 , wherein a number of layers of the DNN is dependent on a size of the first boundary region.

19. A non-transitory computer-readable storage medium storing a program executable by at least one processor to perform:

reconstructing blocks of an image that is to be reconstructed from a coded video bitstream;

decoding first deblocking information in the coded video bitstream including a first deblocking parameter of a deep neural network (DNN) in the video decoder, wherein the first deblocking parameter of the DNN is an updated parameter that has been previously determined by a content adaptive training process;

determining the DNN in the video decoder for a first boundary region comprising a subset of samples in the reconstructed blocks based on the first deblocking parameter included in the first deblocking information; and

deblocking the first boundary region comprising the subset of samples in the reconstructed blocks based on the determined DNN corresponding to the first deblocking parameter.

20. The non-transitory computer-readable storage medium of claim 19 , wherein

the reconstructed blocks include first neighboring reconstructed blocks that have a first shared boundary and include the first boundary region of samples on both sides of the first shared boundary;

the first neighboring reconstructed blocks further include non-boundary regions that are outside the first boundary region; and

the first boundary region in the first neighboring reconstructed blocks is replaced with the deblocked first boundary region.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 27, 2022
From: DING, DING; JIANG, WEI; WANG, WEI; LIU, SHAN
To: TENCENT AMERICA LLC
Reel/Frame 060039/0541 →
Continuity (2)
Provisional Application 63211408 · Jun 16, 2021
Related Publication 20220405979A1 · Dec 22, 2022
References Cited (22)
US 20210012537A1 · Xu et al. · 2021 [cited by applicant]
US 20210099710A1 · Salehifar et al. · 2021 [cited by applicant]
US 20210160522A1 · Lee et al. · 2021 [cited by applicant]
US 20210295474A1 · Wang · 2021 [cited by examiner]
US 20210409783A1 · Wan et al. · 2021 [cited by applicant]
US 20220191483A1 · Li et al. · 2022 [cited by applicant]
WO 2019117646A1 · 2019 [cited by applicant]
Office Action received for Japanese Patent Application No. 2023-524439, mailed on May 13, 2024, 10 pages (5 pages of English Translation and 5 pages of Original Document). [cited by applicant]
Yang et al., “Improved method of deblocking filter based on convolutional neural network in VVC”, 2020 IEEE/CIC International Conference on Communications in China (ICCC), IEEE, 2020, pp. 764-769. [cited by applicant]
D. Minnen, J. Ballé, G. Toderici, “Joint Autoregressive and Hierarchical Priors for Learned Image Compression”, Proceedings of the 32nd International Conference on Neural Information Processing Systems, Dec. 2018, pp. 1… [cited by applicant]
D. Liu, Z. Chen, S. Liu, F. Wu, “Deep Learning-Based Technology in Responses to the Joint Call for Proposals on Video Compression with Capability beyond HEVC”, IEEE Transactions on Circuits and Systems for Video Technol… [cited by applicant]
Y. Li, S. Liu, K. Kawamura, “Methodology and reporting template for neural network coding tool testing”, ISO/IEC JTC1/SC29/WG11 JVET-M1006, pp. 1-4. [cited by applicant]
S. Liu, L. Wang, p. Wu, H. Yang, “JVET AHG report 9: Neural Networks in Video Coding (AHG9)”, ISO/IEC JTC1/SC29/WG11 JVET-J0009, pp. 1-3. [cited by applicant]
S. Liu, X. Zhang, S. Lei, “Rectangular partitioning for Intra prediction in HEVC”, Visual Communications and Image Processing (VCIP), IEEE, Jan. 2012, pp. 1-6. [cited by applicant]
Ballé J, Minnen D, Singh S, et al. Variational image compression with a scale hyperprior[J]. arXiv preprint arXiv:1802.01436, 2018, pp. 1-23. [cited by applicant]
Toderici G, Vincent D, Johnston N, et al. “Full resolution image compression with recurrent neural networks”, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 2017: 5306-5314. [cited by applicant]
Cheng Z, Sun H, Takeuchi M, et al. Learned image compression with discretized gaussian mixture likelihoods and attention modules, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2020: … [cited by applicant]
Supplementary European Search Report issued Jul. 24, 2023 in Application No. 22797612.3.(10 pages). [cited by applicant]
Kee-Koo Kwon: “Adaptive postprocessing algorithm in block-coded images using block classification and MLP”, IEICE Transactions on Fundamentals of Electronics, Communications and Computer Sciences, Engineering Sciences S… [cited by applicant]
Yat Hong Lam et al: “Compressing Weight-updates for Image Artifacts Removal Neural Networks”, arxiv.org, Cornell University Library, 201 Olin Library Cornell University Ithaca, NY 14853, May 10, 2019. [cited by applicant]
Kee-Koo Kwon et al: “Deblocking Algorithm in Mpeg-4 Video Coding Using Block Boundary Characteristics and Adaptive Filtering”, Image Processing, 2005. ICIP 2005. IEEE International Conference on, IEEE, Piscataway, NJ, U… [cited by applicant]
International Search Report and Written Opinion issued Aug. 19, 2022 in Application No. PCT/US2022/072748.(9 pages). [cited by applicant]