IP Library › Granted Patent US 12,238,318
Granted Patent B2
US 12,238,318 · App. 18/466,623 · Granted Feb 25, 2025

Determination of secondary transform for intra prediction

Inventors: Xin Zhao (San Jose, CA); Liang Zhao (Sunnyvale, CA); Shan Liu (San Jose, CA)
Assignee: Tencent America LLC
H04N19/44H04N19/13H04N19/159H04N19/176H04N19/60
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,238,318
App. No.
18/466,623
Granted
Feb 25, 2025
Kind
B2
Abstract

A method for video decoding in a decoder is provided. Coding information of a block to be reconstructed is decoded from a coded video bitstream. The coding information indicates intra prediction information for the block. Responsive to the block being coded with a directional mode, the directional mode is determined based on a nominal mode and an angular offset, the coding information indicating the nominal mode and the angular offset, a first non-separable transform set of one or more non-separable transforms for the block is determined based on the nominal mode, a non-separable transform in the first non-separable transform set is determined based on a non-separable transform index indicated by the coding information, and the block is reconstructed based on the directional mode and the non-separable transform.

Claims (63)

1. A method for video decoding in a decoder, comprising:

decoding coding information of a block to be reconstructed from a coded video bitstream, the coding information indicating intra prediction information for the block; and

when the block is coded with a directional mode,

determining the directional mode based on a nominal mode and an angular offset, the coding information indicating the nominal mode and the angular offset,

determining a first non-separable transform set of one or more non-separable transforms for the block from a plurality of non-separable transform sets based on a nominal direction indicated by the nominal mode, each of the non-separable transform sets corresponding to a respective one of 8 nominal directions, the 8 nominal directions are associated with V_PRED, H_PRED, D45_PRED, D135_PRED, D113_PRED, D157_PRED, D203_PRED, and D67_PRED,

determining a non-separable transform in the first non-separable transform set based on a non-separable transform index indicated by the coding information, and

reconstructing the block based on the directional mode and the non-separable transform.

2. The method of claim 1 , wherein the determining the first non-separable transform set further comprises:

determining a transform set mode that is associated with the nominal mode,

the transform set mode indicating the first non-separable transform set that includes the non-separable transform.

3. The method of claim 1 , wherein the non-separable transform is a non-separable secondary transform.

4. The method of claim 3 , wherein the non-separable secondary transform does not apply to a PAETH mode or a recursive filtering mode.

5. The method of claim 3 , wherein the reconstructing the block based on the directional mode and the non-separable secondary transform further comprises:

applying the non-separable secondary transform only to N first transform coefficients along a scanning order used for entropy coding the first transform coefficients of the block.

6. The method of claim 3 , wherein the reconstructing the block based on the directional mode and the non-separable secondary transform further comprises:

applying the non-separable secondary transform only to first transform coefficients in the block, each of the first transform coefficients having a coordinate (x, y) and a sum of the respective x and y coordinates being less than a threshold value.

7. The method of claim 3 , wherein a horizontal transform and a vertical transform in a primary transform for the block are included in a subset of a set of line graph transforms.

8. The method of claim 3 , wherein

the block includes first transform coefficients obtained with the non-separable secondary transform and second transform coefficients obtained without the non-separable secondary transform; and

the reconstructing the block further includes entropy decoding the first transform coefficients and the second transform coefficients separately.

9. The method of claim 1 , wherein

non-directional modes include a DC mode, a PAETH mode, a SMOOTH mode, a SMOOTH_V mode, a SMOOTH_H mode, recursive filtering modes, and a chroma from luma (CfL) mode, the DC mode, the PAETH mode, the SMOOTH mode, the SMOOTH_V mode, and the SMOOTH_H mode being based on averaging of neighboring samples of the block, and

when the block is coded with one of the non-directional modes,

determining a second non-separable transform set of one or more non-separable transforms associated with the one of the non-directional modes, one of (a) at least another one of the non-directional modes and (b) a nominal mode being associated with the second non-separable transform set,

determining a non-separable transform in the second non-separable transform set based on the non-separable transform index indicated by the coding information, and

reconstructing the block based on the non-directional mode and the non-separable transform.

10. The method of claim 9 , wherein the one of the non-directional modes and the one of the at least another one of the non-directional modes and the nominal mode include one of (a) the recursive filtering modes and one of the DC mode and the SMOOTH mode, (b) the SMOOTH mode, the SMOOTH_H mode, and the SMOOTH_V mode, (c) the SMOOTH mode, the SMOOTH_H mode, the SMOOTH_V mode, and the PAETH mode, (d) the recursive filtering modes, the SMOOTH mode, and the PAETH mode, (e) a vertical mode for the nominal mode and the SMOOTH_V mode, (f) a horizontal mode for the nominal mode and the SMOOTH_H mode, and (v) the CfL mode and one of the DC mode, the SMOOTH mode, and the PAETH mode.

11. A method for video encoding in an encoder, comprising:

when a block is coded with a directional mode,

determining the directional mode based on a nominal mode and an angular offset,

determining a first non-separable transform set of one or more non-separable transforms for the block from a plurality of non-separable transform sets based on a nominal direction indicated by the nominal mode, each of the non-separable transform sets corresponding to a respective one of 8 nominal directions, the 8 nominal directions are associated with V_PRED, H_PRED, D45_PRED, D135_PRED, D113_PRED, D157_PRED, D203_PRED, and D67 PRED,

determining a non-separable transform in the first non-separable transform set associated with a non-separable transform index,

reconstructing the block based on the directional mode and the non-separable transform, and

encoding coding information indicating the nominal mode, the angular offset, and the non-separable transform index in a bitstream.

12. The method of claim 11 , wherein the determining the first non-separable transform set further comprises:

determining a transform set mode that is associated with the nominal mode,

the transform set mode indicating the first non-separable transform set that includes the non-separable transform.

13. The method of claim 11 , wherein the non-separable transform is a non-separable secondary transform.

14. The method of claim 3 , wherein the non-separable secondary transform does not apply to a PAETH mode or a recursive filtering mode.

15. The method of claim 13 , wherein the reconstructing the block based on the directional mode and the non-separable secondary transform further comprises:

applying the non-separable secondary transform only to N first transform coefficients along a scanning order used for entropy coding the first transform coefficients of the block.

16. The method of claim 13 , wherein the reconstructing the block based on the directional mode and the non-separable secondary transform further comprises:

applying the non-separable secondary transform only to first transform coefficients in the block, each of the first transform coefficients having a coordinate (x, y) and a sum of the respective x and y coordinates being less than a threshold value.

17. The method of claim 13 , wherein a horizontal transform and a vertical transform in a primary transform for the block are included in a subset of a set of line graph transforms.

18. The method of claim 13 , wherein

the block includes first transform coefficients obtained with the non-separable secondary transform and second transform coefficients obtained without the non-separable secondary transform; and

the reconstructing the block further includes entropy decoding the first transform coefficients and the second transform coefficients separately.

19. The method of claim 11 , wherein

non-directional modes include a DC mode, a PAETH mode, a SMOOTH mode, a SMOOTH_V mode, a SMOOTH_H mode, recursive filtering modes, and a chroma from luma (CfL) mode, the DC mode, the PAETH mode, the SMOOTH mode, the SMOOTH_V mode, and the SMOOTH_H mode being based on averaging of neighboring samples of the block, and

when the block is coded with one of the non-directional modes,

determining a second non-separable transform set of one or more non-separable transforms associated with the one of the non-directional modes, one of (a) at least another one of the non-directional modes and (b) a nominal mode being associated with the second non-separable transform set,

determining a non-separable transform in the second non-separable transform set associated with the non-separable transform index,

reconstructing the block based on the non-directional mode and the non-separable transform, and

encoding coding information indicating the non-separable transform index in the bitstream.

20. A method of processing visual media data, the method comprising:

processing a bitstream that includes the visual media data according to a format rule, wherein

the bitstream includes coding information of a block to be reconstructed, the coding information indicating intra prediction information for a block; and

the format rule specifies that:

when the block is coded with a directional mode,

the directional mode is determined based on a nominal mode and an angular offset, the coding information indicating the nominal mode and the angular offset,

a first non-separable transform set of one or more non-separable transforms for the block is determined from a plurality of non-separable transform sets based on a nominal direction indicated by the nominal mode, each of the non-separable transform sets corresponding to a respective one of 8 nominal directions, the 8 nominal directions are associated with V_PRED, H_PRED, D45_PRED, D135_PRED, D113_PRED, D157_PRED, D203_PRED, and D67_PRED,

a non-separable transform in the first non-separable transform set is determined based on a non-separable transform index indicated by the coding information, and

the block is processed based on the directional mode and the non-separable transform.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 29, 2024
From: ZHAO, XIN; ZHAO, LIANG; LIU, SHAN
To: TENCENT AMERICA LLC
Reel/Frame 067546/0502 →
Continuity (4)
Continuation 17828919 · May 31, 2022
Continuation 17031272 · Sep 24, 2020
Provisional Application 62941359 · Nov 27, 2019
Related Publication 20230421790A1 · Dec 28, 2023
References Cited (43)
US 9473777B2 · Kim · 2016 [cited by examiner]
US 10038917B2 · Zhou · 2018 [cited by examiner]
US 10142623B2 · Bang · 2018 [cited by examiner]
US 20100086049A1 · Ye · 2010 [cited by examiner]
US 20110280304A1 · Jeon · 2011 [cited by examiner]
US 20120014436A1 · Segall · 2012 [cited by examiner]
US 20120014440A1 · Segall · 2012 [cited by examiner]
US 20130028317A1 · Parfenov · 2013 [cited by examiner]
US 20130108185A1 · Kenji · 2013 [cited by examiner]
US 20130114696A1 · Liu · 2013 [cited by examiner]
US 20130182773A1 · Seregin · 2013 [cited by examiner]
US 20160353103A1 · Park · 2016 [cited by applicant]
US 20170094313A1 · Zhao · 2017 [cited by examiner]
US 20170238014A1 · Said · 2017 [cited by examiner]
US 20170272759A1 · Seregin · 2017 [cited by examiner]
US 20170332084A1 · Seregin et al. · 2017 [cited by applicant]
US 20170347094A1 · Su · 2017 [cited by examiner]
US 20180103252A1 · Hsieh · 2018 [cited by examiner]
US 20190124339A1 · Young · 2019 [cited by examiner]
US 20190335199A1 · Joshi · 2019 [cited by examiner]
US 20200280717A1 · Li · 2020 [cited by examiner]
JP 2018530245A · 2018 [cited by applicant]
KR 1020180063186A · 2018 [cited by applicant]
WO 2017058614A1 · 2017 [cited by applicant]
WO 2019185883A1 · 2019 [cited by applicant]
Anonymous, “Abstract,” AOMedia, Jan. 1, 2019, pp. 1-12. [cited by applicant]
Bross et al., “CE3: Multiple reference line intra prediction Test 1.1.1, 1.1.2, 1.1.3, and 1.1.4,” Joint Video Exploration Team JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-L0283-v2, 12th Meeting: Macao,… [cited by applicant]
Bross et al., “Versatile Video Coding Draft 7,” Joint Video Exploration Team JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, CH, JVET-P2001-vE, 16th Meeting: Geneva, Oct. 1-11, 2019, 492 pages. [cited by applicant]
Chen et al., “Algoirthm description for Versatile Video Coding and Test Model 7 VTM 7,” The Joint Video Exploration Team of ISO/IEC JTC 1/SC 29/WG 11 and ITU-T SG 16, JVET-P2002-v1, 16th Meeting: Geneva, CH, Oct. 1-11, … [cited by applicant]
Chen et al., “Algorithm Description of Joint Exploration Test Model 2,” Joint Video Exploration Team JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-B1001-v3, Feb. 20-26, 2016, 11 pages. [cited by applicant]
Chen et al., “Algorithm Description of Joint Exploration Test Model 7 JEM7,” Joint Video Experts Team JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-G1001-v1, N17055, 7th Meeting: Torino, IT, Jul. 13-21, 2… [cited by applicant]
Chen et al., “An Overview of Core Coding Tools in the AV1 Video Codec,” PCS, 2018, pp. 41-45. [cited by applicant]
Extended European Search Report in EP20892824.2, mailed Dec. 22, 2022, 12 pages. [cited by applicant]
Fernandez, “Intelligent Resampling Methods for Video Compression,” Dissertation, University of Bristol, May 2019, 194 pages. [cited by applicant]
International Search Report in PCT/US20/53056, mailed Jan. 11, 2021, 2 pages. [cited by applicant]
Jeong et al., “A Fast Intra Mode Decisioning Based on Accuracy of Rate Distortion Model for Av1 Intra Encoding,” 2019 34th International Technical Conference on Circuits/Systems, Computers and Communications (ITC-CSCC),… [cited by applicant]
Lu et al., “Symmetric Line Graph Transforms for Inter Predictive Video Coding,” 2016 Picture Coding Symposium PCS, Dec. 4, 2016, pp. 1-5. [cited by applicant]
Moonmo et al., “CE6: Reduced Secondary Transform RST test 6.5.1,” Joint Video Experts Team JVET of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-M0292, 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, pp. 1-14. [cited by applicant]
Office Action in JP2021-558852, mailed Oct. 17, 2022, 24 pages. [cited by applicant]
Rivaz et al., “AV1 Bitstream & Decoding Process Specification,” Jan. 8, 2019, 681 pages. [cited by applicant]
Written Opinion in PCT/US20/53056, mailed Jan. 11, 2021, 8 pages. [cited by applicant]
Zhao et al., “Unified Secondary Transform for Intra Coding Beyond Av1,” 2020 IEEE International Conference on Image Processing ICIP, Sep. 30, 2020, pp. 3393-3397. [cited by applicant]
Office Action received for Korean Patent Application No. 10-2021-7032959, mailed on Sep. 30, 2024, 15 pages (8 pages of English Translation and 7 pages of Original Document). [cited by applicant]