IP Library Granted Patent US 12,598,313
Granted Patent B2
US 12,598,313 · App. 18/532,808 · Granted Apr 7, 2026

Extended low-frequency non-separable transform (LFNST) designs with worst-case complexity handling

Inventors: Hilmi Enes Egilmez (San Diego, CA); Vadim Seregin (San Diego, CA); Marta Karczewicz (San Diego, CA)
Assignee: QUALCOMM Incorporated
H04N19/18H04N19/117H04N19/159H04N19/176
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,598,313
App. No.
18/532,808
Granted
Apr 7, 2026
Kind
B2
Abstract

A video decoder can be configured to determine a number of allowed non-zero coefficients for a block of video data based on a size of the block; obtain a set of dequantized coefficients for the block, wherein the set of dequantized coefficients comprises a first subset of dequantized coefficients that includes non-zero dequantized coefficients and a second subset of dequantized coefficients that includes all zero coefficients, wherein a number of coefficients in the first subset of dequantized coefficients is equal to the number of allowed non-zero coefficients for the block of video data; apply an inverse low-frequency non-separable transform (LFNST) to the first subset of dequantized coefficients to determine a first intermediate subset of coefficients; and apply an inverse separable transform to the first intermediate subset of coefficients and at least a portion of the second subset of coefficients to determine a block of reconstructed residual values.

Claims (51)

1 . A method of decoding video data, the method comprising:

determining a number of allowed non-zero coefficients for a block of reconstructed values according to the equation NZ=floor((WCM×nTbW×nTbH)/(TD1×TD2)), wherein nTbH represents a height of the block, nTbW represents a width of the block, NZ equals the number of allowed non-zero coefficients, TD1 represents a first dimension of an inverse low-frequency non-separable transform (LFNST), TD2 represents a second dimension of the LFNST, and WCM represents a constant value;

obtaining a set of dequantized coefficients for the block of video data, wherein the set of dequantized coefficients comprises a first subset of dequantized coefficients that includes non-zero dequantized coefficients and a second subset of dequantized coefficients that includes all zero coefficients, wherein a number of coefficients in the first subset of dequantized coefficients is equal to the number of allowed non-zero coefficients for the block of video data;

applying an inverse low-frequency non-separable transform (LFNST) to the first subset of dequantized coefficients to determine a first intermediate subset of coefficients; and

applying an inverse separable transform to the first intermediate subset of coefficients and at least a portion of the second subset of coefficients to determine a block of reconstructed residual values.

2 . The method of claim 1 , further comprising:

determining an intra prediction mode for the block of video data;

based on the intra prediction mode, determining a set of inverse LFNST candidates from a plurality of sets; and

selecting the inverse LFNST from the determined set of inverse LFNST candidates.

3 . The method of claim 2 , wherein the set of inverse LFNST candidates includes 3 candidates.

4 . The method of claim 2 , wherein the plurality of sets includes 35 sets.

5 . The method of claim 1 , wherein the first subset of dequantized coefficients comprises 16 dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

6 . The method of claim 1 , wherein the first subset of dequantized coefficients comprises 64 dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

7 . The method of claim 1 , wherein the first subset of dequantized coefficients comprises NZ dequantized coefficients.

8 . The method of claim 1 , wherein the first subset of dequantized coefficients comprises 2×NZ dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

9 . The method of claim 1 , further comprising:

adding the block of reconstructed residual values to a prediction block to form a reconstructed block;

applying one or more filters to the reconstructed block to determine a filtered reconstructed block; and

outputting decoded video data that includes the filtered reconstructed block.

10 . The method of claim 1 , wherein the method is performed as part of a video encoding process.

11 . A device for decoding video data, the device comprising:

a memory configured to store video data;

one or more processors implemented in circuitry and configured to:

determine a number of allowed non-zero coefficients for a block of reconstructed values according to the equation NZ=floor((WCM×nTbW×nTbH)/(TD1×TD2)), wherein nTbH represents a height of the block, nTbW represents a width of the block, NZ equals the number of allowed non-zero coefficients, TD1 represents a first dimension of an inverse low-frequency non-separable transform (LFNST), TD2 represents a second dimension of the LFNST, and WCM represents a constant value;

obtain a set of dequantized coefficients for the block of video data, wherein the set of dequantized coefficients comprises a first subset of dequantized coefficients that includes non-zero dequantized coefficients and a second subset of dequantized coefficients that includes all zero coefficients, wherein a number of coefficients in the first subset of dequantized coefficients is equal to the number of allowed non-zero coefficients for the block of video data;

apply an inverse low-frequency non-separable transform (LFNST) to the first subset of dequantized coefficients to determine a first intermediate subset of coefficients; and

apply an inverse separable transform to the first intermediate subset of coefficients and at least a portion of the second subset of coefficients to determine a block of reconstructed residual values.

12 . The device of claim 11 , wherein the one or more processors are further configured to:

determine an intra prediction mode for the block of video data;

based on the intra prediction mode, determine a set of inverse LFNST candidates from a plurality of sets; and

select the inverse LFNST from the determined set of inverse LFNST candidates.

13 . The device of claim 12 , wherein the set of inverse LFNST candidates includes 3 candidates.

14 . The device of claim 12 , wherein the plurality of sets includes 35 sets.

15 . The device of claim 11 , wherein the first subset of dequantized coefficients comprises 16 dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

16 . The device of claim 11 , wherein the first subset of dequantized coefficients comprises 64 dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

17 . The device of claim 11 , wherein the first subset of dequantized coefficients comprises NZ dequantized coefficients.

18 . The device of claim 11 , wherein the first subset of dequantized coefficients comprises 2×NZ dequantized coefficients, and wherein first intermediate subset of coefficients comprises 64 coefficients.

19 . The device of claim 11 , wherein the one or more processors are further configured to:

add the block of reconstructed residual values to a prediction block to form a reconstructed block;

apply one or more filters to the reconstructed block to determine a filtered reconstructed block; and

output decoded video data that includes the filtered reconstructed block.

20 . The device of claim 11 , wherein the device comprises a video encoder.

21 . The device of claim 11 , wherein the device comprises a wireless communication device, further comprising a receiver configured to receive encoded video data.

22 . The device of claim 21 , wherein the wireless communication device comprises a telephone handset and wherein the receiver is configured to demodulate, according to a wireless communication standard, a signal comprising the encoded video data.

23 . The device of claim 11 , further comprising:

a display configured to display decoded video data.

24 . The device of claim 11 , wherein the device comprises one or more of a camera, a computer, a mobile device, a broadcast receiver device, or a set-top box.

25 . The device of claim 11 , further comprising:

a camera configured to capture video data.

26 . The device of claim 11 , wherein the device comprises a wireless communication device, further comprising a transmitter configured to transmit encoded video data.

27 . The device of claim 26 , wherein the wireless communication device comprises a telephone handset and wherein the transmitter is configured to modulate, according to a wireless communication standard, a signal comprising the encoded video data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 7, 2023
From: EGILMEZ, HILMI ENES; SEREGIN, VADIM; KARCZEWICZ, MARTA
To: QUALCOMM INCORPORATED
Reel/Frame 065803/0642 →
Continuity (3)
Continuation 17326588 · May 21, 2021
Provisional Application 63086888 · Oct 2, 2020
Related Publication 20240114152A1 · Apr 4, 2024
References Cited (29)
US 10306229B2 · Zhao et al. · 2019 [cited by applicant]
US 10349085B2 · Said et al. · 2019 [cited by applicant]
US 10448053B2 · Said et al. · 2019 [cited by applicant]
US 10491922B2 · Zhao et al. · 2019 [cited by applicant]
US 10863199B2 · Said et al. · 2020 [cited by applicant]
US 10972733B2 · Zhao et al. · 2021 [cited by applicant]
US 10986340B2 · Egilmez et al. · 2021 [cited by applicant]
US 11032572B2 · Egilmez et al. · 2021 [cited by applicant]
US 20090052791A1 · Gou et al. · 2009 [cited by applicant]
US 20210092381A1 · Egilmez et al. · 2021 [cited by applicant]
US 20210314619A1 · Jung et al. · 2021 [cited by applicant]
US 20220109858A1 · Egilmez et al. · 2022 [cited by applicant]
CN 110636313A · 2019 [cited by applicant]
EP 3806475A1 · 2021 [cited by applicant]
WO 2020009556A1 · 2020 [cited by applicant]
WO 2020116961A1 · 2020 [cited by applicant]
WO 2020162737A1 · 2020 [cited by applicant]
Bross B., et al., “Versatile Video Coding (Draft 10)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 131, MPEG Meeting, 19th Meeting, by Teleconference, Jun. 22-Jul. 1, 2020, Jun. 29… [cited by applicant]
Bross B., et al., “Versatile Video Coding (Draft 6)”, 127th MPEG, Jul. 8, 2019-Jul. 12, 2019, 15th JVET Meeting, Gothenburg, SE, Jul. 3, 2019-Jul. 12, 2019, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IE… [cited by applicant]
Chang Y-J., et al., “Compression Efficiency Methods Beyond VVC”, JVET-U0100, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 21st Meeting, by Teleconference, Jan. 6-15, 2021, XP030293237, De… [cited by applicant]
Chen J., et al., “Algorithm Description for Versatile Video Coding and Test Model 10 (VTM 10)”, JVET-S2002-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 131. MPEG Meeting, 19th M… [cited by applicant]
Chen J., et al., “Algorithm Description of Joint Exploration Test Model 7 (JEM7)”, 119 . MPEG Meeting, 7. JVET Meeting, Jul. 13, 2017-Jul. 21, 2017, JVET-G1001-V1, Joint Video Exploration Team (JVET) of ITU-T SG 16 WP3 … [cited by applicant]
“Departments”, Fraunhofer Heinrich Hertz Institute, 2021, https://www.hhi.fraunhofer.de/en/departments.html [retrieved on Sep. 21, 2021], pp. 1-4. [cited by applicant]
International Search Report and Written Opinion—PCT/US2021/052255—ISA/EPO—Jan. 4, 2022 18 Pages. [cited by applicant]
ITU-T H.265: “Series H: Audiovisual and Multimedia Systems Infrastructure of Audiovisual Services—Coding of Moving Video”, High Efficiency Video Coding, The International Telecommunication Union, Jun. 2019, 696 Pages. [cited by applicant]
Jain A.K., “A Sinusoidal Family of Unitary Transforms”, IEEE Transactions on Pattern Analysis and Machine Intelligence, IEEE Service Center, vol. PAMI-1, No. 4, Oct. 1, 1979, XP011242370, pp. 356-365, ISSN: 0162-8828. [cited by applicant]
Koo (LGE) M., et al., “CE6: Reduced Secondary Transform (RST) (CE6-3.1)”, 14th JVET Meeting, Mar. 19, 2019-Mar. 27, 2019, Geneva, CH (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG. 16 WP 3), No… [cited by applicant]
Lainema (Nokia) J: “CE6-Related: Simplified LFNST”, 127. MPEG Meeting, 15th JVET Meeting, Jul. 8, 2019-Jul. 12, 2019, Gothenburg, (Motion Picture Expert Group or Joint Video Experts Team (JVET) of ITU-T SG16 WP3 and ISO… [cited by applicant]
Taiwan Search Report—TW110136197—TIPO—Feb. 7, 2025. [cited by applicant]