IP Library › Granted Patent US 12,695,899
Granted Patent B2
US 12,695,899 · App. 18/958,554 · Granted Jul 28, 2026

Video image processing method and device

Inventors: Xiaozhen Zheng (Shenzhen, CN); Suhong Wang (Shenzhen, CN); Shanshe Wang (Shenzhen, CN); Siwei Ma (Shenzhen, CN)
Assignee: SZ DJI TECHNOLOGY CO., LTD.
H04N19/513H04N19/12H04N19/136H04N19/176H04N19/587
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,695,899
App. No.
18/958,554
Filed
Nov 25, 2024
Granted
Jul 28, 2026
Kind
B2
Art Unit
2482
USPC
375/240.16
Abstract

An encoder includes a memory storing program instructions and a processor configured to execute the program instructions to determine a current image block, turn off a temporal motion vector prediction (TMVP) operation in response to a size of the current image block meeting a preset condition so that a temporal candidate motion vector of the current image block is not determined according to the TMVP operation, and encode the current image block. The TMVP operation includes determining a relevant block of the current image block in a temporal neighboring image, and determining the temporal candidate motion vector of the current image block according to a motion vector of the relevant block.

Claims (56)

1 . An encoder comprising:

a memory storing program instructions; and

a processor configured to execute the program instructions to:

determine a current image block;

in response to a size of the current image block meeting a preset condition, turn off a temporal motion vector prediction (TMVP) operation so that a temporal candidate motion vector of the current image block is not determined according to the TMVP operation; and

encode the current image block;

wherein the TMVP operation includes:

determining a relevant block of the current image block in a temporal neighboring image; and

determining the temporal candidate motion vector of the current image block according to a motion vector of the relevant block.

2 . The encoder of claim 1 , wherein the preset condition includes that a number of pixels of the current image block is:

less than a preset number, or

less than or equal to a preset number.

3 . The encoder of claim 2 , wherein the preset number is 32.

4 . The encoder of claim 1 , wherein the processor is further configured to execute the program instructions to, when turning off the TMVP operation so that the temporal candidate motion vector of the current image block is not determined according to the TMVP operation:

turn off the TMVP operation in a motion information candidate list of a first type of mode or in a motion information candidate list of a second type of mode; or

turn off the TMVP operation in the motion information candidate list of the first type of mode and in the motion information candidate list of the second type of mode.

5 . The encoder of claim 4 , wherein:

the first type of mode includes merge mode and/or affine merge mode; and/or

the second type of mode includes advanced motion vector prediction (AMVP) mode.

6 . A decoder comprising:

a memory storing program instructions; and

a processor configured to execute the program instructions to:

determine a current image block;

in response to a size of the current image block meeting a preset condition, turn off a temporal motion vector prediction (TMVP) operation so that a temporal candidate motion vector of the current image block is not determined according to the TMVP operation; and

decode the current image block;

wherein the TMVP operation includes:

determining a relevant block of the current image block in a temporal neighboring image; and

determining the temporal candidate motion vector of the current image block according to a motion vector of the relevant block.

7 . The decoder of claim 6 , wherein the preset condition includes that a number of pixels of the current image block is:

less than a preset number, or

less than or equal to a preset number.

8 . The decoder of claim 7 , wherein the preset number is 32.

9 . The decoder of claim 6 , wherein the processor is further configured to execute the program instructions to, when turning off the TMVP operation so that the temporal candidate motion vector of the current image block is not determined according to the TMVP operation:

turn off the TMVP operation in a motion information candidate list of a first type of mode or in a motion information candidate list of a second type of mode; or

turn off the TMVP operation in the motion information candidate list of the first type of mode and in the motion information candidate list of the second type of mode.

10 . The decoder of claim 9 , wherein:

the first type of mode includes merge mode and/or affine merge mode; and/or

the second type of mode includes advanced motion vector prediction (AMVP) mode.

11 . A bitstream generation method comprising:

determining a current image block;

in response to a size of the current image block meeting a preset condition, turning off a temporal motion vector prediction (TMVP) operation so that a temporal candidate motion vector of the current image block is not determined according to the TMVP operation;

encoding the current image block; and

generating a bitstream;

wherein the TMVP operation includes:

determining a relevant block of the current image block in a temporal neighboring image; and

determining the temporal candidate motion vector of the current image block according to a motion vector of the relevant block.

12 . The method of claim 11 , wherein the preset condition includes that a number of pixels of the current image block is:

less than a preset number, or

less than or equal to a preset number.

13 . The method of claim 12 , wherein the preset number is 32.

14 . The method of claim 11 , wherein turning off the TMVP operation so that the temporal candidate motion vector of the current image block is not determined according to the TMVP operation includes:

turning off the TMVP operation in a motion information candidate list of a first type of mode or in a motion information candidate list of a second type of mode; or

turning off the TMVP operation in the motion information candidate list of the first type of mode and in the motion information candidate list of the second type of mode.

15 . The method of claim 14 , wherein:

the first type of mode includes merge mode and/or affine merge mode; and/or

the second type of mode includes advanced motion vector prediction (AMVP) mode.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 25, 2024
From: ZHENG, XIAOZHEN; WANG, SUHONG; WANG, SHANSHE; MA, SIWEI
To: SZ DJI TECHNOLOGY CO., LTD.
Reel/Frame 069397/0815 →
Priority Claims (1)
WO PCT/CN2019/070315 · Jan 3, 2019 · international
Continuity (5)
Continuation 18341246 · Jun 26, 2023
Continuation 17645143 · Dec 20, 2021
Continuation 17060011 · Sep 30, 2020
Continuation PCTCN2019077893 · Mar 12, 2019
Related Publication 20250088655A1 · Mar 13, 2025
References Cited (56)
US 6480546B1 · Kim et al. · 2002 [cited by applicant]
US 10362330B1 · Li et al. · 2019 [cited by applicant]
US 11178420B2 · Wang et al. · 2021 [cited by applicant]
US 11206422B2 · Zheng et al. · 2021 [cited by applicant]
US 12155856B2 · Zheng · 2024 [cited by examiner]
US 20070110156A1 · Ji et al. · 2007 [cited by applicant]
US 20160219278A1 · Chen et al. · 2016 [cited by applicant]
US 20160286232A1 · Li et al. · 2016 [cited by applicant]
US 20160366435A1 · Chien et al. · 2016 [cited by applicant]
US 20170034512A1 · Casula · 2017 [cited by applicant]
US 20170238005A1 · Chien · 2017 [cited by applicant]
US 20180084260A1 · Chien et al. · 2018 [cited by applicant]
US 20180199057A1 · Chuang et al. · 2018 [cited by applicant]
US 20200077115A1 · Li et al. · 2020 [cited by applicant]
US 20200195948A1 · Li et al. · 2020 [cited by applicant]
US 20210266589A1 · Chen et al. · 2021 [cited by applicant]
CN 101188772A · 2008 [cited by applicant]
CN 101350928A · 2009 [cited by applicant]
CN 101573985A · 2009 [cited by applicant]
CN 101841712A · 2010 [cited by applicant]
CN 101873500A · 2010 [cited by applicant]
CN 102148990A · 2011 [cited by applicant]
CN 102215395A · 2011 [cited by applicant]
CN 102625094A · 2012 [cited by applicant]
CN 102939751A · 2013 [cited by applicant]
CN 104079944A · 2014 [cited by applicant]
CN 104935938A · 2015 [cited by applicant]
CN 106537915A · 2017 [cited by applicant]
CN 107211156A · 2017 [cited by applicant]
CN 108886616A · 2018 [cited by applicant]
CN 109005407A · 2018 [cited by applicant]
CN 109076236A · 2018 [cited by applicant]
JP 2007060693A · 2007 [cited by applicant]
JP 2021516502A · 2021 [cited by applicant]
WO 2016089933A1 · 2016 [cited by applicant]
WO 2018048904A1 · 2018 [cited by applicant]
WO 2018058526A1 · 2018 [cited by applicant]
WO 2019089933A1 · 2019 [cited by applicant]
WO 2020004990A1 · 2020 [cited by applicant]
The World Intellectual Property Organization (WIPO) International Search Report for PCT/CN2019/070315 Sep. 27, 2019 5 Pages. [cited by applicant]
The World Intellectual Property Organization (WIPO) International Search Report for PCT/CN2019/077893 Sep. 16, 2019 5 Pages. [cited by applicant]
Zhichu He, et al., “Framework of AVS2-VIDEO Coding”, ICIP2013, Feb. 13, 2014, p. 1515-1519. [cited by applicant]
Yuan Yuan, et al., “A New Transform Structure for Geometry Motion Partitioning in Video Coding”, Journal of Shanghai University, vol. 19, No. 3, Jun. 2013, p. 240-244. [cited by applicant]
Tamse (Samsung) A et al: “CE4-related: Redundant Removal for ATMVP”, 12. JVET Meeting; Oct. 3, 2018-Oct. 12, 2018; Macao; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16 ), No. JVET-L0055 Oct… [cited by applicant]
Leehet Al: “CE4-related: Fixed sub-block size and restriction for ATMVP”, 12. JVET Meeting; Oct. 3, 2018-Oct. 12, 2018; Macao; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16 ), No. JVET-L046… [cited by applicant]
Xiu X et al: “Draft text for advanced temporal motion vector prediction (ATMVP) and adaptive motion vector resolution (AMVR)”, 11. JVET Meeting; Jul. 11, 2018-Jul. 18, 2018; Ljubljana; (The Joint Video Exploration Team … [cited by applicant]
Benjamin Brossetal: “Versatile Video Coding (Draft 2)”, The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16, No. JVET-K1001 Jul. 18, 2018 (Juk, 18, 2018), pp. 1-139. [cited by applicant]
C-C Chen etal: “CE4-related: A simplification algorithm for ATMVP”, 12. JVET Meeting; Oct. 3, 2018-Oct. 12, 2018; Macao; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG.16 ), No. JVET-L0092 Oct.… [cited by applicant]
Wang (Peking University) SHet al: “CE4-related: Simplification of ATMVP candidate derivation”, 12. JVET Meeting; Oct. 3, 2018-Oct. 12, 2018; Macao; (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T S… [cited by applicant]
Hahyun Lee, et al. , CE4-related: Fixed sub-block size and restriction for ATMVP , Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-L0468-v2 , 12th Meeting: Macao, CN, Oct. 2018, … [cited by applicant]
Jianle Chen, et al. , Algorithm Description of Joint Exploration Test Model 7 (JEM 7) , Joint Video Exploration Team (JVET) of ITU-T SG 16 WP3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-G1001-v1 , 7th Meeting: Torino, IT, Aug… [cited by applicant]
Suhong Wang, et al. , CE4-related: Simplification of ATMVP candidate derivation , Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-L0198_r2 , 12th Meeting: Macao, CN, Jul. 2019, p… [cited by applicant]
Hyeongmun Jang, et al. , [CE4-2.6 related] Simplified A TMVP with fixed sub-block size. , Joint Video Experts Team (JVET) of ITUTSG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-K0080-v1 , Joint Video Experts Team (JVET)… [cited by applicant]
M-Wen Chen, and Xianglin Wang , AHG5: Reducing VVC worst-case memory bandwidth by restricting bi-directional 4×4 inter CUs/Sub-blocks , Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IECJTC 1/SC 29/WG 11 , … [cited by applicant]
Suhong Wang et al. , CE4-related: Remove redundancy between TMVP and ATMVP , Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-M0345_v2, 13th Meeting: Marrakech, MA, Jan. 2019, pp.… [cited by applicant]
Suhong Wang, et al. , CE4-1.5: Remove TMVP merge candidate for the specified blocksizes , Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 , JVET-No. 212_V2 , 14th Meeting: Geneva, CN, M… [cited by applicant]