IP Library Granted Patent US 12,457,321
Granted Patent B2
US 12,457,321 · App. 16/494,330 · Granted Oct 28, 2025

Prediction method and device using reference block

Inventors: Jin-Ho Lee (Daejeon, KR); Sung-Chang Lim (Daejeon, KR); Jung-Won Kang (Daejeon, KR); Hyunsuk Ko (Daejeon, KR); Ha-Hyun Lee (Seoul, KR); Dong-San Jun (Daejeon, KR); Seung-Hyun Cho (Daejeon, KR); Hui-Yong Kim (Daejeon, KR); Jae-Gon Kim (Goyang-si, KR); Do-Hyeon Park (Goyang-si, KR); Hae-Chul Choi (Daejeon, KR)
Assignees: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE; INDUSTRY-UNIVERSITY COOPERATION FOUNDATION KOREA AEROSPACE UNIVERSITY; HANBAT NATIONAL UNIVERSITY INDUSTRY-ACADEMIC COOPERATION FOUNDATION
H04N19/105H04N19/139H04N19/176H04N19/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,457,321
App. No.
16/494,330
Granted
Oct 28, 2025
Kind
B2
Abstract

Disclosed herein are a video decoding method and apparatus and a video encoding method and apparatus. In the encoding and decoding of a video, a prediction mode may be selected from among multiple prediction modes using prediction mode information for a target block, and prediction may be performed on the target block based on the selected prediction mode, wherein the prediction mode information includes information related to a reference block, and wherein the reference block comprises one or more of a spatially neighboring block that is spatially adjacent to the target block and a temporally neighboring block that is temporally adjacent to the target block. When the selected prediction mode includes a plurality of prediction modes, a prediction mode to be used for prediction of the target block may be decided on among the plurality of selected prediction modes through signaling of the prediction mode decision information.

Claims (93)

1. A video decoding method, comprising:

generating a merge list for a target block; and

performing inter prediction for the target block using the merge list,

wherein the merge list is generated based on information of a plurality of adjacent blocks spatially adjacent to the target block,

wherein the target block is composed of a plurality of sub-blocks corresponding to spatially partitioned regions of the target block, respectively,

wherein a size of the target block is greater than a size of each of the plurality of sub-blocks,

wherein the merge list is commonly used to derive initial motion vectors of the plurality of sub-blocks,

wherein a motion vector of a sub-block of the plurality of sub-blocks is derived based on an initial motion vector of the sub-block using a block matching between two blocks determined based on the initial motion vector,

wherein whether bi-directional inter prediction is used for the target block or not is checked to determine a type of prediction for the target block,

wherein the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed in a case that the bi-directional inter prediction is used for the target block, and

wherein whether the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed or not is determined using a difference between a Picture Order Count (POC) of a target picture comprising the target block and a POC of a reference picture for the target block.

2. A video encoding method, comprising:

generating a merge list for a target block; and

generating an index for the merge list,

wherein the merge list is generated based on information of a plurality of adjacent blocks spatially adjacent to the target block,

wherein the target block is composed of a plurality of sub-blocks corresponding to spatially partitioned regions of the target block, respectively,

wherein a size of the target block is greater than a size of each of the plurality of sub-blocks,

wherein the index indicates a candidate of the merge list which is commonly used to derive initial motion vectors of the plurality of sub-blocks in decoding of the target block,

wherein a motion vector of a sub-block of the plurality of sub-blocks is derived based on an initial motion vector of the sub-block using a block matching between two blocks determined based on the initial motion vector in the decoding of the target block,

wherein whether bi-directional inter prediction is used for the target block or not is checked to determine a type of prediction for the target block,

wherein the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed in a case that the bi-directional inter prediction is used for the target block, and

wherein whether the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed or not is determined using a difference between a Picture Order Count (POC) of a target picture comprising the target block and a POC of a reference picture for the target block.

3. A non-transitory computer-readable medium storing a bitstream configured by a video encoding apparatus performing a video encoding method, the video encoding method comprising:

generating a merge list for a target block;

generating an index for the merge list; and

storing the bitstream comprising index information indicating the index in the non-transitory computer-readable medium,

wherein the merge list is generated based on information of a plurality of adjacent blocks spatially adjacent to the target block,

wherein the target block is composed of a plurality of sub-blocks corresponding spatially partitioned regions of the target block, respectively,

wherein a size of the target block is greater than a size of each of the plurality of sub-blocks,

wherein the index indicates a candidate of the merge list which is commonly used to derive initial motion vectors of the plurality of sub-blocks in decoding of the target block,

wherein a motion vector of a sub-block of the plurality of sub-blocks is derived based on an initial motion vector of the sub-block using a block matching between two blocks determined based on the initial motion vector,

wherein whether bi-directional inter prediction is used for the target block or not is checked to determine a type of prediction for the target block,

the inter prediction to derive the initial motion vectors of the plurality of sub-blocks using the merge list is performed in a case that the bi-directional inter prediction is used for the target block, and

wherein whether the inter prediction to derive the initial motion vectors of the plurality of sub-blocks using the merge list is performed or not is determined using a difference between a Picture Order Count (POC) of a target picture comprising the target block and a POC of a reference picture for the target block.

4. A method for sending a bitstream, the method comprising:

sending the bitstream comprising an index for a merge list;

wherein the index indicates a candidate of the merge list,

wherein the merge list is generated based on information of a plurality of adjacent blocks spatially adjacent to the target block,

wherein the target block is composed of a plurality of sub-blocks corresponding to spatially partitioned regions of the target block, respectively,

wherein a size of the target block is greater than a size of each of the plurality of sub-blocks,

wherein the candidate of the merge list is commonly used to derive initial motion vectors of the plurality of sub-blocks,

wherein a motion vector of a sub-block of the plurality of sub-blocks is derived based on an initial motion vector of the sub-block using a block matching between two blocks determined based on the initial motion vector,

wherein whether bi-directional inter prediction is used for the target block or not is checked to determine a type of prediction for the target block,

wherein the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed in a case that the bi-directional inter prediction is used for the target block, and

wherein whether the inter prediction derives the initial motion vectors of the plurality of sub-blocks using the merge list is performed or not is determined using a difference between a Picture Order Count (POC) of a target picture comprising the target block and a POC of a reference picture for the target block.

5. The video decoding method of claim 1 ,

wherein the plurality of adjacent blocks comprise a first block, a second block, a third block, a fourth block, a fifth block, a sixth block and a seventh block,

wherein the first block is a left-most block among a plurality of blocks adjacent to an upper side of the target block,

wherein the second block is a right-most block among a plurality of blocks adjacent to an upper side of the target block,

wherein the third block is an upper-most block among a plurality of blocks adjacent to a left side of the target block,

wherein the fourth block is a bottom-most block among a plurality of blocks adjacent to a left side of the target block,

wherein the fifth block is diagonally adjacent to an upper-left corner of the target block,

wherein the sixth block is diagonally adjacent to an upper-right corner of the target block, and

wherein the seventh block is diagonally adjacent to a bottom-left corner of the target block.

6. The video decoding method of claim 1 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for four reference blocks, and

wherein x-coordinates of rightmost pixels of the four reference blocks are smaller than x-coordinates of leftmost pixels of the target block.

7. The video decoding method of claim 1 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for seven different blocks spatially adjacent to the target block.

8. The video encoding method of claim 2 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for four reference blocks, and

wherein x-coordinates of rightmost pixels of the four reference blocks are smaller than x-coordinates of leftmost pixels of the target block.

9. The video encoding method of claim 2 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for seven different blocks spatially adjacent to the target block.

10. The method of claim 4 ,

wherein the plurality of adjacent blocks comprise a first block, a second block, a third block, a fourth block, a fifth block, a sixth block and a seventh block,

wherein the first block is a left-most block among a plurality of blocks adjacent to an upper side of the target block,

wherein the second block is a right-most block among a plurality of blocks adjacent to an upper side of the target block,

wherein the third block is an upper-most block among a plurality of blocks adjacent to a left side of the target block,

wherein the fourth block is a bottom-most block among a plurality of blocks adjacent to a left side of the target block,

wherein the fifth block is diagonally adjacent to an upper-left corner of the target block,

wherein the sixth block is diagonally adjacent to an upper-right corner of the target block, and

wherein the seventh block is diagonally adjacent to a bottom-left corner of the target block.

11. The method of claim 4 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for four reference blocks, and

wherein x-coordinates of rightmost pixels of the four reference blocks are smaller than x-coordinates of leftmost pixels of the target block.

12. The method of claim 4 ,

wherein a motion vector for the inter prediction for the target block is derived based on information for seven different blocks spatially adjacent to the target block.

13. The video decoding method of claim 1 ,

wherein the number of rows of the plurality of sub-blocks in the target block is 4, and

wherein the number of columns of the plurality of sub-blocks in the target block is 4.

14. The video encoding method of claim 2 ,

wherein the number of rows of the plurality of sub-blocks in the target block is 4, and

wherein the number of columns of the plurality of sub-blocks in the target block is 4.

15. The method of claim 4 ,

wherein the number of rows of the plurality of sub-blocks in the target block is 4, and

wherein the number of columns of the plurality of sub-blocks in the target block is 4.

16. The video decoding method of claim 1 ,

wherein the block matching is performed using a Sum of Absolute Differences (SAD) between the two blocks.

17. The video encoding method of claim 2 ,

wherein the block matching is performed using a Sum of Absolute Differences (SAD) between the two blocks.

18. The method of claim 4 ,

wherein the block matching is performed using a Sum of Absolute Differences (SAD) between the two blocks.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 16, 2019
From: LEE, JIN-HO; LIM, SUNG-CHANG; KANG, JUNG-WON; KO, HYUNSUK; LEE, HA-HYUN; JUN, DONG-SAN; CHO, SEUNG-HYUN; KIM, HUI-YONG; KIM, JAE-GON; PARK, DO-HYEON; CHOI, HAE-CHUL
To: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE; INDUSTRY-UNIVERSITY COOPERATION FOUNDATION KOREA AEROSPACE UNIVERSITY; HANBAT NATIONAL UNIVERSITY INDUSTRY-ACADEMIC COOPERATION FOUNDATION
Reel/Frame 050378/0621 →
Priority Claims (2)
KR 10-2017-0036267 · Mar 22, 2017 · national
KR 10-2018-0033279 · Mar 22, 2018 · national
Continuity (1)
Related Publication 20200084441A1 · Mar 12, 2020
References Cited (41)
US 9025893B2 · Lee et al. · 2015 [cited by applicant]
US 9641844B2 · Kim et al. · 2017 [cited by applicant]
US 20040086047A1 · Kondo · 2004 [cited by examiner]
US 20080310502A1 · Kim et al. · 2008 [cited by applicant]
US 20090016443A1 · Kim et al. · 2009 [cited by applicant]
US 20110080954A1 · Bossen · 2011 [cited by examiner]
US 20110176611A1 · Huang et al. · 2011 [cited by applicant]
US 20120082210A1 · Chien · 2012 [cited by examiner]
US 20120243609A1 · Zheng · 2012 [cited by examiner]
US 20120269271A1 · Chen · 2012 [cited by examiner]
US 20130077689A1 · Lim · 2013 [cited by examiner]
US 20130142259A1 · Lim · 2013 [cited by examiner]
US 20130163669A1 · Lim · 2013 [cited by examiner]
US 20130272412A1 · Seregin · 2013 [cited by examiner]
US 20140153647A1 · Nakamura · 2014 [cited by examiner]
US 20140205014A1 · Nakamura · 2014 [cited by examiner]
US 20140301461A1 · Jeon · 2014 [cited by examiner]
US 20140376638A1 · Nakamura · 2014 [cited by examiner]
US 20150264351A1 · Miyoshi · 2015 [cited by examiner]
US 20150382008A1 · Lim · 2015 [cited by examiner]
US 20160021389A1 · Suzuki · 2016 [cited by examiner]
US 20160112717A1 · Samuelsson et al. · 2016 [cited by applicant]
US 20160286230A1 · Li et al. · 2016 [cited by applicant]
US 20160286232A1 · Li · 2016 [cited by examiner]
US 20160366435A1 · Chien · 2016 [cited by examiner]
US 20160373688A1 · Imajo · 2016 [cited by examiner]
US 20170180738A1 · Park · 2017 [cited by examiner]
US 20170214932A1 · Huang · 2017 [cited by examiner]
US 20180192071A1 · Chuang · 2018 [cited by examiner]
US 20180324454A1 · Lin · 2018 [cited by examiner]
US 20180359470A1 · Lee · 2018 [cited by examiner]
US 20190191171A1 · Ikai · 2019 [cited by examiner]
US 20200260102A1 · Lin · 2020 [cited by examiner]
US 20210195185A1 · Raut · 2021 [cited by examiner]
KR 1020120005927A · 2012 [cited by applicant]
KR 101379185B1 · 2014 [cited by applicant]
WO WO2011142815A1 · 2011 [cited by applicant]
International Search Report issued on Jun. 28, 2018 in counterpart International Patent Application No. PCT/KR2018/003393 (3 pages in English and 3 pages in Korean). [cited by applicant]
Bross, et al. “High efficiency video coding (HEVC) text specification draft 10 (for FDIS & last call).” [cited by applicant]
Chen, et al. “Algorithm Description of Joint Exploration Test Model 3.” [cited by applicant]
Satoshi Shimada, et al. “ [cited by applicant]