IP Library › Granted Patent US 12,231,634
Granted Patent B2
US 12,231,634 · App. 17/621,151 · Granted Feb 18, 2025

Method and apparatus for image encoding and image decoding using area segmentation

Inventors: Gun Bang (Daejeon, KR); Jin-Young Lee (Suwon-si, KR); Woong Lim (Daejeon, KR); Hui-Yong Kim (Daejeon, KR)
Assignees: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE; INDUSTRY-ACADEMIA COOPERATION GROUP OF SEJONG UNIVERSITY
H04N19/119H04N19/117H04N19/172H04N19/86H04N19/96
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,231,634
App. No.
17/621,151
Granted
Feb 18, 2025
Kind
B2
Abstract

Disclosed herein are a method and apparatus for image encoding and image decoding using region segmentation. A picture is divided into multiple sub-pictures, and encoding and/or decoding are performed on the multiple sub-pictures resulting from division. Encoding operations and/or decoding operations on the sub-pictures may be independently performed. Merging and filtering are applied to reconstructed sub-pictures generated by decoding, and a reconstructed picture is generated after processing of merging and filtering.

Claims (58)

1. A video decoding method, comprising:

determining a target sub-picture of a target picture; and

performing decoding on the target picture comprising the target sub-picture, wherein

a position of the target sub-picture in the target picture is configured using a start location, a width and a height of the target sub-picture,

the start location, the width and the height are determined using partitioning information,

the partitioning information is comprised in a Sequence Parameter Set (SPS) from a bitstream,

a plurality of boundaries of the target sub-picture are boundaries between the target sub-picture and a plurality of other sub-pictures adjacent to the target sub-picture,

a first flag in the SPS indicates whether filtering for all sub-picture boundaries in the target picture is disabled or not,

a plurality of second flags for the plurality of boundaries in the SPS are used to determine whether to apply filtering for the plurality of boundaries, respectively, and

filtering for a target sub-picture boundary of the plurality of sub-picture boundaries is performed in a case that the first flag does not indicate that filtering for all sub-picture boundaries in the target picture is disabled and a second flag of the plurality of flags corresponding to the target sub-picture boundary indicates that filtering is applied to the target sub-picture boundary.

2. The video decoding method of claim 1 , wherein:

a size of the target sub-picture is determined in units of a Coding Tree Unit (CTU).

3. The video decoding method of claim 1 , wherein:

a size of the target sub-picture is defined according to a number of rows of CTUs.

4. The video decoding method of claim 1 , wherein:

the filtering is applied to a reconstructed sub-picture.

5. The video decoding method of claim 1 , wherein:

a filter of the filtering is a deblocking filter.

6. The video decoding method of claim 1 , wherein:

a filter of the filtering is a Sample Adaptive Offset (SAO) filter.

7. The video decoding method of claim 1 , wherein:

a filter of the filtering is an Adaptive Loop Filter (ALF).

8. The video decoding method of claim 1 , wherein:

the target sub-picture is decoded independently.

9. The video decoding method of claim 1 , wherein:

filtering information is used for the filtering, and

the filtering information is used to determine whether to apply the filtering on the target sub-picture.

10. The video decoding method of claim 9 , wherein:

the filtering information is comprised in the SPS, and

the filtering information is applied to a picture referring the SPS.

11. The video decoding method of claim 1 , wherein:

the target sub-picture comprises one or more slices.

12. A video encoding method, comprising:

determining a target sub-picture of a target picture; and

performing encoding on the target picture comprising the target sub-picture, wherein

a position of the target sub-picture in the target picture is defined using a start location, a width and a height of the target sub-picture,

partitioning information describing the start location, the width and the height is generated,

a Sequence Parameter Set (SPS) comprising the partitioning information is generated

a bitstream comprising the SPS is generated, and

a plurality of boundaries of the target sub-picture are boundaries between the target sub-picture and a plurality of other sub-pictures adjacent to the target sub-picture,

the SPS comprises a first flag and a plurality of second flags for the plurality of boundaries,

the first flag indicates whether filtering for all sub-picture boundaries in the target picture is disabled or not,

the plurality of second flags for the plurality of boundaries are used to determine whether to apply filtering for the plurality of sub-picture boundaries, respectively, and

filtering for a target sub-picture boundary of the plurality of sub-picture boundaries is performed in a case that the first flag does not indicate that filtering for all sub-picture boundaries in the target picture is disabled and a second flag of the plurality of flags corresponding to the target sub-picture boundary indicates that filtering is applied to the target sub-picture boundary.

13. A non-transitory computer-readable medium storing the bitstream generated by the video encoding method of claim 12 .

14. A non-transitory computer-readable medium storing a bitstream, the bitstream comprising:

a Sequence Parameter Set (SPS), wherein

a target sub-picture of a target picture is determined

decoding on the target picture comprising the target sub-picture is performed,

a position of the target sub-picture in the target picture is configured using a start location, a width and a height of the target sub-picture,

the start location, the width and the height are determined using partitioning information,

the partitioning information is comprised in the SPS,

a plurality of boundaries of the target sub-picture are boundaries between the target sub-picture and a plurality of other sub-pictures adjacent to the target sub-picture,

a first flag in the SPS indicates whether filtering for all sub-picture boundaries in the target picture is disabled or not,

a plurality of second flags for the plurality of boundaries in the SPS are used to determine whether to apply filtering for the plurality of sub-picture boundaries, respectively, and

filtering for a target sub-picture boundary of the plurality of sub-picture boundaries is performed in a case that the first flag does not indicate that filtering for all sub-picture boundaries in the target picture is disabled and a second flag of the plurality of flags corresponding to the target sub-picture boundary indicates that filtering is applied to the target sub-picture boundary.

15. The non-transitory computer-readable medium of claim 14 , wherein:

a size of the target sub-picture is determined in units of a Coding Tree Unit (CTU).

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2021
From: BANG, GUN; LEE, JIN-YOUNG; LIM, WOONG; KIM, HUI-YONG
To: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE; INDUSTRY-ACADEMIA COOPERATION GROUP OF SEJONG UNIVERSITY
Reel/Frame 058436/0401 →
Priority Claims (2)
KR 10-2019-0073694 · Jun 20, 2019 · national
KR 10-2020-0075841 · Jun 22, 2020 · national
Continuity (1)
Related Publication 20220312009A1 · Sep 29, 2022
References Cited (20)
US 9584819B2 · Wang et al. · 2017 [cited by applicant]
US 10225565B2 · Lee et al. · 2019 [cited by applicant]
US 20130094772A1 · Deshpande · 2013 [cited by examiner]
US 20130094773A1 · Misra · 2013 [cited by examiner]
US 20130107953A1 · Chen · 2013 [cited by examiner]
US 20130202051A1 · Zhou · 2013 [cited by examiner]
US 20160127728A1 · Tanizawa et al. · 2016 [cited by applicant]
US 20160269735A1 · Kim et al. · 2016 [cited by applicant]
US 20190082178A1 · Kim et al. · 2019 [cited by applicant]
US 20190141352A1 · Kim et al. · 2019 [cited by applicant]
US 20190174141A1 · Zhou · 2019 [cited by applicant]
US 20210195186A1 · Wu · 2021 [cited by examiner]
US 20210409785A1 · Wang · 2021 [cited by examiner]
JP 2016092837A · 2016 [cited by applicant]
JP 2016195416A · 2016 [cited by applicant]
KR 101644539B1 · 2016 [cited by applicant]
KR 101673021B1 · 2016 [cited by applicant]
KR 1020190050714A · 2019 [cited by applicant]
Boyce et al. (Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 14th Meeting: Geneva, CH, 19-27 Document: JVET-N0275, Marc (Year: 2019). [cited by examiner]
Chirag Pujara et al., “[AHG12/AHG17] Signaling of virtual boundaries,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Mar. 2019, pp. 1-20, Geneva. [cited by applicant]