IP Library Granted Patent US 12,244,930
Granted Patent B2
US 12,244,930 · App. 18/672,265 · Granted Mar 4, 2025

Method and apparatus for reconstructing 360-degree image according to projection format

Inventor: Ki Baek Kim (Daejeon, KR)
Assignee: B1 INSTITUTE OF IMAGE TECHNOLOGY, INC.
H04N23/698G06T3/40H04N19/103H04N19/11H04N19/119H04N19/124H04N19/129H04N19/13H04N19/134H04N19/159H04N19/174H04N19/176H04N19/44H04N19/51H04N19/625H04N19/45
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,244,930
App. No.
18/672,265
Granted
Mar 4, 2025
Kind
B2
Abstract

Disclosed are methods and apparatuses for image data encoding/decoding. A method for decoding a 360-degree image includes the steps of: receiving a bitstream obtained by encoding a 360-degree image; generating a prediction image by making reference to syntax information obtained from the received bitstream; adding the generated prediction image to a residual image obtained by dequantizing and inverse-transforming the bitstream, so as to obtain a decoded image; and reconstructing the decoded image into a 360-degree image according to a projection format. Therefore, the performance of image data compression can be improved.

Claims (36)

1. A method for decoding a 360-degree image, the method comprising:

receiving a bitstream in which the 360-degree image is encoded, the bitstream including data of an extended 2-dimensional image, the extended 2-dimensional image including a 2-dimensional image and a predetermined extension region, and the 2-dimensional image being from an image with a 3-dimensional projection structure and including one or more faces;

generating a prediction image by referring to syntax information obtained from the received bitstream;

obtaining a decoded image by adding the generated prediction image to a residual image, the residual image being obtained by inverse-quantizing and inverse transforming quantized transform coefficients from the bitstream; and

reconstructing the decoded image into the 360-degree image according to a projection format,

wherein the projection format is selectively determined based on identification information, among a plurality of pre-defined projection formats including an ERP format in which the 360-degree image is projected in a two-dimensional plane or a CMP format in which the 360-degree image is projected in a cube,

wherein a size of the extension region is variably determined based on at least one of first information indicating a width on a left side of the extension region or second information indicating a width on a right side of the extension region, independently of a size of the 2-dimensional image,

wherein the extension region is not included in the image with the 3-dimensional projection structure, and

wherein at least one of the identification information, the first information or the second information is obtained from the bitstream.

2. The method of claim 1 , wherein reconstructing the decoded image comprises:

obtaining, from the bitstream, arrangement information according to a regional packing, the arrangement information specifying one of a plurality of conversion types pre-defined in a decoding apparatus; and

rearranging at least one of a plurality of faces in the decoded image according to the arrangement information.

3. The method of claim 2 , wherein the pre-defined conversion types include at least one of 90-degree rotation, 180-degree rotation, 270-degree rotation, horizontal flipping, or vertical flipping.

4. The method of claim 1 , wherein the extension region is adjacent to a boundary of at least one of the faces.

5. The method of claim 4 , wherein the boundary of the face is discontinuous in the 2-dimensional image and continuous in the image with the 3-dimensional projection structure.

6. A method for encoding a 360-degree image, the method comprising:

generating a bitstream in which the 360-degree image is encoded, the bitstream including data of an extended 2-dimensional image, the extended 2-dimensional image including a 2-dimensional image and a predetermined extension region, and the 2-dimensional image being from an image with a 3-dimensional projection structure and including one or more faces;

wherein the bitstream comprises syntax information, the syntax information being referred to for generating a prediction image,

the bitstream comprises quantized transform coefficients, the quantized transform coefficients being inverse-quantized and inverse transformed for obtaining a residual image,

the prediction image is added to the residual image for obtaining a decoded image, and

the decoded image is reconstructed into the 360-degree image according to a projection format,

wherein the projection format is selectively determined based on identification information, among a plurality of pre-defined projection formats including an ERP format in which the 360-degree image is projected in a two-dimensional plane or a CMP format in which the 360-degree image is projected in a cube,

wherein a size of the extension region is variably determined based on at least one of first information indicating a width on a left side of the extension region or second information indicating a width on a right side of the extension region, independently of a size of the 2-dimensional image,

wherein the extension region is not included in the image with the 3-dimensional projection structure, and

wherein at least one of the identification information, the first information or the second information is encoded into the bitstream.

7. A method of transmitting a bitstream, the method comprising:

transmitting the bitstream to an image decoding apparatus;

wherein the bitstream is generated by encoding a 360-degree image, the bitstream including data of an extended 2-dimensional image, the extended 2-dimensional image including a 2-dimensional image and a predetermined extension region, and the 2-dimensional image being from an image with a 3-dimensional projection structure and including one or more faces;

wherein the bitstream comprises syntax information, the syntax information being referred to for generating a prediction image,

the bitstream comprises quantized transform coefficients, the quantized transform coefficients being inverse-quantized and inverse transformed for obtaining a residual image,

the prediction image is added to the residual image for obtaining a decoded image, and

the decoded image is reconstructed into the 360-degree image according to a projection format,

wherein the projection format is selectively determined based on identification information, among a plurality of pre-defined projection formats including an ERP format in which the 360-degree image is projected in a two-dimensional plane or a CMP format in which the 360-degree image is projected in a cube,

wherein a size of the extension region is variably determined based on at least one of first information indicating a width on a left side of the extension region or second information indicating a width on a right side of the extension region, independently of a size of the 2-dimensional image,

wherein the extension region is not included in the image with the 3-dimensional projection structure, and

wherein at least one of the identification information, the first information or the second information is encoded into the bitstream.

Priority Claims (3)
KR 10-2016-0127878 · Oct 4, 2016 · national
KR 10-2016-0129382 · Oct 6, 2016 · national
KR 10-2017-0090612 · Jul 17, 2017 · national
Continuity (5)
Continuation 18466442 · Sep 13, 2023
Continuation 17487277 · Sep 28, 2021
Continuation 16372237 · Apr 1, 2019
Continuation PCTKR2017011143 · Oct 10, 2017
Related Publication 20240314441A1 · Sep 19, 2024
References Cited (57)
US 6141034A · McCutchen · 2000 [cited by applicant]
US 6157396A · Margulis et al. · 2000 [cited by applicant]
US 7126630B1 · Lee et al. · 2006 [cited by applicant]
US 7292722B2 · Lelescu et al. · 2007 [cited by applicant]
US 7623682B2 · Park et al. · 2009 [cited by applicant]
US 8217988B2 · Park · 2012 [cited by applicant]
US 8264524B1 · Davey · 2012 [cited by applicant]
US 11165958B2 · Kim · 2021 [cited by examiner]
US 20040105597A1 · Lelescu et al. · 2004 [cited by applicant]
US 20060034367A1 · Park · 2006 [cited by applicant]
US 20060034374A1 · Park et al. · 2006 [cited by applicant]
US 20060034529A1 · Park et al. · 2006 [cited by applicant]
US 20060210145A1 · Lee et al. · 2006 [cited by applicant]
US 20060251336A1 · Lelescu et al. · 2006 [cited by applicant]
US 20070133954A1 · Maeda et al. · 2007 [cited by applicant]
US 20110274166A1 · Jeon et al. · 2011 [cited by applicant]
US 20120057788A1 · Fukuhara et al. · 2012 [cited by applicant]
US 20120176367A1 · Genova et al. · 2012 [cited by applicant]
US 20130022113A1 · Chen et al. · 2013 [cited by applicant]
US 20130194255A1 · Lee et al. · 2013 [cited by applicant]
US 20130308709A1 · Norkin et al. · 2013 [cited by applicant]
US 20140146873A1 · Lee et al. · 2014 [cited by applicant]
US 20140184727A1 · Xiao et al. · 2014 [cited by applicant]
US 20140301463A1 · Rusanovskyy et al. · 2014 [cited by applicant]
US 20140307774A1 · Minoo et al. · 2014 [cited by applicant]
US 20140341549A1 · Hattori · 2014 [cited by applicant]
US 20150172544A1 · Deng et al. · 2015 [cited by applicant]
US 20150201204A1 · Chen et al. · 2015 [cited by applicant]
US 20160142697A1 · Budagavi et al. · 2016 [cited by applicant]
US 20160142740A1 · Sharman · 2016 [cited by applicant]
US 20160241837A1 · Cole et al. · 2016 [cited by applicant]
US 20160249056A1 · Tsukuba et al. · 2016 [cited by applicant]
US 20160328824A1 · Kim et al. · 2016 [cited by applicant]
US 20170336705A1 · Zhou et al. · 2017 [cited by applicant]
US 20170339391A1 · Zhou et al. · 2017 [cited by applicant]
US 20170347026A1 · Hannuksela · 2017 [cited by applicant]
US 20170366796A1 · Goldentouch · 2017 [cited by applicant]
US 20180025467A1 · Macmillan et al. · 2018 [cited by applicant]
US 20180027226A1 · Abbas et al. · 2018 [cited by applicant]
US 20180040164A1 · Newman et al. · 2018 [cited by applicant]
US 20180084257A1 · Abbas · 2018 [cited by applicant]
US 20180091735A1 · Wang et al. · 2018 [cited by applicant]
US 20190082161A1 · Cole et al. · 2019 [cited by applicant]
US 20190158849A1 · Yu et al. · 2019 [cited by applicant]
US 20190222862A1 · Shin et al. · 2019 [cited by applicant]
US 20200036997A1 · Li et al. · 2020 [cited by applicant]
US 20200322585A1 · Tsunashima · 2020 [cited by applicant]
US 20240048765A1 · Kim · 2024 [cited by applicant]
US 20240334062A1 · Kim · 2024 [cited by examiner]
US 20240340539A1 · Kim · 2024 [cited by examiner]
US 20240406565A1 · Kim · 2024 [cited by examiner]
ITU-T Recommendation H.265: High efficiency video coding (Apr. 2015) (Year: 2015). [cited by applicant]
Ohm, Jens-Rainer, et al. “Comparison of the coding efficiency of video coding standards—including high efficiency video coding (HEVC).” [cited by applicant]
Sze, Vivienne, et al. “High efficiency video coding (HEVC).” [cited by applicant]
Sjoberg, Rickard, et al. “Overview of HEVC high-level syntax and reference picture management.” [cited by applicant]
Bordes, Philippe, et al. “An overview of the emerging HEVC standard.” [cited by applicant]
Series, H. Series H: Audiovisual and Multimedia Systems, [cited by applicant]