IP Library › Granted Patent US 10,555,103
Granted Patent B2
US 10,555,103 · App. 15/860,934 · Granted Feb 4, 2020

Method for outputting audio signal using scene orientation information in an audio decoder, and apparatus for outputting audio signal using the same

Inventors: Tung Chin Lee (Seoul, KR); Sejin Oh (Seoul, KR)
Assignee: LG ELECTRONICS INC.
H04S7/30G06F3/011H04N21/439H04S2400/11H04S2420/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,555,103
App. No.
15/860,934
Granted
Feb 4, 2020
Kind
B2
Abstract

A method for outputting an audio signal corresponding to scene orientation information is disclosed. The method includes receiving an audio signal interacting with a video, generating a decoded audio signal, object metadata, and scene orientation information through decoding, receiving external control information, and generating modified object metadata suitable for a playback environment by modifying the object metadata based on the received external control information, rendering the decoded audio signal using the modified object metadata, and modifying the rendered audio signal according to the scene orientation information.

Claims (28)

1. A method for decoding an audio bitstream by a decoding apparatus, the method comprising:

obtaining extension element configuration information from the audio bitstream;

based on the extension element configuration information indicating that scene orientation information is present in the audio bitstream, obtaining a decoded audio signal, object metadata, and the scene orientation information from the audio bitstream;

receiving external control information, and generating modified object metadata by modifying the object metadata based on the received external control information;

rendering the decoded audio signal using the modified object metadata; and

modifying the rendered audio signal based on the scene orientation information,

wherein the scene orientation information is information indicating a camera capture direction during generation of a video scene associated with the decoded audio signal.

2. The method according to claim 1 , wherein the scene orientation information is used for execution of a random access function for the video scene associated with the decoded audio signal.

3. The method according to claim 1 , wherein the scene orientation information includes yaw information, pitch information, and roll information,

the yaw information indicating an angle of rotating the camera capture direction in a z-axis,

the pitch information indicating an angle of rotating the camera capture direction in an x-axis, and

the roll information indicating an angle of rotating the camera capture direction in a y-axis.

4. The method according to claim 1 , wherein the modified object metadata includes a relative position of an audio object and a gain in a space corresponding to a user location.

5. The method according to claim 1 , further comprising performing binaural rendering on the rendered audio signal, using a Binaural Room Impulse Response (BRIR) to output the rendered audio signal as a 2-channel surround audio signal.

6. An apparatus for decoding an audio bitstream, the apparatus comprising:

an audio decoder configured to obtain extension element configuration information from the audio bitstream, and

based on the extension element configuration information indicating that scene orientation information is present in the audio bitstream, obtain a decoded audio signal, object metadata, and the scene orientation information from the audio bitstream;

a metadata processor configured to receive external control information, and generate modified object metadata by modifying the object metadata based on the received external control information; and

a renderer configured to render the decoded audio signal using the modified object metadata,

wherein the renderer modifies the rendered audio signal based on the scene orientation information, and

wherein the scene orientation information is information indicating a camera capture direction during generation of a video scene associated with the decoded audio signal.

7. The apparatus according to claim 6 , wherein the scene orientation information is used for execution of a random access function for the video scene associated with the decoded audio signal.

8. The apparatus according to claim 6 , wherein the scene orientation information includes yaw information, pitch information, and roll information,

the yaw information indicating an angle of rotating the camera capture direction in a z-axis,

the pitch information indicating an angle of rotating the camera capture direction in an x-axis, and

the roll information indicating an angle of rotating the camera capture direction in a y-axis.

9. The apparatus according to claim 6 , wherein the modified object metadata includes a relative position of an audio object and a gain in a space corresponding to a user location.

10. The apparatus according to claim 6 , further comprising a binaural renderer configured to perform binaural rendering on the rendered audio signal, using a Binaural Room Impulse Response (BRIR) to output the rendered audio signal as a 2-channel surround audio signal.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 4, 2018
From: LEE, TUNG CHIN; OH, SEJIN
To: LG ELECTRONICS INC.
Reel/Frame 044536/0208 →
Continuity (2)
Provisional Application 62479323 · Mar 31, 2017
Related Publication 20180288553A1 · Oct 4, 2018