IP Library › Granted Patent US 12,204,815
Granted Patent B2
US 12,204,815 · App. 17/828,755 · Granted Jan 21, 2025

Adaptive audio delivery and rendering

Inventors: Shan Liu (San Jose, CA); Jun Tian (Belle Mead, NJ); Xiaozhong Xu (State College, PA)
Assignee: Tencent America LLC
G06F3/165G10L15/22H04R3/04
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,204,815
App. No.
17/828,755
Granted
Jan 21, 2025
Kind
B2
Abstract

Aspects of the disclosure provide methods and apparatuses (e.g., client devices and server devices) for audio processing. In some examples, a client device includes processing circuitry. The processing circuitry transmits, to a server device, a selection signal indicative of an audio encoding configuration for encoding audio content in an audio input. The processing circuitry receives, from the server device, an encoded bitstream in response to the transmitting of the selection signal. The encoded bitstream includes the audio content that is encoded according to the audio encoding configuration. The processing circuitry renders audio signals based on the encoded bitstream.

Claims (44)

1. A method of audio processing at a client device, comprising:

transmitting, to a server device, a selection signal indicative of an audio encoding configuration for encoding audio content in an audio input, wherein the selection signal is indicative of at least one categorization layer of a plurality of categorization layers, each of the plurality of categorization layers including a different subset of the audio content in the audio input and being assigned a respective layer identifier (ID);

receiving, from the server device, an encoded bitstream including the audio content that is encoded according to the audio encoding configuration in response to the transmitting of the selection signal; and

rendering audio signals based on the encoded bitstream.

2. The method of claim 1 , wherein the transmitting the selection signal further comprises:

transmitting the selection signal indicative of a bitrate for encoding the audio content.

3. The method of claim 2 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream including one or more audio channels that are encoded according to the bitrate.

4. The method of claim 2 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream including one or more audio objects that are encoded according to the bitrate.

5. The method of claim 2 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream including audio higher order ambisonics (HOA) signals that are encoded according to the bitrate.

6. The method of claim 1 , wherein

each of the plurality of categorization layers includes a different combination of audio channels, audio objects, or audio higher order ambisonics (HOA) signals.

7. The method of claim 6 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream that is encoded based on a subset of audio channels in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio channels.

8. The method of claim 6 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream that is encoded based on a subset of audio objects in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio objects.

9. The method of claim 6 , wherein the receiving the encoded bitstream further comprises:

receiving the encoded bitstream that is encoded based on a reduced order set of higher order ambisonics (HOA) signals in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio objects.

10. The method of claim 1 , wherein the transmitting the selection signal further comprises:

transmitting the layer ID associated with the audio encoding configuration.

11. The method of claim 1 , further comprising:

determining the selection signal according to at least one of a media processing capability of the client device, a network connection of the client device, and a preference input.

12. An apparatus for audio processing, comprising processing circuitry configured to:

transmit, to a server device, a selection signal indicative of an audio encoding configuration for encoding audio content in an audio input, wherein the selection signal is indicative of at least one categorization layer of a plurality of categorization layers, each of the plurality of categorization layers including a different subset of the audio content in the audio input and being assigned a respective layer identifier (ID);

receive, from the server device, an encoded bitstream including the audio content that is encoded according to the audio encoding configuration in response to the transmitting of the selection signal; and

render audio signals based on the encoded bitstream.

13. The apparatus of claim 12 , wherein the processing circuitry configured to:

transmit the selection signal indicative of a bitrate for encoding the audio content.

14. The apparatus of claim 13 , wherein the processing circuitry configured to:

receive the encoded bitstream including one or more audio channels that are encoded according to the bitrate.

15. The apparatus of claim 13 , wherein the processing circuitry configured to:

receive the encoded bitstream including one or more audio objects that are encoded according to the bitrate.

16. The apparatus of claim 13 , wherein the processing circuitry configured to:

receive the encoded bitstream including audio higher order ambisonics (HOA) signals that are encoded according to the bitrate.

17. The apparatus of claim 12 , wherein

each of the plurality of categorization layers includes a different combination of audio channels, audio objects, or audio higher order ambisonics (HOA) signals.

18. The apparatus of claim 17 , wherein the processing circuitry configured to:

receive the encoded bitstream that is encoded based on a subset of audio channels in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio channels.

19. The apparatus of claim 17 , wherein the processing circuitry configured to:

receive the encoded bitstream that is encoded based on a subset of audio objects in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio channels.

20. The apparatus of claim 17 , wherein the processing circuitry configured to:

receive the encoded bitstream that is encoded based on a reduced order set of higher order ambisonics (HOA) signals in the audio content of the audio input, each of the plurality of categorization layers including a different subset of the audio channels.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 31, 2022
From: LIU, SHAN; TIAN, JUN; XU, XIAOZHONG
To: TENCENT AMERICA LLC
Reel/Frame 060058/0557 →
Continuity (2)
Provisional Application 63196066 · Jun 2, 2021
Related Publication 20220391167A1 · Dec 8, 2022
References Cited (8)
US 20170063960A1 · Stockhammer et al. · 2017 [cited by applicant]
US 20220263883A1 · Lee · 2022 [cited by examiner]
US 20220383881A1 · Shahbazi Mirzahasanloo · 2022 [cited by examiner]
JP 2018532146A · 2018 [cited by applicant]
KR 20200078537A · 2020 [cited by applicant]
WO 2021015484A1 · 2021 [cited by applicant]
International Search Report and Written Opinion in PCT/US2022/072731, mailed Aug. 30, 2022, 7 pages. [cited by applicant]
Japanese Office Action issued Dec. 5, 2023 in Application No. 2022-566186. (10 pages). [cited by applicant]