IP Library › Granted Patent US 12,633,300
Granted Patent B2
US 12,633,300 · App. 18/080,537 · Granted May 19, 2026

Processing of audio data using a plurality of distributed computer devices

Inventor: Jorge Cenzano Ferret (Seattle, WA)
Assignee: Meta Platforms, Inc.
G10L19/173H04L25/49G10L19/167
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,633,300
App. No.
18/080,537
Granted
May 19, 2026
Kind
B2
Abstract

According to examples, a system for using to processing of audio data using a plurality of distributed computer manner is described. The system may include a processor and a memory storing instructions. The processor may cause the system to receive audio data associated with a content item in an initial format, process the audio data to generate one or more audio segments for distributed processing, and decode the one or more audio segments from the audio data in the initial format to generate decoded audio data in a decoding format. The processor may then encode the decoded audio data in a decoding format to encoded audio data in an encoding format and trim a segment of the encoded audio data in the encoded format to generate a trimmed segment of audio data that may be utilized to enable continuous playback by a receiving device.

Claims (41)

1 . A system for implementing audio transcoding in a distributed fashion, comprising:

a first computer system comprising one or more processors and memory storing instructions, which when executed by the one or more processors, cause the one or more processors to:

receive audio data associated with a content item in an initial format;

decode the audio data associated with the content item in the initial format to generate decoded audio data in a decoding format;

encode the decoded audio data in the decoding format to encoded audio data in an encoding format; and

trim, by a first computer system, a segment of the encoded audio data in the encoding format to generate a trimmed segment of audio data; and

a second computer system discrete relative to the first computer system, the second computer system configured to:

determine a preceding audio segment while the first computer system trims the segment;

receive the trimmed segment from the first computer system;

concatenate the trimmed segment of audio data with a preceding audio segment to generate a concatenated trimmed segment of audio data and preceding audio segment; and

transmit the concatenated trimmed segment of audio data and preceding audio segment to a user device for playback.

2 . The system of claim 1 , wherein the initial format is a publication format utilized by a remote device, and the publication format is Advanced Audio Coding (AAC).

3 . The system of claim 1 , wherein to receive the audio data associated with the content item in the initial format, the instructions when executed by the one or more processors further cause the one or more processors to determine a length of an audio segment for processing.

4 . The system of claim 1 , wherein the decoding format is pulse-code modulation (PCM).

5 . The system of claim 1 , wherein the decoded audio data in the decoding format comprises an audio segment for playback, a preceding data element, and a following data element.

6 . The system of claim 5 , wherein the audio segment for playback comprises five (5) frames of audio data.

7 . A method for processing of audio data using a plurality of distributed computer devices, comprising:

receiving audio data associated with a content item in an initial format;

decoding the audio data associated with the content item in the initial format to generate decoded audio data in a decoding format;

encoding the decoded audio data in the decoding format to encoded audio data in an encoding format;

trimming, by a first computer system, a segment of the encoded audio data in the encoding format to generate a trimmed segment of audio data;

determining, by a second computer system discrete relative to the first computer system, a preceding audio segment while the first computer system trims the segment;

concatenating, by the second computer system, the trimmed segment of audio data with a preceding audio segment to generate a concatenated trimmed segment of audio data and preceding audio segment; and

transmitting the concatenated trimmed segment of audio data and preceding audio segment to a user device for playback.

8 . The method of claim 7 , wherein receiving the audio data associated with the content item in the initial format comprises determining an audio segment for playback.

9 . The method of claim 7 , wherein the decoded audio data in the decoding format comprises an audio segment for playback, a preceding data element, and a following data element.

10 . The method of claim 9 , wherein the audio segment for playback comprises five (5) frames of audio data.

11 . The method of claim 9 , wherein the preceding data element comprises two (2) frames of audio data.

12 . The method of claim 9 , wherein the following data element comprises one (1) frame of audio data.

13 . A non-transitory computer-readable storage medium for implementing audio transcoding in a distributed fashion and having an executable stored thereon, wherein the executable when executed instructs a processor to:

receive audio data associated with a content item in an initial format;

decode the audio data associated with the content item in the initial format to generate decoded audio data in a decoding format;

encode the decoded audio data in the decoding format to encoded audio data in an encoding format;

trim, by a first computer system, a segment of the encoded audio data in the encoding format to generate a trimmed segment of audio data;

determine, by a second computer system discrete relative to the first computer system, a preceding audio segment while the first computer system trims the segment;

concatenate, by the second computer system, the trimmed segment of audio data with a preceding audio segment to generate a concatenated trimmed segment of audio data and preceding audio segment; and

transmit the concatenated trimmed segment of audio data and preceding audio segment to a user device for playback.

14 . The non-transitory computer readable storage medium of claim 13 , wherein the initial format is a publication format utilized by a remote device, and the publication format is Advanced Audio Coding (AAC).

15 . The non-transitory computer-readable storage medium of claim 13 , wherein the decoding format is pulse-code modulation (PCM).

16 . The non-transitory computer-readable storage medium of claim 13 , wherein the encoded audio data in the encoding format comprises an audio segment for playback, a preceding data element, and a following data element.

17 . The non-transitory computer-readable storage medium of claim 13 , wherein the encoded audio data in the encoding format further comprises a priming data element.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 19, 2023
From: CENZANO FERRET, JORGE
To: META PLATFORMS, INC.
Reel/Frame 062427/0115 →
Continuity (2)
Provisional Application 63322905 · Mar 23, 2022
Related Publication 20230306977A1 · Sep 28, 2023
References Cited (10)
US 8494866B2 · Stewart et al. · 2013 [cited by applicant]
US 9324332B2 · Doehla et al. · 2016 [cited by applicant]
US 10218826B2 · Eswaran et al. · 2019 [cited by applicant]
US 20120046956A1 · Stewart · 2012 [cited by examiner]
US 20180069950A1 · Eswaran · 2018 [cited by examiner]
US 20210352342A1 · Thoma et al. · 2021 [cited by applicant]
EP 1299879B1 · 2006 [cited by examiner]
EP 1855271A1 · 2007 [cited by examiner]
Torii, et al. “Asymmetric Multi-Processing Mobile Application Processor MP211,” NEC J. of Adv. Tech., 2005. (Viewable but no permission for download—see citation at Advisory Action) (Year: 2005). [cited by examiner]
Rodrigues R., et al., “MPEG DASH-Some QoE-based Insights into the Tradeoff between Audio and Video for Live Music Concert Streaming Under Congested Network Conditions,” Eighth International Conference on Quality of Mult… [cited by applicant]