IP Library Granted Patent US 12,512,097
Granted Patent B2
US 12,512,097 · App. 17/623,379 · Granted Dec 30, 2025

Methods to employ compaction in ASR service usage to reduce transcription charges

Inventors: Ankur Anil Aher (Maharashtra, IN); Jeffry Copps Robert Jose (Tamil Nadu, IN)
Assignee: Adeia Guides Inc.
G10L15/22G10L15/26G10L15/30G10L15/063
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,512,097
App. No.
17/623,379
Granted
Dec 30, 2025
Kind
B2
Abstract

Systems and methods for processing audio streams are disclosed herein. An audio stream including speech content is received. The audio stream is compacted to generate a compacted audio stream and the compacted audio stream is transmitted to an automatic speech recognition (ASR) service for transcription of the speech content to text content. In response to transmitting the compacted audio stream for transcription, text content, a transcription of the audio stream, is received from the ASR service, and transcription charges are determined.

Claims (28)

1 . A method of processing audio streams comprising:

receiving a first audio stream and a second audio stream including speech content;

compacting the first audio stream and the second audio stream to generate a compacted audio stream, wherein compacting the first audio stream and the second audio stream comprises removing information from a header of the first audio stream and removing excess speech content from the second audio stream;

transmitting to an automated speech recognition (ASR) service the compacted audio stream for transcription of the speech content to text content, wherein transcription charges for transcription of the compacted audio stream are transaction-based; and

in response to transmitting the compacted audio stream for transcription, receiving text content that is a transcription of the first audio stream and the second audio stream.

2 . The method of claim 1 , wherein compacting the first audio stream and the second audio stream includes removing non meaningful voice and/or silence from the speech content.

3 . The method of claim 2 , wherein the compacted audio stream is a first compacted audio stream of a plurality of compacted audio streams, and wherein compacting the first audio stream and the second audio stream increases a frequency of transmission of the plurality of compacted audio streams.

4 . The method of claim 1 , further comprising:

in response to the receiving, determining whether the first audio stream or the second audio stream requires time-based speech-to-text services; and

in response to determining the second audio stream does not require time-based speech-to-text services, executing steps for effectuating transaction-based speech-to-text services.

5 . The method of claim 1 , wherein compacting the first audio stream and the second audio stream further comprises trimming the first audio stream to remove excess speech content from the speech content of the first audio stream.

6 . The method of claim 1 , wherein the first audio stream and the second audio stream are stored in a storage and the compacted audio stream is based on contents of the storage.

7 . The method of claim 1 , wherein removing the information from the header of the first audio stream comprises removing timing information from the header of the first audio stream.

8 . A system for processing audio streams comprising:

input/output (I/O) circuitry configured to:

receive a first audio stream and a second audio stream including speech content;

control circuitry configured to:

compact the first audio stream and the second audio stream to generate a compacted audio stream, wherein the control circuitry is configured to compact the first audio stream and the second audio stream by removing information from a header of the first audio stream and removing excess speech content from the second audio stream;

transmit to an automated speech recognition (ASR) service the compacted audio stream for transcription of the speech content to text content, wherein transcription charges for transcription of the compacted audio stream are transaction-based; and

in response to transmitting the compacted audio stream for transcription, receive text content that is a transcription of the first audio stream and the second audio stream.

9 . The system of claim 8 , wherein the control circuitry is configured to compact the first audio stream and the second audio stream by removing non meaningful voice and/or silence from the speech content.

10 . The system of claim 9 , wherein the compacted audio stream is a first compacted audio stream of a plurality of compacted audio streams, and wherein compacting the first audio stream and the second audio stream increases a frequency of transmission of the plurality of compacted audio streams.

11 . The system of claim 8 , wherein the control circuitry is further configured to:

in response to the receiving, determining whether the first audio stream or the second audio stream requires time-based speech-to-text services; and

in response to determining the second audio stream does not require time-based speech-to-text services, executing steps for effectuating transaction-based speech-to-text services.

12 . The system of claim 8 , wherein the control circuitry is configured to compact the first audio stream and the second audio stream by trimming the first audio stream to remove excess speech content from the speech content of the first audio stream.

13 . The system of claim 8 , wherein the first audio stream and the second audio stream are stored in a storage and the compacted audio stream is based on contents of the storage.

14 . The system of claim 8 , wherein the control circuitry is configured to remove the information from the header of the first audio stream by removing timing information from the header of the first audio stream.

Assignments (3)
CHANGE OF NAME Recorded Oct 3, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069106/0238 →
SECURITY INTEREST Recorded May 3, 2023
From: ADEIA GUIDES INC.; ADEIA IMAGING LLC; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA SOLUTIONS LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063529/0272 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 28, 2021
From: AHER, ANKUR ANIL; ROBERT JOSE, JEFFRY COPPS
To: ROVI GUIDES, INC.
Reel/Frame 058492/0591 →
Continuity (1)
Related Publication 20220375465A1 · Nov 24, 2022
References Cited (17)
US 6157911A · Kuroda · 2000 [cited by examiner]
US 6691089B1 · Su et al. · 2004 [cited by applicant]
US 10133538B2 · Mclaren et al. · 2018 [cited by applicant]
US 10176809B1 · Piérard · 2019 [cited by applicant]
US 11017778B1 · Thomson et al. · 2021 [cited by applicant]
US 11545134B1 · Federico et al. · 2023 [cited by applicant]
US 11790916B2 · Agarwal · 2023 [cited by examiner]
US 20120072211A1 · Edgington et al. · 2012 [cited by applicant]
US 20130166279A1 · Dines · 2013 [cited by examiner]
US 20170053643A1 · Ben-David · 2017 [cited by applicant]
US 20190341052A1 · Allibhai · 2019 [cited by applicant]
US 20210183374A1 · Thomson · 2021 [cited by examiner]
JP 2007074608A · 2005 [cited by examiner]
WO WO2016169329A1 · 2016 [cited by examiner]
Hain et al., “Segment Generation and Clustering in the HTK Broadcast News Transcription System,” Proc. Darpa BNT US, 133-137 (1998) (5 pages). [cited by applicant]
PCT International Search Report and Written Opinion for International Application No. PCT/IB2019/001349, dated Sep. 16, 2020 (16 Pages). [cited by applicant]
PCT International Search Report and Written Opinion for International Application No. PCT/IB2019/001348, dated Sep. 8, 2020 (14 Pages). [cited by applicant]