IP Library Granted Patent US 10,446,138
Granted Patent B2
US 10,446,138 · App. 15/791,953 · Granted Oct 15, 2019

System and method for assessing audio files for transcription services

Inventors: Eric Shellef (Givaatayim, IL); Kobi Ben Tzvi (Ramat Hasharon, IL); Tom Livne (Ramat Gan, IL)
Assignee: Verbit Software Ltd.
G10L15/1815G06N3/08G06Q30/0611G10L15/005G10L15/16G10L21/0272G10L25/51G10L25/60G10L25/84G06N5/003G06N7/005G06N20/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,446,138
App. No.
15/791,953
Granted
Oct 15, 2019
Kind
B2
Abstract

A system and method for assessing transcription costs based on an audio file are provided. The method includes method for assessing an audio file for transcription includes accessing at least one audio file for transcription assessment; analyzing the at least one audio file to determine at least one transcription characteristic based on the at least one audio file; and calculating, based on the at least one determined transcription characteristic, an initial bid value for transcription of the audio file.

Claims (24)

1. A method for assessing an audio file for transcription, comprising:

generating signatures of voices determined to be unique within the audio file;

determining number of speakers captured within the audio file based on the signatures;

determining accent of each speaker from the speakers utilizing at least one of support vector machines using Gaussian mixture models (GMM-SVM), Gaussian mixture model with universal background model (GMM-UBM), and iVectors for speech processing applications;

identifying a main speaker and another speaker within the audio file;

identifying characteristics of noise in a recording of the main speaker, and characteristics of noise in a recording of the other speaker;

calculating a bid value for transcription of the audio file based on the number of speakers, the accent of each speaker, the characteristics of noise in the recording of the main speaker, and the characteristics of noise in the recording of the other speaker; and

identifying transcription candidates by a server to provide services based on matching of specific skillsets with the accent of each speaker.

2. The method of claim 1 , wherein the bid value depends on time of introducing the transcription of the audio file to a marketplace.

3. The method of claim 1 , wherein the bid value depends on current demand for transcriptions when introducing the transcription of the audio file to a marketplace.

4. The method of claim 1 , wherein the bid value depends on current availability of transcription services when introducing the transcription of the audio file to a marketplace.

5. A system for assessing an audio file for transcription, comprising:

a processing circuitry; and

a memory, the memory containing instructions that, when executed by the processing circuitry, configure the system to:

generate signatures of voices determined to be unique within the audio file;

determine number of speakers captured within the audio file based on the signatures;

determine accent of each speaker from the speakers utilizing at least one of support vector machines using Gaussian mixture models (GMM-SVM), Gaussian mixture model with universal background model (GMM-UBM), and iVectors for speech processing applications;

identify a main speaker and another speaker within the audio file;

identify characteristics of noise in a recording of the main speaker, and characteristics of noise in a recording of the other speaker;

calculate a bid value for transcription of the audio file based on the number of speakers, the accent of each speaker, the characteristics of noise in the recording of the main speaker, and the characteristics of noise in the recording of the other speaker; and

identify transcription candidates by a server to provide services based on matching of specific skillsets with the accent of each speaker.

6. The system of claim 5 , wherein the bid value depends on time of introducing the transcription of the audio file to a marketplace.

7. The system of claim 5 , wherein the bid value depends on current demand for transcriptions when introducing the transcription of the audio file to a marketplace.

8. The system of claim 5 , wherein the bid value depends on current availability of transcription services when introducing the transcription of the audio file to a marketplace.

Assignments (4)
AMENDED AND RESTATED INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Feb 3, 2021
From: VERBIT SOFTWARE LTD
To: SILICON VALLEY BANK
Reel/Frame 055209/0474 →
SECURITY INTEREST Recorded Jun 4, 2020
From: VERBIT SOFTWARE LTD.
To: SILICON VALLEY BANK
Reel/Frame 052836/0856 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded May 24, 2019
From: VERBIT SOFTWARE LTD
To: SILICON VALLEY BANK
Reel/Frame 049285/0338 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 24, 2017
From: SHELLEF, ERIC; BEN TZVI, KOBI; LIVNE, TOM
To: VERBIT SOFTWARE LTD.
Reel/Frame 043937/0044 →
Continuity (2)
Provisional Application 62509853 · May 23, 2017
Related Publication 20180342240A1 · Nov 29, 2018