IP Library › Granted Patent US 12,380,877
Granted Patent B2
US 12,380,877 · App. 17/521,713 · Granted Aug 5, 2025

Training of speech recognition systems

Inventors: David Thomson (North Salt Lake, UT); Jadie Adams (Salt Lake City, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/063G10L15/26G10L2015/0631
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,380,877
App. No.
17/521,713
Granted
Aug 5, 2025
Kind
B2
Abstract

A method may include obtaining first audio data of a first communication session between a first and second device and during the first communication session, obtaining a first text string that is a transcription of the first audio data and training a model of an automatic speech recognition system using the first text string and the first audio data. The method may further include in response to completion of the training, deleting the first audio data and the first text string and after deleting the first audio data and the first text string, obtaining second audio data of a second communication session between a third and fourth device and during the second communication session obtaining a second text string that is a transcription of the second audio data and further training the model of the automatic speech recognition system using the second text string and the second audio data.

Claims (40)

1. A method comprising:

obtaining a model of an automatic speech recognition system;

obtaining first audio data of a first communication session between a first device of a first user and a second device of a second user;

training a first copy of the model based on the first audio data;

obtaining second audio data of a second communication session between a third device of a third user and a fourth device of a fourth user;

training a second copy of the model based on the second audio data;

determining a set of acoustic parameters using both the trained first copy of the model and the trained second copy of the model;

updating the model using the set of acoustic parameters;

after updating the model, obtaining third audio data of a third communication session between a fifth device of a fifth user and a sixth device of a sixth user, wherein the third user and the fourth user are both separate and distinct from the first user and the second user and the fifth and sixth users are both separate and distinct from the first, second, third, and fourth users; and

generating, during the third communication session, a transcription of the third audio data by applying the updated model.

2. The method of claim 1 , wherein the model includes an acoustic model, a language model, a confidence model, and/or classification model of the automatic speech recognition system.

3. The method of claim 1 , further comprising obtaining a connected graph that includes a plurality of word combinations, the plurality of word combinations derived from the first audio data using automatic speech recognition, wherein the first copy of the model is trained using the connected graph.

4. The method of claim 1 , further comprising obtaining a plurality of phonemes from the first audio data, wherein the first copy of the model is trained using the phonemes.

5. The method of claim 1 , wherein the training of the first copy of the model of the automatic speech recognition system based on the first audio data completes after the first communication session.

6. The method of claim 1 , further comprising in response to completion of the training of the first copy of the model, deleting the first audio data.

7. The method of claim 6 , wherein the first audio data is deleted during the first communication session.

8. The method of claim 1 , wherein the training of the second copy of the model occurs during the training of the first copy of the model.

9. At least one non-transitory computer-readable media configured to store one or more instructions that in response to being executed by at least one computing system cause performance of the method of claim 1 .

10. The method of claim 1 , further comprising determining a classification for the first audio data, the classification indicating an intent of a user when speaking words in the first audio data, wherein the training the model is based on the classification of the first audio data.

11. A system comprising:

one or more processors; and

one or more computer-readable media configured to store one or more instructions that in response to being executed by the one or more processors cause or direct performance of operations, the operations comprising:

obtaining a model of an automatic speech recognition system;

obtaining first audio data of a first communication session between a first device of a first user and a second device of a second user;

training a first copy of the model based on the first audio data;

obtaining second audio data of a second communication session between a third device of a third user and a fourth device of a fourth user;

training a second copy of the model based on the second audio data;

determining a set of acoustic parameters using both the trained first copy of the model and the trained second copy of the model;

updating the model using the set of acoustic parameters;

after updating the model, obtaining third audio data of a third communication session between a fifth device of a fifth user and a sixth device of a sixth user, wherein the third user and the fourth user are both separate and distinct from the first user and the second user and the fifth and sixth users are both separate and distinct from the first, second, third, and fourth users; and

generating, during the third communication session, a transcription of the third audio data by applying the updated model.

12. The system of claim 11 , wherein the model includes an acoustic model, a language model, a confidence model, and/or classification model of the automatic speech recognition system.

13. The system of claim 11 , wherein the operations further comprise obtaining a connected graph that includes a plurality of word combinations, the plurality of word combinations derived from the first audio data using automatic speech recognition, wherein the first copy of the model is trained using the connected graph.

14. The system of claim 11 , wherein the operations further comprise obtaining a plurality of phonemes from the first audio data, wherein the first copy of the model is trained using the phonemes.

15. The system of claim 11 , wherein the training of the first copy of the model of the automatic speech recognition system based on the first audio data completes after the first communication session.

16. The system of claim 11 , wherein the operations further comprise in response to completion of the training of the first copy of the model, deleting the first audio data.

17. The system of claim 16 , wherein the first audio data is deleted during the first communication session.

18. The system of claim 11 , wherein the training of the second copy of the model occurs during the training of the first copy of the model.

19. The system of claim 11 , wherein the training the model of the automatic speech recognition system based on the first audio data is performed during the first communication session.

20. The system of claim 11 , wherein the operations further comprise: determining a classification for the first audio data, the classification indicating an intent of a user when speaking words in the first audio data, wherein the training the model is based on the classification of the first audio data.

Assignments (3)
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 19, 2021
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 058168/0538 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 19, 2021
From: HOLM, MICHAEL; BLACK, DAVID; BAROCIO, JESSE; THOMSON, DAVID; BOEKWEG, SCOTT; ROYLANCE, SHANE; CLEMENTS, KIERSTEN; BOEHME, KENNETH; ADAMS, JADIE; SKAGGS, JONATHAN; ORZECHOWSKI, GRZEGORZ; MCCLELLAN, JOSHUA
To: CAPTIONCALL, LLC
Reel/Frame 058168/0584 →
Continuity (2)
Continuation 16209524 · Dec 4, 2018
Related Publication 20220122587A1 · Apr 21, 2022
References Cited (236)
US 5606643A · Balasubramanian et al. · 1997 [cited by applicant]
US 5649060A · Ellozy et al. · 1997 [cited by applicant]
US 5724405A · Engelke et al. · 1998 [cited by applicant]
US 5855000A · Waibel et al. · 1998 [cited by applicant]
US 5883986A · Kopec et al. · 1999 [cited by applicant]
US 6122613A · Baker · 2000 [cited by applicant]
US 6208964B1 · Sabourin · 2001 [cited by applicant]
US 6208970B1 · Ramanan · 2001 [cited by applicant]
US 6366882B1 · Bijl et al. · 2002 [cited by applicant]
US 6374221B1 · Haimi-Cohen · 2002 [cited by applicant]
US 6385582B1 · Iwata · 2002 [cited by applicant]
US 6457031B1 · Hanson · 2002 [cited by applicant]
US 6535848B1 · Ortega et al. · 2003 [cited by applicant]
US 6704709B1 · Kahn et al. · 2004 [cited by applicant]
US 6728677B1 · Kannan et al. · 2004 [cited by applicant]
US 6813603B1 · Groner et al. · 2004 [cited by applicant]
US 6816468B1 · Cruickshank · 2004 [cited by applicant]
US 6832189B1 · Kanevsky et al. · 2004 [cited by applicant]
US 6947896B2 · Hanson · 2005 [cited by applicant]
US 7003463B1 · Maes et al. · 2006 [cited by applicant]
US 7016844B2 · Othmer et al. · 2006 [cited by applicant]
US 7130790B1 · Flanagan et al. · 2006 [cited by applicant]
US 7174299B2 · Fujii et al. · 2007 [cited by applicant]
US 7191130B1 · Leggetter et al. · 2007 [cited by applicant]
US 7191135B2 · O'Hagan · 2007 [cited by applicant]
US 7228275B1 · Endo et al. · 2007 [cited by applicant]
US 7236932B1 · Grajski · 2007 [cited by applicant]
US 7519536B2 · Maes et al. · 2009 [cited by applicant]
US 7606718B2 · Cloran · 2009 [cited by applicant]
US 7613610B1 · Zimmerman et al. · 2009 [cited by applicant]
US 7660715B1 · Thambiratnam · 2010 [cited by applicant]
US 7836412B1 · Zimmerman · 2010 [cited by applicant]
US 7844454B2 · Coles et al. · 2010 [cited by applicant]
US 7907705B1 · Huff et al. · 2011 [cited by applicant]
US 7930181B1 · Goffin et al. · 2011 [cited by applicant]
US 7957970B1 · Gorin et al. · 2011 [cited by applicant]
US 7962339B2 · Pieraccini et al. · 2011 [cited by applicant]
US 8019608B2 · Carraux et al. · 2011 [cited by applicant]
US 8180639B2 · Pieraccini et al. · 2012 [cited by applicant]
US 8223944B2 · Cloran et al. · 2012 [cited by applicant]
US 8249878B2 · Carraux et al. · 2012 [cited by applicant]
US 8286071B1 · Zimmerman et al. · 2012 [cited by applicant]
US 8332227B2 · Maes et al. · 2012 [cited by applicant]
US 8332231B2 · Cloran · 2012 [cited by applicant]
US 8335689B2 · Wittenstein et al. · 2012 [cited by applicant]
US 8379801B2 · Romriell et al. · 2013 [cited by applicant]
US 8407052B2 · Hager · 2013 [cited by applicant]
US 8484031B1 · Yeracaris et al. · 2013 [cited by applicant]
US 8484042B2 · Cloran · 2013 [cited by applicant]
US 8504372B2 · Carraux et al. · 2013 [cited by applicant]
US 8537979B1 · Pollock · 2013 [cited by applicant]
US 8560321B1 · Yeracaris et al. · 2013 [cited by applicant]
US 8605682B2 · Efrati et al. · 2013 [cited by applicant]
US 8626520B2 · Cloran · 2014 [cited by applicant]
US 8743003B2 · De Lustrac et al. · 2014 [cited by applicant]
US 8744848B2 · Hoepfinger et al. · 2014 [cited by applicant]
US 8781510B2 · Gould et al. · 2014 [cited by applicant]
US 8812321B2 · Gilbert et al. · 2014 [cited by applicant]
US 8868425B2 · Maes et al. · 2014 [cited by applicant]
US 8874070B2 · Basore et al. · 2014 [cited by applicant]
US 8892447B1 · Srinivasan et al. · 2014 [cited by applicant]
US 8898065B2 · Newman et al. · 2014 [cited by applicant]
US 8930194B2 · Newman et al. · 2015 [cited by applicant]
US 9002713B2 · Ljolje et al. · 2015 [cited by applicant]
US 9076450B1 · Sadek et al. · 2015 [cited by applicant]
US 9117450B2 · Cook et al. · 2015 [cited by applicant]
US 9153231B1 · Salvador · 2015 [cited by examiner]
US 9183843B2 · Fanty et al. · 2015 [cited by applicant]
US 9191789B2 · Pan · 2015 [cited by applicant]
US 9197745B1 · Chevrier et al. · 2015 [cited by applicant]
US 9215409B2 · Montero et al. · 2015 [cited by applicant]
US 9245522B2 · Hager · 2016 [cited by applicant]
US 9245525B2 · Yeracaris et al. · 2016 [cited by applicant]
US 9247052B1 · Walton · 2016 [cited by applicant]
US 9318110B2 · Roe · 2016 [cited by applicant]
US 9324324B2 · Knighton · 2016 [cited by applicant]
US 9336689B2 · Romriell et al. · 2016 [cited by applicant]
US 9344562B2 · Moore et al. · 2016 [cited by applicant]
US 9374536B1 · Nola et al. · 2016 [cited by applicant]
US 9380150B1 · Bullough et al. · 2016 [cited by applicant]
US 9386152B2 · Riahi et al. · 2016 [cited by applicant]
US 9444934B2 · Nelson et al. · 2016 [cited by applicant]
US 9460719B1 · Antunes et al. · 2016 [cited by applicant]
US 9472185B1 · Yeracaris · 2016 [cited by examiner]
US 9502033B2 · Carraux et al. · 2016 [cited by applicant]
US 9514747B1 · Bisani et al. · 2016 [cited by applicant]
US 9525830B1 · Roylance et al. · 2016 [cited by applicant]
US 9535891B2 · Raheja et al. · 2017 [cited by applicant]
US 9548048B1 · Solh et al. · 2017 [cited by applicant]
US 9571638B1 · Knighton et al. · 2017 [cited by applicant]
US 9576498B1 · Zimmerman et al. · 2017 [cited by applicant]
US 9621732B2 · Olligschlaeger · 2017 [cited by applicant]
US 9628620B1 · Rae et al. · 2017 [cited by applicant]
US 9632997B1 · Johnson et al. · 2017 [cited by applicant]
US 9633657B2 · Svendsen et al. · 2017 [cited by applicant]
US 9641681B2 · Nuta et al. · 2017 [cited by applicant]
US 9653076B2 · Kim · 2017 [cited by applicant]
US 9654628B2 · Warren et al. · 2017 [cited by applicant]
US 9704111B1 · Antunes et al. · 2017 [cited by applicant]
US 9710819B2 · Cloran et al. · 2017 [cited by applicant]
US 9715876B2 · Hager · 2017 [cited by applicant]
US 9741347B2 · Yeracaris et al. · 2017 [cited by applicant]
US 9761241B2 · Maes et al. · 2017 [cited by applicant]
US 9842587B2 · Hakkani-Tur et al. · 2017 [cited by applicant]
US 9858256B2 · Hager · 2018 [cited by applicant]
US 9886956B1 · Antunes et al. · 2018 [cited by applicant]
US 9922654B2 · Chang et al. · 2018 [cited by applicant]
US 9947322B2 · Kang et al. · 2018 [cited by applicant]
US 9953653B2 · Newman et al. · 2018 [cited by applicant]
US 9990925B2 · Weinstein et al. · 2018 [cited by applicant]
US 10032455B2 · Newman et al. · 2018 [cited by applicant]
US 10044854B2 · Rae et al. · 2018 [cited by applicant]
US 10049669B2 · Newman et al. · 2018 [cited by applicant]
US 10049676B2 · Yeracaris et al. · 2018 [cited by applicant]
US 10147419B2 · Yeracaris et al. · 2018 [cited by applicant]
US 20020152071A1 · Chaiken et al. · 2002 [cited by applicant]
US 20030050777A1 · Walker, Jr. · 2003 [cited by applicant]
US 20040083105A1 · Jaroker · 2004 [cited by applicant]
US 20050049868A1 · Busayapongchai · 2005 [cited by applicant]
US 20050094777A1 · McClelland · 2005 [cited by applicant]
US 20050119897A1 · Bennett et al. · 2005 [cited by applicant]
US 20050226394A1 · Engelke et al. · 2005 [cited by applicant]
US 20050226398A1 · Bojeun · 2005 [cited by applicant]
US 20060072727A1 · Bantz et al. · 2006 [cited by applicant]
US 20060074623A1 · Tankhiwale · 2006 [cited by applicant]
US 20060089857A1 · Zimmerman et al. · 2006 [cited by applicant]
US 20070153989A1 · Howell et al. · 2007 [cited by applicant]
US 20070208570A1 · Bhardwaj et al. · 2007 [cited by applicant]
US 20070225970A1 · Kady et al. · 2007 [cited by applicant]
US 20080040111A1 · Miyamoto et al. · 2008 [cited by applicant]
US 20080049908A1 · Doulton · 2008 [cited by applicant]
US 20080133245A1 · Proulx et al. · 2008 [cited by applicant]
US 20090018833A1 · Kozat et al. · 2009 [cited by applicant]
US 20090037171A1 · McFarland et al. · 2009 [cited by applicant]
US 20090177461A1 · Ehsani et al. · 2009 [cited by applicant]
US 20090187410A1 · Wilpon et al. · 2009 [cited by applicant]
US 20090248416A1 · Gorin et al. · 2009 [cited by applicant]
US 20090299743A1 · Rogers · 2009 [cited by applicant]
US 20100027765A1 · Schultz et al. · 2010 [cited by applicant]
US 20100063815A1 · Cloran et al. · 2010 [cited by applicant]
US 20100076752A1 · Zweig et al. · 2010 [cited by applicant]
US 20100076843A1 · Ashton · 2010 [cited by applicant]
US 20100121637A1 · Roy et al. · 2010 [cited by applicant]
US 20100145729A1 · Katz · 2010 [cited by applicant]
US 20100312556A1 · Ljolje et al. · 2010 [cited by applicant]
US 20100318355A1 · Li et al. · 2010 [cited by applicant]
US 20100323728A1 · Gould et al. · 2010 [cited by applicant]
US 20110087491A1 · Wittenstein et al. · 2011 [cited by applicant]
US 20110112833A1 · Frankel et al. · 2011 [cited by applicant]
US 20110123003A1 · Romriell et al. · 2011 [cited by applicant]
US 20110128953A1 · Wozniak et al. · 2011 [cited by applicant]
US 20110295603A1 · Meisel · 2011 [cited by applicant]
US 20120016671A1 · Jaggi et al. · 2012 [cited by applicant]
US 20120178064A1 · Katz · 2012 [cited by applicant]
US 20120214447A1 · Russell et al. · 2012 [cited by applicant]
US 20120245934A1 · Talwar et al. · 2012 [cited by applicant]
US 20120316882A1 · Fiumi · 2012 [cited by applicant]
US 20130035937A1 · Webb et al. · 2013 [cited by applicant]
US 20130060572A1 · Garland et al. · 2013 [cited by applicant]
US 20130132084A1 · Stonehocker et al. · 2013 [cited by applicant]
US 20130132086A1 · Xu et al. · 2013 [cited by applicant]
US 20130151250A1 · Van Blon · 2013 [cited by applicant]
US 20130317818A1 · Bigham et al. · 2013 [cited by applicant]
US 20140018045A1 · Tucker · 2014 [cited by applicant]
US 20140067390A1 · Webb · 2014 [cited by applicant]
US 20140163977A1 · Hoffmeister · 2014 [cited by examiner]
US 20140163981A1 · Cook et al. · 2014 [cited by applicant]
US 20140207451A1 · Topiwala et al. · 2014 [cited by applicant]
US 20140278402A1 · Charugundla · 2014 [cited by applicant]
US 20140314220A1 · Charugundla · 2014 [cited by applicant]
US 20150051908A1 · Romriell et al. · 2015 [cited by applicant]
US 20150073790A1 · Steuble et al. · 2015 [cited by applicant]
US 20150058006A1 · Bisani et al. · 2015 [cited by applicant]
US 20150094105A1 · Pan · 2015 [cited by applicant]
US 20150095026A1 · Bisani et al. · 2015 [cited by applicant]
US 20150106091A1 · Wetjen et al. · 2015 [cited by applicant]
US 20150130887A1 · Thelin et al. · 2015 [cited by applicant]
US 20150149162A1 · Melamed · 2015 [cited by examiner]
US 20150287408A1 · Svendsen et al. · 2015 [cited by applicant]
US 20150288815A1 · Charugundla · 2015 [cited by applicant]
US 20150332670A1 · Akbacak et al. · 2015 [cited by applicant]
US 20150341486A1 · Knighton · 2015 [cited by applicant]
US 20160012751A1 · Hirozawa · 2016 [cited by applicant]
US 20160062970A1 · Sadkin et al. · 2016 [cited by applicant]
US 20160078860A1 · Paulik et al. · 2016 [cited by applicant]
US 20160259779A1 · Absky et al. · 2016 [cited by applicant]
US 20160379626A1 · Deisher et al. · 2016 [cited by applicant]
US 20170085506A1 · Gordon · 2017 [cited by applicant]
US 20170116993A1 · Miglietta et al. · 2017 [cited by applicant]
US 20170148433A1 · Catanzaro et al. · 2017 [cited by applicant]
US 20170187876A1 · Hayes et al. · 2017 [cited by applicant]
US 20170201613A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206808A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206888A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206890A1 · Tapuhi et al. · 2017 [cited by applicant]
US 20170206891A1 · Lev-Tov · 2017 [cited by examiner]
US 20170206914A1 · Engelke et al. · 2017 [cited by applicant]
US 20170208172A1 · Engelke et al. · 2017 [cited by applicant]
US 20170236527A1 · Dimitriadis et al. · 2017 [cited by applicant]
US 20170270929A1 · Aleksic et al. · 2017 [cited by applicant]
US 20170294186A1 · Pinto et al. · 2017 [cited by applicant]
US 20180013886A1 · Rae et al. · 2018 [cited by applicant]
US 20180034961A1 · Engelke et al. · 2018 [cited by applicant]
US 20180052823A1 · Scally et al. · 2018 [cited by applicant]
US 20180081869A1 · Hager · 2018 [cited by applicant]
US 20180197545A1 · Willett et al. · 2018 [cited by applicant]
US 20180270350A1 · Engelke · 2018 [cited by examiner]
US 20190037072A1 · Engelke et al. · 2019 [cited by applicant]
US 20190088251A1 · Mun · 2019 [cited by examiner]
US 20190103095A1 · Singaraju · 2019 [cited by examiner]
US 20200066281A1 · Grancharov · 2020 [cited by examiner]
US 20200194006A1 · Grancharov · 2020 [cited by examiner]
US 20210125603A1 · Liang · 2021 [cited by examiner]
EP 0555545A1 · 1993 [cited by applicant]
EP 0645757A1 · 1995 [cited by applicant]
EP 2587478A2 · 2011 [cited by applicant]
EP 2372707A1 · 2011 [cited by applicant]
JP 2011002656A · 2011 [cited by applicant]
JP 2015018238A · 2015 [cited by applicant]
KR 20170134115A · 2017 [cited by applicant]
WO 1998034217A1 · 1998 [cited by applicant]
WO 0225910A2 · 2002 [cited by applicant]
WO 2014022559A1 · 2014 [cited by applicant]
WO 2014049944A1 · 2014 [cited by applicant]
WO 2014176489A2 · 2014 [cited by applicant]
WO 2012165529A1 · 2015 [cited by applicant]
WO 2015131028A1 · 2015 [cited by applicant]
WO 2015148037A1 · 2015 [cited by applicant]
WO 2017054122A1 · 2017 [cited by applicant]
Peter Hays, CEO, VTCSecure, LLC, Nov. 7, 2017, Notice of Ex Parte in CG Docket Nos. 03-123 and 10-51, 2 pgs. [cited by applicant]
Lambert Mathias, Statistical Machine Translation and Automatic Speech Recognition under Uncertainty, Dissertation submitted to Johns Hopkins University, Dec. 2007, 131 pages. [cited by applicant]
Mike Wald, Captioning for Deaf and Hard of Hearing People by Editing Automatic Speech Recognition in Real Time, CCHP'06 Proceedings of the 10th International Conference on Computers Helping People with Special Needs, Ju… [cited by applicant]
Jeff Adams, Kenneth Basye, Alok Parlikar, Andrew Fletcher, & Jangwon Kim, Automated Speech Recognition for Captioned Telephone Conversations, Faculty Works, Nov. 3, 2017, pp. 1-12. [cited by applicant]
Lasecki et al., Real-Time Captioning by Groups of Non-Experts, Rochester Institute of Technology, Oct. 2012, USA. [cited by applicant]
Benjamin Lecouteux, Georges Linares, Stanislas Oger, Integrating imperfect transcripts into speech recognition systems for building high-quality corpora, ScienceDirect, Jun. 2011. [cited by applicant]
International Search Report and Written Opinion, as issued in connection with International Patent Application No. PCT/US2019/062863, dated Mar. 26, 2020. [cited by applicant]