IP Library Granted Patent US 12,640,149
Granted Patent B2
US 12,640,149 · App. 18/103,037 · Granted May 26, 2026

Techniques for wake-up work recognition and related systems and methods

Inventors: Meik Pfeffinger (Ulm, DE); Timo Matheja (Neo-Ulm, DE); Tobias Herbig (Ulm, DE); Tim Haulick (Blaubeuren, DE)
Assignee: CERENCE OPERATING COMPANY
G10L15/22G06F3/167G10L15/08G10L17/22G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,640,149
App. No.
18/103,037
Granted
May 26, 2026
Kind
B2
Abstract

A system for detection of at least one designated wake-up word for at least one speech-enabled application. The system comprises at least one microphone; and at least one computer hardware processor configured to perform: receiving an acoustic signal generated by the at least one microphone at least in part as a result of receiving an utterance spoken by a speaker; obtaining information indicative of the speaker's identity; interpreting the acoustic signal at least in part by determining, using the information indicative of the speaker's identity and automated speech recognition, whether the utterance spoken by the speaker includes the at least one designated wake-up word; and interacting with the speaker based, at least in part, on results of the interpreting.

Claims (42)

1 . A system for detecting at least one designated wake-up word for at least one speech-enabled application, the system comprising:

at least one computer hardware component configured to perform:

receiving an acoustic signal generated by at least one microphone at least in part as a result of receiving an utterance spoken by a speaker;

digitizing the acoustic signal to create a digitized acoustic signal;

sending the digitized acoustic signal to at least one computer hardware processor, the at least one computer hardware processor configured to perform:

obtaining information indicative of a speaker's identity by processing, at least in part, the digitized acoustic signal;

using the information indicative of the speaker's identity to determine whether the utterance spoken by the speaker includes at least one or more wake-up words associated with the speaker's identity;

in response to determining that the utterance spoken by the speaker includes the at least one designated wake-up word, interacting with the speaker, wherein the at least one designated wake-up word includes a first designated wake-up word for a first speech-enabled application of the at least one speech-enabled application, and wherein the first designated wake-up word is specific to the speaker such that no other speaker can use the first designated wake-up word to interact with the at least one speech-enabled application.

2 . The system of claim 1 , wherein interacting with the speaker comprises allowing the speaker to control the at least one speech-enabled application.

3 . The system of claim 1 , wherein the at least one computer hardware processor is programmed to use the information indicative of the speaker's identity to determine whether the speaker is authorized to control the at least one speech-enabled application, and to allow the speaker to control the at least one speech-enabled application if it is determined that the speaker is authorized to control the at least one speech-enabled application, and not allow the speaker to control the at least one speech-enabled application if it is determined that the speaker is not authorized to control the at least one speech-enabled application.

4 . The system of claim 1 , wherein obtaining the speaker's identity comprises:

obtaining speech characteristics from the digitized acoustic signal;

comparing the obtained speech characteristics against stored speech characteristics for each of multiple speakers registered with the system.

5 . The system of claim 1 , wherein determining whether the utterance spoken by the speaker includes the at least one designated wake-up word comprises:

using automated speech recognition to determine whether the utterance spoken by the speaker includes a wake-up word in the one or more wake-up words, wherein the automated speech recognition is performed using the one or more wake-up words associated with the speaker identity.

6 . The system of claim 1 , wherein obtaining information indicative of the speaker's identity comprises determining a position of the speaker in an environment.

7 . The system of claim 6 , wherein the at least one computer hardware processor is configured to determine, using the position of the speaker in the environment, whether the speaker is authorized to control the at least one speech-enabled application, and to allow the speaker to control the at least one speech-enabled application if it is determined that the speaker is authorized to control the at least one speech-enabled application, and not allow the speaker to control the at least one speech-enabled application if it is determined that the speaker is not authorized to control the at least one speech-enabled application.

8 . The system of claim 6 , wherein the at least one computer hardware processor is configured to determine the position of the speaker inside a vehicle based, at least in part, on information gathered by at least one sensor in the vehicle.

9 . The system of claim 6 , wherein the at least one computer hardware processor receives digitized acoustic signals from a plurality of microphones, and wherein the position of the speaker is determined using the digitized acoustic signals received from the plurality of microphones.

10 . The system of claim 1 , wherein the at least one microphone comprises a plurality of microphones installed in a respective plurality of acoustic zones inside of a vehicle, wherein each of the plurality of acoustic zones comprises a seating area for a passenger in the vehicle.

11 . The system of claim 1 , wherein interacting with the speaker comprises inferring, based at least in part on the information indicative of the speaker's identity, at least one action to take when interacting with the speaker.

12 . The system of claim 1 , wherein obtaining information indicative of the speaker's identity comprises obtaining the speaker's identity; and

wherein determining whether the utterance spoken by the speaker includes the at least one designated wake-up word comprises:

accessing a list of wake-up words associated with the speaker's identity; and

determining whether the utterance includes any wake-up word in the list of wake-up words associated with the speaker's identity.

13 . The system of claim 1 , wherein determining whether the utterance spoken by the speaker includes the at least one designated wake-up word comprises:

compensating for interference received by the at least one microphone by using the information associated with the speaker's identity.

14 . The system of claim 1 , wherein the at least one computer hardware processor is further configured to store the information about the speaker's identity in at least one data store.

15 . The system of claim 14 , wherein the at least one data store comprises a plurality of data records including a first data record, the first data record comprising information selected from the group consisting of an identity of a particular speaker, a position of the particular speaker in an environment, a list of one or more wake-up words associated with the particular speaker, a list of one or more speech-enabled applications that the particular speaker is allowed to control, a list of one or more speech-enabled applications that the particular speaker is not allowed to control, and information obtained from one or more sensors.

16 . A method for detecting at least one designated wake-up word for at least one speech-enabled application, the method comprising:

receiving a digitized acoustic signal generated by a computer hardware component based on an utterance spoken by a speaker and received at at least one microphone;

obtaining information indicative of the speaker's identity;

using the information indicative of the speaker's identity to determine whether the utterance spoken by the speaker includes the at least one designated wake-up word associated with the speaker's identity; and

in response to determining that the utterance spoken by the speaker includes the at least one designated wake-up word, interacting with the speaker.

17 . The method of claim 16 , wherein obtaining information indicative of the speaker's identity comprises determining a position of the speaker in an environment.

18 . At least one non-transitory computer-readable storage medium storing processor-executable instructions that, when executed by at least one computer hardware processor, cause the at least one computer hardware processor to perform a method for detecting at least one designated wake-up word for at least one speech- enabled application, the method comprising:

receiving a digitized acoustic signal generated by at least one microphone at least in part as a result of receiving an utterance spoken by a speaker and processed to generate the digitized acoustic signal;

obtaining information indicative of the speaker's identity;

using the information indicative of the speaker's identity to determine whether the utterance spoken by the speaker includes the at least one designated wake-up word associated with the speaker's identity; and

in response to determining that the utterance spoken by the speaker includes the at least one designated wake-up word, interacting with the speaker.

19 . The medium of claim 18 , wherein method is further configured to store the information about the speaker's identity in at least one data store.

20 . The medium of claim 19 , wherein the at least one data store comprises a plurality of data records including a first data record, the first data record comprising information selected from the group consisting of an identity of a particular speaker, a position of the particular speaker in an environment, a list of one or more wake-up words associated with the particular speaker, a list of one or more speech-enabled applications that the particular speaker is allowed to control, a list of one or more speech-enabled applications that the particular speaker is not allowed to control, and information obtained from one or more sensors.

Assignments (2)
RELEASE (REEL 067417 / FRAME 0303) Recorded Jan 2, 2025
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: CERENCE OPERATING COMPANY
Reel/Frame 069797/0422 →
SECURITY AGREEMENT Recorded Apr 15, 2024
From: CERENCE OPERATING COMPANY
To: WELLS FARGO BANK, N.A., AS COLLATERAL AGENT
Reel/Frame 067417/0303 →
Continuity (2)
Continuation 16308849
Related Publication 20230178077A1 · Jun 8, 2023
References Cited (123)
US 5428707A · Gould · 1995 [cited by applicant]
US 5774859A · Houser · 1998 [cited by applicant]
US 5802305A · McKaughan et al. · 1998 [cited by applicant]
US 5983186A · Miyazawa et al. · 1999 [cited by applicant]
US 6006175A · Holzrichter · 1999 [cited by applicant]
US 6070140A · Tran · 2000 [cited by applicant]
US 6092043A · Squires et al. · 2000 [cited by applicant]
US 6397186B1 · Bush et al. · 2002 [cited by applicant]
US 6408396B1 · Forbes · 2002 [cited by applicant]
US 6411926B1 · Chang · 2002 [cited by applicant]
US 6449496B1 · Beith et al. · 2002 [cited by applicant]
US 6594630B1 · Zlokarnik et al. · 2003 [cited by applicant]
US 6756700B2 · Zeng · 2004 [cited by applicant]
US 6859776B1 · Cohen et al. · 2005 [cited by applicant]
US 6941265B2 · Bi et al. · 2005 [cited by applicant]
US 6965786B2 · Qu et al. · 2005 [cited by applicant]
US 7114090B2 · Kardach et al. · 2006 [cited by applicant]
US 7567827B2 · Kim · 2009 [cited by applicant]
US 7574361B2 · Yeager et al. · 2009 [cited by applicant]
US 7720683B1 · Vermeulen et al. · 2010 [cited by applicant]
US 7774204B2 · Mozer et al. · 2010 [cited by applicant]
US 8056070B2 · Goller et al. · 2011 [cited by applicant]
US 8181046B2 · Marcu et al. · 2012 [cited by applicant]
US 8190420B2 · Kadirkamanathan et al. · 2012 [cited by applicant]
US 8285545B2 · Lee et al. · 2012 [cited by applicant]
US 8548176B2 · Bright · 2013 [cited by applicant]
US 8620389B2 · Schrager · 2013 [cited by applicant]
US 8666751B2 · Murthi et al. · 2014 [cited by applicant]
US 8977255B2 · Freeman et al. · 2015 [cited by applicant]
US 9087520B1 · Salvador · 2015 [cited by applicant]
US 9112984B2 · Sejnoha et al. · 2015 [cited by applicant]
US 9361885B2 · Ganong, III et al. · 2016 [cited by applicant]
US 9558749B1 · Secker-Walker · 2017 [cited by applicant]
US 9646610B2 · Macho · 2017 [cited by applicant]
US 9747899B2 · Pogue · 2017 [cited by applicant]
US 9940936B2 · Sejnoha et al. · 2018 [cited by applicant]
US 9992642B1 · Rapp · 2018 [cited by applicant]
US 10057421B1 · Chiu · 2018 [cited by applicant]
US 10332525B2 · Secker-Walker · 2019 [cited by applicant]
US 10373612B2 · Parthasarathi · 2019 [cited by applicant]
US 10521512B2 · Alders · 2019 [cited by applicant]
US 10536773B2 · Metheja · 2020 [cited by applicant]
US 10783899B2 · Graf · 2020 [cited by applicant]
US 11232788B2 · Yavagal · 2022 [cited by applicant]
US 11437020B2 · Premont · 2022 [cited by applicant]
US 11600269B2 · Pfeffinger · 2023 [cited by examiner]
US 20020193989A1 · Geilhufe et al. · 2002 [cited by applicant]
US 20030040339A1 · Chang · 2003 [cited by applicant]
US 20030120486A1 · Brittan et al. · 2003 [cited by applicant]
US 20030216909A1 · Davis et al. · 2003 [cited by applicant]
US 20070129949A1 · Alberth et al. · 2007 [cited by applicant]
US 20080118080A1 · Gratke et al. · 2008 [cited by applicant]
US 20090055178A1 · Coon · 2009 [cited by applicant]
US 20100009719A1 · Oh et al. · 2010 [cited by applicant]
US 20100121636A1 · Burke et al. · 2010 [cited by applicant]
US 20100124896A1 · Kumar · 2010 [cited by applicant]
US 20100185448A1 · Meisel · 2010 [cited by applicant]
US 20110054899A1 · Phillips et al. · 2011 [cited by applicant]
US 20120034904A1 · LeBeau et al. · 2012 [cited by applicant]
US 20120035924A1 · Jitkoff et al. · 2012 [cited by applicant]
US 20120127072A1 · Kim · 2012 [cited by applicant]
US 20120197637A1 · Gratke et al. · 2012 [cited by applicant]
US 20120281885A1 · Syrdal · 2012 [cited by applicant]
US 20120310646A1 · Hu et al. · 2012 [cited by applicant]
US 20120329389A1 · Royston et al. · 2012 [cited by applicant]
US 20130080167A1 · Mozer · 2013 [cited by applicant]
US 20130080171A1 · Mozer et al. · 2013 [cited by applicant]
US 20130289994A1 · Newman et al. · 2013 [cited by applicant]
US 20130339028A1 · Rosner et al. · 2013 [cited by applicant]
US 20140012573A1 · Hung et al. · 2014 [cited by applicant]
US 20140012586A1 · Rubin et al. · 2014 [cited by applicant]
US 20140039888A1 · Taubman · 2014 [cited by applicant]
US 20140163978A1 · Basye · 2014 [cited by applicant]
US 20140249817A1 · Hart et al. · 2014 [cited by applicant]
US 20140274203A1 · Ganong, III et al. · 2014 [cited by applicant]
US 20140274211A1 · Sejnoha et al. · 2014 [cited by applicant]
US 20140278435A1 · Ganong, III et al. · 2014 [cited by applicant]
US 20140365225A1 · Haiut · 2014 [cited by applicant]
US 20150006176A1 · Pogue · 2015 [cited by applicant]
US 20150053779A1 · Adamek · 2015 [cited by applicant]
US 20150106085A1 · Lindahl · 2015 [cited by applicant]
US 20150340042A1 · Sejnoha et al. · 2015 [cited by applicant]
US 20160039356A1 · Talwar et al. · 2016 [cited by applicant]
US 20160078869A1 · Syrdal · 2016 [cited by applicant]
US 20160189706A1 · Zopf et al. · 2016 [cited by applicant]
US 20160314782A1 · Klimanis · 2016 [cited by applicant]
US 20160358605A1 · Ganong, III et al. · 2016 [cited by applicant]
US 20170116983A1 · Furukawa et al. · 2017 [cited by applicant]
US 20170270919A1 · Parthasarathi · 2017 [cited by applicant]
US 20180114531A1 · Kumar · 2018 [cited by applicant]
US 20180182380A1 · Fritz · 2018 [cited by applicant]
US 20180310144A1 · Rapp · 2018 [cited by applicant]
US 20180366114A1 · Anbazhagan · 2018 [cited by applicant]
US 20190073999A1 · Premont · 2019 [cited by applicant]
US 20190287526A1 · Ren et al. · 2019 [cited by applicant]
US 20190311715A1 · Pfeffinger · 2019 [cited by examiner]
US 20190355365A1 · Kim · 2019 [cited by applicant]
US 20200020329A1 · Gordon · 2020 [cited by applicant]
US 20200035231A1 · Parthasarathi · 2020 [cited by applicant]
US 20200184966A1 · Yavagal · 2020 [cited by applicant]
CN 101650943A · 2010 [cited by applicant]
CN 103021409A · 2013 [cited by applicant]
CN 103632668A · 2014 [cited by applicant]
CN 104575504A · 2015 [cited by applicant]
CN 105009204A · 2015 [cited by applicant]
CN 105575395A · 2016 [cited by applicant]
CN 106098059A · 2016 [cited by applicant]
EP 1511010A1 · 2005 [cited by applicant]
EP 2899955A1 · 2015 [cited by applicant]
EP 2932500B1 · 2017 [cited by applicant]
WO 2014066192A1 · 2014 [cited by applicant]
WO WO2015196063 · 2014 [cited by examiner]
WO 2015196063A1 · 2015 [cited by applicant]
M. Adamski and B. Von Solms, “An open speaker recognition enabled identification and authentication system,” 2014 IST-Africa Conference Proceedings, Pointe aux Piments, Mauritius, 2014, pp. 1-8. (Year: 2014). [cited by examiner]
Chines Office Action and Translation thereof for Chinese Application No. 201480013903.1 dated Jul. 28, 2017. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2014/024270 mailed Sep. 24, 2015. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2016/037495 mailed Dec. 27, 2018. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/CN2016/105343 mailed May 23, 2019. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2016/017317 mailed Aug. 23, 2018. [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2016/017317 mailed May 12, 2016. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2014/024270 mailed Jun. 16, 2014. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/CN2016/105343 mailed Sep. 21, 2017. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2016/037495 mailed Dec. 5, 2016. [cited by applicant]