IP Library Patent Application 16532751
Patent Application
App. No. 16/532,751

AUTOMATED SPEECH RECOGNITION SYSTEM

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
16/532,751
Abstract

There is provided an automated speech recognition system that applies weights to grapheme-to-phoneme models, and interpolates pronunciations from combinations of the models, to recognize utterances of foreign named entities for naive, informed, and in-between pronunciations.

Claims (23)

1 . An automated speech recognition (ASR) system, comprising:

a microphone;

a recognition dictionary storage that contains:

(a) a first recognition dictionary that stores a first pronunciation of a token that was generated from a first grapheme-to-phoneme model (G2P) for said token; and

(b) a second recognition dictionary that stores a second interpretation of said token that was generated from a second G2P model for said token;

a G2P weight storage that contains:

(a) a first G2P weight that is applicable to said first G2P model to yield said first pronunciation for said token; and

(b) a second G2P weight that is applicable to said second G2P model to yield said second pronunciation for said token;

a processor that receives an utterance containing a spoken form of said token from said microphone; and

a memory that contains instructions that are readable by said processor to control said processor to:

obtain metadata concerning said token;

modify said first G2P weight and said second G2P weight based on said metadata, thus yielding a first weighted G2P model and a second weighted G2P model;

interpolate said first weighted G2P model and said second weighted G2P model to yield a resultant pronunciation for said token; and

provide an output based on said resultant pronunciation.

2 . The ASR system of claim 1 ,

wherein said utterance is spoken by a user, and

wherein said metadata identifies a characteristic of said user.

3 . The ASR system of claim 2 , wherein said characteristic of said user is a native language of said user.

4 . The ASR system of claim 1 , further comprising:

a user device; and

a global positioning system that identifies a present location of said user device,

wherein said metadata comprises said present location.

5 . The ASR system of claim 1 , wherein said output comprises a signal to control a device.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 5, 2020
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 052114/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 6, 2019
From: HAHN, STEFAN CHRISTOF; GEORGALA, EFTHYMIA; DIVAY, OLIVIER STEPHANE JEROME; MARSHALL, ERIC JOSEPH
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 049970/0567 →