IP Library Granted Patent US 11,114,101
Granted Patent B2
US 11,114,101 · App. 16/962,734 · Granted Sep 7, 2021

Speech recognition with image signal

Inventors: Olaf Petrus Quirinus Mossinkoff (The Hague, NL); Johannes Leonardus Jozef Meijer (Haarlem, NL)
Assignee: IEBM B.V.
G10L15/25G10L15/22H04N5/2256H04N5/23219G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,114,101
App. No.
16/962,734
Granted
Sep 7, 2021
Kind
B2
Abstract

A method of speech recognition and person identification based thereon, comprising: recording speech from a speech signal using a microphone; illuminating a speaking mouth; recording a degree of light reflected by the mouth from a reflection signal using a sensor; and recording combined parameters of the speech signal and of the reflection signal, and coupling them to letters associated therewith, per predetermined time duration; comparing a combination occurring in speech of parameters of the speech signal and of the reflection signal to the recorded combined parameters of the speech signal and of the reflection signal which are coupled to letters; and deciding on the basis of the comparison to which letter the combination occurring in the speech of parameters of the speech signal and of the reflection signal corresponds, using block-width modulation of the reflection signal.

Claims (28)

1. A method of speech recognition, comprising:

recording speech from a speech signal using a microphone;

illuminating a speaking mouth;

recording a degree of light reflected by the mouth from a reflection signal using a sensor;

for speech recognition, in one time duration of the predetermined length of speech to be recognized:

comparing a combination of parameters of the speech signal and of the reflection signal with pre-recorded combined parameters of a speech signal obtained during a training phase and of a reflection signal obtained during said training phase of at least some of a plurality of pre-recorded time durations, wherein said combination of parameters includes information regarding said degree of light reflected by the mouth, and wherein said pre-recorded combined parameters include information regarding said degree of light reflected by the mouth determined during said training phase;

wherein the training speech signal and the training reflection signal are coupled to letters associated therewith per time duration; and

deciding on the basis of the comparison to which letter the combination occurring in the speech of parameters of the speech signal and of the reflection signal corresponds,

wherein block-width modulation is applied to the reflection signal.

2. The method of claim 1 , wherein the parameter of the speech signal is taken from a group comprising within the one time duration of speech to be recognized at least:

volume dispersion of a difference between a highest and a lowest value of a volume of the speech signal; and

a ratio of sound of the speech signal within and outside a noise level (signal-to-noise ratio).

3. The method of claim 1 , wherein the parameter of the reflection signal is taken from a group comprising within the one time duration at least:

an average of the reflection signal; and

a degree of increase or decrease of the reflection signal.

4. The method of claim 3 , further comprising determining the average of the reflection signal as an average over the one time duration of half a block duration of a block wave used in block wave modulation.

5. The method of claim 3 , further comprising determining the degree of increase or decrease of the reflection signal as a measurement in degrees.

6. The method of claim 1 , wherein the time duration has a predetermined length of 1, 2, 3, 4 or 5 milliseconds.

7. The method of claim 1 , further comprising subdividing the speech signal into portions corresponding to letters, and indicating at least one of:

starting and ending times of letters in the speech signal; and

time durations expressed in a number of times of the one time duration.

8. The method of claim 1 , further comprising determining a maximum and minima of a degree of light reflected by the mouth in the reflection signal, and normalizing the reflection signal on the basis of the maximum and minima.

9. The method of claim 1 , further comprising:

forming a preselection of at least one potential subsequent letter on the basis of the decision to which letter the combination occurring in the speech of parameters of the speech signal and of the reflection signal corresponds to.

10. The method of claim 1 , further comprising:

synchronously registering speech and registering a degree of light reflected by the mouth in the reflection signal.

11. The method of claim 1 , further comprising:

person recognition based on recognition from the speech signal and the reflection signal of viseme and phoneme combinations.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 20, 2020
From: MOSSINKOFF, OLAF PETRUS QUIRINUS; MEIJER, JOHANNES LEONARDUS JOZEF
To: IEBM B.V.
Reel/Frame 053549/0089 →
Priority Claims (2)
NL 2020358 · Jan 31, 2018 · national
NL 2021041 · Jun 1, 2018 · national
Continuity (1)
Related Publication 20200357407A1 · Nov 12, 2020
Cited By (7)
US 12,204,627 US 12,205,595 US 12,216,749 US 12,216,750 US 12,254,882 US 12,340,808 US 12,505,190