IP Library Granted Patent US 9,502,033
Granted Patent B2
US 9,502,033 · App. 14/627,560 · Granted Nov 22, 2016

Distributed speech recognition using one way communication

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,502,033
App. No.
14/627,560
Granted
Nov 22, 2016
Kind
B2
Abstract

A speech recognition client sends a speech stream and control stream in parallel to a server-side speech recognizer over a network. The network may be an unreliable, low-latency network. The server-side speech recognizer recognizes the speech stream continuously. The speech recognition client receives recognition results from the server-side recognizer in response to requests from the client. The client may remotely reconfigure the state of the server-side recognizer during recognition.

Claims (12)

1. A method performed by at least one computer processor executing computer program instructions stored on at least one non-transitory computer-readable medium, the method comprising:

(A) receiving a speech stream and a control stream from a client, the speech stream including a minimum configuration state identification number required to begin recognition of a first portion of the speech stream from a client;

(B) determining whether a configuration state identification number associated with a state of an automatic speech recognition engine is at least as great as the received minimum configuration state identification number;

(C) if the configuration state identification number associated with the state of the automatic speech recognition engine is determined to be at least as great as the received minimum configuration state identification number, then using the automatic speech recognition engine to recognize the first portion of the speech stream and thereby to produce a first speech recognition result; and

(D) if the configuration state identification number associated with the state of the automatic speech recognition engine is not determined to be at least as great as the received minimum configuration state identification number, then incrementing the configuration state identification number associated with the state of the automatic speech recognition engine until the configuration state identification number associated with the state of the automatic speech recognition engine is determined to be at least as great as the received minimum configuration state identification number before using the automatic speech recognition engine to recognize the first portion of the speech stream and thereby to produce the first speech recognition result.

2. The method of claim 1 , wherein (A) further comprises receiving, within the speech stream, a tag comprising the indication of the minimum configuration state identification number.

3. A non-transitory computer-readable medium comprising computer program instructions stored thereon, wherein the computer program instructions are executable by at least one processor to perform a method, the method comprising:

(A) receiving a speech stream and a control stream from a client, the speech stream including a minimum configuration state identification number required to begin recognition of a first portion of the speech stream from a client;

(B) determining whether a configuration state identification number associated with a state of an automatic speech recognition engine is at least as great as the received minimum configuration state identification number; and

(C) if the configuration state identification number associated with the state of the automatic speech recognition engine is determined to be at least as great as the received minimum configuration state identification number, then using the automatic speech recognition engine to recognize the first portion of the speech stream and thereby to produce a first speech recognition result; and

(D) if the configuration state identification number associated with the state of the automatic speech recognition engine is not determined to be at least as great as the received minimum configuration state identification number, then incrementing the configuration state identification number associated with the state of the automatic speech recognition engine until the configuration state identification number associated with the state of the automatic speech recognition engine is determined to be at least as great as the received minimum configuration state identification number before using the automatic speech recognition engine to recognize the first portion of the speech stream and thereby to produce the first speech recognition result.

4. The non-transitory computer-readable medium of claim 3 , wherein (A) further comprises receiving, within the speech stream, a tag comprising the indication of the minimum configuration state identification number.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 1, 2024
From: 3M INNOVATIVE PROPERTIES COMPANY
To: SOLVENTUM INTELLECTUAL PROPERTIES COMPANY
Reel/Frame 066435/0347 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 22, 2021
From: MMODAL IP LLC
To: 3M INNOVATIVE PROPERTIES COMPANY
Reel/Frame 057883/0129 →
CHANGE OF ADDRESS Recorded Apr 14, 2017
From: MMODAL IP LLC
To: MMODAL IP LLC
Reel/Frame 042271/0858 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 27, 2016
From: CARRAUX, ERIC; KOLL, DETLEF
To: MMODAL IP LLC
Reel/Frame 039560/0149 →