IP Library Patent Application 11317392
Patent Application
App. No. 11/317,392

Multi dimensional confidence

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
11/317,392
Abstract

A method for managing interactive dialog between a machine and a user is claimed. In one embodiment, an interaction between the machine and the user is managed in response to a confidence value, wherein the confidence value is dependent upon speech recognition confidence and at least one non-acoustic confidence value. The non-acoustic confidence values can be turn-taking confidence, speech duration confidence, state-completion confidence or mode confidence. In multiple embodiments, the non-acoustic confidence value can be dependent upon a timing position of the possible speech onset. The non-acoustic confidence value can be dependent upon a duration of an audio input. The non-acoustic confidence value can be dependent upon a model of user attention. The non-acoustic confidence value can be dependent upon a history of exit conditions associated with interactions during a course of a session.

Claims (28)

1 . A method for managing interactive dialog between a machine and a user comprising:

verbalizing at least one desired sequence of one or more spoken phrases;

enabling a user to hear the at least one desired sequence of one or more spoken phrases;

receiving audio input from the user or an environment of the user; and

managing an interaction between the at least one desired sequence of spoken phrases and the audio input in response to at least one confidence value, wherein the confidence value is dependent upon speech recognition confidence and at least one non-acoustic confidence value.

2 . The method of claim 1 , further comprising determining a timing position of a possible speech onset from the audio input, wherein the at least one non-acoustic confidence value is dependent upon the timing position of the possible speech onset.

3 . The method of claim 1 , wherein the non-acoustic measure of confidence is turn-taking confidence.

4 . The method of claim 1 , wherein the at least one non-acoustic confidence value is dependent upon speech duration confidence.

5 . The method of claim 1 , wherein the at least one non-acoustic confidence value is dependent upon state-completion confidence.

6 . The method of claim 1 , wherein the at least one non-acoustic confidence value is dependent upon mode confidence.

7 . The method of claim 1 , wherein the at least one confidence value is dependent upon a model that is dependent on a history of exit conditions associated with a plurality of interactions during a course of a session.

8 . The method of claim 1 , further comprising determining a duration of audio input, wherein the at least one non-acoustic confidence value is dependent upon the duration of audio input.

9 . The method of claim 1 , wherein the at least one confidence value is dependent upon a model of user attention.

10 . The method of claim 1 , wherein the at least one confidence value is further dependent upon a plurality of non-acoustic confidence values.

11 . The method of claim 1 , wherein the at least one confidence value comprises a plurality of discrete states, wherein at least two discrete states of the plurality of discrete states are associated with a different confidence value.

12 . A system for managing interactive dialog between a machine and a user comprising:

means for verbalizing at least one desired sequence of one or more spoken phrases;

means for enabling a user to hear the at least one desired sequence of one or more spoken phrases;

means for receiving audio input from the user or an environment of the user; and

means for managing an interaction between the at least one desired sequence of spoken phrases and the audio input in response to at least one confidence value, wherein the confidence value is dependent upon speech recognition confidence and at least one non-acoustic confidence value.

13 . The system of claim 12 , further comprising means for determining a timing position of a possible speech onset from the audio input, wherein the at least one non-acoustic confidence value is dependent upon the timing position of the possible speech onset.

14 . The system of claim 12 , wherein the non-acoustic measure of confidence is turn-taking confidence.

15 . The system of claim 12 , wherein the at least one non-acoustic confidence value is dependent upon speech duration confidence.

16 . The system of claim 12 , wherein the at least one non-acoustic confidence value is dependent upon state-completion confidence.

17 . The system of claim 12 , wherein the at least one non-acoustic confidence value is dependent upon mode confidence.

18 . The system of claim 12 , wherein the at least one confidence value is dependent upon a model that is dependent on a history of exit conditions associated with a plurality of interactions during a course of a session.

19 . The system of claim 12 , further comprising means for determining a duration of audio input, wherein the at least one non-acoustic confidence value is dependent upon the duration of audio input.

20 . The system of claim 12 , wherein the at least one confidence value is dependent upon a model of user attention.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2006
From: ATTWATER, DAVID; BALENTINE, BRUCE
To: ENTERPRISE INTEGRATION GROUP, INC.
Reel/Frame 018051/0750 →