IP Library Granted Patent US 10,437,934
Granted Patent B2
US 10,437,934 · App. 15/650,561 · Granted Oct 8, 2019

Translation with conversational overlap

Inventors: Jeffrey Baker (Newbury Park, CA); Sal Gregory Garcia (Camarillo, CA); Paul Anthony Long (Tarzana, CA)
Assignee: Dolby Laboratories Licensing Corporation
G06F17/2836G06F17/275G06F17/289G10L13/08G10L15/26G10L13/00H04R1/1083H04R3/005H04R2201/107
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,437,934
App. No.
15/650,561
Granted
Oct 8, 2019
Kind
B2
Abstract

A plurality of utterances of a first user from the language of the first user is translated into a language of a second user. The confidence scores associated with the translated utterances are compared with a confidence threshold. A predetermined utterance gap is adjusted based on the comparison. The predetermined utterance gap is a duration of time that occurs between utterances.

Claims (44)

1. A method, comprising:

receiving a plurality of utterances of a first person in a first language;

detecting an utterance gap between sequential utterances of the plurality of utterances;

determining, prior to translating an utterance, whether the utterance will be translated by comparing the utterance gap after the utterance is completed to a threshold utterance gap and, if it is determined that the utterance will be translated:

translating the utterance from the first language to a second language to produce a translated utterance;

determining a translation confidence score for the translated utterance;

determining whether the confidence score is greater than or equal to a confidence level;

determining, based on whether the confidence score is greater than or equal to the confidence level, whether the confidence score is great enough to output the translated utterance; and

determining accrued translation confidence scores for a plurality of utterances, wherein the threshold utterance gap is increased if a percentage of the accrued translation confidence scores is less than a confidence threshold.

2. The method of claim 1 , wherein the plurality of utterances include a first utterance of the first person and a second utterance of the first person, wherein the first utterance of the first person and the second utterance of the first person are received at a device associated with second person.

3. The method of claim 2 , wherein the first utterance of the first person and the second utterance of the first person are data associated with spoken utterances transmitted from a device associated with the first person.

4. The method of claim 2 , wherein the first utterance of the first person and the second utterance of the first person are spoken utterances of the first person.

5. The method of claim 1 , wherein the threshold utterance gap is less than a turn threshold duration.

6. The method of claim 1 , further comprising outputting the translated utterance at a device associated with a second person.

7. The method of claim 6 , wherein the device associated with the second person includes a pair of earbuds.

8. The method of claim 7 , wherein the pair of earbuds are configured to occlude a direct sound path associated with the plurality of utterances of the first person by attenuating the plurality of utterances.

9. The method of claim 8 , wherein an amount of attenuation of the plurality of utterances is adjustable.

10. The method of claim 6 , wherein the translated utterance is outputted to appear to come from a predetermined spatial location.

11. The method of claim 6 , wherein the translated utterance is outputted to appear to come from a spatial location of the first person.

12. The method of claim 1 , wherein the threshold utterance gap is adjustable based at least in part on a speech pattern of the first person.

13. The method of claim 1 , wherein the threshold utterance gap is adjustable based at least in part on a cadence of the first person's speech.

14. A system, comprising:

a processor configured for:

receiving a plurality of utterances of a first person in a first language;

detecting an utterance gap between sequential utterances of the plurality of utterances;

determining, prior to translating an utterance, whether the utterance will be translated by comparing the utterance gap after the utterance is completed to a threshold utterance gap and, if the processor determines that the utterance will be translated:

translating the utterance from the first language to a second language to produce a translated utterance;

determining a translation confidence score for the translated utterance;

determining whether the confidence score is greater than or equal to a confidence level;

determining, based on whether the confidence score is greater than or equal to the confidence level, whether the confidence score is great enough to output the translated utterance; and

determining accrued translation confidence scores for a plurality of utterances, wherein the threshold utterance gap is decreased if a percentage of the accrued translation confidence scores is greater than or equal to a confidence threshold and wherein the confidence threshold corresponds with a percentage of translations that are accurate.

15. The system of claim 14 , wherein the threshold utterance gap is increased if a percentage of the accrued translation confidence scores is less than the confidence threshold.

16. The system of claim 14 , wherein the processor is further configured to output the translated plurality of utterances at a device associated with a second person.

17. A computer program product, the computer program product being embodied in a non-transitory computer readable storage medium and comprising computer instructions for:

receiving a plurality of utterances of a first person in a first language;

detecting an utterance gap between sequential utterances of the plurality of utterances;

determining, prior to translating an utterance, whether the utterance will be translated by comparing the utterance gap after the utterance is completed to a threshold utterance gap and, if it is determined that the utterance will be translated:

translating the utterance from the first language to a second language to produce a translated utterance;

determining a translation confidence score for the translated utterance;

determining whether the confidence score is greater than or equal to a confidence level;

determining, based on whether the confidence score is greater than or equal to the confidence level, whether the confidence score is great enough to output the translated utterance; and

determining accrued translation confidence scores for a plurality of utterances, wherein the threshold utterance gap is increased if a percentage of the accrued translation confidence scores is less than a confidence threshold.

18. The system of claim 14 , wherein the processor is further configured for determining whether a maximum condition is satisfied if the processor determines that the confidence score is not great enough to output the translated utterance.

19. The system of claim 18 , wherein the processor is further configured for combining the translated utterance with a subsequent utterance if the processor determines that the maximum condition is not satisfied.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 3, 2018
From: BAKER, JEFFREY; GARCIA, SAL GREGORY; LONG, PAUL ANTHONY
To: DOPPLER LABS, INC.
Reel/Frame 045711/0811 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 23, 2018
From: DOPPLER LABS, INC.
To: DOLBY LABORATORIES LICENSING CORPORATION
Reel/Frame 044703/0475 →
Continuity (2)
Continuation 15277897 · Sep 27, 2016
Related Publication 20180089179A1 · Mar 29, 2018