IP Library Granted Patent US 9,892,115
Granted Patent B2
US 9,892,115 · App. 14/589,658 · Granted Feb 13, 2018

Translation training with cross-lingual multi-media support

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,892,115
App. No.
14/589,658
Granted
Feb 13, 2018
Kind
B2
Abstract

An improved lecture support system integrates multi-media presentation materials with spoken content so that the listener can follow with both the speech and the supporting materials that accompany the presentation to provide additional understanding. Computer-based systems and methods are disclosed for translation of a spoken presentation (e.g., a lecture, a video) along with the accompanying presentation materials. The content of the presentation materials can be used to improve presentation translation, as it extracts supportive material from the presentation materials as they relate to the speech.

Claims (70)

1. A method comprising:

receiving, by an interface of a translation system, speech by a speaker in a first language;

recognizing, by an automatic speech recognition module of the translation system, the received speech;

transcribing the recognized speech in the first language;

translating, by a machine translation module of the translation system, the speech transcription into a second language, the translated speech transcription stored in a file of a database of the translation system;

receiving, by the interface of the translation system, presentation materials associated with the speech in the first language;

extracting text in the first language from the presentation materials;

translating, by the machine translation module, the extracted text into the second language;

generating translated presentation materials in the second language based on the text in the second language;

modifying the file storing the translated speech transcription by importing at least one formatting feature from the original or translated presentation materials into the translated speech transcription; and

displaying, by a display device, the modified translated speech transcription.

2. The method of claim 1 , further comprising modifying an automatic speech recognition language model of the translation system by:

identifying a first unknown word in the extracted text;

generating a pronunciation for the first unknown word; and

modifying an automatic speech recognition language model probability associated with the first unknown word.

3. The method of claim 2 , wherein modifying the automatic speech recognition language model further comprises:

receiving, based on an internet search, materials related to the extracted text;

identifying a second unknown word in the related materials;

generating a pronunciation for the second unknown word; and

modifying an automatic speech recognition language model probability associated with the second unknown word.

4. The method of claim 1 , further comprising modifying a machine translation language model of the translation system by:

identifying a third unknown word in the extracted original text;

receiving, based on an internet search, a translation of the third unknown word; and

adding the translation to the machine translation language model.

5. The method of claim 1 , wherein the original presentation materials comprise slides, the method further comprising generating a time-stamp associated with a transition from a first slide to a second slide.

6. The method of claim 5 , wherein the modification comprises determining a paragraph break in the translated speech based on the time-stamp.

7. The method of claim 5 , wherein the modification comprises inserting a punctuation mark in the transcription based on the time-stamp.

8. The method of claim 1 , wherein the modification comprises:

identifying a mathematical formula in the translated speech; and

generating an associated transcription using mathematical notation.

9. The method of claim 1 , wherein the modification comprises:

identifying a first element in the transcription of the translated speech;

identifying a second element in the translated presentation materials, the second element corresponding to the first element; and

generating a hyperlink between the first element and the second element.

10. The method of claim 1 , wherein the speaker is a user of the translation system.

11. A computer program product for translating a multimedia presentation, the computer program product comprising a non-transitory computer-readable storage medium containing computer program code for:

receiving, by an interface of a translation system, speech by a speaker in a first language;

recognizing, by an automatic speech recognition module of the translation system, the received speech

transcribing the recognized speech in the first language;

translating, by a machine translation module of the translation system, the speech transcription into a second language, the translated speech transcription stored in a file of a database of the translation system;

receiving presentation materials associated with the speech in the first language;

extracting text in the first language from the presentation materials;

translating, by the machine translation module, the extracted text into the second language;

generating translated presentation materials in the second language based on the text in the second language;

modifying the file storing the translated speech transcription by importing at least one formatting feature from the original or translated presentation materials into the translated speech transcription; and

displaying, by a display device, the modified translated speech transcription.

12. The computer program product of claim 11 , further comprising modifying an automatic speech recognition language model of the translation system by:

identifying a first unknown word in the extracted original text;

generating a pronunciation for the first unknown word; and

modifying an automatic speech recognition language model probability associated with the first unknown word.

13. The computer program product of claim 12 , wherein modifying the automatic speech recognition language model further comprises:

receiving, based on an internet search, materials related to the extracted original text;

identifying a second unknown word in the related materials;

generating a pronunciation for the second unknown word; and

modifying an automatic speech recognition language model probability associated with the second unknown word.

14. The computer program product of claim 11 , further comprising modifying a machine translation language model of the translation system by:

identifying a third unknown word in the extracted original text;

receiving, based on an internet search, a translation of the third unknown word; and

adding the translation to the machine translation language model.

15. The computer program product of claim 11 , wherein the original presentation materials comprise slides, the method further comprising generating a time-stamp associated with a transition from a first slide to a second slide.

16. The computer program product of claim 15 , wherein the modification comprises determining a paragraph break in the transcription based on the time-stamp.

17. The computer program product of claim 15 , wherein the modification comprises inserting a punctuation mark in the transcription based on the time-stamp.

18. The computer program product of claim 11 , wherein the modification comprises:

identifying a mathematical formula in the translated speech; and

generating an associated transcription using mathematical notation.

19. The computer program product of claim 11 , wherein the modification comprises:

identifying a first element in the transcription of the translated speech;

identifying a second element in the translated presentation materials, the second element corresponding to the first element; and

generating a hyperlink between the first element and the second element.

20. The computer program product of claim 11 , wherein the speaker is a user of the translation system.

Assignments (2)
CHANGE OF NAME Recorded Nov 18, 2021
From: FACEBOOK, INC.
To: META PLATFORMS, INC.
Reel/Frame 058897/0824 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 16, 2015
From: WAIBEL, ALEXANDER
To: FACEBOOK, INC.
Reel/Frame 037118/0453 →