IP Library › Granted Patent US 12,520,014
Granted Patent B2
US 12,520,014 · App. 19/027,088 · Granted Jan 6, 2026

Systems and methods for generating translated media streams with synthesized laughter

Inventors: Ben Avi Ingel (Binyamina, IL); Ron Zass (Kiryat Tivon, IL)
H04N21/8126G06F40/58G10L13/00G10L13/033G10L13/086G10L13/10H04N21/2668H04N21/458H04N21/4755
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,520,014
App. No.
19/027,088
Granted
Jan 6, 2026
Kind
B2
Abstract

Systems, methods and non-transitory computer readable media for generating media streams are provided. The generation of the media streams involves receiving an input media stream including an individual speaking in a first language; determining from the input media stream a set of speech-related characteristics associated with the individual; generating, using the set of speech-related characteristics associated with the individual, a synthesized voice for the individual; and producing an output media stream comprising for each set of words spoken by the individual in the first language in the input media stream a translation set of words in a target language using the synthesized voice.

Claims (43)

1 . A method for generating media streams, the method comprising:

receiving an input media stream of an individual speaking in a first language;

determining from the input media stream a set of speech-related characteristics associated with the individual;

identifying in the input media stream a set of words spoken by the individual in the first language;

identifying in the input media stream non-verbal vocalizations articulated by the individual;

generating, using the set of speech-related characteristics associated with the individual, a synthesized voice for the individual; and

generating an output media stream that includes, for the identified set of words spoken by the individual in the first language a corresponding translated set of words in a target language using the synthesized voice, and for the identified non-verbal vocalizations articulated by the individual corresponding non-verbal vocalizations using the synthesized voice, wherein the non-verbal vocalizations generated using the synthesized voice include laughter.

2 . The method of claim 1 , wherein the set of speech-related characteristics includes voice characteristics of the individual while speaking the set of words in the first language.

3 . The method of claim 2 , wherein the voice characteristics includes at least one of prosodic characteristics of a voice of the individual, characteristics of a pitch of the voice of the individual, characteristics of a loudness of the voice of the individual, characteristics of an intonation of the voice of the individual, or characteristics of a stress of the voice of the individual.

4 . The method of claim 1 , wherein the set of speech-related characteristics includes articulation characteristics of the individual while speaking the set of words in the first language.

5 . The method of claim 4 , wherein the articulation characteristics includes at least one of characteristics of speech rhythm, characteristics of speech tempo, characteristics of a linguistic tone of a speech, characteristics of pauses within the speech, characteristics of an accent of the speech, characteristics of a language register of the speech, characteristics of a language of the speech, or a form of the speech.

6 . The method of claim 1 , wherein the set of speech-related characteristics includes characteristics of emotional states of the individual while speaking the set of words in the first language.

7 . The method of claim 1 , further comprising determining a desired level of accent to introduce in the synthesized voice, and wherein the output media stream includes articulations in the target language using the synthesized voice with the desired level of accent.

8 . The method of claim 1 , further comprising: processing visual data in the input media stream to identify text written in the first language and including in the output media stream a translation in the target language of the identified text.

9 . The method of claim 1 , further comprising:

causing a display in a graphical user interface (GUI) of information indicative of a plurality of available target languages; and

receiving, via the GUI, a selection of a preferred target language.

10 . The method of claim 1 , further comprising: receiving a user selection to enable selective manipulation of at least one voice characteristic of the synthesized voice.

11 . The method of claim 1 , wherein the non-verbal vocalizations generated using the synthesized voice further include cheering.

12 . The method of claim 1 , wherein the non-verbal vocalizations generated using the synthesized voice further include crying.

13 . The method of claim 1 , further comprising storing data associated with the determined set of speech-related characteristics for future generation of other media streams using the synthesized voice.

14 . The method of claim 1 , wherein the input media stream further includes speech in a second language, and the method further includes analyzing the input media stream to identify first words in the first language and second words in the second language, wherein the output media stream includes a first plurality of words corresponding to the first words articulated using the synthesized voice and a second plurality of words corresponding to the second words articulated using a second synthesized voice.

15 . The method of claim 14 , wherein, in the output media stream, both the first plurality of words and the second plurality of words are spoken in the target language, and the second plurality the second plurality of words is spoken in the target language with an accent of the second language.

16 . The method of claim 1 , wherein the input media stream is associated with a real-time conversation between the individual and at least one other individual, and the method further includes determining the target language based on an identity of the at least one other individual.

17 . The method of claim 1 , wherein the input media stream is associated with a real-time conversation between the individual and at least one other individual, and the method further includes receiving user selection indicative of the target language.

18 . The method of claim 1 , wherein the input media stream is associated with a real-time conversation between the individual and at least one other individual, and the method further includes changing the synthesized voice as the real-time conversation progresses to improve how the individual sounds when speaking the target language.

19 . The method of claim 1 , wherein the input media stream is associated with a real-time conversation between the individual and at least one other individual, and the output media stream is generated in a manner that takes into account a gender of the at least one other individual.

20 . A system for generating media streams, comprising:

a microphone for recording an input media stream of an individual speaking in a first language;

at least one processing device configured to:

determine from the input media stream a set of speech-related characteristics associated with the individual;

identify in the input media stream a set of words spoken by the individual in the first language;

identify in the input media stream non-verbal vocalizations articulated by the individual;

generate, using the set of speech-related characteristics associated with the individual, a synthesized voice for the individual; and

generate an output media stream that includes, for the identified set of words spoken by the individual in the first language a corresponding translated set of words in a target language using the synthesized voice; and

a transmitter for transmitting the output media stream as part of a real-time conversation, and for the identified non-verbal vocalizations articulated by the individual corresponding non-verbal vocalizations using the synthesized voice, wherein the non-verbal vocalizations generated using the synthesized voice include laughter.

21 . A computer program product for generating media streams, the computer program product embodied in a non-transitory computer-readable medium and including instructions that when executed by at least one processor cause the at least one processor to execute a method, the method comprising:

receiving an input media stream of an individual speaking in a first language;

determining from the input media stream a set of speech-related characteristics associated with the individual;

identifying in the input media stream a set of words spoken by the individual in the first language;

identifying in the input media stream non-verbal vocalizations articulated by the individual;

generating, using the set of speech-related characteristics associated with the individual, a synthesized voice for the individual; and

producing an output media stream that includes, for the identified set of words spoken by the individual in the first language a corresponding translated set of words in a target language using the synthesized voice, and for the identified non-verbal vocalizations articulated by the individual corresponding non-verbal vocalizations using the synthesized voice, wherein the non-verbal vocalizations generated using the synthesized voice include laughter.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 26, 2026
From: INGEL, BEN AVI; ZASS, RON
To: VIDUBLY LTD
Reel/Frame 075094/0907 →
Continuity (7)
Continuation 18643486 · Apr 23, 2024
Continuation 18097900 · Jan 17, 2023
Continuation 17460644 · Aug 30, 2021
Continuation 16813984 · Mar 10, 2020
Provisional Application 62822856 · Mar 23, 2019
Provisional Application 62816137 · Mar 10, 2019
Related Publication 20250168465A1 · May 22, 2025
References Cited (155)
US 4305131A · Best · 1981 [cited by applicant]
US 6077085A · Parry · 2000 [cited by applicant]
US 6778252B2 · Moulton et al. · 2004 [cited by applicant]
US 8140322B2 · Simonsen et al. · 2012 [cited by applicant]
US 9280973B1 · Soyannwo et al. · 2016 [cited by applicant]
US 9747282B1 · Baker et al. · 2017 [cited by applicant]
US 9864933B1 · Cosic · 2018 [cited by applicant]
US 10360716B1 · van der Meulen et al. · 2019 [cited by applicant]
US 10423999B1 · Doctor · 2019 [cited by applicant]
US 10433052B2 · Zass et al. · 2019 [cited by applicant]
US 10467792B1 · Roche · 2019 [cited by applicant]
US 10516938B2 · Zass · 2019 [cited by applicant]
US 10607134B1 · Cosic · 2020 [cited by applicant]
US 10827024B1 · Schissel et al. · 2020 [cited by applicant]
US 11024194B1 · Beigman Klebanov · 2021 [cited by applicant]
US 11140459B2 · Ingel · 2021 [cited by applicant]
US 11159597B2 · Ingel · 2021 [cited by applicant]
US 11195542B2 · Zass · 2021 [cited by applicant]
US 11202131B2 · Zass · 2021 [cited by applicant]
US 11232645B1 · Roche et al. · 2022 [cited by applicant]
US 11244385B1 · Fraser · 2022 [cited by applicant]
US 11520079B2 · Zass · 2022 [cited by applicant]
US 11837249B2 · Zass · 2023 [cited by applicant]
US 11966688B1 · Ehrlich · 2024 [cited by applicant]
US 12010399B2 · Ingel et al. · 2024 [cited by applicant]
US 20020087317A1 · Lee et al. · 2002 [cited by applicant]
US 20020161578A1 · Saindon et al. · 2002 [cited by applicant]
US 20020161579A1 · Saindon et al. · 2002 [cited by applicant]
US 20040068410A1 · Mohamed et al. · 2004 [cited by applicant]
US 20040172257A1 · Liqin et al. · 2004 [cited by applicant]
US 20040186712A1 · Coles et al. · 2004 [cited by applicant]
US 20050255431A1 · Baker · 2005 [cited by applicant]
US 20050272013A1 · Knight · 2005 [cited by applicant]
US 20060285654A1 · Nesvadba et al. · 2006 [cited by applicant]
US 20070124166A1 · Van Luchene · 2007 [cited by applicant]
US 20070130529A1 · Shrubsole · 2007 [cited by applicant]
US 20070208569A1 · Subramanian et al. · 2007 [cited by applicant]
US 20070220575A1 · Cooper et al. · 2007 [cited by applicant]
US 20080015968A1 · Van Luchene · 2008 [cited by applicant]
US 20080195386A1 · Proidl · 2008 [cited by applicant]
US 20080295130A1 · Worthen · 2008 [cited by applicant]
US 20090037179A1 · Liu et al. · 2009 [cited by applicant]
US 20090175596A1 · Hirai · 2009 [cited by applicant]
US 20100082326A1 · Bangalore · 2010 [cited by applicant]
US 20100100907A1 · Chang · 2010 [cited by applicant]
US 20100238179A1 · Kelly · 2010 [cited by applicant]
US 20110064388A1 · Brown et al. · 2011 [cited by applicant]
US 20110076992A1 · Chou · 2011 [cited by applicant]
US 20120054619A1 · Spooner et al. · 2012 [cited by applicant]
US 20130038737A1 · Yehezkel et al. · 2013 [cited by applicant]
US 20130110513A1 · Jhunja et al. · 2013 [cited by applicant]
US 20130188862A1 · Lievens · 2013 [cited by applicant]
US 20140142918A1 · Dotterer et al. · 2014 [cited by applicant]
US 20140164507A1 · Tesch et al. · 2014 [cited by applicant]
US 20140303958A1 · Lee et al. · 2014 [cited by applicant]
US 20140358518A1 · Wu et al. · 2014 [cited by applicant]
US 20150092007A1 · Koborita et al. · 2015 [cited by applicant]
US 20150319518A1 · Wilson · 2015 [cited by applicant]
US 20150356967A1 · Byron et al. · 2015 [cited by applicant]
US 20160005436A1 · Axen et al. · 2016 [cited by applicant]
US 20160021334A1 · Rossano et al. · 2016 [cited by applicant]
US 20160042766A1 · Kummer · 2016 [cited by applicant]
US 20160049146A1 · Chang · 2016 [cited by applicant]
US 20160132578A1 · Allen · 2016 [cited by applicant]
US 20160254795A1 · Ballard · 2016 [cited by applicant]
US 20160328391A1 · Choi · 2016 [cited by applicant]
US 20160365087A1 · Freud · 2016 [cited by applicant]
US 20170004820A1 · Lv et al. · 2017 [cited by applicant]
US 20170011745A1 · Navaratnam · 2017 [cited by applicant]
US 20170075877A1 · Lepeltier · 2017 [cited by applicant]
US 20170076749A1 · Kanevsky · 2017 [cited by applicant]
US 20170116186A1 · Usami · 2017 [cited by examiner]
US 20170255616A1 · Yun et al. · 2017 [cited by applicant]
US 20180143809A1 · Zang et al. · 2018 [cited by applicant]
US 20180174577A1 · Jothilingam et al. · 2018 [cited by applicant]
US 20180174595A1 · Dirac et al. · 2018 [cited by applicant]
US 20180240458A1 · Zass · 2018 [cited by applicant]
US 20180253992A1 · Koul et al. · 2018 [cited by applicant]
US 20180260448A1 · Osotio et al. · 2018 [cited by applicant]
US 20180322875A1 · Adachi · 2018 [cited by applicant]
US 20180357215A1 · Austin et al. · 2018 [cited by applicant]
US 20180374461A1 · Serletic · 2018 [cited by applicant]
US 20190065478A1 · Tsujikawa et al. · 2019 [cited by applicant]
US 20190114302A1 · Bequet · 2019 [cited by applicant]
US 20190164533A1 · Lawrenson et al. · 2019 [cited by applicant]
US 20190166176A1 · Jain et al. · 2019 [cited by applicant]
US 20190250891A1 · Kumar et al. · 2019 [cited by applicant]
US 20190317739A1 · Turek et al. · 2019 [cited by applicant]
US 20190354592A1 · Musham et al. · 2019 [cited by applicant]
US 20200005796A1 · Ziv et al. · 2020 [cited by applicant]
US 20200007946A1 · Olkha · 2020 [cited by applicant]
US 20200042601A1 · Doggett · 2020 [cited by applicant]
US 20200043112A1 · Brinkley, II · 2020 [cited by applicant]
US 20200051566A1 · Shin · 2020 [cited by applicant]
US 20200058289A1 · Gabryjelski et al. · 2020 [cited by applicant]
US 20200066304A1 · Chen · 2020 [cited by applicant]
US 20200105245A1 · Gupta · 2020 [cited by applicant]
US 20200111474A1 · Kumar et al. · 2020 [cited by applicant]
US 20200143813A1 · Nakagawa · 2020 [cited by applicant]
US 20200150981A1 · Westberg et al. · 2020 [cited by applicant]
US 20200159862A1 · Kleiner et al. · 2020 [cited by applicant]
US 20200174776A1 · Vinod et al. · 2020 [cited by applicant]
US 20200221176A1 · Hwang et al. · 2020 [cited by applicant]
US 20200265829A1 · Liu · 2020 [cited by examiner]
US 20200296510A1 · Li et al. · 2020 [cited by applicant]
US 20200311120A1 · Zhao et al. · 2020 [cited by applicant]
US 20210019373A1 · Freitag · 2021 [cited by applicant]
US 20210042110A1 · Basyrov et al. · 2021 [cited by applicant]
US 20210063363A1 · Kaminski et al. · 2021 [cited by applicant]
US 20210081101A1 · Speare et al. · 2021 [cited by applicant]
US 20210097976A1 · Chicote et al. · 2021 [cited by applicant]
US 20210182468A1 · Co et al. · 2021 [cited by applicant]
US 20210192824A1 · Chen · 2021 [cited by applicant]
US 20210224319A1 · Ingel · 2021 [cited by applicant]
US 20210225365A1 · Sinha et al. · 2021 [cited by applicant]
US 20210232759A1 · Schick et al. · 2021 [cited by applicant]
US 20210264369A1 · Zass · 2021 [cited by applicant]
US 20210271815A1 · Li et al. · 2021 [cited by applicant]
US 20210279822A1 · Bellaish · 2021 [cited by applicant]
US 20210287150A1 · Zass · 2021 [cited by applicant]
US 20210303318A1 · Raghavan · 2021 [cited by applicant]
US 20210397418A1 · Nikumb et al. · 2021 [cited by applicant]
US 20210400101A1 · Ingel · 2021 [cited by applicant]
US 20220070550A1 · Ingel · 2022 [cited by applicant]
US 20220382524A1 · Ansari et al. · 2022 [cited by applicant]
US 20230048149A1 · Zass · 2023 [cited by applicant]
US 20230049015A1 · Zass · 2023 [cited by applicant]
US 20230052442A1 · Zass · 2023 [cited by applicant]
US 20230057835A1 · Zass · 2023 [cited by applicant]
US 20230069088A1 · Zass · 2023 [cited by applicant]
US 20230095089A1 · Kaliyaperumal et al. · 2023 [cited by applicant]
US 20230115185A1 · Huang et al. · 2023 [cited by applicant]
US 20230252224A1 · Tran · 2023 [cited by applicant]
US 20230409298A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20230418459A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20230418571A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20230418572A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20230418632A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20230418633A1 · Ciminelli et al. · 2023 [cited by applicant]
US 20240055014A1 · Zass · 2024 [cited by applicant]
US 20240086051A1 · Ciminelli et al. · 2024 [cited by applicant]
US 20240220521A1 · Ehrlich · 2024 [cited by applicant]
US 20240220712A1 · Zass · 2024 [cited by applicant]
US 20240220714A1 · Ehrlich · 2024 [cited by applicant]
US 20240256583A1 · Zass · 2024 [cited by applicant]
US 20240256767A1 · Zass · 2024 [cited by applicant]
US 20240265191A1 · Zass · 2024 [cited by applicant]
US 20240265197A1 · Zass · 2024 [cited by applicant]
US 20240273305A1 · Zass · 2024 [cited by applicant]
US 20240276072A1 · Ingel et al. · 2024 [cited by applicant]
CN 102422639A · 2012 [cited by applicant]
EP 1928189A1 · 2008 [cited by applicant]
WO 2017088136A1 · 2017 [cited by applicant]
WO 2019169686A1 · 2019 [cited by applicant]
K. Nurgaliyev et al.; “Improved Multi-user Interaction in a Smart Environment through a Preference-Based Conflict Resolution Virtual Assistant,” Nov. 23, 2017 International Conference on Intelligent Environments (IE), 2… [cited by applicant]