IP Library › Granted Patent US 12,518,774
Granted Patent B2
US 12,518,774 · App. 18/105,847 · Granted Jan 6, 2026

Identifying optimal articulatory event-types for computer analysis of speech

Inventors: Raziel Haimi-Cohen (Springfield, NJ); Ilan D. Shallom (Gedera, IL)
Assignee: Cordio Medical Ltd.
G10L25/30G10L25/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,518,774
App. No.
18/105,847
Filed
Feb 5, 2023
Granted
Jan 6, 2026
Kind
B2
Art Unit
2654
USPC
704/200
Abstract

A method includes, based on one or more representations of an articulatory event-type, computing, by a processor, a score quantifying an estimated degree to which an instance of the articulatory event-type indicates a state, with respect to a disease, in which the instance was produced. The method further includes storing or communicating the score for subsequent use in evaluating the state of a subject based on a test utterance produced by the subject. Other embodiments are also described.

Claims (125)

1 . A system, comprising:

a memory configured to store program instructions; and

a processor, configured to:

load the program instructions from the memory, and

by executing the program instructions:

obtain one or more representations of a type of sound produced by articulatory organs,

compute, based on the representations, a score quantifying an estimated degree to which an instance of the type of sound, when produced by someone having a disease, indicates a state, with respect to the disease, in which the instance was produced, and

store or communicate the score for subsequent use in evaluating the state of a subject with respect to the disease, based on a test utterance produced by the subject.

2 . A method, comprising:

obtaining on one or more representations of a type of sound produced by articulatory organs;

computing, by a processor, based on the representations, a score quantifying an estimated degree to which an instance of the type of sound, when produced by someone having a disease, indicates a state, with respect to the disease, in which the instance was produced; and

storing or communicating the score for subsequent use in evaluating the state of a subject with respect to the disease, based on a test utterance produced by the subject.

3 . The method according to claim 2 , wherein the type of sound includes one or more phonemes.

4 . The method according to claim 2 , wherein the score quantifies the estimated degree to which any instance of the type of sound indicates the state.

5 . The method according to claim 4 , wherein the representations are respective segments of one or more speech samples.

6 . The method according to claim 5 ,

wherein the speech samples are first-state speech samples produced while in a first state with respect to the disease, and the segments are first-state segments, and

wherein computing the score comprises computing the score based on:

one or more same-state distances quantifying a same-state similarity of the first-state segments to each other, and

one or more cross-state distances quantifying a cross-state similarity of the first-state segments to one or more second-state segments of at least one second-state speech sample produced while in a second state with respect to the disease.

7 . The method according to claim 6 , wherein computing the score based on the same-state distances and cross-state distances comprises:

computing respective counts for multiple segments, each of which is one of the first-state segments or one of the second-state segments, by, for each of the segments:

identifying, from a set S of distances including those of the same-state distances associated with the segment and those of the cross-state distances associated with the segment, a subset S′, which includes, for a positive integer q, q smallest ones of the distances, and

computing the count for the segment as (i) a number ν of the distances in S′ that are same-state distances, or (ii) q−ν; and

computing the score based on the counts.

8 . The method according to claim 6 , wherein computing the score based on the same-state distances and cross-state distances comprises computing the score by comparing a same-state statistic of the same-state distances to a cross-state statistic of the cross-state distances.

9 . The method according to claim 5 ,

wherein the speech samples are first-state speech samples produced while in a first state with respect to the disease, and the segments are first-state segments, and

wherein computing the score comprises computing the score based on (i) multiple same-state distances between respective ones of the first-state segments and a first-state model representing the type of sound as produced while in the first state, and (ii) multiple cross-state distances between respective ones of the first-state segments and a second-state model representing the type of sound as produced while in the second state.

10 . The method according to claim 5 , wherein computing the score comprises:

obtaining respective outputs for the segments from a discriminator, each of the outputs estimating, for a different respective one of the segments, the state in which the segment was produced; and

computing the score based on a measure of accuracy of the outputs.

11 . The method according to claim 5 , wherein computing the score comprises computing the score based on respective neuronal outputs produced by a neural-network discriminator in response to evaluating, based on each of the segments, the state in which the segment was produced.

12 . The method according to claim 11 , wherein computing the score based on the neuronal outputs comprises:

computing respective score-components for the segments, each of the score-components having (i) a magnitude that is an increasing function of the neuronal output produced by the neural-network discriminator in response to the segment to which the score-component belongs, and (ii) a sign that depends on whether the evaluation, based on the segment to which the score-component belongs, is correct; and

computing the score based on the score-components.

13 . The method according to claim 4 , wherein the representations include a first-state model representing the type of sound as produced while in a first state with respect to the disease.

14 . The method according to claim 13 , wherein computing the score comprises computing the score based on a distance between the first-state model and a second-state model representing the type of sound as produced while in a second state with respect to the disease.

15 . The method according to claim 14 ,

wherein the first-state model and second-state model represent the type of sound as produced by one or more other subjects,

wherein the distance is a first distance, and

wherein computing the score comprises computing the score based on a second distance between:

a subject-specific model representing the type of sound as produced by the subject while in the first state, and

the first-state model or the second-state model.

16 . The method according to claim 13 , wherein the first-state model represents the type of sound as produced by the subject, and wherein computing the score comprises computing the score based on:

a same-state distance quantifying a same-state similarity of the first-state model to another first-state model representing the type of sound as produced by one or more other subjects while in the first state, and

a cross-state distance quantifying a cross-state similarity of the first-state model to a second-state model representing the type of sound as produced by one or more other subjects while in the second state.

17 . The method according to claim 4 ,

wherein the type of sound is a target-language type of sound in a target language,

wherein the score is a target-language score, and

wherein computing the target-language score comprises:

inputting one or more instances of the target-language type of sound to a tool configured to facilitate computing a source-language score for any instance of a source-language type of sound in a source language different from the target language, the source-language score quantifying another estimated degree to which the instance of the source-language type of sound indicates the state, with respect to the disease, in which the instance of the source-language type of sound was produced, and

computing the target-language score based on computations performed by the tool in response to the inputting.

18 . The method according to claim 17 ,

wherein the tool includes a neural network trained to predict an accuracy with which a discriminator would estimate the state based on the instance of the source-language type of sound, and

wherein computing the target-language score comprises computing the target-language score based on respective predicted accuracies output by the neural network in response to the instances of the target-language type of sound.

19 . The method according to claim 17 ,

wherein the test utterance is a target-language test utterance in the target language,

wherein the tool includes a neural-network discriminator configured to process a source-language test utterance in the source language so as to evaluate the state in which the source-language test utterance was produced, and

wherein computing the target-language score comprises computing the target-language score based on neuronal outputs produced by the neural-network discriminator in response to processing the instances of the target-language type of sound.

20 . The method according to claim 2 , wherein the representations include a segment of a speech sample representing the instance, which speech sample was produced by the subject.

21 . The method according to claim 20 , further comprising:

obtaining respective outputs, from a discriminator, for multiple training instances of the type of sound, each of the outputs estimating the state in which a different respective one of the training instances was produced; and

using the training instances, the outputs, and respective actual states in which the training instances were produced, training a neural network to predict an accuracy of the discriminator for any instance of the type of sound,

wherein computing the score comprises computing the score based on the accuracy predicted for the segment by the neural network.

22 . The method according to claim 20 , wherein computing the score comprises computing the score based on a neuronal output produced by a neural-network discriminator in response to processing the segment.

23 . A computer software product comprising a tangible non-transitory computer-readable medium in which program instructions are stored, which instructions, when read by a processor, cause the processor to:

obtain one or more representations of a type of sound produced by articulatory organs,

compute, based on the representations, a score quantifying an estimated degree to which an instance of the type of sound, when produced by someone having a disease, indicates a state, with respect to the disease, in which the instance was produced, and

store or communicate the score for subsequent use in evaluating the state of a subject with respect to the disease, based on a test utterance produced by the subject.

24 . The computer software product according to claim 23 , wherein the type of sound includes one or more phonemes.

25 . The computer software product according to claim 23 , wherein the score quantifies the estimated degree to which any instance of the type of sound indicates the state.

26 . The computer software product according to claim 25 , wherein the representations are respective segments of one or more speech samples.

27 . The computer software product according to claim 26 ,

wherein the speech samples are first-state speech samples produced while in a first state with respect to the disease, and the segments are first-state segments, and

wherein the instructions cause the processor to compute the score based on:

one or more same-state distances quantifying a same-state similarity of the first-state segments to each other, and

one or more cross-state distances quantifying a cross-state similarity of the first-state segments to one or more second-state segments of at least one second-state speech sample produced while in a second state with respect to the disease.

28 . The computer software product according to claim 27 , wherein the instructions cause the processor to compute the score based on the same-state distances and cross-state distances by:

computing respective counts for multiple segments, each of which is one of the first-state segments or one of the second-state segments, by, for each of the segments:

identifying, from a set S of distances including those of the same-state distances associated with the segment and those of the cross-state distances associated with the segment, a subset S′, which includes, for a positive integer q, q smallest ones of the distances, and

computing the count for the segment as (i) a number ν of the distances in S′ that are same-state distances, or (ii) q−ν, and

computing the score based on the counts.

29 . The computer software product according to claim 27 , wherein the instructions cause the processor to compute the score based on the same-state distances and cross-state distances by comparing a same-state statistic of the same-state distances to a cross-state statistic of the cross-state distances.

30 . The computer software product according to claim 26 ,

wherein the speech samples are first-state speech samples produced while in a first state with respect to the disease, and the segments are first-state segments, and

wherein the instructions cause the processor to compute the score based on (i) multiple same-state distances between respective ones of the first-state segments and a first-state model representing the type of sound as produced while in the first state, and (ii) multiple cross-state distances between respective ones of the first-state segments and a second-state model representing the type of sound as produced while in the second state.

31 . The computer software product according to claim 26 , wherein the instructions cause the processor to compute the score by:

obtaining respective outputs for the segments from a discriminator, each of the outputs estimating, for a different respective one of the segments, the state in which the segment was produced, and

computing the score based on a measure of accuracy of the outputs.

32 . The computer software product according to claim 26 , wherein the instructions cause the processor to compute the score based on respective neuronal outputs produced by a neural-network discriminator in response to evaluating, based on each of the segments, the state in which the segment was produced.

33 . The computer software product according to claim 32 , wherein the instructions cause the processor to compute the score based on the neuronal outputs by:

computing respective score-components for the segments, each of the score-components having (i) a magnitude that is an increasing function of the neuronal output produced by the neural-network discriminator in response to the segment to which the score-component belongs, and (ii) a sign that depends on whether the evaluation, based on the segment to which the score-component belongs, is correct, and

computing the score based on the score-components.

34 . The computer software product according to claim 25 , wherein the representations include a first-state model representing the type of sound as produced while in a first state with respect to the disease.

35 . The computer software product according to claim 34 , wherein the instructions cause the processor to compute the score based on a distance between the first-state model and a second-state model representing the type of sound as produced while in a second state with respect to the disease.

36 . The computer software product according to claim 35 ,

wherein the first-state model and second-state model represent the type of sound as produced by one or more other subjects,

wherein the distance is a first distance, and

wherein the instructions cause the processor to compute the score based on a second distance between:

a subject-specific model representing the type of sound as produced by the subject while in the first state, and

the first-state model or the second-state model.

37 . The computer software product according to claim 34 , wherein the first-state model represents the type of sound as produced by the subject, and wherein the instructions cause the processor to compute the score based on:

a same-state distance quantifying a same-state similarity of the first-state model to another first-state model representing the type of sound as produced by one or more other subjects while in the first state, and

a cross-state distance quantifying a cross-state similarity of the first-state model to a second-state model representing the type of sound as produced by one or more other subjects while in the second state.

38 . The computer software product according to claim 25 ,

wherein the type of sound is a target-language type of sound in a target language,

wherein the score is a target-language score, and

wherein the instructions cause the processor to compute the target-language score by:

inputting one or more instances of the target-language type of sound to a tool configured to facilitate computing a source-language score for any instance of a source-language type of sound in a source language different from the target language, the source-language score quantifying another estimated degree to which the instance of the source-language type of sound indicates the state, with respect to the disease, in which the instance of the source-language type of sound was produced, and

computing the target-language score based on computations performed by the tool in response to the inputting.

39 . The computer software product according to claim 38 ,

wherein the tool includes a neural network trained to predict an accuracy with which a discriminator would estimate the state based on the instance of the source-language type of sound, and

wherein the instructions cause the processor to compute the target-language score based on respective predicted accuracies output by the neural network in response to the instances of the target-language type of sound.

40 . The computer software product according to claim 38 ,

wherein the test utterance is a target-language test utterance,

wherein the tool includes a neural-network discriminator configured to process a source-language test utterance in the source language so as to evaluate the state in which the source-language test utterance was produced, and

wherein the instructions cause the processor to compute the target-language score based on neuronal outputs produced by the neural-network discriminator in response to processing the instances of the target-language type of sound.

41 . The computer software product according to claim 23 , wherein the representations include a segment of a speech sample representing the instance, which speech sample was produced by the subject.

42 . The computer software product according to claim 41 ,

wherein the instructions further cause the processor to:

obtain respective outputs, from a discriminator, for multiple training instances of the type of sound, each of the outputs estimating the state in which a different respective one of the training instances was produced, and

using the training instances, the outputs, and respective actual states in which the training instances were produced, train a neural network to predict an accuracy of the discriminator for any instance of the type of sound, and

wherein the instructions cause the processor to compute the score based on the accuracy predicted for the segment by the neural network.

43 . The computer software product according to claim 41 , wherein the instructions cause the processor to compute the score based on a neuronal output produced by a neural-network discriminator in response to processing the segment.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 6, 2023
From: HAIMI-COHEN, RAZIEL; SHALLOM, ILAN D.
To: CORDIO MEDICAL LTD.
Reel/Frame 062593/0941 →
Continuity (1)
Related Publication 20240265937A1 · Aug 8, 2024
References Cited (277)
US 4292471A · Kuhn et al. · 1981 [cited by applicant]
US 4838275A · Lee · 1989 [cited by applicant]
US 5853005A · Scanlon · 1998 [cited by applicant]
US 5864810A · Digalakis et al. · 1999 [cited by applicant]
US 6006188A · Bogdashevsky et al. · 1999 [cited by applicant]
US 6168568B1 · Gavriely · 2001 [cited by applicant]
US 6241683B1 · Macklem et al. · 2001 [cited by applicant]
US 6289313B1 · Heinonen et al. · 2001 [cited by applicant]
US 6389393B1 · Gong · 2002 [cited by applicant]
US 6396416B1 · Kuusela et al. · 2002 [cited by applicant]
US 6527729B1 · Turcott · 2003 [cited by applicant]
US 6600949B1 · Turcott · 2003 [cited by applicant]
US 7092874B2 · Clavbo · 2006 [cited by applicant]
US 7225013B2 · Geva et al. · 2007 [cited by applicant]
US 7226422B2 · Hatlestsad et al. · 2007 [cited by applicant]
US 7267652B2 · Coyle et al. · 2007 [cited by applicant]
US 7283962B2 · Meyerhoif et al. · 2007 [cited by applicant]
US 7363226B2 · Shiomi et al. · 2008 [cited by applicant]
US 7398213B1 · Levanon et al. · 2008 [cited by applicant]
US 7457753B2 · Moran et al. · 2008 [cited by applicant]
US 7529670B1 · Michaelis · 2009 [cited by applicant]
US 7762264B1 · Raming et al. · 2010 [cited by applicant]
US 8478596B2 · Schultz · 2013 [cited by applicant]
US 8591430B2 · Amurthur et al. · 2013 [cited by applicant]
US 8684900B2 · Tran · 2014 [cited by applicant]
US 8689606B2 · Schellekens et al. · 2014 [cited by applicant]
US 8784311B2 · Shrivastav et al. · 2014 [cited by applicant]
US 8903725B2 · Pilz · 2014 [cited by applicant]
US 9070357B1 · Kennedy et al. · 2015 [cited by applicant]
US 9138167B1 · Leydon · 2015 [cited by applicant]
US 9153231B1 · Salvador et al. · 2015 [cited by applicant]
US 9445763B2 · Davis et al. · 2016 [cited by applicant]
US 9492096B2 · Brockway et al. · 2016 [cited by applicant]
US 9579056B2 · Rosenbek et al. · 2017 [cited by applicant]
US 9685174B2 · Karam et al. · 2017 [cited by applicant]
US 9922641B1 · Chun · 2018 [cited by applicant]
US 10311980B2 · Kim et al. · 2019 [cited by applicant]
US 10796205B2 · Shi et al. · 2020 [cited by applicant]
US 10796805B2 · Lotan et al. · 2020 [cited by applicant]
US 10847177B2 · Shallom · 2020 [cited by applicant]
US 10896765B2 · Kim et al. · 2021 [cited by applicant]
US 10991384B2 · Eyben · 2021 [cited by examiner]
US 11011188B2 · Shallom · 2021 [cited by applicant]
US 11024327B2 · Shallom · 2021 [cited by applicant]
US 11417342B2 · Shallom · 2022 [cited by applicant]
US 11484211B2 · Shallom · 2022 [cited by applicant]
US 11538490B2 · Shallom · 2022 [cited by applicant]
US 11610600B2 · Shallom · 2023 [cited by applicant]
US 12046238B2 · Khaleghi · 2024 [cited by applicant]
US 20020059029A1 · Todder et al. · 2002 [cited by applicant]
US 20030115054A1 · Iso-Sipila · 2003 [cited by applicant]
US 20030220790A1 · Kepuska · 2003 [cited by applicant]
US 20040097822A1 · Muz et al. · 2004 [cited by applicant]
US 20050038635A1 · Klefenz et al. · 2005 [cited by applicant]
US 20050060153A1 · Gable et al. · 2005 [cited by applicant]
US 20060058697A1 · Mochizuki et al. · 2006 [cited by applicant]
US 20060116878A1 · Nagamine · 2006 [cited by applicant]
US 20060167385A1 · Guion · 2006 [cited by applicant]
US 20060293609A1 · Stahmann et al. · 2006 [cited by applicant]
US 20070005357A1 · Moran et al. · 2007 [cited by applicant]
US 20070100623A1 · Hentschel et al. · 2007 [cited by applicant]
US 20070225975A1 · Imoto · 2007 [cited by applicant]
US 20070288183A1 · Bulkes et al. · 2007 [cited by applicant]
US 20080013747A1 · Tran · 2008 [cited by applicant]
US 20080275349A1 · Halperin et al. · 2008 [cited by applicant]
US 20090036777A1 · Zhang et al. · 2009 [cited by applicant]
US 20090043586A1 · MacAuslan · 2009 [cited by applicant]
US 20090099848A1 · Lerner et al. · 2009 [cited by applicant]
US 20090227888A1 · Salmi et al. · 2009 [cited by applicant]
US 20090326937A1 · Chitsaz et al. · 2009 [cited by applicant]
US 20100201807A1 · Mcpherson · 2010 [cited by applicant]
US 20110021940A1 · Chu et al. · 2011 [cited by applicant]
US 20110092779A1 · Chang et al. · 2011 [cited by applicant]
US 20110125044A1 · Rhee · 2011 [cited by applicant]
US 20110184250A1 · Schmidt et al. · 2011 [cited by applicant]
US 20120041279A1 · Freeman et al. · 2012 [cited by applicant]
US 20120116186A1 · Shrivastav et al. · 2012 [cited by applicant]
US 20120220899A1 · Oh et al. · 2012 [cited by applicant]
US 20120265024A1 · Shrivastav et al. · 2012 [cited by applicant]
US 20120283598A1 · Horii et al. · 2012 [cited by applicant]
US 20130018274A1 · O'Neill · 2013 [cited by applicant]
US 20130158434A1 · Shen et al. · 2013 [cited by applicant]
US 20130166279A1 · Dines et al. · 2013 [cited by applicant]
US 20130218582A1 · Lalonde · 2013 [cited by applicant]
US 20140005564A1 · Vanovic et al. · 2014 [cited by applicant]
US 20140073993A1 · Poellabauer et al. · 2014 [cited by applicant]
US 20140153794A1 · Varaklis et al. · 2014 [cited by applicant]
US 20140249424A1 · Fan et al. · 2014 [cited by applicant]
US 20140294188A1 · Rini et al. · 2014 [cited by applicant]
US 20140302472A1 · Fletcher · 2014 [cited by applicant]
US 20140314212A1 · Bentley et al. · 2014 [cited by applicant]
US 20150073306A1 · Abeyratne et al. · 2015 [cited by applicant]
US 20150126888A1 · Patel et al. · 2015 [cited by applicant]
US 20150127350A1 · Agiomyrgiannakis · 2015 [cited by applicant]
US 20150216448A1 · Lotan et al. · 2015 [cited by applicant]
US 20150265205A1 · Rosenbek et al. · 2015 [cited by applicant]
US 20160015289A1 · Simon et al. · 2016 [cited by applicant]
US 20160045161A1 · Alshaer et al. · 2016 [cited by applicant]
US 20160081611A1 · Hampton et al. · 2016 [cited by applicant]
US 20160095545A1 · Levanon · 2016 [cited by applicant]
US 20160113618A1 · Su et al. · 2016 [cited by applicant]
US 20160249842A1 · Ohana Lubelchick · 2016 [cited by applicant]
US 20160302003A1 · Rahman et al. · 2016 [cited by applicant]
US 20170069312A1 · Sundararajan et al. · 2017 [cited by applicant]
US 20170084295A1 · Tsiartas et al. · 2017 [cited by applicant]
US 20170262606A1 · Abdullah · 2017 [cited by examiner]
US 20170280239A1 · Sekiya et al. · 2017 [cited by applicant]
US 20170325779A1 · Spina et al. · 2017 [cited by applicant]
US 20170354363A1 · Quatieri et al. · 2017 [cited by applicant]
US 20180004913A1 · Ghasemzadeh et al. · 2018 [cited by applicant]
US 20180108440A1 · Stevens et al. · 2018 [cited by applicant]
US 20180125444A1 · Kahlman et al. · 2018 [cited by applicant]
US 20180214061A1 · Knoth et al. · 2018 [cited by applicant]
US 20180296092A1 · Hassan et al. · 2018 [cited by applicant]
US 20190130910A1 · Kariya et al. · 2019 [cited by applicant]
US 20190221317A1 · Kempanna et al. · 2019 [cited by applicant]
US 20190311815A1 · Kim et al. · 2019 [cited by applicant]
US 20190385711A1 · Shriberg et al. · 2019 [cited by applicant]
US 20200077940A1 · Srivastava et al. · 2020 [cited by applicant]
US 20200098384A1 · Nematihosseinabadi · 2020 [cited by examiner]
US 20200152226A1 · Anushiravani et al. · 2020 [cited by applicant]
US 20200168230A1 · Roh et al. · 2020 [cited by applicant]
US 20210065676A1 · Abrami · 2021 [cited by examiner]
US 20210110894A1 · Shriberg et al. · 2021 [cited by applicant]
US 20210193169A1 · Faizakof et al. · 2021 [cited by applicant]
US 20210256992A1 · Shallom · 2021 [cited by applicant]
US 20220130415A1 · Garrison et al. · 2022 [cited by applicant]
US 20220165295A1 · Shallom · 2022 [cited by applicant]
US 20220328064A1 · Shriberg et al. · 2022 [cited by applicant]
US 20220409063A1 · Shallom · 2022 [cited by applicant]
US 20220415308A1 · Berisha et al. · 2022 [cited by applicant]
US 20230072242A1 · Kim et al. · 2023 [cited by applicant]
US 20230080870A1 · Shallom · 2023 [cited by applicant]
US 20230190177A1 · Shor · 2023 [cited by examiner]
US 20230352013A1 · Khaleghi · 2023 [cited by applicant]
CN 102125427A · 2011 [cited by applicant]
CN 102423262A · 2012 [cited by applicant]
CN 202261466U · 2012 [cited by applicant]
CN 102497472A · 2012 [cited by applicant]
CN 107622797A · 2018 [cited by applicant]
DE 102015218948A1 · 2017 [cited by applicant]
EP 1091689A1 · 2001 [cited by applicant]
EP 1855594A1 · 2007 [cited by applicant]
EP 2124223A1 · 2009 [cited by applicant]
EP 2438863A1 · 2012 [cited by applicant]
EP 3365057A1 · 2018 [cited by applicant]
FR 2672793A1 · 1992 [cited by applicant]
GB 1219618A · 1971 [cited by applicant]
GB 2493458A · 2013 [cited by applicant]
JP 04082538A · 1992 [cited by applicant]
JP 09173320A · 1997 [cited by applicant]
JP 2003044078A · 2003 [cited by applicant]
JP 2004302786A · 2004 [cited by applicant]
JP 2006230548A · 2006 [cited by applicant]
JP 2016006504A · 2016 [cited by applicant]
JP 2017191166A · 2017 [cited by applicant]
JP 6263308B1 · 2018 [cited by applicant]
SE 508439C2 · 1998 [cited by applicant]
WO 03068062A1 · 2003 [cited by applicant]
WO 2005074799A1 · 2005 [cited by applicant]
WO 2006033044A3 · 2006 [cited by applicant]
WO 2006079062A1 · 2006 [cited by applicant]
WO 2010004025A1 · 2010 [cited by applicant]
WO 2010015865A1 · 2010 [cited by applicant]
WO 2010123483A2 · 2010 [cited by applicant]
WO 2012038903A2 · 2012 [cited by applicant]
WO 2012104743A2 · 2012 [cited by applicant]
WO 2013043847A1 · 2013 [cited by applicant]
WO 2013170131A1 · 2013 [cited by applicant]
WO 2014037843A1 · 2014 [cited by applicant]
WO 2014045257A1 · 2014 [cited by applicant]
WO 2014188408A1 · 2014 [cited by applicant]
WO 2016028495A1 · 2016 [cited by applicant]
WO 2017060828A1 · 2017 [cited by applicant]
WO 2017068582A1 · 2017 [cited by applicant]
WO 2017147221A1 · 2017 [cited by applicant]
WO 2018021920A1 · 2018 [cited by applicant]
WO 2019089830A1 · 2019 [cited by applicant]
WO 2019210261A1 · 2019 [cited by applicant]
R. Voleti, J. M. Liss and V. Berisha, “A Review of Automated Speech and Language Features for Assessment of Cognitive and Thought Disorders,” in IEEE Journal of Selected Topics in Signal Processing, vol. 14, No. 2, pp. … [cited by examiner]
Hickey., “App lets you monitor lung health using only a smartphone”, pp. 1-5, Sep. 18, 2012. [cited by applicant]
Gandler et al., “Mobile FEV: Evaluation of iPhone Spirometer”, pp. 1-1, Feb. 14, 2013. [cited by applicant]
Abushakra et al., “Lung capacity estimation through acoustic signal of breath”, 13th IEEE International Conference on BioInformatics and BioEngineering, pp. 386-391, Nov. 11-Nov. 13, 2012. [cited by applicant]
G.P. Imports, Inc., “Spirometer Pro”, pp. 1-3, Jan. 8, 2010. [cited by applicant]
Murton et al., “Acoustic speech analysis of patients with decompensated heart failure: A pilot study”, The Journal of the Acoustical Society of America, vol. 142, Issue 4, pp. 1-28, Oct. 24, 2017. [cited by applicant]
Gillespie et al., “The Effects of Hyper- and Hypocapnia on Phonatory Laryngeal Airway Resistance in Women”, Research Article, Journal of Speech, Language, and 638 Hearing Research, vol. 58, pp. 638-652, Jun. 2015. [cited by applicant]
Wang et al., “Accuracy of perceptual and acoustic methods for the detection of inspiratory loci in spontaneous speech”, Behavior Research Methods, vol. 44, Issue 4, pp. 1121-1128, Dec. 2012. [cited by applicant]
Larson et al., “SpiroSmart: using a microphone to measure lung function on a mobile phone”, Proceedings of the 2012 ACM Conference on Ubiquitous Computing (UbiComp '12), pp. 280-289, Sep. 5-8, 2012. [cited by applicant]
Abushakra et al., “An Automated Approach Towards Estimating Lung Capacity from Respiration Sounds”, IEEE Healthcare Innovations Conference (HIC'12), pp. 1-5, Jan. 2012. [cited by applicant]
Williammson et al., “Vocal and Facial Biomarkers of Depression Based on Motor Incoordination and Timing”, 4th International Audio/Visual Emotion Challenge and Workshop: Depression Challenge, Orlando, Florida, USA , pp. … [cited by applicant]
Ciccarelli et al., “Neurophysiological Vocal Source Modeling for Biomarkers of Disease”, Interspeech 2016: Understanding Speech Processing in Humans and Machines, Technical Program, San Francisco, USA, pp. 1-7, Sep. 8-1… [cited by applicant]
Helfer et al., “Classification of depression state based on articulatory precision”, Proceedings of the 14th Annual Conference of the International Speech Communication Association (Interspeech), pp. 2172-2176, year 201… [cited by applicant]
Horwitz., “Vocal Modulation Features in the Prediction of Major Depressive Disorder Severity”, Master Thesis, Massachusetts Institute of Technology, pp. 1-115, Sep. 2014. [cited by applicant]
Hillel., “Using phonation time to estimate vital capacity in amyotrophic lateral sclerosis”, Arch Phys Med Rehabil, vol. 70, pp. 618-620, Aug. 1989. [cited by applicant]
Yanagihara., “Phonation and Respiration”, Folia Phoniat, vol. 18, pp. 323-340, year 1966. [cited by applicant]
Dewar et al., “Chronic obstructive pulmonary disease: diagnostic considerations.”, American Academy of Family Physicians, vol. 73, pp. 669-676, Feb. 2006. [cited by applicant]
Solomon et al., “Respiratory and laryngeal contributions to maximum phonation duration”, Journal of voice, vol. 14, No. 3, pp. 331-340, Sep. 2000. [cited by applicant]
Dogan et al., “Subjective and objective evaluation of voice quality in patients with asthma”, Journal of voice, vol. 21, No. 2, pp. 224-230, Mar. 2007. [cited by applicant]
Orenstein et al., “Measuring ease of breathing in young patients with cystic fibrosis”, Pediatric Pulmonology, vol. 34, No. 6, pp. 473-477, Aug. 8, 2002. [cited by applicant]
Lee et al., “Speech Segment Durations Produced by Healthy and Asthmatic Subjects”, Journal of Speech and Hearing Disorders, vol. 653, pp. 186-193, May 31, 1988. [cited by applicant]
Rabiner, L., “A tutorial on hidden Markov models and selected applications in speech recognition,” Proceedings of the IEEE, vol. 77, issue 2, pp. 257-286, Feb. 1989. [cited by applicant]
Rabiner et al., “Fundamentals of Speech Recognition”, Prentice Hall, pp. 1-18 (related section 6.4.3.), year 1993. [cited by applicant]
Sakoe et al., “Dynamic Programming Algorithm Optimization for Spoken Word Recognition”, IEEE Transactions on Acoustics, Speech and Signal Processing, vol. ASSP-26, No. 1, pp. 43-49, Feb. 1978. [cited by applicant]
Lee et al., “Consistency of acoustic and aerodynamic measures of voice production over 28 days under various testing conditions”, Journal of Voice, vol. 13, issue 4, pp. 477-483, Dec. 1, 1999. [cited by applicant]
Walia et al., “Level of Asthma: A Numerical Approach based on Voice Profiling”, IJEDR(International Journal of Engineering Development and Research), vol. 4, Issue 4, pp. 717-722, year 2016. [cited by applicant]
Mulligan et al., “Detecting regional lung properties using audio transfer functions of the respiratory system”, 31st Annual International Conference of the IEEE EMBS, pp. 5697-5700, Sep. 2-6, 2009. [cited by applicant]
Christina et al., “HMM-based speech recognition system for the dysarthric speech evaluation of articulatory subsystem”, International Conference on Recent Trends in Information Technology, pp. 54-59, Apr. 1, 2012. [cited by applicant]
Wang et al., “Vocal folds disorder detection using pattern recognition methods”, 29th Annual International Conference of the IEEE Engineering in Medicine and Biology Society, pp. 3253-3256, Aug. 22-26, 2007. [cited by applicant]
Masada et al., “Feature Extraction by ICA and Clustering for Lung Sound Classification”, IPSJ Symposium Series, vol. 2007, pp. 1-9, year 2007. [cited by applicant]
De La Torre et al., “Discriminative feature weighting for HMM-based continuous speech recognizers”, Speech Communication, vol. 38, pp. 267-286, year 2001. [cited by applicant]
Mswanathan et al., “Complexity Measures of Voice Recordings as a Discriminative Tool for Parkinson's Disease”, MDPI Journal Biosensors 2020, vol. 10, No. 1, pp. 1-16, Dec. 20, 2019. [cited by applicant]
Williamson et al., “Segment-dependent dynamics in predicting Parkinson's disease”, ISCA Conference Interspeech 2015, pp. 518-522, Dresden, Germany, Sep. 6-10, 2015. [cited by applicant]
Valente et al., “Maximum Entropy Discrimination (MED) Feature subset Selection for Speech Recognition”, IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU), ResearchGate publication, pp. 327-332, Nov.… [cited by applicant]
Ramirez et al., “Voice activity detection. Fundamentals and speech recognition system robustness”, Robust Speech Recognition and Understanding, I-Tech, Vienna, Austria, pp. 1-24, Jun. 2007. [cited by applicant]
Bachu et al., “Separation of Voiced and Unvoiced Speech Signals using Energy and Zero Crossing Rate”, ASEE Regional Conference, pp. 1-7, year 2008. [cited by applicant]
International Application # PCT/IB2024/050483 Search Report dated May 7, 2024. [cited by applicant]
EP Application # 21832054.7 Search Report dated Mar. 11, 2024. [cited by applicant]
Eden et al., “Measuring phonological distance between languages”, Doctor Thesis, UCL Department of Linguistics, pp. 1-254, year 2018. [cited by applicant]
Falkhausen et al., “Calculation of Distance Measures Between Hidden Markov Models”, 4th European Conference on Speech Communication and Technology, pp. 1487-1490, Sep. 18-21, 1995. [cited by applicant]
Kempton et al., “Cross-language phone recognition when the target language phoneme inventory is not known”, 12th Annual Conference of the International Speech Communication Association, pp. 3165-3168, Aug. 27-31, 2011. [cited by applicant]
Kienappel et al., “Cross-Language Transfer of Multilingual Phoneme Models”, ASR2000—Automatic Speech Recognition Challenges for the New Millennium, pp. 1-5, Sep. 18-20, 2000. [cited by applicant]
Gupta et al., “Characterizing Exhaled Airflow from Breathing and Talking,” Indoor Air, vol. 20, pp. 31-39, year 2010. [cited by applicant]
Mirza et al., “Analytical Modeling and Simulation of a CMOS-MEMS Cantilever Based CO2 Sensor for Medical Applications,” Proceedings IEEE Regional Symposium on Micro and Nanoelectronics, pp. 70-73, Sep. 27, 2013. [cited by applicant]
Bhagya et al., “Speed of Sound-Based Capnographic Sensor with Second-Generation CNN for Automated Classification of Cardiorespiratory Abnormalities,” IEEE Sensors Journal, vol. 19, issue 19, pp. 8887-8894, Oct. 1, 2019. [cited by applicant]
Kohler, “Multi-Lingual Phoneme Recognition Exploiting Acoustic-Phonetic Similarities of Sounds”, 4th International Conference on Spoken Language Processing (ICSLP 96), pp. 1-4, Oct. 3-6, 1996. [cited by applicant]
Pitts et al., “Using Voluntary Cough to Detect Penetration and Aspiration During Oropharyngeal Swallowing in Patients With Parkinson Disease”, Chest, vol. 138, issue 6, pp. 1426-1431, Dec. 2010. [cited by applicant]
Sooful et al., “An acoustic distance measure for automatic cross-language phoneme mapping”, PRASA proceedings, pp. 1-4, year 2001. [cited by applicant]
Sooful et al., “Comparison of acoustic distance measures for automatic cross-language phoneme mapping”, 7th International Conference on Spoken Language Processing, pp. 1-4, Sep. 16-20, 2002. [cited by applicant]
Sterling et al., “Automated Cough Assessment on a Mobile Platform”, Journal of Medical Engineering, pp. 1-10, year 2014. [cited by applicant]
Vaswani et al., “Attention is all you need”, 31st Conference on Neural Information Processing systems, pp. 1-11, year 2017. [cited by applicant]
Zeng et al., “A new distance measure for hidden Markov models”, Expert Systems with Applications, vol. 37, pp. 1550-1555, year 2010. [cited by applicant]
Rabiner et al., “Fundamentals of Speech Recognition”, Prentice Hall, chapters 7-8, pp. 390-481, year 1993. [cited by applicant]
Wikipedia, “Breathing,” pp. 1-13, last edited Oct. 17, 2021, as downloaded from https://en.wikipedia.org/wiki/Breathing. [cited by applicant]
“Sound Speed in Gases,” Sound and Hearing, HyperPhysics, Department of Physics and Astronomy, Georgia State University, USA, pp. 1-3, year 2017, as downloaded from http://hyperphysics.phy-astr.gsu.edu/hbase/Sound/souspe… [cited by applicant]
“Echo Devices,” Amazon.com, Inc, Interest-Based Ads, pp. 1-6, year 2021, as downloaded from https://www.amazon.com/echo-devices/s?k=echo+devices. [cited by applicant]
“The Best Google Home Speakers in 2021,” Tom's Guide, Future US Inc., pp. 1-21, year 2021, as downloaded from https://www.tomsguide.com/best-picks/best-google-home-speakers. [cited by applicant]
West et al., “Measurements of Pulmonary Gas Exchange Efficiency using Expired Gas and Oximetry: Results in Normal Subjects,” American Journal of Physiology—Lung Cellular and Molecular Physiology, vol. 314, No. 4, pp. L6… [cited by applicant]
West et al., “A New Method for Noninvasive Measurement of Pulmonary Gas Exchange Using Expired Gas,” Respiratory Physiology & Neurobiology, vol. 247, pp. 112-115, year 2018. [cited by applicant]
Huang et al., “An Accurate Air Temperature Measurement System Based on an Envelope Pulsed Ultrasonic Time-of-Flight Technique,” Review of Scientific Instruments, vol. 78, pp. 115102-1-115102-9, year 2007. [cited by applicant]
Jedrusyna, “An Ultrasonic Air Temperature Meter”, Book “Recent Advances in Mechatronics”, Springer, Berlin, Heidelberg, pp. 85-89, year 2010. [cited by applicant]
Cramer, “The Variation of the Specific Heat Ratio and the Speed of Sound in Air with Temperature, Pressure, Humidity, and CO2 Concentration,” Journal of the Acoustical Society of America, vol. 93, No. 5, pp. 2510-2516, … [cited by applicant]
Rao et al., “Acoustic Methods for Pulmonary Diagnosis,” HHS Public Access, Author manuscript, pp. 1-39, year 2020 (final version published in IEEE Reviews in Biomedical Engineering, vol. 12, pp. 221-239, year 2019). [cited by applicant]
Cohen, “Signal processing methods for upper airway and pulmonary dysfunction diagnosis,” IEEE Engineering in Medicine and Biology Magazine, vol. 9, No. 1, pp. 72-75, Mar. 1, 1990. [cited by applicant]
Ney, “The Use of a One-Stage Dynamic Programming Algorithm for Connected Word Recognition,” IEEE Transactions on Acoustics, Speech, and Signal Processing, vol. ASSP-32, No. 2, pp. 263-271, Apr. 1984. [cited by applicant]
Haimi-Cohen et al., U.S. Appl. No. 18/105,848, filed Feb. 5, 2023. [cited by applicant]
Sakran et al., “A Review: Automatic Speech Segmentation”, International Journal of Computer Science and Mobile Computing (IJCSMC), vol. 6, issue 4, pp. 308-315, Apr. 2017. [cited by applicant]
Nicora et al., “Evaluating pointwise reliability of machine learning prediction”, Journal of Biomedical Informatics, vol. 127, pp. 1-15, Mar. 2022. [cited by applicant]
JP Application # 2021-517971 Office Action dated May 16, 2023. [cited by applicant]
EP Application # 21209891.7 Office Action dated May 19, 2023. [cited by applicant]
Indian Application # 202247066856 Office Action dated Mar. 29, 2023. [cited by applicant]
EP Application # 19201720.0 Office Action dated Mar. 30, 2023. [cited by applicant]
EP Application # 20158058.6 Summons to Oral Proceedings dated Apr. 19, 2023. [cited by applicant]
Katsir et al., U.S. Appl. No. 18/319,518, filed May 18, 2023. [cited by applicant]
Haimi-Cohen et al., U.S. Appl. No. 18/328,738, filed Jun. 4, 2023. [cited by applicant]
Haimi-Cohen et al., U.S. Appl. No. 18/328,739, filed Jun. 4, 2023. [cited by applicant]
International Application # PCT/IB2024/054360 Search Report dated Jun. 28, 2024. [cited by applicant]
JP Application # 2022576351 Office Action dated Jul. 2, 2024. [cited by applicant]
U.S. Appl. No. 17/902,836 Office Action, filed Jul. 8, 2024. [cited by applicant]
International Application # PCT/IB2024/054359 Search Report dated Jul. 9, 2024. [cited by applicant]
AU Application # 2021384028 Office Action Aug. 15, 2024. [cited by applicant]
EP Application # 24181539.8 Search Report dated Sep. 4, 2024. [cited by applicant]
CN Application # 2020800180012 Office Action dated Jan. 30, 2024. [cited by applicant]
IN Application # 202347030550 Office Action dated Dec. 13, 2023. [cited by applicant]
CN Application # 2019800670875 Office Action dated Dec. 20, 2023. [cited by applicant]
JP Application # 2021549583 Office Action dated Dec. 25, 2023. [cited by applicant]
JP Application # 2021551893 Office Action dated Dec. 25, 2023. [cited by applicant]
CN Application # 202080017839X Office Action dated Jan. 27, 2024. [cited by applicant]
JP Application # 2022548568 Office Action dated Oct. 29, 2024. [cited by applicant]
U.S. Appl. No. 17/531,828 Office Action dated Jan. 15, 2025. [cited by applicant]
U.S. Appl. No. 18/105,848 Office Action dated Jan. 21, 2025. [cited by applicant]
CN Application # 202180045274.0 Office Action dated Feb. 28, 2025. [cited by applicant]
U.S. Appl. No. 18/328,738 Office Action dated Apr. 11, 2025. [cited by applicant]
U.S. Appl. No. 18/328,739 Office Action dated Apr. 25, 2025. [cited by applicant]
CN Application # 202180017631.2 Office Action dated Mar. 31, 2025. [cited by applicant]
Canadian Examination report, application # 3,114,864, dated Oct. 16, 2025. [cited by applicant]
Korean Office Action, application # 10-2021-7032027, dated Oct. 24, 2025. [cited by applicant]
Korean Office Action, application # 10-2021-7032025, dated Sep. 26, 2025. [cited by applicant]
Cited By (1)
US 12,645,940