SYSTEM AND METHOD FOR DETECTING ARTICULATION ERRORS
System and method for analyzing audio data are provided. The audio data may be analyzed to detect articulation errors. For example, the audio data may be analyzed to detect articulation errors of a selected speaker, such as articulation errors of a wearer of a wearable audio sensor, of a speaker engaged in conversation with the wearer of the wearable audio sensor, and so forth. Feedbacks and reports may be provided based on the detected articulation errors.
1 . A system for processing audio, the system comprising:
at least one processing unit configured to:
obtain audio data captured by one or more audio sensors; and
analyze the audio data to detect an articulation error.
2 . The system of claim 1 , wherein the at least one processing unit is further configured to:
analyze the audio data to obtain textual information; and
analyze the textual information to detect the articulation error.
3 . The system of claim 1 , wherein the at least one processing unit is further configured to identify the type of the articulation error.
4 . The system of claim 3 , wherein the type of the articulation error is one of:
substitution articulation error, omission articulation error, distortion articulation error, and addition articulation error.
5 . The system of claim 1 , wherein the at least one processing unit is further configured to:
obtain additional audio data captured by the one or more audio sensors after the analysis of the audio data;
analyze the additional audio data to detect one or more additional articulation errors; and
provide one or more reports to a user based on information associated with the articulation error and information associated with the one or more additional articulation errors.
6 . The system of claim 1 , wherein the one or more audio sensors are included in a wearable apparatus; the system includes the wearable apparatus; and wherein obtaining the audio data comprises capturing the audio data from an environment of a wearer of the wearable apparatus using the one or more audio sensors; and wherein the at least one processing unit is further configured to provide feedback to the wearer based on the detect the articulation error.
7 . The system of claim 6 , wherein the at least one processing unit is further configured to:
after providing the feedback, obtain additional audio data captured by the one or more audio sensors;
analyze the additional audio data to detect an additional articulation error;
determine that the additional articulation error is within a selected time period after the articulation error; and
based on said determination, withhold providing feedback associated with the additional articulation error.
8 . The system of claim 1 , wherein the one or more audio sensors are included in a wearable apparatus; the system includes the wearable apparatus; and wherein obtaining the audio data comprises capturing the audio data from an environment of a wearer of the wearable apparatus using the one or more audio sensors; and wherein the at least one processing unit is further configured to:
determine that the articulation error is an articulation error of the wearer; and
provide feedback to the wearer based on said determination.
9 . The system of claim 1 , wherein the at least one processing unit is further configured to:
analyze the audio data to determine a context associated with the detected articulation error; and
provide information to a user based on the detected articulation error and the determined context.
10 . A method for processing audio, the method comprising:
obtaining audio data captured by one or more audio sensors; and
analyzing the audio data to detect an articulation error.
11 . The method of claim 10 , further comprising:
analyzing the audio data to obtain textual information; and
analyzing the textual information to detect the articulation error.
12 . The method of claim 10 , wherein the detection of the articulation error is based on at least one rule, the at least one rule is a result of one or more machine learning algorithms trained on training examples.
13 . The method of claim 10 , further comprising identifying the type of the articulation error.
14 . The method of claim 13 , wherein the type of the articulation error is one of: substitution articulation error, omission articulation error, distortion articulation error, and addition articulation error.
15 . The method of claim 10 , further comprising:
obtaining additional audio data captured by the one or more audio sensors after the analysis of the audio data;
analyzing the additional audio data to detect one or more additional articulation errors; and
providing one or more reports to a user based on information associated with the articulation error and information associated with the one or more additional articulation errors.
16 . The method of claim 10 , wherein the one or more audio sensors are included in a wearable apparatus; obtaining the audio data comprises capturing the audio data from an environment of a wearer of the wearable apparatus using the one or more audio sensors; and wherein the method further comprising providing feedback to the wearer based on said determination.
17 . The method of claim 16 , further comprising:
after providing the feedback, obtaining additional audio data captured by the one or more audio sensors;
analyzing the additional audio data to detect an additional articulation error;
determining that the additional articulation error is within a selected time period after the articulation error; and
based on said determination, withhold providing feedback associated with the additional articulation error.
18 . The method of claim 10 , wherein the one or more audio sensors are included in a wearable apparatus; obtaining the audio data comprises capturing the audio data from an environment of a wearer of the wearable apparatus using the one or more audio sensors; and wherein the method further comprising:
determining that the articulation error is an articulation error of the wearer; and
providing feedback to the wearer based on said determination.
19 . The method of claim 10 , further comprising:
analyzing the audio data to determine a context associated with the detected articulation error; and
providing information to a user based on the detected articulation error and the determined context.
20 . A non-transitory computer readable medium storing data and computer implementable instructions for carrying out the method of claim 10 .