IP Library Granted Patent US 10,546,582
Granted Patent B2
US 10,546,582 · App. 15/515,010 · Granted Jan 28, 2020

Information processing device, method of information processing, and program

Inventors: Yuhei Taki (Kanagawa, JP); Shinichi Kawano (Tokyo, JP); Takashi Shibuya (Tokyo, JP); Emiru Tsunoo (Tokyo, JP)
Assignee: SONY CORPORATION
G10L15/22G06F3/167G10L15/1822G10L15/28G10L25/84
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,546,582
App. No.
15/515,010
Granted
Jan 28, 2020
Kind
B2
Abstract

There is provided an information processing device technology that enables an improvement in precision of sound recognition processing based on collected sound information, the information processing device including: a recognition controller that causes a speech recognition processing portion to execute sound recognition processing based on collected sound information obtained by a sound collecting portion; and an output controller that generates an output signal to output a recognition result obtained through the sound recognition processing. The output controller causes an output portion to output an evaluation result regarding a type of sound based on the collected sound information prior to the recognition result.

Claims (62)

1. An information processing device comprising:

a recognition controller configured to cause a speech recognition processing portion to execute sound recognition processing based on collected sound information obtained by a sound collecting portion; and

an output controller configured to

generate an output signal to output a recognition result obtained through the sound recognition processing,

wherein the output controller causes an output portion to output an evaluation result regarding a type of sound based on the collected sound information prior to the recognition result, and

wherein the output portion comprises a display,

cause the output portion to display an evaluation result object corresponding to the evaluation result, and

move the displayed evaluation result object to a predetermined target position based on the recognition result,

wherein at least one display parameter of the displayed evaluation result object is changed according to an evaluated likelihood that the collected sound information on which the evaluation result is based contains speech, and

wherein the at least one display parameter of the displayed evaluation result object includes a shape of the displayed evaluation result object.

2. The information processing device according to claim 1 ,

wherein the output controller is further configured to cause the output portion to display a sound collection notification object for providing notification about sound collection when the collected sound information is obtained, and cause the sound collection notification object to be changed to the displayed evaluation result object in accordance with the evaluation result when the evaluation result is obtained.

3. The information processing device according to claim 2 ,

wherein the output controller is further configured to cause the output portion to display the sound collection notification object corresponding to a volume of the collected sound information when the collected sound information is obtained.

4. The information processing device according to claim 1 ,

wherein the at least one display parameter of the displayed evaluation result object further includes at least one of transparency, a color, a size, and motion of the displayed evaluation result object.

5. The information processing device according to claim 1 ,

wherein the output controller is further configured to cause the output portion to output different evaluation result objects when the evaluation result is greater than a threshold value and when the evaluation result is less than the threshold value.

6. The information processing device according to claim 5 ,

wherein the output controller is further configured to cause the output portion to output the threshold value.

7. The information processing device according to claim 5 ,

wherein the recognition controller is further configured to determine termination of a part serving as a target of speech recognition processing on the basis of timing when a period of time during which the evaluation result is less than the threshold value exceeds a predetermined period of time in the collected sound information.

8. The information processing device according to claim 5 ,

wherein the recognition controller is further configured to determine termination of a part serving as a target of speech recognition processing on the basis of timing when a period of time during which a volume is less than a predetermined volume exceeds a predetermined period of time in the collected sound information.

9. The information processing device according to claim 5 ,

wherein the recognition controller is further configured to add or change a condition for determining termination of a part serving as a target of speech recognition processing when the evaluation result is less than the threshold value after an utterance by a user and a volume of the collected sound information is greater than a predetermined volume.

10. The information processing device according to claim 1 ,

wherein the recognition controller is further configured to cause speech recognition processing based on the collected sound information to be performed when the evaluation result is greater than a threshold value.

11. The information processing device according to claim 1 ,

wherein the recognition controller is further configured to refrain from causing speech recognition processing based on the collected sound information to be performed when the evaluation result is less than a threshold value.

12. The information processing device according to claim 1 ,

wherein the output controller is further configured to determine the evaluation result object to be output by the output portion on the basis of a history of the evaluation result.

13. The information processing device according to claim 1 ,

wherein the sound recognition processing includes processing of specifying a character string on the basis of the collected sound information.

14. The information processing device according to claim 1 ,

wherein the output controller is further configured to cause the output portion to output different evaluation result objects when a first evaluation result regarding a type of sound based on the collected sound information is greater than a first threshold value and when a predetermined second evaluation result of the collected sound information is greater than a second threshold value.

15. The information processing device according to claim 1 ,

wherein the sound recognition processing includes speech recognition processing based on the collected sound information.

16. The information processing device according to claim 1 ,

wherein the shape of the displayed evaluation result object is changed, and at least one of a transparency, a color, a size, or a motion of the displayed evaluation result object is also changed, based on the evaluated likelihood that the collected sound information on which the evaluation result is based contains speech.

17. The information processing device according to claim 1 ,

wherein the shape of the displayed evaluation result object is changed according to the evaluated likelihood such that

when the evaluated likelihood is greater than a threshold value, the displayed evaluation result object has a first contour, and

when the evaluated likelihood is less than or equal to the threshold value, the displayed evaluation result object has a second contour that is different from the first contour.

18. A method of information processing, comprising:

causing a speech recognition processing portion to execute sound recognition processing based on collected sound information obtained by a sound collecting portion; and

generating an output signal to output a recognition result obtained through the sound recognition processing,

wherein an output portion is caused to output an evaluation result regarding a type of sound based on the collected sound information prior to the recognition result,

wherein the output portion is a display,

causing the output portion to display an evaluation result object corresponding to the evaluation result, and

moving the displayed evaluation result object to a predetermined target position based on the recognition result,

wherein at least one display parameter of the displayed evaluation result object is changed according to an evaluated likelihood that the collected sound information on which the evaluation result is based contains speech, and

wherein the at least one display parameter of the displayed evaluation result object includes a shape of the displayed evaluation result object.

19. A non-transitory computer-readable medium having embodied thereon a program, which when executed by a computer causes the computer to execute a method, the method comprising:

causing a speech recognition processing portion to execute sound recognition processing based on collected sound information obtained by a sound collecting portion;

generating an output signal to output a recognition result obtained through the sound recognition processing,

wherein an output portion is caused to output an evaluation result regarding a type of sound based on the collected sound information prior to the recognition result, and

wherein the output portion comprises a display;

causing the output portion to display an evaluation result object corresponding to the evaluation result; and

moving the displayed evaluation result object to a predetermined target position based on the recognition result,

wherein at least one display parameter of the displayed evaluation result object is changed according to an evaluated likelihood that the collected sound information on which the evaluation result is based contains speech, and

wherein the at least one display parameter of the displayed evaluation result object includes a shape of the displayed evaluation result object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2017
From: TAKI, YUHEI; KAWANO, SHINICHI; SHIBUYA, TAKASHI; TSUNOO, EMIRU
To: SONY CORPORATION
Reel/Frame 042107/0846 →
Priority Claims (1)
JP 2014-266615 · Dec 26, 2014 · national
Continuity (1)
Related Publication 20170229121A1 · Aug 10, 2017