IP Library › Granted Patent US 10,937,415
Granted Patent B2
US 10,937,415 · App. 16/089,132 · Granted Mar 2, 2021

Information processing device and information processing method for presenting character information obtained by converting a voice

Inventors: Ayumi Kato (Kanagawa, JP); Shinichi Kawano (Tokyo, JP); Yuhei Taki (Kanagawa, JP); Yusuke Nakagawa (Kanagawa, JP)
Assignee: SONY CORPORATION
G10L15/1815G06F3/16G10L15/07G10L15/08G10L15/187G10L15/1807G10L15/22G10L15/26G10L15/30G10L21/0216G10L25/84G10L2015/221G10L2015/223G10L2015/227G10L2015/228
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,937,415
App. No.
16/089,132
Granted
Mar 2, 2021
Kind
B2
Abstract

There is provided an information processing device to further improve the operability of user interfaces that use a voice as an input, the information processing device including: an acquisition unit configured to acquire context information in a period for collection of a voice; and a control unit configured to cause a predetermined output unit to present a candidate for character information obtained by converting the voice in a mode in accordance with the context information.

Claims (53)

1. An information processing device comprising:

an acquisition unit configured to acquire context information in a period during which a voice is collected; and

a control unit configured to cause a predetermined output unit to present a plurality of candidates for character information obtained by converting the voice in a mode in accordance with the context information,

wherein the context information includes information regarding a mode of an utterance of the voice,

wherein the information regarding the mode of the utterance of the voice includes a speed of the utterance of the voice,

wherein the control unit limits a number of the plurality of candidates presented by the output unit in accordance with the speed of the utterance of the voice, and

wherein the acquisition unit and the control unit are each implemented via at least one processor.

2. The information processing device according to claim 1 , wherein the context information includes information regarding a state of an environment in which the voice is collected.

3. The information processing device according to claim 1 , wherein the context information includes information regarding a state of input information or an input situation of the input information.

4. The information processing device according to claim 1 , wherein the context information includes information regarding a state of a user who has uttered the voice.

5. The information processing device according to claim 1 , wherein the control unit controls levels of priority among the plurality of candidates for the character information obtained by converting the voice in accordance with the context information acquired in the period.

6. The information processing device according to claim 5 ,

wherein the context information includes information regarding an influence of noise, and

the control unit controls the levels of priority in accordance with levels of similarity based on pronunciation of the voice in a case in which the influence of the noise is less than a threshold value.

7. The information processing device according to claim 5 ,

wherein the context information includes information regarding an influence of noise, and

the control unit controls the levels of priority in accordance with a co-occurrence relationship between a plurality of pieces of character information obtained by converting the voice in a case in which the influence of the noise is higher than a threshold value.

8. The information processing device according to claim 5 ,

wherein the control unit causes the output unit to present one or more candidates for character information indicating pronunciation of the voice in accordance with an analysis result of sound that is the collected voice, and controls the levels of priority on a basis of a result of selection with respect to at least some of the presented one or more candidates indicating the pronunciation.

9. The information processing device according to claim 5 , wherein the control unit switches data referred to for controlling the levels of priority in accordance with the context information.

10. The information processing device according to claim 5 ,

wherein the control unit controls the levels of priority for each of a plurality of different algorithms on a basis of the algorithms in accordance with the context information, and

wherein the control unit causes the output unit to present the plurality of candidates for the character information for each of the algorithms in accordance with the levels of priority controlled on the basis of the algorithms.

11. The information processing device according to claim 10 , wherein the control unit presents the plurality of candidates for the character information in different modes for each of the algorithms.

12. The information processing device according to claim 1 ,

wherein the control unit receives designation of a condition for the character information obtained by converting the voice and limits the plurality of candidates for the character information presented by the output unit on a basis of the condition.

13. The information processing device according to claim 1 ,

wherein the control unit presents information indicating a number of the plurality of candidates for the character information obtained by converting the voice in association with the character information.

14. The information processing device according to claim

wherein the acquisition unit successively acquires a result of a voice recognition process executed on a basis of successively collected voice, and

the control unit causes the output unit to present the plurality of one or more candidates for character information obtained by converting at least part of the successively collected voice on a basis of the successively acquired result of the voice recognition process.

15. The information processing device according to claim 1 , wherein the control unit further causes the output unit to present character information indicating pronunciation of the voice as the character information obtained by converting the voice in accordance with the context information.

16. An information processing device comprising:

a transmission unit configured to transmit context information to an external device in a period during which a voice is collected, wherein the context information is acquired by a predetermined acquisition unit; and

an output unit configured to present a plurality of candidates for character information obtained by converting the voice transmitted from the external device,

wherein the voice is converted in a mode in accordance with the context information,

wherein the context information includes information regarding a mode of an utterance of the voice,

wherein the information regarding the mode of the utterance of the voice includes a speed of the utterance of the voice,

wherein a number of the plurality of candidates presented by the output unit is limited in accordance with the speed of the utterance of the voice, and

wherein the transmission unit and the output unit are each implemented via at least one processor.

17. An information processing method comprising, by a computer system:

acquiring context information in a period during which a voice is collected; and

causing a predetermined output unit to present a plurality of candidates for character information obtained by converting the voice in a mode in accordance with the acquired context information,

wherein the context information includes information regarding a mode of an utterance of the voice,

wherein the information regarding the mode of the utterance of the voice includes a speed of the utterance of the voice, and

wherein a number of the presented plurality of candidates is limited in accordance with the speed of the utterance of the voice.

18. An information processing method comprising, by a computer system:

transmitting context information to an external device in a period during which a voice is collected, wherein the context information is acquired by a predetermined acquisition unit; and

presenting a plurality of candidates for character information obtained by converting the voice transmitted from the external device,

wherein the voice is converted in a mode in accordance with the context information,

wherein the context information includes information regarding a mode of an utterance of the voice,

wherein the information regarding the mode of the utterance of the voice includes a speed of the utterance of the voice, and

wherein a number of the presented plurality of candidates is limited in accordance with the speed of the utterance of the voice.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 27, 2018
From: KATO, AYUMI; KAWANO, SHINICHI; TAKI, YUHEI; NAKAGAWA, YUSUKE
To: SONY CORPORATION
Reel/Frame 047157/0908 →
Priority Claims (1)
JP JP2016-118622 · Jun 15, 2016 · national
Continuity (1)
Related Publication 20190130901A1 · May 2, 2019