IP Library › Granted Patent US 9,685,152
Granted Patent B2
US 9,685,152 · App. 14/892,624 · Granted Jun 20, 2017

Technology for responding to remarks using speech synthesis

Inventors: Hiroaki Matsubara (Hamamatsu, JP); Junya Ura (Hamamatsu, JP); Takehiko Kawahara (Hamamatsu, JP); Yuji Hisaminato (Hamamatsu, JP); Katsuji Yoshimura (Hamamatsu, JP)
Assignee: YAMAHA CORPORATION
G10L13/0335G10L13/027G10L15/18G10L13/033G10L13/06G10L13/10G10L25/90H04M2201/39
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,685,152
App. No.
14/892,624
Granted
Jun 20, 2017
Kind
B2
Abstract

The present invention is provided with: a voice input section that receives a remark (a question) via a voice signal; a reply creation section that creates a voice sequence of a reply (response) to the remark; a pitch analysis section that analyzes the pitch of a first segment (e.g., word ending) of the remark; and a voice generation section (a voice synthesis section, etc.) that generates a reply, in the form of voice, represented by the voice sequence. The voice generation section controls the pitch of the entire reply in such a manner that the pitch of a second segment (e.g., word ending) of the reply assumes a predetermined pitch (e.g., five degrees down) with respect to the pitch of the first segment of the remark. Such arrangements can realize synthesis of replying voice capable of giving a natural feel to the user.

Claims (36)

1. A voice synthesis apparatus comprising:

a voice input section configured to receive a voice signal of a remark;

a pitch analysis section configured to analyze a pitch of a first segment of the remark;

an acquisition section configured to acquire a reply to the remark; and

a voice generation section configured to generate voice of the reply acquired by said acquisition section, said voice generation section controlling a pitch of the voice of the reply in such a manner that a second segment of the reply has a pitch associated with the pitch of the first segment analyzed by said pitch analysis section,

wherein said voice generation section controls the pitch of the voice of the reply in such a manner that an interval of the pitch of said second segment relative to the pitch of said first segment becomes a consonant interval except in a case where the pitch of said second segment and the pitch of said first segment are in perfect unison.

2. A voice synthesis apparatus comprising:

a voice input section configured to receive a voice signal of a remark;

a pitch analysis section configured to analyze a pitch of a first segment of the remark;

an acquisition section configured to acquire a reply to the remark; and

a voice generation section configured to generate voice of the reply acquired by said acquisition section, said voice generation section controlling a pitch of the voice of the reply in such a manner that a second segment of the reply has a pitch associated with the pitch of the first segment analyzed by said pitch analysis section, wherein said voice generation section controls the pitch of the voice of the reply in such a manner that an interval of the pitch of said second segment relative to the pitch of said first segment becomes any one of intervals, except in a case where the pitch of said second segment and the pitch of said first segment are in perfect unison, within a range of one octave up and one octave down from the pitch of said first segment.

3. The voice synthesis apparatus as claimed in claim 2 , wherein any one of a first mode and a second mode is settable as an operation mode of said voice generation section, and

wherein, in said first mode, said voice generation section controls the pitch of the voice of the reply in such a manner that the interval of the pitch of said second segment relative to the pitch of said first segment becomes a consonant interval except in a case where the pitch of said second segment and the pitch of said first segment are in perfect unison, and

in said second mode, said voice generation section controls the pitch of the voice of the reply in such a manner that the interval of the pitch of said second segment relative to the pitch of said first segment becomes a dissonant interval.

4. The voice synthesis apparatus as claimed in claim 1 , wherein said voice generation section controls the pitch of the voice of the reply in such a manner that the interval of the pitch of said second segment relative to the pitch of said first segment becomes a consonant interval of five degrees lower than the pitch of said first segment.

5. The voice synthesis apparatus as claimed in claim 1 , wherein said voice generation section provisionally sets the pitch of the second segment of the voice of the reply at the pitch associated with the pitch of the first segment, and

said voice generation section is further configured to perform at least one of:

an operation of, if the provisionally-set pitch of the second segment is lower than a predetermined first threshold value, changing the provisionally-set pitch of the second segment to a pitch shifted one octave up; and

an operation of, if the provisionally-set pitch of the second segment is higher than a predetermined second threshold value, changing the provisionally-set pitch of the second segment to a pitch one octave down.

6. The voice synthesis apparatus as claimed in claim 1 , wherein said voice generation section provisionally sets the pitch of the second segment of the voice of the reply at the pitch associated with the pitch of the first segment, and

said voice generation section is further configured to change the provisionally-set pitch to a pitch shifted one octave up or down in accordance with a designated attribute.

7. The voice synthesis apparatus as claimed in claim 1 , wherein any one of a first mode and a second mode is settable as an operation mode of said voice generation section, and

wherein, in said first mode, said voice generation section controls the pitch of the voice of the reply in such a manner that the interval of the pitch of said second segment relative to the pitch of said first segment becomes a consonant interval except in a case where the pitch of said second segment and the pitch of said first segment are in perfect unison, and

in said second mode, said voice generation section controls the pitch of the voice of the reply in such a manner that the interval of the pitch of said second segment relative to the pitch of said first segment becomes a dissonant interval.

8. A computer-implemented method comprising:

receiving a voice signal of a remark;

analyzing a pitch of a first segment of the remark;

acquiring a reply to the remark;

synthesizing voice of the acquired reply; and

controlling a pitch of the reply in such a manner that a pitch of a second segment of the voice of the reply has a pitch associated with the analyzed pitch of the first segment and an interval of the pitch of the second segment relative to the pitch of the first segment becomes a consonant interval except in a case where the pitch of the second segment and the pitch of the first segment are in perfect unison.

9. A computer-implemented method comprising:

receiving a voice signal of a remark;

analyzing a pitch of a first segment of the remark;

acquiring a reply to the remark;

synthesizing voice of the acquired reply; and

controlling a pitch of the reply in such a manner that a pitch of a second segment of the voice of the reply has a pitch associated with the analyzed pitch of the first segment and an interval of the pitch of the second segment relative to the pitch of the first segment becomes any one of intervals, except in a case where the pitch of the second segment and the pitch of the first segment are in perfect unison, within a range of one octave up and one octave down from the pitch of said first segment.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 23, 2015
From: MATSUBARA, HIROAKI; URA, JUNYA; KAWAHARA, TAKEHIKO; HISAMINATO, YUJI; YOSHIMURA, KATSUJI
To: YAMAHA CORPORATION
Reel/Frame 037113/0026 →
Priority Claims (9)
JP 2013-115111 · May 31, 2013 · national
JP 2013-198217 · Sep 25, 2013 · national
JP 2013-198218 · Sep 25, 2013 · national
JP 2013-198219 · Sep 25, 2013 · national
JP 2013-203839 · Sep 30, 2013 · national
JP 2013-203840 · Sep 30, 2013 · national
JP 2013-205260 · Sep 30, 2013 · national
JP 2013-205261 · Sep 30, 2013 · national
JP 2014-048636 · Mar 12, 2014 · national
Continuity (1)
Related Publication 20160086597A1 · Mar 24, 2016