IP Library Granted Patent US 12,573,369
Granted Patent B2
US 12,573,369 · App. 17/765,668 · Granted Mar 10, 2026

Method for controlling utterance device, server, utterance device, and program

Inventors: Sara Asai (Osaka, JP); Satoru Matsunaga (Osaka, JP); Hiroki Urabe (Osaka, JP); Masahiro Ishii (Hyogo, JP)
Assignee: PANASONIC INTELLECTUAL PROPERTY MANAGEMENT CO., LTD.
G10L13/033
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,573,369
App. No.
17/765,668
Filed
Jan 22, 2024
Granted
Mar 10, 2026
Kind
B2
Art Unit
2656
USPC
704/258
Abstract

A method for controlling an utterance device, a server ( 10 ), an utterance device ( 20 ), and a program control the utterance device ( 20 ). The server ( 10 ) receives utterance source information from an information source device ( 40 ), and set the utterance device ( 20 ) based on the utterance source information. Then, the server ( 10 ) provides an utterance sound source that has a sound source characteristic according to the utterance device ( 20 ) to the utterance device ( 20 ), and causes the utterance device ( 20 ) to utter using the utterance sound source.

Claims (70)

1 . A method for controlling an utterance device, comprising:

receiving utterance source information from an information source device;

setting an utterance device based on the utterance source information;

providing an utterance sound source that has a sound source characteristic according to the utterance device to the utterance device; and

causing the utterance device to utter using the utterance sound source, wherein at least one of the following conditions (a), (b), and (c) is satisfied:

(a) the sound source characteristic includes a sampling frequency, and the sampling frequency is set according to a frequency component that attenuates by being blocked by the utterance device due to arrangement of a speaker of the utterance device;

(b) providing an utterance sound source to the utterance device includes:

receiving an inquiry using the set sound source characteristic from the utterance device;

selecting a sound source, as the utterance sound source, that has the sound source characteristic in the inquiry from a plurality of sound sources; and

transmitting an access destination corresponding to the utterance sound source to the utterance device so as to cause the utterance device to download the utterance sound source; and

(c) providing an utterance sound source to the utterance device includes:

selecting a plurality of candidate sound sources according to the sound source characteristic from a plurality of sound sources;

transmitting access destinations corresponding to the plurality of candidate sound sources to the utterance device; and

providing the utterance sound source to the utterance device, via an access destination corresponding to an utterance sound source selected from the plurality of candidate sound sources.

2 . The method for controlling an utterance device according to claim 1 , wherein the sound source characteristic is set based on at least one of a type, an identifier, utterance performance, an operating state, a location, and a distance to a user of the utterance device; user information of a user of the utterance device; and arrangement of a speaker of the utterance device.

3 . The method for controlling an utterance device according to claim 1 ,

wherein the sound source characteristic includes at least one of a format of voice data, a timbre characteristic, a sound quality characteristic, a volume, and utterance content.

4 . The method for controlling an utterance device according to claim 1 , wherein the sound source characteristic includes a sampling frequency;

wherein a sampling frequency is set according to utterance performance of the utterance device.

5 . The method for controlling an utterance device according to claim 1 , wherein the sound source characteristic includes a sound volume;

wherein a volume is set according to a distance between the utterance device and a user, or

in a case where the utterance device is determined to be in an operating state, a volume is set to be larger than that in a case where the utterance device is determined not to be in the operating state.

6 . The method for controlling an utterance device according to claim 1 , wherein the sound source characteristic includes at least one of a volume, a speaking speed, and a frequency component;

wherein in a case where an age of a user as an utterance target of the utterance device is determined to be a predetermined age or more, a volume is set to be larger, a speaking speed is set to be slower, and/or a larger number of high frequency components are set to be included than in a case where the age is determined to be less than the predetermined age.

7 . The method for controlling an utterance device according to claim 1 , wherein providing an utterance sound source to the utterance device includes:

setting a sound source characteristic according to the utterance device;

selecting a sound source, as the utterance sound source, that has the set sound source characteristic from a plurality of sound sources; and

transmitting an access destination corresponding to the utterance sound source to the utterance device so as to cause the utterance device to download the utterance sound source.

8 . A server that controls an utterance device, the server comprising:

a server storage that stores sound sources providable to the utterance device; and

a server controller configured to:

receive utterance source information from an information source device, set an utterance device based on the utterance source information,

provide an utterance sound source that has a sound source characteristic according to the utterance device to the utterance device, and

cause the utterance device to utter using the utterance sound source,

wherein at least one of the following conditions (a), (b) and (c) is satisfied:

(a) the sound source characteristic includes a sampling frequency, and the sampling frequency is set according to a frequency component that attenuates by being blocked by the utterance device due to arrangement of a speaker of the utterance device;

(b) when providing an utterance sound source to the utterance device, the server controller is further configured to:

receive an inquiry using the set sound source characteristic from the utterance device;

select a sound source, as the utterance sound source, that has the sound source characteristic in the inquiry from a plurality of sound sources; and

transmit an access destination corresponding to the utterance sound source to the utterance device so as to cause the utterance device to download the utterance sound source; and

(c) when providing an utterance sound source to the utterance device, the server controller is further configured to:

select a plurality of candidate sound sources according to the sound source characteristic from a plurality of sound sources;

transmit access destinations corresponding to the plurality of candidate sound sources to the utterance device; and

provide the utterance sound source to the utterance device, via an access destination corresponding to an utterance sound source selected from the plurality of candidate sound sources.

9 . The server that controls an utterance device according to claim 8 , wherein the sound source characteristic is set based on at least one of a type, an identifier, utterance performance, an operating state, a location, and a distance to a user of the utterance device; user information of a user of the utterance device; and arrangement of a speaker of the utterance device.

10 . The server that controls an utterance device according to claim 8 , wherein the sound source characteristic includes at least one of a format of voice data, a timbre characteristic, a sound quality characteristic, a volume, and utterance content.

11 . The server that controls an utterance device according to claim 8 , wherein the sound source characteristic includes a sampling frequency;

wherein a sampling frequency is set according to utterance performance of the utterance device.

12 . The server that controls an utterance device according to claim 8 , wherein the sound source characteristic includes a sound volume;

wherein a volume is set according to a distance between the utterance device and a user, or

in a case where the utterance device is determined to be in an operating state, a volume is set to be larger than that in a case where the utterance device is determined not to be in the operating state.

13 . The server that controls an utterance device according to claim 8 , wherein the sound source characteristic includes at least one of a volume, a speaking speed, and a frequency component;

wherein in a case where an age of a user as an utterance target of the utterance device is determined to be a predetermined age or more, a volume is set to be larger, a speaking speed is set to be slower, and/or a larger number of high frequency components are set to be included than in a case where the age is determined to be less than the predetermined age.

14 . The server that controls an utterance device according to claim 8 , wherein when providing an utterance sound source to the utterance device, the server controller is further configured to:

set a sound source characteristic according to the utterance device;

select a sound source, as the utterance sound source, that has the set sound source characteristic from a plurality of sound sources; and

transmit an access destination corresponding to the utterance sound source to the utterance device so as to cause the utterance device to download the utterance sound source.

15 . An utterance device capable of making utterance, comprising:

a device storage that stores at least one of a type, an identifier, utterance performance, an operating state, a location, and a distance to a user of the utterance device, user information of a user of the utterance device, and arrangement of a speaker of the utterance device; and

a device controller configured to:

set a sound source characteristic suitable for the utterance device based on at least one of the type, the identifier, the utterance performance, the operating state, the location, and the distance to a user of the utterance device, the user information of a user of the utterance device, and the arrangement of a speaker of the utterance device;

make an inquiry to a server by using the set sound source characteristic;

acquire an utterance sound source that has the sound source characteristic from the server; and

utter using the utterance sound source;

wherein at least one of the following conditions (a), (b), and (c) is satisfied:

(a) the sound source characteristic includes a sampling frequency, and the sampling frequency is set according to a frequency component that attenuates by being blocked by the utterance device due to arrangement of a speaker of the utterance device;

(b) the utterance sound source has the sound source characteristic which is set according to the utterance device, and the utterance sound source is selected from a plurality of sound sources, and

the device controller is further configured to receive, from the server, an access destination corresponding to the utterance sound source and download the utterance sound source;

(c) the device controller is further configured to receive, from the server, access destinations corresponding to a plurality of candidate sound sources according to the sound source characteristic, and

acquire the utterance sound source, via the access destinations.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 27, 2022
From: ASAI, SARA; MATSUNAGA, SATORU; URABE, HIROKI; ISHII, MASAHIRO
To: PANASONIC INTELLECTUAL PROPERTY MANAGEMENT CO., LTD.
Reel/Frame 060648/0417 →
Priority Claims (1)
JP 2021-066716 · Apr 9, 2021 · national
Continuity (1)
Related Publication 20240221720A1 · Jul 4, 2024
References Cited (21)
US 9123339B1 · Shaw · 2015 [cited by examiner]
US 10565989B1 · Wheeler · 2020 [cited by examiner]
US 20200126566A1 · Wang · 2020 [cited by examiner]
US 20200388268A1 · Saito · 2020 [cited by applicant]
JP 2006126548 · 2006 [cited by applicant]
JP 2006126548A · 2006 [cited by examiner]
JP 2009139390 · 2009 [cited by applicant]
JP 2010048959 · 2010 [cited by applicant]
JP 2010048959A · 2010 [cited by examiner]
JP 2015164251 · 2015 [cited by applicant]
JP 2016062077 · 2016 [cited by applicant]
JP 2016062077A · 2016 [cited by examiner]
JP 6640266 · 2020 [cited by applicant]
JP 2021002062 · 2021 [cited by applicant]
JP 2021002062A · 2021 [cited by examiner]
WO 2015129523 · 2015 [cited by applicant]
WO 2019138652 · 2019 [cited by applicant]
Notice of Reasons for Refusal issued May 14, 2024 in counterpart Japanese Patent Application No. 2023-060786, with English machine translation. [cited by applicant]
International Search Report issued Oct. 19, 2021 in International (PCT) Application No. PCT/JP2021/030644. [cited by applicant]
Office Action issued Feb. 14, 2023 in counterpart Japanese Application No. 2022-519353, with machine translation. [cited by applicant]
Translation of the International Preliminary Report on Patentability issued Oct. 19, 2023 in International Application No. PCT/JP2021/030644. [cited by applicant]