IP Library Granted Patent US 11,205,440
Granted Patent B2
US 11,205,440 · App. 16/691,713 · Granted Dec 21, 2021

Sound playback system and output sound adjusting method thereof

Inventors: Kuo-Ping Yang (Taipei, TW); Kuo-Wei Kao (Taipei, TW); Kai-Yuan Hsiao (Taipei, TW); Jian-Ying Li (Taipei, TW)
Assignee: PIXART IMAGING INC.
G10L21/0364G06K9/00228G10L17/00G10L17/06G10L21/013G06K2009/00322
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,205,440
App. No.
16/691,713
Granted
Dec 21, 2021
Kind
B2
Abstract

A sound playback system and an output sound adjusting method thereof are disclosed. The method includes the following steps: receiving an input sound signal from a user, wherein the input sound signal includes a voice signal indicating the age of the user; transmitting the input sound signal to a remote voice system; performing a voice recognition process according to the voice signal of the input sound signal to obtain a voice recognition result; adjusting a gain value of each frequency band of an output sound signal according to the voice recognition result; and transmitting the output sound signal to a near-end electronic device to output the output sound signal to be heard by the user.

Claims (31)

1. An output sound adjusting method used for allowing a user to operate a near-end electronic device to adjust an output sound signal emitted by a remote voice system, the method comprising the following steps:

receiving an input sound signal emitted by the user, wherein the input sound signal comprises a voice signal representing the user's age;

transmitting the input sound signal to the remote voice system;

executing a voice recognition process according to the voice signal of the input sound signal to obtain a voice recognition result;

executing a voiceprint analysis process based on the input sound signal to obtain an age analysis result;

capturing a facial image of the user;

executing an image analysis process based on the facial image to obtain a facial image analysis result;

referring to the facial image analysis result, the voice recognition result, and the age analysis result at the same time so as to adjust a gain value of each frequency band of the output sound signal; and

transmitting the output sound signal to the near-end electronic device so as to emit the output sound signal to be heard by the user.

2. The output sound adjusting method as claimed in claim 1 , further comprising the following steps:

when the facial image analysis result, the voice recognition result, and the age analysis result are different, the gain value of each frequency band of the output sound signal is adjusted according to the facial image analysis result or the age analysis result.

3. The output sound adjusting method as claimed in claim 1 , further comprising the following steps:

when the voice recognition result and the facial image analysis result are different, the gain value of each frequency band of the output sound signal is adjusted according to the facial image analysis result.

4. A sound playback system, comprising:

a near-end electronic device, comprising:

a sound receiving module for receiving a input sound signal emitted by a user, wherein the input sound signal comprises a voice signal representing the user's age;

a transmission module, electrically connected to the sound receiving module for transmitting the input sound signal to a network; and

a sound module, electrically connected to the transmission module for emitting an output sound signal to be heard by the user; and a remote voice system, connected to the near-end electronic device via the network for transmitting the output sound signal to the near-end electronic device to emit the output sound signal from the sound module, the remote voice system comprising:

a recognition module, for receiving the input sound signal to execute a voice recognition process according to the voice signal of the input sound signal so as to obtain a voice recognition result;

an equalizer, for adjusting a gain value of each frequency band of the output sound signal; and

a processing module, electrically connected to the recognition module and the equalizer, for controlling the equalizer to adjust the gain value of each frequency band of the output sound signal according to the voice recognition result so as to transmit the output sound signal to the near-end electronic device to emit the output sound signal from the sound module.

5. The sound playback system as claimed in claim 4 , wherein the remote voice system further comprises a voiceprint analysis module used for executing a voiceprint analysis process based on the input sound signal to obtain an age analysis result, such that the processing module compares the voice recognition result and the age analysis result so as to control the equalizer to adjust the gain value of each frequency band of the output sound signal.

6. The sound playback system as claimed in claim 5 , wherein when the voice recognition result and the age analysis result are different, the processing module controls the equalizer to adjust the gain value of each frequency band of the output sound signal according to the age analysis result.

7. The sound playback system as claimed in claim 5 , wherein the near-end electronic device further comprises a capturing module used for capturing a facial image of the user; the remote voice system further comprises an image analysis module used for executing an image analysis process based on the facial image to obtain a facial image analysis result, such that the processing module compares the facial image analysis result, the voice recognition result, and the age analysis result to control the equalizer to adjust the gain value of each frequency band of the output sound signal.

8. The sound playback system as claimed in claim 7 , wherein when the facial image analysis result, the voice recognition result, and the age analysis result are different, the processing module controls the equalizer to adjust the gain value of each frequency band of the output sound signal according to the facial image analysis result or the age analysis result.

9. The sound playback system as claimed in claim 4 , wherein the near-end electronic device further comprises a capturing module used for capturing a facial image of the user and the remote voice system further comprises an image analysis module used for executing an image analysis process based on the facial image to obtain a facial image analysis result, such that the processing module compares the facial image analysis result and the voice recognition result to control the equalizer to adjust the gain value of each frequency band of the output sound signal.

10. The sound playback system as claimed in claim 9 , wherein when the voice recognition result and the facial image analysis result are different, the processing module controls the equalizer to adjust the gain value of each frequency band of the output sound signal according to the facial image analysis result.

11. A remote voice system, for receiving an input sound signal from a near-end electronic and transmitting an output sound signal to the near-end electronic correspondingly, the remote voice system comprising:

a recognition module, for receiving the input sound signal to execute a voice recognition process according to a voice signal of the input sound signal so as to obtain a voice recognition result;

an equalizer, for adjusting a gain value of each frequency band of the output sound signal; and

a processing module, electrically connected to the recognition module and the equalizer, for controlling the equalizer to adjust the gain value of each frequency band of the output sound signal according to the voice recognition result so as to transmit the output sound signal to the near-end electronic device.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 6, 2022
From: PIXART IMAGING INC.
To: AIROHA TECHNOLOGY CORP.
Reel/Frame 060591/0264 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 6, 2020
From: UNLIMITER MFA CO., LTD.
To: PIXART IMAGING INC.
Reel/Frame 053985/0983 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 22, 2019
From: YANG, KUO-PING; KAO, KUO-WEI; HSIAO, KAI-YUAN; LI, JIAN-YING
To: UNLIMITER MFA CO., LTD.
Reel/Frame 051083/0804 →
Priority Claims (1)
TW 107147837 · Dec 28, 2018 · national
Continuity (1)
Related Publication 20200211579A1 · Jul 2, 2020
Cited By (1)
US 12,222,430