IP Library Granted Patent US 8,645,133
Granted Patent B2
US 8,645,133 · App. 13/761,307 · Granted Feb 4, 2014

Adaptation of voice activity detection parameters based on encoding modes

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,645,133
App. No.
13/761,307
Granted
Feb 4, 2014
Kind
B2
Abstract

Encoding audio signals with selecting an encoding mode for encoding the signal categorizing the signal into active segments having voice activity and non-active segments having substantially no voice activity by using categorization parameters depending on the selected encoding mode and encoding at least the active segments using the selected encoding mode.

Claims (43)

1. A method comprising:

dividing an audio signal into a plurality of segments;

categorizing each of the plurality of segments as an active segment or a non-active segment based at least in part on one or more categorization parameters, at least one of the one or more categorization parameters being dependent upon a selected encoding mode for encoding the segments;

encoding at least those segments of the plurality of segments categorized as active segments using the selected mode for encoding.

2. The method of claim 1 , wherein the at least one of the one or more categorization parameters is such that for a low quality of the selected encoding mode a lower number of temporal sections are detected as active sections than for a high quality of the selected encoding mode.

3. The method of claim 1 , wherein:

the one or more categorization parameters include at least one parameter that comprises an energy threshold value; and

categorizing each of the plurality of segments comprises comparing energy information of the audio signal to at least the energy threshold value.

4. The method of claim 1 , wherein:

the one or more categorization parameters include at least one parameter that comprises a signal-to-noise threshold value; and

categorizing each of the plurality of segments comprises comparing signal-to-noise information of the audio signal to at least the signal-to-noise threshold value.

5. The method of claim 1 , wherein:

the one or more categorization parameters include at least one parameter that comprises pitch information; and

categorizing each of the plurality of segments comprises comparing the pitch of the audio signal to at least the pitch information.

6. The method of claim 1 , wherein:

the one or more categorization parameters include at least one parameter that comprises tone information; and

categorizing each of the plurality of segments comprises comparing the tone of the audio signal to at least the tone information.

7. The method of claim 1 , further comprising creating spectral sub-bands from the audio signal.

8. The method of claim 7 , wherein categorizing each of the plurality of segments comprises categorizing selected sub-bands.

9. The method of claim 1 , wherein the one or more categorization parameters include at least one parameter that is dependent upon noise information.

10. The method of claim 1 , wherein the one or more categorization parameters include at least one parameter that is dependent upon traffic information.

11. An apparatus comprising:

a division unit arranged for dividing an audio signal into a plurality of segments;

an adaptive categorization unit arranged for categorizing each of the plurality of segments as an active segment or a non-active based at least in part on one or more categorization parameters, at least one of the one or more categorization parameters being dependent upon a selected encoding mode for encoding the segments; and

an encoding unit arranged for encoding at least those segments of the plurality of segments categorized as active segments using the selected mode for encoding.

12. The apparatus of claim 11 , wherein the at least one of the one or more categorization parameters depends on an encoding bitrate of the encoding mode.

13. The apparatus of claim 11 , wherein the one or more categorization parameters include one or more of:

at least one parameter that comprises an energy threshold value;

at least one parameter that comprises a signal-to-noise threshold value;

at least one parameter that comprises pitch information; and

at least one parameter that comprises tone information.

14. The apparatus of claim 11 , wherein the one or more categorization parameters include at least one parameter that is dependent upon noise information.

15. The apparatus of claim 11 , wherein the one or more categorization parameters include at least one parameter that is dependent upon traffic information.

16. A system comprising:

a transmission network;

a transmitter comprising an audio encoder with a division unit arranged for dividing an audio signal into a plurality of segments;

an adaptive categorization unit arranged for categorizing the plurality of segments into active segments and non-active segments based at least in part on one or more categorization parameters, at least one of the one or more categorization parameters being dependent upon a selected encoding mode for encoding the segments; and

an encoding unit arranged for encoding at least those segments of the plurality of segments categorized as active segments using the selected mode for encoding; and

a receiver for receiving the encoded audio signal.

17. A chipset comprising:

a division unit arranged for dividing an audio signal into a plurality of segments;

an adaptive categorization unit arranged for categorizing each of the plurality of segments as an active segment or a non-active segment based at least in part on one or more categorization parameters, at least one of the one or more categorization parameters being dependent upon a selected encoding mode for encoding the segments; and

an encoding unit arranged for encoding at least the active segments using the selected encoding mode.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 12, 2022
From: PIECE FUTURE PTE LTD.
To: NOKIA TECHNOLOGIES OY
Reel/Frame 062115/0779 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 17, 2022
From: NOKIA TECHNOLOGIES OY
To: PIECE FUTURE PTE LTD
Reel/Frame 058673/0912 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 12, 2018
From: CONVERSANT WIRELESS LICENSING S.A R.L.
To: NOKIA TECHNOLOGIES OY
Reel/Frame 046851/0302 →
CHANGE OF NAME Recorded Oct 20, 2017
From: CORE WIRELESS LICENSING S.A.R.L.
To: CONVERSANT WIRELESS LICENSING S.A R.L.
Reel/Frame 044242/0401 →