کد مقاله | کد نشریه | سال انتشار | مقاله انگلیسی | نسخه تمام متن |
---|---|---|---|---|
566025 | 1452025 | 2015 | 14 صفحه PDF | دانلود رایگان |
• A method of speech enhancement that reconstructs clean speech from acoustic features.
• Features estimated by a statistical method incorporating noise and speaker adaptation.
• Listening tests find enhancement highly effective in reducing background noise.
This paper proposes a method of speech enhancement where a clean speech signal is reconstructed from a sinusoidal model of speech production and a set of acoustic speech features. The acoustic features are estimated from noisy speech and comprise, for each frame, a voicing classification (voiced, unvoiced or non-speech), fundamental frequency (for voiced frames) and spectral envelope. Rather than using different algorithms to estimate each parameter, a single statistical model is developed. This comprises a set of acoustic models and has similarity to the acoustic modelling used in speech recognition. This allows noise and speaker adaptation to be applied to acoustic feature estimation to improve robustness. Objective and subjective tests compare reconstruction-based enhancement with other methods of enhancement and show the proposed method to be highly effective at removing noise.
Journal: Speech Communication - Volume 75, December 2015, Pages 62–75