Dempster-Shafer theory for enhanced statistical model-based voice activity detection

Article ID	Journal	Published Year	Pages	File Type
4973638	Computer Speech & Language	2018	12 Pages	PDF

Abstract

â¢We develop the voice activity detection based on DS theory.â¢Three statistical model-based VADs are used as the baseline systems.â¢Probabilities from the three VADs are combined to DS theory.â¢Proposed system works well over the existing methods.

In this paper, we propose to combine the posterior probabilities of voice activity derived from different statistical model-based algorithms for enhanced voice activity detection. For this, the Dempster-Shafer (DS) theory of evidence is employed to represent and combine the different probabilities estimated by three different statistical model-based VAD algorithms including the Sohn's likelihood ratio test (LRT)-based method, smoothed LRT-based method, and multiple observation LRT-based method. By considering a generalization of the Bayesian framework and permitting the characterization of uncertainty and ignorance through the DS theory, the probability of an ignorant state is eliminated through the orthogonal sum of several speech presence probabilities, which results in the performance improvement when detecting voice activity. According to objective test results, it is discovered the proposed DS theory-based VAD method offers significant improvements over the conventional approaches.

Keywords

Likelihood ratio test Dempster-Shafer theory Voice activity detection