By Topic

Unsupervised Speech Activity Detection Using Voicing Measures and Perceptual Spectral Flux

Sign In

Cookies must be enabled to login.After enabling cookies , please use refresh or reload or ctrl+f5 on the browser for the login options.

The purchase and pricing options are temporarily unavailable. Please try again later.
2 Author(s)
Sadjadi, S.O. ; Dept. of Electr. Eng., Univ. of Texas at Dallas, Richardson, TX, USA ; Hansen, J.H.L.

Effective speech activity detection (SAD) is a necessary first step for robust speech applications. In this letter, we propose a robust and unsupervised SAD solution that leverages four different speech voicing measures combined with a perceptual spectral flux feature, for audio-based surveillance and monitoring applications. Effectiveness of the proposed technique is evaluated and compared against several commonly adopted unsupervised SAD methods under simulated and actual harsh acoustic conditions with varying distortion levels. Experimental results indicate that the proposed SAD scheme is highly effective and provides superior and consistent performance across various noise types and distortion levels.

Published in:

Signal Processing Letters, IEEE  (Volume:20 ,  Issue: 3 )