Features for voice activity detection: a comparative analysisReportar como inadecuado

Features for voice activity detection: a comparative analysis - Descarga este documento en PDF. Documentación en PDF para descargar gratis. Disponible también para leer online.

EURASIP Journal on Advances in Signal Processing

, 2015:91

Advances in adaptive filtering theory and applications to acoustic and speech signal processing


In many speech signal processing applications, voice activity detection VAD plays an essential role for separating an audio stream into time intervals that contain speech activity and time intervals where speech is absent. Many features that reflect the presence of speech were introduced in literature. However, to our knowledge, no extensive comparison has been provided yet. In this article, we therefore present a structured overview of several established VAD features that target at different properties of speech. We categorize the features with respect to properties that are exploited, such as power, harmonicity, or modulation, and evaluate the performance of some dedicated features. The importance of temporal context is discussed in relation to latency restrictions imposed by different applications. Our analyses allow for selecting promising VAD features and finding a reasonable trade-off between performance and complexity.

KeywordsSpeech detection Speech properties Feature selection  Download fulltext PDF

Autor: Simon Graf - Tobias Herbig - Markus Buck - Gerhard Schmidt

Fuente: https://link.springer.com/

Documentos relacionados