Title
Modified LTSE-VAD Algorithm for Applications Requiring Reduced Silence Frame Misclassification
Abstract
The LTSE-VAD is one of the best known algorithms for voice activity detection. In this paper we present a modified version of this algorithm, that makes the VAD decision not taking into account account the estimated background noise level, but the signal to noise ratio (SNR). This makes the algorithm robust not only to noise level changes, but also to signal level changes. We compare the modified algorithm with the original one, and with three other standard VAD systems. The results show that the modified version gets the lowest silence misclassification rate, while maintaining a reasonably low speech misclassification rate. As a result, this algorithm is more suitable for identification tasks, such as speaker or emotion recognition, where silence misclassification can be very harmful. A series of automatic emotion identification experiments are also carried out, proving that the modified version of the algorithm helps increasing the correct emotion classification rate.
Year
Venue
Field
2010
LREC 2010 - SEVENTH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION
Computer science,Emotion recognition,Ambient noise level,Signal level,Voice activity detection,Signal-to-noise ratio,Noise level,Algorithm,Emotion classification,Speech recognition,Silence
DocType
Citations 
PageRank 
Conference
1
0.39
References 
Authors
8
7
Name
Order
Citations
PageRank
Iker Luengo1948.58
Eva Navas230328.48
Igor Odriozola3104.55
Ibon Saratxaga416914.42
Inmaculada Hernáez510511.88
Iñaki Sainz6426.45
Daniel Erro713915.90