View article

[PDF] from qut.edu.au

The QUT-NOISE-SRE protocol for the evaluation of noisy speaker recognition

Authors

David Dean, Ahilan Kanagasundaram, Houman Ghaemmaghami, Md Hafizur Rahman, Sridha Sridharan

Publication date

2015

Journal

Proceedings of the 16th Annual Conference of the International Speech Communication Association, Interspeech 2015

Pages

3456-3460

Publisher

International Speech Communication Association

Description

The QUT-NOISE-SRE protocol is designed to mix the large QUT-NOISE database, consisting of over 10 hours of back- ground noise, collected across 10 unique locations covering 5 common noise scenarios, with commonly used speaker recognition datasets such as Switchboard, Mixer and the speaker recognition evaluation (SRE) datasets provided by NIST. By allowing common, clean, speech corpora to be mixed with a wide variety of noise conditions, environmental reverberant responses, and signal-to-noise ratios, this protocol provides a solid basis for the development, evaluation and benchmarking of robust speaker recognition algorithms, and is freely available to download alongside the QUT-NOISE database. In this work, we use the QUT-NOISE-SRE protocol to evaluate a state-of-the-art PLDA i-vector speaker recognition system, demonstrating the importance of designing voice-activity-detection front-ends specifically for speaker recognition, rather than aiming for perfect coherence with the true speech/non-speech boundaries.

Total citations

Cited by 38

2016201720182019202020212022202320246 2 1 4 3 1 6 5 10

Scholar articles

The QUT-NOISE-SRE protocol for the evaluation of noisy speaker recognition

D Dean, A Kanagasundaram, H Ghaemmaghami… - Proceedings of the 16th Annual Conference of the …, 2015