View article

[PDF] from isca-archive.org

A New Feature for Automatic Speaker Verification Anti-Spoofing: Constant Q Cepstral Coefficients

Authors

Massimiliano Todisco, Héctor Delgado, Nicholas Evans

Publication date

2016

Conference

Odyssey 2016, The Speaker and Language Recognition Workshop

Description

Efforts to develop new countermeasures in order to protect automatic speaker verification from spoofing have intensified over recent years. The ASVspoof 2015 initiative showed that there is great potential to detect spoofing attacks, but also that the detection of previously unforeseen spoofing attacks remains challenging. This paper argues that there is more to be gained from the study of features rather than classifiers and introduces a new feature for spoofing detection based on the constant Q transform, a perceptually-inspired time-frequency analysis tool popular in the study of music. Experimental results obtained using the standard ASVspoof 2015 database show that, when coupled with a standard Gaussian mixture model-based classifier, the proposed constant Q cepstral coefficients (CQCCs) outperform all previously reported results by a significant margin. In particular, those for a subset of unknown spoofing attacks (for which no matched training data was used) is 0.46%, a relative improvement of 72% over the best, previously reported results.

Total citations

Cited by 421

20162017201820192020202120222023202410 45 48 56 56 67 59 52 28

Scholar articles

A New Feature for Automatic Speaker Verification Anti-Spoofing: Constant Q Cepstral Coefficients.

M Todisco, H Delgado, NWD Evans - Odyssey, 2016