Cosine metric learning for speaker verification in the i-vector space

Zhongxin Bai, Xiao Lei Zhang, Jingdong Chen

科研成果: 期刊稿件会议文章同行评审

4 引用 (Scopus)

摘要

It is known that the equal-error-rate (EER) performance of a speaker verification system is determined by the overlap region of the decision scores of true and imposter trials. Also, the cosine similarity scores of the true or imposter trials produced by the state-of-the-art i-vector front-end approximate to a Gaussian distribution, and the overlap region of the two classes of trials depends mainly on their between-class distance. Motivated by the above facts, this paper presents a cosine similarity learning (CML) framework for speaker verification, which combines classical compensation techniques and the cosine similarity scoring for improving the EER performance. CML minimizes the overlap region by enlarging the between-class distance while introducing a regularization term to control the with-in class variance, which is initialized by a traditional channel compensation technique such as linear discriminant analysis. Experiments are carried out to compare the proposed CML framework with several traditional channel compensation baselines on the NIST speaker recognition evaluation data sets. The results show that CML outperforms all the studied initialization compensation techniques.

源语言英语
页(从-至)1126-1130
页数5
期刊Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
2018-September
DOI
出版状态已出版 - 2018
活动19th Annual Conference of the International Speech Communication, INTERSPEECH 2018 - Hyderabad, 印度
期限: 2 9月 20186 9月 2018

指纹

探究 'Cosine metric learning for speaker verification in the i-vector space' 的科研主题。它们共同构成独一无二的指纹。

引用此