Design Guidelines for Inclusive Speaker Verification Evaluation Datasets

Hutiri, Wiebke Toussaint; Gorce, Lauriane; Ding, Aaron Yi

Electrical Engineering and Systems Science > Audio and Speech Processing

arXiv:2204.02281 (eess)

[Submitted on 5 Apr 2022 (v1), last revised 13 Sep 2022 (this version, v2)]

Title:Design Guidelines for Inclusive Speaker Verification Evaluation Datasets

Authors:Wiebke Toussaint Hutiri, Lauriane Gorce, Aaron Yi Ding

View PDF

Abstract:Speaker verification (SV) provides billions of voice-enabled devices with access control, and ensures the security of voice-driven technologies. As a type of biometrics, it is necessary that SV is unbiased, with consistent and reliable performance across speakers irrespective of their demographic, social and economic attributes. Current SV evaluation practices are insufficient for evaluating bias: they are over-simplified and aggregate users, not representative of real-life usage scenarios, and consequences of errors are not accounted for. This paper proposes design guidelines for constructing SV evaluation datasets that address these short-comings. We propose a schema for grading the difficulty of utterance pairs, and present an algorithm for generating inclusive SV datasets. We empirically validate our proposed method in a set of experiments on the VoxCeleb1 dataset. Our results confirm that the count of utterance pairs/speaker, and the difficulty grading of utterance pairs have a significant effect on evaluation performance and variability. Our work contributes to the development of SV evaluation practices that are inclusive and fair.

Comments:	Accepted to INTERSPEECH 2022 (submitted version)
Subjects:	Audio and Speech Processing (eess.AS); Computers and Society (cs.CY); Machine Learning (cs.LG)
Cite as:	arXiv:2204.02281 [eess.AS]
	(or arXiv:2204.02281v2 [eess.AS] for this version)
	https://doi.org/10.48550/arXiv.2204.02281

Submission history

From: Wiebke Toussaint Hutiri [view email]
[v1] Tue, 5 Apr 2022 15:28:26 UTC (253 KB)
[v2] Tue, 13 Sep 2022 13:05:52 UTC (276 KB)

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:Design Guidelines for Inclusive Speaker Verification Evaluation Datasets

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:Design Guidelines for Inclusive Speaker Verification Evaluation Datasets

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators