Investigating Active-learning-based Training Data Selection for Speech Spoofing Countermeasure

Wang, Xin; Yamagishi, Junich

Electrical Engineering and Systems Science > Audio and Speech Processing

arXiv:2203.14553 (eess)

[Submitted on 28 Mar 2022 (v1), last revised 7 Oct 2022 (this version, v3)]

Title:Investigating Active-learning-based Training Data Selection for Speech Spoofing Countermeasure

Authors:Xin Wang, Junich Yamagishi

View PDF

Abstract:Training a spoofing countermeasure (CM) that generalizes to various unseen data is desired but challenging. While methods such as data augmentation and self-supervised learning are applicable, the imperfect CM performance on diverse test sets still calls for additional strategies. This study took the initiative and investigated CM training using active learning (AL), a framework that iteratively selects useful data from a large pool set and fine-tunes the CM. This study compared a few methods to measure the data usefulness and the impact of using different pool sets collected from various sources. The results showed that the AL-based CMs achieved better generalization than our strong baseline on multiple test tests. Furthermore, compared with a top-line CM that simply used the whole data pool set for training, the AL-based CMs achieved similar performance using less training data. Although no single best configuration was found for AL, the rule of thumb is to include diverse spoof and bona fide data in the pool set and to avoid any AL data selection method that selects the data that the CM feels confident in.

Comments:	To appear in Proc. SLT 2022, modified based on a paper rejected by Interspeech 2022
Subjects:	Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2203.14553 [eess.AS]
	(or arXiv:2203.14553v3 [eess.AS] for this version)
	https://doi.org/10.48550/arXiv.2203.14553

Submission history

From: Xin Wang [view email]
[v1] Mon, 28 Mar 2022 07:53:03 UTC (1,563 KB)
[v2] Sun, 19 Jun 2022 12:24:32 UTC (1,561 KB)
[v3] Fri, 7 Oct 2022 12:45:47 UTC (496 KB)

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:Investigating Active-learning-based Training Data Selection for Speech Spoofing Countermeasure

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:Investigating Active-learning-based Training Data Selection for Speech Spoofing Countermeasure

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators