Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech

Katrin Tomanek; Vicky Zayats; Dirk Padfield; Kara Vaillancourt; Fadi Biadsy

Conference ProceedingsOPEN ACCESS

Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech

EMNLP 2021 - 2021 Conference on Empirical Methods in Natural Language Processing, Proceedings (2021) 6751-6760

DOI: 10.18653/v1/2021.emnlp-main.541

34Citations

71Readers

Abstract

Automatic Speech Recognition (ASR) systems are often optimized to work best for speakers with canonical speech patterns. Unfortunately, these systems perform poorly when tested on atypical speech and heavily accented speech. It has previously been shown that personalization through model fine-tuning substantially improves performance. However, maintaining such large models per speaker is costly and difficult to scale. We show that by adding a relatively small number of extra parameters to the encoder layers via so-called residual adapter, we can achieve similar adaptation gains compared to model fine-tuning, while only updating a tiny fraction (less than 0.5%) of the model parameters. We demonstrate this on two speech adaptation tasks (atypical and accented speech) and for two state-of-the-art ASR architectures.

Cite

CITATION STYLE

APA

Tomanek, K., Zayats, V., Padfield, D., Vaillancourt, K., & Biadsy, F. (2021). Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech. In EMNLP 2021 - 2021 Conference on Empirical Methods in Natural Language Processing, Proceedings (pp. 6751–6760). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2021.emnlp-main.541

Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech

Abstract

Cite

Register to see more suggestions