English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System

Cámbara, Guillermo; Peiró-Lilja, Alex; Farrús, Mireia; Luque, Jordi

Electrical Engineering and Systems Science > Audio and Speech Processing

arXiv:2105.05041 (eess)

[Submitted on 9 May 2021]

Title:English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System

Authors:Guillermo Cámbara, Alex Peiró-Lilja, Mireia Farrús, Jordi Luque

View PDF

Abstract:Nowadays, research in speech technologies has gotten a lot out thanks to recently created public domain corpora that contain thousands of recording hours. These large amounts of data are very helpful for training the new complex models based on deep learning technologies. However, the lack of dialectal diversity in a corpus is known to cause performance biases in speech systems, mainly for underrepresented dialects. In this work, we propose to evaluate a state-of-the-art automatic speech recognition (ASR) deep learning-based model, using unseen data from a corpus with a wide variety of labeled English accents from different countries around the world. The model has been trained with 44.5K hours of English speech from an open access corpus called Multilingual LibriSpeech, showing remarkable results in popular benchmarks. We test the accuracy of such ASR against samples extracted from another public corpus that is continuously growing, the Common Voice dataset. Then, we present graphically the accuracy in terms of Word Error Rate of each of the different English included accents, showing that there is indeed an accuracy bias in terms of accentual variety, favoring the accents most prevalent in the training corpus.

Comments:	2 pages, 1 figure, 1 table. To be published in Phonetics and Phonology in Europe 2021
Subjects:	Audio and Speech Processing (eess.AS); Computation and Language (cs.CL); Machine Learning (cs.LG); Sound (cs.SD)
Cite as:	arXiv:2105.05041 [eess.AS]
	(or arXiv:2105.05041v1 [eess.AS] for this version)
	https://doi.org/10.48550/arXiv.2105.05041

Submission history

From: Guillermo Cámbara [view email]
[v1] Sun, 9 May 2021 08:24:33 UTC (215 KB)

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Audio and Speech Processing

Title:English Accent Accuracy Analysis in a State-of-the-Art Automatic Speech Recognition System

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators