Where Are We at with Automatic Speech Recognition for the Bambara Language?

Seydou DIALLO; Yacouba Diarra; Panga Azazia Kamaté; Aboubacar Ouattara; Mamadou K. KEITA; Adam Bouno Kampo

Where Are We at with Automatic Speech Recognition for the Bambara Language?

Seydou DIALLO, Yacouba Diarra, Panga Azazia Kamaté, Aboubacar Ouattara, Mamadou K. KEITA, Adam Bouno Kampo

Published: 27 Jan 2026, Last Modified: 17 Feb 2026AfricaNLP 2026EveryoneRevisionsBibTeXCC BY 4.0

Abstract: This paper introduces the first standardized benchmark for evaluating Automatic Speech Recognition (ASR) in the Bambara language, utilizing one hour of professionally recorded Malian constitutional text. Designed as a controlled reference set under near-optimal acoustic and linguistic conditions, the benchmark was used to evaluate 37 models, ranging from Bambara-trained systems to large-scale commercial models. Our findings reveal that current ASR performance remains significantly below deployment standards; the top-performing system in terms of Word Error Rate (WER) achieved 46.76\% and the best Character Error Rate (CER) of 13.00\% was set by another model, while several prominent multilingual models exceeded 100\% WER due to severe hallucinations. These results suggest that multilingual pre-training and model scaling alone are insufficient for underrepresented languages. Furthermore, because this dataset represents a best-case scenario of the most simplified and formal form of spoken Bambara, these figures likely establish an upper bound for performance in practical, real-world settings. We provide the benchmark and an accompanying public leaderboard to facilitate transparent evaluation and future research in Bambara speech technology.

Submission Number: 49

Loading