This dataset contains parallel text in three languages: Bambara (Bamanankan), French, and English. I
Configuration Train (hours) Dev (hours) Test (hours) Total (hours) bm-to-bm-weak 136.37 23.40 11.28
The Bambara-Texts dataset is a collection of monolingual Bambara text designed for pretraining langu
The Djelia Bambara Audio Dataset is a comprehensive resource aimed at supporting research and develo