This dataset contains Bemba-to-English sentences which is intended to machine translation task. This dataset was gained from Tatoeba-Translations repository by Yasmin Moslem.
Tatoeba is a dataset of sentences and its translations [1].
There are several preprocessing processes done with this dataset.
Take for about ~300k English data from Tatoeba.
Translating that English data using our MT model and keep the translating score.