AFRIMGSM is an evaluation dataset comprising translations of a subset of the GSM8k dataset into 16 African languages.
It includes test sets across all 18 languages, maintaining an English and French subsets from the original GSM8k dataset.
There are 18 languages available :
The examples look like this for English:
from datasets import load_dataset