Logo Lanfrica

Afri-XNLI+: Extending the XNLI Dataset with New African Language Translations

Domain:

natural language processing

Record type:

dataset
Creator:
KonIbeAmoKes
Publisher:
Ton
Host:avatar
Afri-XNLI+ is a multilingual dataset consisting of translated premise-hypothesis sentence pairs from the XNLI development and test sets. The resource provides aligned translations for multiple African languages while preserving the original entailment, contradiction, and neutral labels. The dataset was created using a hybrid workflow combining machine translation assistance, human translation, and validation.