Yoruba Text C3 is the largest Yoruba texts collected and used to train FastText embeddings in the Yo
A parallel corpus of Asante Twi and English sentence pairs compiled by Ghana NLP Community. Source
Fasttext word embedding models for Yoruba and Twi languages based on the paper Massive vs. Cura
Transformer-based language models have been changing the modern Natural Language Processing (NLP) la
This dataset contains 166156 parallel speech-text pairs for Twi, a language spoken primarily in Ghan