Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Yoruba Twi Text C3

Domain:

natural language processing

Record type:

dataset
Yoruba Text C3 is the largest Yoruba texts collected and used to train FastText embeddings in the YorubaTwi Embedding paper: aclweb.org

Visit

github.com

Connected records

paper

Tasks

language modelingembeddings

Languages

Yoruba

Licenses

Similar

Yorùbá Text C3Twi-English Parallel TextYoruba-Twi FastText Embedding modelsContextual Text Embeddings for TwiTwi Trigrams Speech-Text Parallel Datasetghananlpcommunity/asante-twi-bible-speech-text

Yorùbá Text C3

Yoruba Text C3 is the largest Yoruba texts collected and used to train FastText embeddings in the Yo

Twi-English Parallel Text

A parallel corpus of Asante Twi and English sentence pairs compiled by Ghana NLP Community. Source

Yoruba-Twi FastText Embedding models

Fasttext word embedding models for Yoruba and Twi languages based on the paper    Massive vs. Cura

Contextual Text Embeddings for Twi

Transformer-based language models have been changing the modern Natural Language Processing (NLP) la

Twi Trigrams Speech-Text Parallel Dataset

This dataset contains 166156 parallel speech-text pairs for Twi, a language spoken primarily in Ghan

ghananlpcommunity/asante-twi-bible-speech-text