Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

multilingual Corpora for Ethiopian Languages

Domain:

natural language processing

Record type:

dataset
Creator:
Ash
Host:

Visit

www.kaggle.com

Languages

Amharic

Licenses

Apache 2.0

Similar

AAUThematic4LT/Parallel-Corpora-for-Ethiopian-LanguagesMultilingual Parallel Text Corpora for East African LanguagesWebCrawl African : A Multilingual Parallel Corpora for African LanguagesGhanaNLP Parallel Corpora: Comprehensive Multilingual Resources for Low-Resource Ghanaian LanguagesDNN-based Multilingual Acoustic Modeling for Four Ethiopian LanguagesGlot500: Scaling Multilingual Corpora and Language Models to 500 Languages

AAUThematic4LT/Parallel-Corpora-for-Ethiopian-Languages

# Parallel Corpora for Ethiopian Languages: This is a repository that consists of parallel corpora b

Multilingual Parallel Text Corpora for East African Languages

This is a partial multilingual parallel corpora of 5 East African languages. The dataset contains an

WebCrawl African : A Multilingual Parallel Corpora for African Languages

WebCrawl African is a mixed domain multilingual parallel corpora for a pool of African languages com

GhanaNLP Parallel Corpora: Comprehensive Multilingual Resources for Low-Resource Ghanaian Languages

Low resource languages present unique challenges for natural language processing due to the limited

DNN-based Multilingual Acoustic Modeling for Four Ethiopian Languages

In this paper, we present the results of experiments conducted on multilingual acoustic modeling in

Glot500: Scaling Multilingual Corpora and Language Models to 500 Languages

The NLP community has mainly focused on scaling Large Language Models (LLMs) vertically, i.e., makin