Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Mada-French Parallel Corpus 1.0

Domain:

natural language processing

Record type:

dataset
Creator:
Ins
Host:
This dataset comprises a parallel corpus of Mada–French literary text translations totalling 2,154 lines. It is designed to support the benchmarking, training and evaluation of machine translation models for Mada, a language spoken in Cameroon. The corpus provides aligned, sentence and paragraph-level translations that capture the stylistic, lexical and syntactic features of literary Mada discourse and how these are rendered in the local variety of French.

Visit

mozilladatacollective.com

Tasks

machine translation

Languages

MadaMada

Tags

mdcmozilla data collectiveMTTSV

Licenses

Nwulite Obodo Open Data Licence 1.0 (NOODL-1.0)

Similar

French-Fongbe Parallel CorpusFrench-Adja Parallel CorpusFrench–Medumba Parallel CorpusEwondo-French Parallel CorpusSFPC (Sango-French Parallel Corpus)besacier/mboshi-french-parallel-corpus

French-Fongbe Parallel Corpus

Ce dataset est un corpus parallèle Français-Fongbe (Bénin) généré par IA et structuré pour l'entraîn

French-Adja Parallel Corpus

The first publicly available parallel text corpus for Adja machine translation, targeting an under-r

French–Medumba Parallel Corpus

A small parallel corpus of French ↔ Medumba (byv) sentence pairs, intended as a seed resource for ma

Ewondo-French Parallel Corpus

This dataset is a parallel corpus of Ewondo and French texts. The text was obtained by transcribing

SFPC (Sango-French Parallel Corpus)

The first quality-filtered, verse-aligned Sango-French parallel corpus, constructed for neural machi

besacier/mboshi-french-parallel-corpus

# mboshi-french-parallel-corpus This repository contains a speech corpus collected during a realist