Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Koumankan4dyula

Domain:

natural language processing

Record type:

dataset
Creator:
uvci
Host:
The Koumankan4Dyula corpus consists of 10929 pairs of Dioula-French sentences. This corpus is part of the Koumankan project, which proposes a scalable and cost-effective method for extending the CommonVoice dataset to the Dyula language and other African languages. Train 73% 8065 Valid 14% 1471 Test 13% 1393 Maintenance

Visit

huggingface.co

Tasks

machine translation

Languages

BamanankanJula

Similar

uvci/koumankan4dyulafrench-datasets/uvci-koumankan4dyula

uvci/koumankan4dyula

The Koumankan4Dyula corpus consists of approximately 15 hours i.e. 10,929 recordings of Dioula langu

french-datasets/uvci-koumankan4dyula

Ce répertoire est vide, il a été créé pour améliorer le référencement du jeu de données uvci/koumank