Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Autshumato English-Sesotho sa Leboa Parallel Corpora

Domaine:

natural language processing

Type de record:

dataset
Créateur:
D.P. SnymanCindy McKellarHandré Groenewald
Éditeur:
North-West UniversityCentre for Text Technology (CTexT)
Hôte:avatar
Parallel corpora aligned on sentence level through a combination of automatic and manual alignment techniques. The parallel corpora were obtained from the SA government domain.

Visit

hdl.handle.net

Tasks

machine translation

Languages

Sotho, NorthernSotho, Southern

Licenses

Creative Commons Attribution-NonCommercial-ShareAlike 2.5 South Africa: http://creativecommons.org/licenses/by-nc-sa/2.5/za/

Similaires

Autshumato English-Sesotho Parallel CorporaAutshumato English-Sesotho sa Leboa Translation MemoryAutshumato English-Xitsonga Parallel CorporaAutshumato English-Setswana Parallel CorporaAutshumato English-Afrikaans Parallel CorporaAutshumato English-Sepedi Parallel Corpora

Autshumato English-Sesotho Parallel Corpora

Aligned parallel corpora for the language pair English-Sesotho. The data is given as two separate UT

Autshumato English-Sesotho sa Leboa Translation Memory

Translation memory from English (EN-GB) to Sesotho sa Leboa, in the government domain for use in the

Autshumato English-Xitsonga Parallel Corpora

Aligned English-Xitsonga parallel corpus. The data is given as two seperate UTF-8 text files; with e

Autshumato English-Setswana Parallel Corpora

Aligned English-Setswana parallel corpus. This set contains data that was translated by professional

Autshumato English-Afrikaans Parallel Corpora

Aligned parallel corpora for the language pair English-Afrikaans. The data is given as two separate

Autshumato English-Sepedi Parallel Corpora

Aligned parallel corpora for the language pair English-Sepedi. The data is given as two separate UTF