Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

A corpus of conversational Luganda

Domain:

natural language processing

Record type:

dataset
Creator:
NamZelWitzlack-Makarevich, AlenaLor
Publisher:
Zenodo
Host:avatar
These are the transcript and translation into English of a corpus of conversational Luganda (or Ganda), a Great Lakes Bantu language of Uganda. The overarching aim of the project is to contribute towards a better linguistic description of Luganda. The files are in the .eaf format (the ELAN Annotation Format, also known as the EUDICO Annotation Format). The corpus was recorded in Kampala in 2019 and 2023. It contains conversations between colleagues, peers, and acquaintances. The audio recordings are available in a separate project. The video recordings of some conversations are available on request. Further annotations and transcriptions will be added in later versions.

Visit

doi.orgzenodo.org

Languages

Ganda

Tags

conversation, corpus, Ganda, Luganda, Bantu, Uganda

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcodeOpen Accessinfo:eu-repo/semantics/openAccess