Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

LESSONS LEARNED AFTER DEVELOPMENT AND USE OF A DATA COLLECTION APP FOR LANGUAGE DOCUMENTATION (LIG-AIKUMA)

Domain:

natural language processing

Record type:

papersoftwaredataset
Creator:
BesGauthier, ElodieVoi
Editor:
GroLabUniIns
Publisher:
CCSD
Host:avatar
International audience Lig-Aikuma is a free Android app running on various mobile phones and tablets. It proposes a range of different speech collection modes (recording, respeaking, translation and elicitation) and offers the possibility to share recordings between users. More than 250 hours of speech in 6 different languages from sub-Saharan Africa (including 3 oral languages in the process of being documented) have already been collected with Lig-Aikuma. This paper presents the lessons learned after 3 years of development and use of Lig-Aikuma. While significant data collections were conducted, this has not been done without difficulties. Some mixed results lead us to stress the importance of design choices, data sharing architecture and user manual. We also discuss other potential uses of the app, discovered during its deployment: data collection for language revitalisation, data collection for speech technology development (ASR) and enrichment of existing corpora through the addition of spoken comments.

Visit

hal.science

Tasks

speech processing

Tags

Eliciting speechSpeech recordingLanguage documentationData collectionMobile applications[INFO.INFO-CL]Computer Science [cs]/Computation and Language [cs.CL][SHS.LANGUE]Humanities and Social Sciences/Linguistics

Licenses

https://about.hal.science/hal-authorisation-v1/info:eu-repo/semantics/OpenAccess