Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Parallel English–Akan Navigation Dataset Parallel English–Akan Navigation Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
IsaIsaDanFii
Éditeur:
Sci
Hôte:avatar
This corpus was developed to support assistive navigation technology for visually impaired Akan-speaking users, including a smart GPS navigation aid that delivers spoken guidance in Akan. It contains a curated bilingual parallel corpus developed to support domain-specific speech and language technologies, primarily, automatic speech recognition (ASR) and text-to-speech (TTS). The corpus was constructed to address the scarcity of navigation and accessibility-specific data for low-resource African languages, particularly Akan, and to support voice-enabled mobility aids for visually impaired users. It comprises 6,600 English–Akan parallel sentences spanning seven functional navigation categories: route instructions, hazard and safety alerts, contextual orientation cues, status and confirmation feedback, landmark and location references, environmental descriptions, and system status prompts. Navigation topics represented include turn-by-turn route guidance, gutter and vehicle hazards, road crossing, cardinal direction and spatial positioning, terrain and elevation cues, on-track and off-track confirmations, campus halls and academic buildings, gates and junctions, indoor and outdoor environmental description, and device status and connectivity prompts. The corpus also contains approximately 9.03 hours of Akan audio recordings (one male (IM), one female voice (HA)) corresponding to the 6,600 sentences. This corpus was developed to support assistive navigation technology for visually impaired Akan-speaking users, including a smart GPS navigation aid that delivers spoken guidance in Akan. It contains a curated bilingual parallel corpus developed to support domain-specific speech and language technologies, primarily, automatic speech recognition (ASR) and text-to-speech (TTS). The corpus was constructed to address the scarcity of navigation and accessibility-specific data for low-resource African languages, particularly Akan, and to support voice-enabled mobility aids for visually impaired users. It comprises 6,600 English–Akan parallel sentences spanning seven functional navigation categories: route instructions, hazard and safety alerts, contextual orientation cues, status and confirmation feedback, landmark and location references, environmental descriptions, and system status prompts. Navigation topics represented include turn-by-turn route guidance, gutter and vehicle hazards, road crossing, cardinal direction and spatial positioning, terrain and elevation cues, on-track and off-track confirmations, campus halls and academic buildings, gates and junctions, indoor and outdoor environmental description, and device status and connectivity prompts. The corpus also contains approximately 9.03 hours of Akan audio recordings (one male (IM), one female voice (HA)) corresponding to the 6,600 sentences.

Visit

doi.org

Tasks

automatic speech recognitionmachine translationspeech processingtext to speech

Languages

Akan

Tags

LinguisticsArtificial intelligenceAkanlow-resource languagesmachine translationnatural language processingparallel corpustext-to-speechassistive navigation systemsvisually impaired

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

Akan–English Maternal Health Parallel Corpus Akan–English Maternal Health Parallel CorpusTwi-English Parallel DatasetEnglish-Giriama Parallel Sentence DatasetPristine Twi-English Parallel DatasetPristine Twi-English Parallel DatasetSinhala-English Parallel Word Dictionary Dataset

Akan–English Maternal Health Parallel Corpus Akan–English Maternal Health Parallel Corpus

This dataset contains a curated bilingual parallel corpus developed to support domain-specific neura

Twi-English Parallel Dataset

This dataset contains a large-scale parallel corpus of Twi-English sentence pairs, featuring synthet

English-Giriama Parallel Sentence Dataset

This dataset consists of sentence pairs in English and their corresponding translations in Giriama (

Pristine Twi-English Parallel Dataset

A large-scale Twi ↔ English parallel dataset derived from the Pristine Twi Dataset by the Ghana NLP

Pristine Twi-English Parallel Dataset

A large-scale Twi ↔ English parallel dataset derived from the Pristine Twi Dataset by the Ghana NLP

Sinhala-English Parallel Word Dictionary Dataset

Parallel datasets are vital for performing and evaluating any kind of multilingual task. However, in