Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

1000 African Voices: Advancing inclusive multi-speaker multi-accent speech synthesis

Domain:

natural language processing

Record type:

papermodel
Creator:
OguOwoOlaAle
Editor:
SpeLul
Publisher:
CCSD
Host:avatar
Accepted at Interspeech 2024 International audience Recent advances in speech synthesis have enabled many useful applications like audio directions in Google Maps, screen readers, and automated content generation on platforms like TikTok. However, these systems are mostly dominated by voices sourced from data-rich geographies with personas representative of their source data. Although 3000 of the world's languages are domiciled in Africa, African voices and personas are under-represented in these systems. As speech synthesis becomes increasingly democratized, it is desirable to increase the representation of African English accents. We present Afro-TTS, the first pan-African accented English speech synthesis system able to generate speech in 86 African accents, with 1000 personas representing the rich phonological diversity across the continent for downstream application in Education, Public Health, and Automated Content Creation. Speaker interpolation retains naturalness and accentedness, enabling the creation of new voices.

Visit

hal.science

Tasks

speech processingtext to speech

Tags

text-to-speechAfrican-accented TTSaccented speechmulti-accent TTSmulti-speaker TTS[INFO.INFO-SD]Computer Science [cs]/Sound [cs.SD][INFO]Computer Science [cs][INFO.INFO-AI]Computer Science [cs]/Artificial Intelligence [cs.AI]

Licenses

https://creativecommons.org/licenses/by/4.0/info:eu-repo/semantics/OpenAccess