Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

ik/akan-tts-wavtokenizer-combined

Domain:

natural language processing

Record type:

dataset
Creator:
IK
Host:
Combined word-aligned dataset for training Akan/Twi TTS models. Audio encoded with WavTokenizer (75Hz, single codebook, codes 0-4095) and word boundaries from MMS forced alignment. Total samples 96,615 Total hours 222.9h Sources 5 Split Samples Hours train 95,165 219.4h validation 966 2.2h test 484 1.2h Source Samples

Visit

huggingface.co

Tasks

speech processingtext to speech

Languages

Akan

Tags

ttsspeechakantwiwavtokenizerword-aligned

Licenses

cc-by-sa-4.0

Similar

ik/asante-twi-wavtokenizer-alignedToadoum/yoruba-tts-combined

ik/asante-twi-wavtokenizer-aligned

Toadoum/yoruba-tts-combined