Logo Lanfrica

inandutu/luganda_dataset

Domain:

natural language processing

Record type:

dataset
Creator:
ina
Host:
# luganda_dataset ## The corpus is helpful in projects related to NLP and luganda language ### The file contains the phonetic file, transcription file, lexicon dictionary and 511 database sentences ### Align the sentences with corresponding wav files ie recording ### Added Luganda Languange and voice github.com ### The complete research is a speech synthesis systems that detects Luganda text to Luganda speech using MARYTTS engine. The Luganda downloadable voice is here goo.gl. # Instructions to Install and Test the Synthesized Voice ### Install Java and add it to system path ### Clone github.com or version 5.2 SNAPSHOT. ### Download Luganda target folder from dropbox link goo.gl ### Unzip it and place it at the root of marytts project ### Navigate to target/marytts-5.2-SNAPSHOT/bin/marytts-server in command line. When the server is running; browse localhost ### github.com . The little dataset used to train the voices. test/.txt and lg.txt # Paper ### Link to the paper arxiv.org