Recognizing the value of large data sets for speech-to-text data sets, and seeing the opportunity that there are many text corpuses for Amharic and Swahili languages, this project aims to design and build a robust, large scale, fault tolerant, highly available Kafka cluster that can be used to post a sentence and receive an audio file. By the end o
# Text-to-Speech-Data-Collection-Kafka-Airflow-Spark