Dataset Description:
This dataset is a large-scale collection of 138,905 hours of processed Swahili (SW) single-channel call center audio recordings, containing 2,467,010 hours of processed call center audio recordings across 32 languages, designed to support the development and training of advanced speech AI and conversational AI systems.