Dataset Description:
This dataset is a large-scale collection of 136,385 hours of processed Swahili (SW) dual-channel call center audio recordings, containing 2,065,026 hours of processed call center audio recordings across 32 languages, designed to support the development and training of advanced speech AI and conversational AI systems.