Dataset Description:
This dataset is a large-scale collection of 323 hours of processed Somali (SO) and 105 hours of processed Somali (UG) single-channel call center audio recordings, containing 2,467,010 hours of processed call center audio recordings across 32 languages, designed to support the development and training of advanced speech AI and conversational AI systems.