A large-scale multilingual speech dataset developed by Data Science Nigeria. Contains more than 3,000 hours of transcribed audio across four Nigerian languages: Hausa, Igbo, Nigerian Pidgin, and Yorùbá. The dataset supports Automatic Speech Recognition (ASR) and speech technology for low-resource African languages. It combines both scripted and spontaneous speech collected through community-centered, ethical protocols that respect linguistic and cultural diversity.