The dataset contains a train and test dataset for training and testing classification algorithms. The train dataset contains recordings from chimpanzees, red-capped mangabeys, mandrills and mixed species of guenons from the Mefou Primate Sanctuary in Cameroon, as well as background forest recordings from a primary wet tropical forest in South-West Gabon.
The test dataset is fully independent from the training dataset and contains recordings from chimpanzees from the semi-natural Sanaga Sanctuary in Cameroon. Semi-natural here implies that enclosures contain a high degree of forest cover and are also surrounded by primary forests directly adjacent to the enclosures. An earlier version of this multi species dataset was submitted for the Interspeech 2021 challenge (Schuller et al. 2021; Zwerts et al. 2021).
All recordings were made using Audiomoth (v1.1.0) recorders [16]. Devices recorded 1- min segments continuously at 48 kHz and 30.6 dB gain, storing the data in one minute WAVE-files, with interruptions from two to five seconds between recordings for the recorder to save the files.
References:
Schuller, B.W., Batliner, A., Bergler, C., Mascolo, C., Han, J., Lefter, I., Kaya, H., Amiriparian, S., Baird, A., Stappen, L., Ottl, S., Gerczuk, M., Tzirakis, P., Brown, C., Chauhan, J., Grammenos, A., Hasthanasombat, A., Spathis, D., Xia, T., Cicuta, P., Rothkrantz, L.J.M., Zwerts, J.A., Treep, J., Kaandorp, C.S. (2021) The INTERSPEECH 2021 Computational Paralinguistics Challenge: COVID-19 Cough, COVID-19 Speech, Escalation & Primates. Proc. Interspeech 2021, 431-435, doi: 10.21437/Interspeech.2021-19
Zwerts, J.A., Treep, J., Kaandorp, C.S., Meewis, F., Koot, A.C., Kaya, H. (2021) Introducing a Central African Primate Vocalisation Dataset for Automated Species Classification. Proc. Interspeech 2021, 466-470, doi: 10.21437/Interspeech.2021-154