Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

brucewlee/LAMA-Music-Genre-Dataset

Record type:

dataset
Creator:
bru
Host:
.wav files, training dataset (MFCC), and graph plots (FFTs, MFCCs, Waveforms) from Latin America, Asia, MiddleEast, and Africa # LAMA : A World Music Genre Dataset > LAMA - LatinAmerica, Asia, MiddleEastern, Africa Genre Dataset This Dataset consists of the .wav files, training dataset (MFCC), and graph plots (FFTs, MFCCs, STFTs, Waveforms) of YouTube videos classified into four categories: LatinAmerica, Asia, MiddleEastern, and Africa. I hope that this work can help in several Deep Learning, Machine Learning projects in *Music Genre Classification*. **Anyone is free to use/change/contribute to this Dataset. Please cite if you use this Dataset in your research/projects.** ## Getting Started > Some data couldn't be uploaded to GitHub because the file size was too large. Instead, I attached a Harvard Dataverse link below to retrieve the data. The data contained in **LAMA** can be classified into three categories: 1. **.wav files** (`Data_Original`, `Data_Cropped`) -> Uploaded in Harvard Dataverse. Link below. 2. **Waveform, FFT, STFT, MFCC graph plots** (`Graphs`, `Sample_Graphs`) -> Uploaded in Harvard Dataverse. Link below. 3. **mfcc training data** (`training_data.json`) ## Statistics - `Data_Original` : 113 (Africa), 83 (Asia), 90 (LatinAmerica), 80 (MiddleEast) - `Data_Cropped` (1 min clips of original) : 108 (Africa), 77 (Asia), 87 (LatinAmerica), 77 (MiddleEast) - `Graphs` -> FFT : 108 (Africa), 77 (Asia), 87 (LatinAmerica), 77 (MiddleEast) Link to the full audio data ### graph plot EXAMPLES: *from `Graphs`* *from `Sample_Graphs`* ## Acknowledgements The classification of genre in this Dataset is mostly from the "AudioSet" project by the Sound and Video Understanding teams at Google Research. I chose the best examples from their website and preprocessed them to create this Dataset. In addition, I obtained much of the knowledge needed to create this Dataset from the YouTube Channel, "Valerio Velardo - The Sound of AI. Mr. Velardo makes amazing videos, and I believe that he deserves more recognition. ## Citing this DataSet ``` @data{DVN/13BPFB_2020, author = {Lee, Bruce …

Visit

github.com

Tags

africaasiaaudio-processingclassificationdatasetgenregenre-classificationgenre-suggestiongenres-classificationharvard-dataverse+6