A speech-recognition dataset of African American English (AAE) utterances spanning
multiple regional varieties. Each record provides an audio clip, its verbatim
reference transcript, and speaker/region metadata intended for evaluating ASR and
in-context-learning approaches on dialectal, low-resource speech.
This release is a stratified sample of utterances drawn across all regional
collections.
Split
Utterances