The first structured digital vocabulary dataset for Sango (ISO 639-1: sg, ISO 639-3: sag), the national and most widely spoken language of the Central African Republic. Sango is a creole language with over 5 million speakers, yet it remains severely underrepresented in NLP research and digital resources.