This dataset is a multilingual Somali lexical resource containing Somali terms with corresponding Italian and English glosses. It is designed to support Natural Language Processing (NLP), translation systems, and Somali language technology development.
The dataset currently consists of approximately 239,000 entries, each stored as a single text string combining abbreviation, Somali term, and translations.