This is the Amharic-only extract from the Aya Dataset, a multilingual instruction fine-tuning dataset.
By @henok
The Aya Dataset is a multilingual instruction fine-tuning dataset curated by an open-science community via Aya Annotation Platform from Cohere For AI. The dataset contains a total of 204k human-annotated prompt-completion pairs along with the demographics data of the annotators.