This amharic text dataset can be used to train/finetune models for the following tasks
classification : using the categories
summarization : using the headlines
Here is a github repo that contains three notebooks that use this dataset to finetune the following models.
xlm-roberta-base : a multilingual transformer model with 280M parameters