Contemporary Amharic Corpus: Automatically Morpho-Syntactically Tagged Amharic Corpus
We introduced the contemporary Amharic corpus, which is automatically tagged for morpho-syntactic information. Texts are collected from 25,199 documents from different domains and about 24 million orthographic words are tokenized. Since it is partly a web corpus, w