This dataset was created to probe blind spots in a small, recently released base language model by focusing on two closely related capabilities:
Factual knowledge capacity – whether the model can correctly answer widely known facts in Ethiopia (e.g., major cities, historical events, landmarks).