This dataset highlights blind spots and errors of the open-source language model Qwen3-0.6B, released on Hugging Face. It contains 10 diverse prompts across multiple categories:
Factual knowledge
Ambiguous or tricky questions
Commonsense reasoning
Hardware and robotics knowledge
Local and African context
Arithmetic
Local language translation (Kenyan Swahili/Slang)
Simple/edge cases
Code generation
Cultural or language understanding