Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering

Domain:

natural language processing

Record type:

paperdataset
Creator:
BahGho
Host:avatar
The rapid advancement of large language models (LLMs) has significantly propelled progress in natural language processing (NLP). However, their effectiveness in specialized, low-resource domains-such as Arabic legal contexts-remains limited. This paper introduces MizanQA (pronounced Mizan, meaning "scale" in Arabic, a universal symbol of justice), a benchmark designed to evaluate LLMs on Moroccan legal question answering (QA) tasks, characterised by rich linguistic and legal complexity. The dataset draws on Modern Standard Arabic, Islamic Maliki jurisprudence, Moroccan customary law, and French legal influences. Comprising over 1,700 multiple-choice questions, including multi-answer formats, MizanQA captures the nuances of authentic legal reasoning. Benchmarking experiments with multilingual and Arabic-focused LLMs reveal substantial performance gaps, highlighting the need for tailored evaluation metrics and culturally grounded, domain-specific LLM development.

Visit

arxiv.org

Tasks

question answering

Tags

Computation and LanguageArtificial IntelligenceInformation Retrieval

Similar

Question Answering Research in African Languages with Large Language Models MedQA-MA: A Moroccan Arabic medical question-answering dataset for virtual healthcare assistants and large language modelsMKG-Rank: Enhancing Large Language Models with Knowledge Graph for Multilingual Medical Question AnsweringLow Resource Question Answering: An Amharic Benchmarking DatasetContext-Based Question Answering Using Large Language BERT Variant Models for Low Resourced Sesotho sa Leboa LanguageTounsiBench: Benchmarking Large Language Models for Tunisian Arabic

Question Answering Research in African Languages with Large Language Models

Native African languages are grossly underrepresented in prevailing large language models (LLMs). Th

MedQA-MA: A Moroccan Arabic medical question-answering dataset for virtual healthcare assistants and large language models

MKG-Rank: Enhancing Large Language Models with Knowledge Graph for Multilingual Medical Question Answering

Large Language Models (LLMs) have shown remarkable progress in medical question answering (QA), yet

Low Resource Question Answering: An Amharic Benchmarking Dataset

Context-Based Question Answering Using Large Language BERT Variant Models for Low Resourced Sesotho sa Leboa Language

TounsiBench: Benchmarking Large Language Models for Tunisian Arabic