Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Pearl: A Multimodal Culturally-Aware Arabic Instruction Dataset

Domain:

natural language processing

Record type:

paperdataset
Creator:
AlwMagMekNac
Host:avatar
Mainstream large vision-language models (LVLMs) inherently encode cultural biases, highlighting the need for diverse multimodal datasets. To address this gap, we introduce PEARL, a large-scale Arabic multimodal dataset and benchmark explicitly designed for cultural understanding. Constructed through advanced agentic workflows and extensive human-in-the-loop annotations by 37 annotators from across the Arab world, PEARL comprises over 309K multimodal examples spanning ten culturally significant domains covering all Arab countries. We further provide two robust evaluation benchmarks (PEARL and PEARL-LITE) along with a specialized subset (PEARL-X) explicitly developed to assess nuanced cultural variations. Comprehensive evaluations on state-of-the-art open and proprietary LVLMs demonstrate that reasoning-centric instruction alignment substantially improves models' cultural grounding compared to conventional scaling methods. PEARL establishes a foundational resource for advancing culturally-informed multimodal modeling research. All datasets and benchmarks are publicly available. github.com

Visit

arxiv.org

Tasks

computer visionimage-text retrieval

Tags

Computation and Language

Similar

CaMMT: Benchmarking Culturally Aware Multimodal Machine TranslationOASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QAPalm: A Culturally Inclusive and Linguistically Diverse Dataset for Arabic LLMsMORAD: A Multimodal Dataset of Authentic Emotional Expressions in Moroccan ArabicAryWiki-Instruct: A high-fidelity instruction tuning dataset for Moroccan Arabic (Darija)MSA-Moroccan Dialect: A Multimodal Sentiment Analysis Dataset for Moroccan Arabic (Darija)

CaMMT: Benchmarking Culturally Aware Multimodal Machine Translation

Translating cultural content poses challenges for machine translation systems due to the differences

OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA

Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA),

Palm: A Culturally Inclusive and Linguistically Diverse Dataset for Arabic LLMs

As large language models (LLMs) become increasingly integrated into daily life, ensuring their cultu

MORAD: A Multimodal Dataset of Authentic Emotional Expressions in Moroccan Arabic

AryWiki-Instruct: A high-fidelity instruction tuning dataset for Moroccan Arabic (Darija)

MSA-Moroccan Dialect: A Multimodal Sentiment Analysis Dataset for Moroccan Arabic (Darija)

This dataset provides the first publicly available multimodal resource for sentiment analysis in Mor