Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Physical Commonsense for Yorùbá & Pidgin

Domain:

natural language processing

Record type:

dataset
Creator:
tar
Host:
This dataset was developed for the MRL 2025 Shared Task on Multilingual Physical Reasoning. For more details, see Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures. It provides a test collection for evaluating physical commonsense reasoning, that is, a model's ability to understand how objects, actions, and outcomes relate in everyday scenarios.

Visit

huggingface.co

Tasks

commonsense reasoning

Languages

Pidgin, NigerianYoruba

Tags

yorubanigerian-pidginpiqa

Licenses

cc-by-4.0

Similar

Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures v0.1XCOPA: A Multilingual Dataset for Causal Commonsense ReasoningDetecting sentiments and combatting hate speech in Hausa, Igbo, Nigerian-Pidgin and Yorùbá - NaijaSenti: a Nigerian Corpus for Multilingual Sentiment AnalysisCommonsense Reasoning in Arab CultureCommonsense Knowledge Augmentation for Low-Resource Languages via Adversarial LearningWDYW.01: Tone Restoration for Standard Yorùbá

Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures v0.1

To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and cultures. In this paper, we present Global PIQA, a participatory commonsense reasoning benchmark for over 100 lan

XCOPA: A Multilingual Dataset for Causal Commonsense Reasoning

In order to simulate human language capacity, natural language processing systems must be able to re

Detecting sentiments and combatting hate speech in Hausa, Igbo, Nigerian-Pidgin and Yorùbá - NaijaSenti: a Nigerian Corpus for Multilingual Sentiment Analysis

The NaijaSenti dataset is the first large-scale human-annotated Twitter sentiment dataset for Hausa,

Commonsense Reasoning in Arab Culture

Despite progress in Arabic large language models, such as Jais and AceGPT, their evaluation on commo

Commonsense Knowledge Augmentation for Low-Resource Languages via Adversarial Learning

Commonsense reasoning is one of the ultimate goals of artificial intelligence research because it si

WDYW.01: Tone Restoration for Standard Yorùbá

Purpose: This paper addresses automatic tone restoration for Standard Yorùbá, a three-tone language