Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

The World Wide recipe: A community-centred framework for fine-grained data collection and regional bias operationalisation

Domaine:

natural language processing

Type de record:

paper

We introduce the World Wide recipe, which sets forth a framework for culturally aware and participatory data collection, and the resultant regionally diverse World Wide Dishes evaluation dataset. We also analyse bias operationalisation to highlight how current systems underperform across several dimensions: (in-)accuracy, (mis-)representation, and cultural (in-)sensitivity, with evidence from qualitative community-based observations and quantitative automated tools.

We find that these T2I models generally do not produce quality outputs of dishes specific to various regions. This is true even for the US, which is typically considered more well-resourced in training data—although the generation of US dishes does outperform that of the investigated African countries. The models demonstrate the propensity to produce inaccurate and culturally misrepresentative, flattening, and insensitive outputs. These representational biases have the potential to further reinforce stereotypes and disproportionately contribute to erasure based on region.

The dataset and code are available at The World Wide Dishes Datas….

Visit

dl.acm.orgarxiv

Connected records

dataset

Tasks

computer visionimage-text retrieval

Languages

AcholiAfrikaansAkaAkanAmazighArabic, Moroccan SpokenAtesoBembaBurjiBwamu, Cwi+60

Tags

world wide dishescommunity-centred data collectionbias operationalisationcommunity-centred evaluation

Licenses

https://github.com/oxai/world-wide-dishes/blob/main/LICENCE.md

Similaires

A Sector-Wide Data Collection Framework for Improving Operational Efficiency in Nigeria’s Retail Logistics EcosystemFine-grained Population Maps for Tanzania and ZambiaThe development of a fine grained class set for Amazigh POS taggingAI-Based End-to-End Solutions for Fine-Grained Classification, Detection and Segmentation in Real-World ScenariosAdi17: A Fine-Grained Arabic Dialect Identification DatasetARCADE: A City-Scale Corpus for Fine-Grained Arabic Dialect Tagging

A Sector-Wide Data Collection Framework for Improving Operational Efficiency in Nigeria’s Retail Logistics Ecosystem

Nigeria's retail logistics ecosystem faces challenges stemming from infrastructural deficits, fragme

Fine-grained Population Maps for Tanzania and Zambia

Official dataset to paper "Fine-grained Population Mapping from Coarse Census Counts and Open Geo

The development of a fine grained class set for Amazigh POS tagging

AI-Based End-to-End Solutions for Fine-Grained Classification, Detection and Segmentation in Real-World Scenarios

This thesis presents solutions to the complex challenges posed by Learning with Noisy L

Adi17: A Fine-Grained Arabic Dialect Identification Dataset

Presenter: Suwon Shon, ICASSP 2020, Virtual Event, May 4-8, 2020 In this paper, we describe a method

ARCADE: A City-Scale Corpus for Fine-Grained Arabic Dialect Tagging

The Arabic language is characterized by a rich tapestry of regional dialects that differ substantial