Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Kurmanji LLM Evaluation Dataset (ChatGPT-4o, Gemini, Claude, Grok)

Domaine:

natural language processing

Type de record:

dataset
Créateur:
uzu
Éditeur:
Zenodo
Hôte:avatar
This dataset contains sample outputs and evaluation scores from the study “A Test of Meaning, Form, and Culture in Kurmanji: An Evaluation of Large Language Models’ Performance.” It includes responses generated by four large language models—ChatGPT-4o, Gemini, Claude, and Grok—in reaction to prompts written in Kurmanji, a morphologically rich and culturally embedded Kurdish language. Each model output was scored by human annotators based on grammatical correctness, semantic accuracy, and cultural appropriateness. The dataset also includes evaluation notes for qualitative insight. This sample represents a subset of the full analysis described in the manuscript. The dataset is shared under the CC-BY 4.0 license, and may be reused for research, teaching, or further model evaluation in low-resource language contexts.

Visit

doi.orgzenodo.org

Tags

KurmanjiKurdishLow-resource languageLLM evaluationLarge Language ModelsAI and linguisticsChatGPTNatural language processingMultilingual NLPLanguage and culture

Licenses

Creative Commons Attribution 4.0 Internationalhttps://creativecommons.org/licenses/by/4.0/legalcode

Similaires

’n Ondersoek na die gehalte van tegniese vertaling deur ChatGPT-4oFrom LLM to NMT: Advancing Low-Resource Machine Translation with ClaudeUse and Application of ChatGPT-4o by Cataloguers and Bibliographers from Botswana, Nigeria, and South AfricaTwi Health Speech Dataset Gemini (500 hours)Zero-shot POS tagging accuracy comparison of Gemini 2.0 Flash Thinking, Llama 3, and Claude 3 on XTREME for low-resourceAmharic LLM Training Dataset

’n Ondersoek na die gehalte van tegniese vertaling deur ChatGPT-4o

An investigation into the quality of technical translation by ChatGPT-4o This article reports on a

From LLM to NMT: Advancing Low-Resource Machine Translation with Claude

We show that Claude 3 Opus, a large language model (LLM) released by Anthropic in March 2024, exhibi

Use and Application of ChatGPT-4o by Cataloguers and Bibliographers from Botswana, Nigeria, and South Africa

Twi Health Speech Dataset Gemini (500 hours)

A domain-specific speech recognition dataset for Twi, one of Ghana's most widely spoken languages, s

Zero-shot POS tagging accuracy comparison of Gemini 2.0 Flash Thinking, Llama 3, and Claude 3 on XTREME for low-resource

Named Entity Recognition (NER) and Part-of-Speech (POS) tagging are critical tasks for Natural Langu

Amharic LLM Training Dataset

Complete production-ready Amharic dataset for large language model training and deployment. from da