Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Can LLMs Faithfully Explain Themselves in Low-Resource Languages? A Case Study on Emotion Detection in Persian

Domaine:

natural language processing

Type de record:

paper
Créateur:
MehYouBeyBah
Hôte:avatar
Large language models (LLMs) are increasingly used to generate self-explanations alongside their predictions, a practice that raises concerns about the faithfulness of these explanations, especially in low-resource languages. This study evaluates the faithfulness of LLM-generated explanations in the context of emotion classification in Persian, a low-resource language, by comparing the influential words identified by the model against those identified by human annotators. We assess faithfulness using confidence scores derived from token-level log-probabilities. Two prompting strategies, differing in the order of explanation and prediction (Predict-then-Explain and Explain-then-Predict), are tested for their impact on explanation faithfulness. Our results reveal that while LLMs achieve strong classification performance, their generated explanations often diverge from faithful reasoning, showing greater agreement with each other than with human judgments. These results highlight the limitations of current explanation methods and metrics, emphasizing the need for more robust approaches to ensure LLM reliability in multilingual and low-resource contexts.

Visit

arxiv.org

Tasks

emotion identification

Tags

Computation and Language

Similaires

Misinformation Detection in COVID-19 News for Low-Resource Languages: A Sesotho Case StudyA Natural Language Processing Framework for Toxicity Detection in Low-Resource Languages: A Case Study on the Twi LanguageDeep Persian sentiment analysis: Cross-lingual training for low-resource languagesMulti-Label Emotion Recognition in Low-Resource Dialects: A Case Study on Algerian Arabic with Large Language ModelsLLM Probe: Evaluating LLMs for Low-Resource LanguagesMultilingual jailbreaking of LLMs using low-resource languages

Misinformation Detection in COVID-19 News for Low-Resource Languages: A Sesotho Case Study

A Natural Language Processing Framework for Toxicity Detection in Low-Resource Languages: A Case Study on the Twi Language

This dataset contains 2,001 text entries labeled for toxicity classification. Each entry represents

Deep Persian sentiment analysis: Cross-lingual training for low-resource languages

With the advent of deep neural models in natural language processing tasks, having a large amount of

Multi-Label Emotion Recognition in Low-Resource Dialects: A Case Study on Algerian Arabic with Large Language Models

LLM Probe: Evaluating LLMs for Low-Resource Languages

Despite rapid advances in large language models (LLMs), their linguistic abilities in low-resource a

Multilingual jailbreaking of LLMs using low-resource languages

Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrai