Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

In Generative AI we Trust: Can Chatbots Effectively Verify Political Information?

Domaine:

natural language processing

Type de record:

paper
Créateur:
KuzMakVziSto
Hôte:avatar
This article presents a comparative analysis of the ability of two large language model (LLM)-based chatbots, ChatGPT and Bing Chat, recently rebranded to Microsoft Copilot, to detect veracity of political information. We use AI auditing methodology to investigate how chatbots evaluate true, false, and borderline statements on five topics: COVID-19, Russian aggression against Ukraine, the Holocaust, climate change, and LGBTQ+ related debates. We compare how the chatbots perform in high- and low-resource languages by using prompts in English, Russian, and Ukrainian. Furthermore, we explore the ability of chatbots to evaluate statements according to political communication concepts of disinformation, misinformation, and conspiracy theory, using definition-oriented prompts. We also systematically test how such evaluations are influenced by source bias which we model by attributing specific claims to various political and social actors. The results show high performance of ChatGPT for the baseline veracity evaluation task, with 72 percent of the cases evaluated correctly on average across languages without pre-training. Bing Chat performed worse with a 67 percent accuracy. We observe significant disparities in how chatbots evaluate prompts in high- and low-resource languages and how they adapt their evaluations to political communication concepts with ChatGPT providing more nuanced outputs than Bing Chat. Finally, we find that for some veracity detection-related tasks, the performance of chatbots varied depending on the topic of the statement or the source to which it is attributed. These findings highlight the potential of LLM-based chatbots in tackling different forms of false information in online environments, but also points to the substantial variation in terms of how such potential is realized due to specific factors, such as language of the prompt or the topic. 22 pages, 8 figures

Visit

arxiv.org

Tasks

text classification

Tags

Computation and LanguageComputers and Society

Similaires

Can we trust remote sensing evapotranspiration products over Africa?Can we trust remote sensing ET products over Africa?Chatbots and Citizen Satisfaction: Examining the Role of Trust in AI-Chatbots as a Moderating VariableTrust norms for generative AI data gathering in the African context“In YouTube We Trust”Scaling Arabic Medical Chatbots Using Synthetic Data: Enhancing Generative AI with Synthetic Patient Records

Can we trust remote sensing evapotranspiration products over Africa?

Abstract. Evapotranspiration (ET) is one of the most important components in the water cycle. Howeve

Can we trust remote sensing ET products over Africa?

Abstract. Evapotranspiration (ET) is one of the most important components in the water cycle. Howeve

Chatbots and Citizen Satisfaction: Examining the Role of Trust in AI-Chatbots as a Moderating Variable

International audience This study seeks to investigate the effect of deploying artifi

Trust norms for generative AI data gathering in the African context

Abstract Can trust norms within the African moral system support data gathering for Generative

“In YouTube We Trust”

YouTube has enabled new forms of political dissent in Arab societies. This chapter examines the deve

Scaling Arabic Medical Chatbots Using Synthetic Data: Enhancing Generative AI with Synthetic Patient Records

The development of medical chatbots in Arabic is significantly constrained by the scarcity of large-