Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Arko98/TALLIP-FakeNews-Dataset

Domaine:

natural language processing

Type de record:

dataset
Créateur:
Ark
Hôte:
Multilingual Fake News Dataset created for the research paper "A Transformer Based Approach to Multilingual Fake News Detection in Low Resource Languages" accepted at the ACM Transactions on Asian and Low-Resource Language Information Processing (ACM TALLIP) # TALLIP-FakeNews-Dataset Multilingual Fake News Dataset created for the research paper "A Transformer Based Approach to Multilingual Fake News Detection in Low Resource Languages" accepted at the ACM Transactions on Asian and Low-Resource Language Information Processing (ACM TALLIP) # Information This Dataset is associated with the Research Work titled "A Transformer Based Approach to Multilingual Fake News Detection" published in ACM Transaction on Asian and Low-Resource Language Information Processing (ACM-TALLIP) Journal by Arkadipta De, Dibyanayan Bandyopadhyay, Baban Gain and Asif Ekbal in a joint research work from IIT Hyderabad and IIT Patna. # Dataset Link The dataset is available in the link: iitp.ac.in # Research Paper Link The paper can be found at: dl.acm.org # Contents: 1) English Version of Dataset (Train and Test) 2) Hindi Version of Dataset (Train and Test) 3) Swahili Version of Dataset (Train and Test) 4) Vietnamese Version of Dataset (Train and Test) 5) Indonesian Version of Dataset (Train and Test) 6) Multilingual Version of Dataset (Train and Test) Each Dataset has Six different domains (Technology, Bussiness, Education, Politics, Celebrity News, Entertainment) # Note: 1. The English Version of the dataset has been collected, cleaned and processed from the research paper "Automatic Detection of Fake News" by Verónica Pérez-Rosas, Bennett Kleinberg, Alexandra Lefevre, Rada Mihalcea published Proceedings of the 27th International Conference on Computational Linguistics (COLING 2018). The Paper URL: aclweb.org. **If you use only the English version of the dataset then please cite the paper given in the URL.** 2. The extension of the dataset has been done by the authors of this paper. **If you use this dataset in any research work, please cite the paper** ``` @article{10.1145/3472619, author = {De, Arkadipta and Bandy …

Visit

github.com

Languages

Swahili

Licenses

MIT

Similaires

erKrishna26/Amharic-language-FakeNews-detectionBOUTEF: A Multilingual Corpus for FakeNews in North Africa -- Language as a Weapon

erKrishna26/Amharic-language-FakeNews-detection

Comparative Analysis of BiLSTM and AfriBERTa for Fake News Detection in Low-Resource Amharic Languag

BOUTEF: A Multilingual Corpus for FakeNews in North Africa -- Language as a Weapon

The rapid spread of fake news on social media has become a major challenge, particularly in multilin