Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

04_Corpus_BurkinaFaso_Extracted_EntitiesAndSentences.jsonl

Domain:

natural language processing

Record type:

dataset
Creator:
JaiTeiVal
Editor:
Tei
Publisher:
Rec
Host:avatar
Contient les informations extraites automatiquement (analysis_result) à partir du contenu des articles du corpus de journaux du Burkina Faso : [1] entités spatiales (label = LOC), [2] organisations (label = ORG) et [3] du lexique expert (spaCy), [4] entités temporelles extraites avec HeidelTime (label = DATE | DURATION) et [5] phrases analysées en sentiment ('polarizedSentences', dont le 'polarity_label' peut être positive, négative ou neutral) avec le modèle Codestral. - Contains the information extracted automatically (analysis_result) from the content of the articles in the corpus of newspapers from Burkina Faso: [1] spatial entities (label = LOC), [2] organisations (label = ORG) and [3] from the expert lexicon (spaCy), [4] temporal entities extracted with HeidelTime (label = DATE | DURATION) and [5] sentences analysed for sentiment (‘polarizedSentences’, whose ‘polarity_label’ can be positive, negative or neutral) with the Codestral model.

Visit

doi.orgentrepot.recherche.data.gouv.fr

Tasks

named entity recognitionsentiment analysisinformation extractiontext classification