Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Deep learning-based extractive and abstractive summarization for the Azerbaijani language

Domaine:

natural language processing

Type de record:

datasetpaper
Créateur:
MirSam
Éditeur:
Pee
Hôte:
This article investigates extractive and abstractive text summarization for the Azerbaijani language, a low-resource and underrepresented language in natural language processing. While the underlying modeling approaches are well established, their application to Azerbaijani summarization remains largely unexplored due to the scarcity of large-scale datasets and prior empirical studies. To address this gap, we conduct a systematic evaluation of both extractive methods based on sentence ranking and an abstractive approach using a fine-tuned mT5-base model. Our experiments are carried out on a large-scale dataset comprising over 115,000 Azerbaijani news articles paired with human-written summaries. The models are evaluated using standard automatic metrics, including Recall-Oriented Understudy for Gisting Evaluation (ROUGE), Bilingual Evaluation Understudy (BLEU), and Metric for Evaluation of Translation with Explicit ORdering (METEOR), yielding strong results that highlight the benefits of task specific fine-tuning for abstractive summarization, while also demonstrating the competitiveness of extractive baselines. In addition, we analyze the impact of long input sequences and discuss architectural and dataset-related limitations affecting performance. Overall, this study provides a comprehensive empirical baseline for Azerbaijani text summarization and serves as a reference point for future research in low-resource summarization and related Azerbaijani Natural Language Processing (NLP) applications.

Visit

doi.org

Tasks

natural language generationsummarization

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

Extractive Text Summarization Using Deep Learning for Tigrigna LanguageMohamed-Qadar/Extractive-and-Abstractive-Summarization-Techniques-in-Somali-Language-Using-NLPAbstractive Tigrigna Text Summarization using Deep Learning ApproachCUET_SSTM at the GEM’24 Summarization Task: Integration of extractive and abstractive method for long text summarization in Swahili languageA Gold-Standard Dataset for Benchmarking Balinese Extractive and Abstractive Text SummarizationAbstractive text summarization of low-resourced languages using deep learning

Extractive Text Summarization Using Deep Learning for Tigrigna Language

Mohamed-Qadar/Extractive-and-Abstractive-Summarization-Techniques-in-Somali-Language-Using-NLP

Abstractive Tigrigna Text Summarization using Deep Learning Approach

Text summarization has become essential due to the vast amounts of text data shared online. It is th

CUET_SSTM at the GEM’24 Summarization Task: Integration of extractive and abstractive method for long text summarization in Swahili language

A Gold-Standard Dataset for Benchmarking Balinese Extractive and Abstractive Text Summarization

1. Research Hypothesis and Data Scope The central hypothesis guiding the creation of the BaliSummari

Abstractive text summarization of low-resourced languages using deep learning

Background Humans must be able to cope with the huge amounts of information produce