Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Aslema at NADI 2026: Augmentation through Fewshot for SLU

Domaine:

natural language processing

Type de record:

paper
Créateur:
ShaBhaChoAla
Éditeur:
arXiv
Hôte:avatar
We present Aslema, our system for NADI 2026 Shared Task 5, which consists of two subtasks: intent recognition and slot filling. We evaluate four omni LLMs in a zero-shot setting and compare them with fine-tuned models. Our results show that fine-tuning consistently outperforms zero-shot inference. We further explore synthetic data augmentation by using an LLM to generate culturally grounded Tunisian Derja utterances, followed by voice cloning to generate synthetic speech. Incorporating this synthetic data improves performance on both tasks. Our final submitted system, based on Qwen3-Omni-30B and trained with a mixture of original and synthetic data, achieves 86.8% intent accuracy and 34.7 WER on the devtest split. On the official test set it ranks 1st in slot filling (59.5 CoER) and 4th among 8 teams in intent recognition (66.1% accuracy). We release our experimental scripts and will soon share the synthetic dataset to support further research in this area. LLMs, Native, Arabic LLMs, Augmentation, Multilingual, Multimodal, Language Diversity, Contextual Understanding, Minority Languages, Culturally Informed, Foundation Models, Large Language Models, Audio Models, Omni Models, Slot Filling

Visit

doi.org

Languages

Arabic, Tunisian Spoken

Tags

Computation and Language (cs.CL)Artificial Intelligence (cs.AI)FOS: Computer and information sciencesF.2.2; I.2.768T50

Licenses

Creative Commons Attribution Non Commercial Share Alike 4.0 Internationalhttps://creativecommons.org/licenses/by-nc-sa/4.0/legalcode

Similaires

Enhancing Automatic Speech Recognition Systems for Amazigh Language Through Data AugmentationElyadata/TARIC-SLUWhen Is TTS Augmentation Through a Pivot Language Useful?Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language UnderstandingTARIC-SLU: A Tunisian Benchmark Dataset for Spoken Language UnderstandingBalancing Data through Data Augmentation Improves the Generality of Transfer Learning for Diabetic Retinopathy Classification

Enhancing Automatic Speech Recognition Systems for Amazigh Language Through Data Augmentation

Elyadata/TARIC-SLU

The primary contributions of this work are as follows: Release of the TARIC-SLU corpus: The very fi

When Is TTS Augmentation Through a Pivot Language Useful?

Developing Automatic Speech Recognition (ASR) for low-resource languages is a challenge due to the s

Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding

Spoken language understanding (SLU) is indispensable for half of all living languages that lack a fo

TARIC-SLU: A Tunisian Benchmark Dataset for Spoken Language Understanding

Balancing Data through Data Augmentation Improves the Generality of Transfer Learning for Diabetic Retinopathy Classification

The incidence of diabetes in Mauritius is amongst the highest in the world. Diabetic retinopathy (DR