Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Cross-lingual Matryoshka Representation Learning across Speech and Text

Domaine:

natural language processing

Type de record:

papermodeldataset
Créateur:
Sy,DouCerIll
Hôte:avatar
Speakers of under-represented languages face both a language barrier, as most online knowledge is in a few dominant languages, and a modality barrier, since information is largely text-based while many languages are primarily oral. We address this for French-Wolof by training the first bilingual speech-text Matryoshka embedding model, enabling efficient retrieval of French text from Wolof speech queries without relying on a costly ASR-translation pipelines. We introduce large-scale data curation pipelines and new benchmarks, compare modeling strategies, and show that modality fusion within a frozen text Matryoshka model performs best. Although trained only for retrieval, the model generalizes well to other tasks, such as speech intent detection, indicating the learning of general semantic representations. Finally, we analyze cost-accuracy trade-offs across Matryoshka dimensions and ranks, showing that information is concentrated only in a few components, suggesting potential for efficiency improvements. Preprint, under review

Visit

arxiv.org

Tasks

information retrieval

Languages

Wolof

Tags

Computation and Language

Similaires

Unsupervised Cross-lingual Representation Learning at ScaleDistilXLSR: A Light Weight Cross-Lingual Speech Representation ModelCross-lingual Representation Learning via Centroid Intervention Fusionabdouaziz/Semantic-Aware-Cross-Lingual-Speech-Translation-Representation-for-WolofGATE: General Arabic Text Embedding for Enhanced Semantic Textual Similarity with Matryoshka Representation Learning and Hybrid Loss TrainingText-to-speech system for low-resource language using cross-lingual transfer learning and data augmentation

Unsupervised Cross-lingual Representation Learning at Scale

This paper shows that pretraining multilingual language models at scale leads to significant performance gains for a wide range of cross-lingual transfer tasks. We train a Transformer-based masked language model on one hundred languages, using more than two terabyt

DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model

Multilingual self-supervised speech representation models have greatly enhanced the speech recogniti

Cross-lingual Representation Learning via Centroid Intervention Fusion

Large language models (LLMs) exhibit uneven multilingual performance, especially when dealing with l

abdouaziz/Semantic-Aware-Cross-Lingual-Speech-Translation-Representation-for-Wolof

# Semantic-Aware Cross-Lingual Speech Representation for Wolof PhD research project that aligns **W

GATE: General Arabic Text Embedding for Enhanced Semantic Textual Similarity with Matryoshka Representation Learning and Hybrid Loss Training

Semantic textual similarity (STS) is a critical task in natural language processing (NLP), enabling

Text-to-speech system for low-resource language using cross-lingual transfer learning and data augmentation

Abstract Deep learning techniques are currently being applied in automated text-to-speech (TTS) sys