Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

Domain:

natural language processing

Record type:

paper

We introduce FLEURS, the Few-shot Learning Evaluation of Universal Representations of Speech benchmark. FLEURS is an n-way parallel speech dataset in 102 languages built on top of the machine translation FLoRes-101 benchmark, with approximately 12 hours of speech supervision per language. FLEURS can be used for a variety of speech tasks, including Automatic Speech Recognition (ASR), Speech Language Identification (Speech LangID), Translation and Retrieval. In this paper, we provide baselines for the tasks based on multilingual pre-trained models like mSLAM. The goal of FLEURS is to enable speech technology in more languages and catalyze research in low-resource speech understanding.

Visit

arxiv.org

Connected records

dataset

Tasks

automatic speech recognitionspeech processing

Languages

AfrikaansAmharicChichewaDholuoFulfulde, AdamawaFulfulde, Central-Eastern NigerFulfulde, Western NigerGandaHausaIgbo+13

Tags

speech benchmarkfleursFLoRes-101

Licenses

cc-by-4.0

Similar

Few-shot Learning with Multilingual Language ModelsFew-Shot Learning Translation from New LanguagesOffline Handwritten Amharic Character Recognition Using Few-shot LearningLearning Robust and Multilingual Speech RepresentationsA Federated Approach to Few-Shot Hate Speech Detection for Marginalized CommunitiesVisually grounded few-shot word learning in low-resource settings

Few-shot Learning with Multilingual Language Models

Large-scale generative language models such as GPT-3 are competitive few-shot learners. While these

Few-Shot Learning Translation from New Languages

Recent work shows strong transfer learning capability to unseen languages in sequence-to-sequence ne

Offline Handwritten Amharic Character Recognition Using Few-shot Learning

Few-shot learning is an important, but challenging problem of machine learning aimed at learning fro

Learning Robust and Multilingual Speech Representations

Unsupervised speech representation learning has shown remarkable success at finding representations

A Federated Approach to Few-Shot Hate Speech Detection for Marginalized Communities

Hate speech online remains an understudied issue for marginalized communities, particularly in the G

Visually grounded few-shot word learning in low-resource settings

We propose a visually grounded speech model that learns new words and their visual depictions from j