Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

End-to-End Automatic Speech Translation of Audiobooks

Domaine:

natural language processing

Type de record:

paperdataset
Créateur:
BérBesKocPie
Hôte:avatar
We investigate end-to-end speech-to-text translation on a corpus of audiobooks specifically augmented for this task. Previous works investigated the extreme case where source language transcription is not available during learning nor decoding, but we also study a midway case where source language transcription is available at training time only. In this case, a single model is trained to decode source speech into target text in a single pass. Experimental results show that it is possible to train compact and efficient end-to-end speech translation models in this setup. We also distribute the corpus and hope that our speech translation baseline on this corpus will be challenged in the future. Accepted to ICASSP 2018 (poster presentation)

Visit

arxiv.org

Tasks

machine translationspeech processingspeech translation

Tags

Computation and Language

Similaires

Towards End-to-End Training of Automatic Speech Recognition for Nigerian PidginTowards a Deep Understanding of Multilingual End-to-End Speech TranslationListen and Translate: A Proof of Concept for End-to-End Speech-to-Text TranslationImproving End-to-End Speech Translation for the Low Resource Language Fongbe to Frenchabu14/end-to-end-speech-to-textJuliusFx131/End-to-End-Machine-Translation-System

Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin

Nigerian Pidgin remains one of the most popular languages in West Africa. With at least 75 million speakers along the West African coast, the language has spread to diasporic communities through Nigerian immigrants in England, Canada, and America, amongst others. I

Towards a Deep Understanding of Multilingual End-to-End Speech Translation

In this paper, we employ Singular Value Canonical Correlation Analysis (SVCCA) to analyze representa

Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation

This paper proposes a first attempt to build an end-to-end speech-to-text translation system, which

Improving End-to-End Speech Translation for the Low Resource Language Fongbe to French

This study addresses the challenges of end-to-end (E2E) Speech-to-Text Translation (STT) for the low

abu14/end-to-end-speech-to-text

end to end amharic text to speech voice recognition project # Amharic Speech Recognition In this p

JuliusFx131/End-to-End-Machine-Translation-System

This repository features an end-to-end Dyula-to-French translation system built with Fairseq, addres