Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

The USYD-JD Speech Translation System for IWSLT 2021

Domaine:

natural language processing

Type de record:

papermodel
Créateur:
Ding, LiangWu,Tao
Hôte:avatar
This paper describes the University of Sydney& JD's joint submission of the IWSLT 2021 low resource speech translation task. We participated in the Swahili-English direction and got the best scareBLEU (25.3) score among all the participants. Our constrained system is based on a pipeline framework, i.e. ASR and NMT. We trained our models with the officially provided ASR and MT datasets. The ASR system is based on the open-sourced tool Kaldi and this work mainly explores how to make the most of the NMT models. To reduce the punctuation errors generated by the ASR model, we employ our previous work SlotRefine to train a punctuation correction model. To achieve better translation performance, we explored the most recent effective strategies, including back translation, knowledge distillation, multi-feature reranking and transductive finetuning. For model structure, we tried auto-regressive and non-autoregressive models, respectively. In addition, we proposed two novel pre-train approaches, i.e. \textit{de-noising training} and \textit{bidirectional training} to fully exploit the data. Extensive experiments show that adding the above techniques consistently improves the BLEU scores, and the final submission system outperforms the baseline (Transformer ensemble model trained with the original parallel data) by approximately 10.8 BLEU score, achieving the SOTA performance. IWSLT 2021 winning system of the low-resource speech translation track

Visit

arxiv.org

Tasks

machine translationspeech processingspeech translation

Languages

Swahili

Tags

Computation and LanguageArtificial Intelligence

Similaires

The USYD-JD Speech Translation System for IWSLT2021 نظام ترجمة الكلام USYD - JD لـ IWSLT2021 Le système de traduction vocale USYD-JD pour IWSLT2021 El sistema de traducción de voz USYD-JD para IWSLT2021IMS' Systems for the IWSLT 2021 Low-Resource Speech Translation TaskGMU Systems for the IWSLT 2025 Low-Resource Speech Translation Shared Task

The USYD-JD Speech Translation System for IWSLT2021 نظام ترجمة الكلام USYD - JD لـ IWSLT2021 Le système de traduction vocale USYD-JD pour IWSLT2021 El sistema de traducción de voz USYD-JD para IWSLT2021

This paper describes the University of Sydney & JD's joint submission of the IWSLT 2021 low resource

IMS' Systems for the IWSLT 2021 Low-Resource Speech Translation Task

This paper describes the submission to the IWSLT 2021 Low-Resource Speech Translation Shared Task by

GMU Systems for the IWSLT 2025 Low-Resource Speech Translation Shared Task

This paper describes the GMU systems for the IWSLT 2025 low-resource speech translation shared task.