Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

BridgeNets: Student-Teacher Transfer Learning Based on Recursive Neural Networks and its Application to Distant Speech Recognition

Domain:

natural language processing

Record type:

papermodel
Creator:
KimEl-Lee
Host:avatar
Despite the remarkable progress achieved on automatic speech recognition, recognizing far-field speeches mixed with various noise sources is still a challenging task. In this paper, we introduce novel student-teacher transfer learning, BridgeNet which can provide a solution to improve distant speech recognition. There are two key features in BridgeNet. First, BridgeNet extends traditional student-teacher frameworks by providing multiple hints from a teacher network. Hints are not limited to the soft labels from a teacher network. Teacher's intermediate feature representations can better guide a student network to learn how to denoise or dereverberate noisy input. Second, the proposed recursive architecture in the BridgeNet can iteratively improve denoising and recognition performance. The experimental results of BridgeNet showed significant improvements in tackling the distant speech recognition problem, where it achieved up to 13.24% relative WER reductions on AMI corpus compared to a baseline neural network without teacher's hints. Accepted to 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP 2018)

Visit

arxiv.org

Tasks

automatic speech recognitionspeech processingtransfer learning

Tags

Computation and LanguageSoundAudio and Speech Processing

Similar

Ethiopian Traffic Sign Recognition Using Customized Convolutional Neural Networks and Transfer LearningTransfer Learning based Speech Affect Recognition in UrduMUST: A Multilingual Student-Teacher Learning approach for low-resource speech recognitionThe Application of Probabilistic Neural Network in Speech Recognition Based on Partition ClusteringDeep Neural Networks Based Automatic Speech Recognition For Four Ethiopian LanguagesThe convolutional neural networks for Amazigh speech recognition system

Ethiopian Traffic Sign Recognition Using Customized Convolutional Neural Networks and Transfer Learning

Intelligent transportation systems rely greatly on their capacity to identify and recognize traffic

Transfer Learning based Speech Affect Recognition in Urdu

It has been established that Speech Affect Recognition for low resource languages is a difficult tas

MUST: A Multilingual Student-Teacher Learning approach for low-resource speech recognition

Student-teacher learning or knowledge distillation (KD) has been previously used to address data sca

The Application of Probabilistic Neural Network in Speech Recognition Based on Partition Clustering

A probabilistic neural network (PNN) speech recognition model based on the partition clustering algo

Deep Neural Networks Based Automatic Speech Recognition For Four Ethiopian Languages

Presenter: Solomon Teferra Abate, ICASSP 2020, Virtual Event, May 4-8, 2020 In this work, we present

The convolutional neural networks for Amazigh speech recognition system