Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Real to H-space Autoencoders for Theme Identification in Telephone Conversations

Domaine:

natural language processing

Type de record:

paper
Créateur:
ParMorBosLin
Éditeur:
LabMcGANR
Éditeur:
CCSDIns
Hôte:avatar
International audience Machine learning (ML) and deep learning with deep neural networks (DNN), have drastically improved the performances of modern systems on numerous spoken language understanding (SLU) related tasks. Since most of current researches focus on new neural architectures to enhance the performances in realistic conditions, few recent works investigated the use of different algebras with neural networks (NN), to better represent the nature of the data being processed. To this extent, quaternion-valued neural networks (QNN) have shown better performances, and an important reduction of the number of neural parameters compared to traditional real-valued neural networks, when dealing with multidimensional signal. Nonetheless, the use of QNNs is strictly limited to quaternion input or output features. This paper introduces a new unsupervised method based on a hybrid autoencoder (AE) called real-to-quaternion autoencoder (R2H), to extract a quaternion-valued input signal from any real-valued data, to be processed by QNNs. The experiments performed to identify the most related theme of a given telephone conversation from a customer care service (CCS), demonstrate that the R2H approach outperforms all the previously established models, either real-or quaternion-valued ones, in term of accuracy and with up to four times fewer neural parameters.

Visit

hal.science

Tags

spoken language understandingquaternion autoencoderquaternion neural networksIndex Terms-Features extraction[INFO]Computer Science [cs]

Licenses

info:eu-repo/semantics/OpenAccess

Similaires

ON THE USE OF LINGUISTIC FEATURES IN AN AUTOMATIC SYSTEM FOR SPEECH ANALYTICS OF TELEPHONE CONVERSATIONSAMHARIC SPEAKER IDENTIFICATION IN TELEPHONE CALLING BY DEEP LEARNING APPROACHVisionArena: 230K Real World User-VLM Conversations with Preference LabelsAutomated multilingual telephone access to financial servicesStaging the body and space in television: Jozi H as a case in pointConditional Random Field Autoencoders for Unsupervised Structured Prediction

ON THE USE OF LINGUISTIC FEATURES IN AN AUTOMATIC SYSTEM FOR SPEECH ANALYTICS OF TELEPHONE CONVERSATIONS

International audience A research on the analysis of human/human conversations in a c

AMHARIC SPEAKER IDENTIFICATION IN TELEPHONE CALLING BY DEEP LEARNING APPROACH

AMHARIC SPEAKER IDENTIFICATION IN TELEPHONE CALLING BY DEEP LEARNING APPROACH

VisionArena: 230K Real World User-VLM Conversations with Preference Labels

With the growing adoption and capabilities of vision-language models (VLMs) comes the need for bench

Automated multilingual telephone access to financial services

A working prototype automated telephone based enquiry and payment system functioning in three Africa

Staging the body and space in television: Jozi H as a case in point

Abstract The medical drama series is uniquely positioned to draw together a technol

Conditional Random Field Autoencoders for Unsupervised Structured Prediction

We introduce a framework for unsupervised learning of structured predictors with overlapping, global