Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Deep Visible and Thermal Image Fusion with Cross-Modality Feature Selection for Pedestrian Detection

Domaine:

mobility

Type de record:

paper
Créateur:
Li,ShaShiGua
Éditeur:
ColBeiAdvXin
Éditeur:
CCSDSpringer International Publishing
Hôte:avatar
Part 2: AI International audience This paper proposes a deep RGB and thermal image fusion method for pedestrian detection. A two-branch structure is designed to learn the features of RGB and thermal images respectively, and these features are fused with a cross-modality feature selection module for detection. It includes the following stages. First, we learn features from paired RGB and thermal images through a backbone network with a residual structure, and add a feature squeeze-excitation module to the residual structure; Then we fuse the learned features from two branches, and a cross-modality feature selection module is designed to strengthen the effective information and compress the useless information during the fusion process; Finally, multi-scale features are fused for pedestrian detection. Two sets of experiments on the public KAIST pedestrian dataset are conducted, and experimental results show that our method is better than the state-of-the-art methods. The robustness of fused features is improved, and the miss rate is reduced obviously.

Visit

inria.hal.science

Tasks

computer vision

Tags

Feature fusionCross-modality featuresPedestrian detection[INFO]Computer Science [cs]

Licenses

http://creativecommons.org/licenses/by/info:eu-repo/semantics/OpenAccess