Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

WildIng: A Wildlife Image Invariant Representation Model for Geographical Domain Shift

Domain:

environment and energygeospatial

Record type:

papermodelsoftware
Creator:
SanIsaGir
Host:avatar
Wildlife monitoring is crucial for studying biodiversity loss and climate change. Camera trap images provide a non-intrusive method for analyzing animal populations and identifying ecological patterns over time. However, manual analysis is time-consuming and resource-intensive. Deep learning, particularly foundation models, has been applied to automate wildlife identification, achieving strong performance when tested on data from the same geographical locations as their training sets. Yet, despite their promise, these models struggle to generalize to new geographical areas, leading to significant performance drops. For example, training an advanced vision-language model, such as CLIP with an adapter, on an African dataset achieves an accuracy of 84.77%. However, this performance drops significantly to 16.17% when the model is tested on an American dataset. This limitation partly arises because existing models rely predominantly on image-based representations, making them sensitive to geographical data distribution shifts, such as variation in background, lighting, and environmental conditions. To address this, we introduce WildIng, a Wildlife image Invariant representation model for geographical domain shift. WildIng integrates text descriptions with image features, creating a more robust representation to geographical domain shifts. By leveraging textual descriptions, our approach captures consistent semantic information, such as detailed descriptions of the appearance of the species, improving generalization across different geographical locations. Experiments show that WildIng enhances the accuracy of foundation models such as BioCLIP by 30% under geographical domain shift conditions. We evaluate WildIng on two datasets collected from different regions, namely America and Africa. The code and models are publicly available at github.com.

Visit

arxiv.org

Tasks

image classificationcomputer vision

Tags

Computer Vision and Pattern RecognitionArtificial Intelligence

Similar

Intra-African Domain Shift in Wildlife Camera Trap AI]% {Intra-African Geographic Domain Shift in Wildlife Camera Trap Species Classification: A Comparative Study of Supervised and Zero-Shot Foundation ModelsMIFR: A Modality-Invariant and Fair Representation Framework for Skin Disease ClassificationCross-lingual NER Model Accuracy Degradation under Extreme Domain ShiftIntra-African Geographic Domain Shift in Wildlife Camera Trap Species Classification: A Comparative Study of Supervised and Zero-Shot Foundation ModelsCross-lingual Ranking Model Robustness to Domain Shift in Low-Resource Legal SettingsRevisiting Invariant Learning for Out-of-Domain Generalization on Multi-Site Mammogram Datasets

Intra-African Domain Shift in Wildlife Camera Trap AI]% {Intra-African Geographic Domain Shift in Wildlife Camera Trap Species Classification: A Comparative Study of Supervised and Zero-Shot Foundation Models

This repository contains al

MIFR: A Modality-Invariant and Fair Representation Framework for Skin Disease Classification

Skin diseases represent a major global public health burden, yet machine learning tools developed to

Cross-lingual NER Model Accuracy Degradation under Extreme Domain Shift

Cross-lingual Named Entity Recognition (NER) leverages knowledge transfer between languages to ident

Intra-African Geographic Domain Shift in Wildlife Camera Trap Species Classification: A Comparative Study of Supervised and Zero-Shot Foundation Models

Abstract Camera trap networks such as Snapshot Safari have generated millions of l

Cross-lingual Ranking Model Robustness to Domain Shift in Low-Resource Legal Settings

Transferring information retrieval (IR) models from a high-resource language (typically English) to

Revisiting Invariant Learning for Out-of-Domain Generalization on Multi-Site Mammogram Datasets

Achieving health equity in Artificial Intelligence (AI) requires diagnostic models that maintain rel