Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Fine-Tuning YOLOv8s for Unified Human and Face Detection in Crowded Environments

Record type:

model
Creator:
OusAhmLat
Publisher:
Eng
Host:
Accurate detection of human bodies and faces in densely populated scenes remains challenging due to occlusions and overlapping instances. This paper presents a lightweight object detection solution built upon the You Only Look Once version 8 small (YOLOv8s) architecture, fine-tuned for challenging urban scenes where occlusion, density, and limited computing resources are common. Leveraging an enhanced dataset with detailed person and face annotations, our model achieves a good mean Average Precision at IoU threshold 0.5 (mAP@0.5) of 57.61%, with particularly robust performance on full-body human detection (Average Precision (AP) = 73.5%). Despite moderate face detection accuracy (AP = 42.1%), qualitative results demonstrate solid performance under real-world constraints. The model's compact size and high inference speed make it ideally suited for deployment on edge devices, such as mobile cameras and embedded Artificial Intelligence (AI) systems. A compelling use case is explored through the lens of crowd monitoring in Jamaa El-Fna square in Marrakech, a bustling and high-density public space that demands real-time situational awareness. This work offers a practical tool for urban analytics and public safety, and it lays the foundation for future improvements in face detection, post-processing, and real-time system integration.

Visit

doi.org

Tasks

computer visionimage classification

Languages

Arabic, Moroccan Spoken

Licenses

https://creativecommons.org/licenses/by/4.0/

Similar

Fine-Tuning Transformers and LLMs for Fake News Detection in Algerian DialectSequential Fine-Tuning Strategies for Cross-Lingual Euphemism Detection in YorubaMultimodal Fine-Tuning for Cross-Lingual Euphemism Detection in mT5 ModelsIntermediate Task Fine-Tuning for Cross-Lingual Euphemism Detection EfficiencyCross-lingual Fine-tuning Robustness in Euphemism Detection BenchmarksIntermediate Language Fine-Tuning for Cross-Lingual Euphemism Detection in XLM-R

Fine-Tuning Transformers and LLMs for Fake News Detection in Algerian Dialect

Sequential Fine-Tuning Strategies for Cross-Lingual Euphemism Detection in Yoruba

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Multimodal Fine-Tuning for Cross-Lingual Euphemism Detection in mT5 Models

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Intermediate Task Fine-Tuning for Cross-Lingual Euphemism Detection Efficiency

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Cross-lingual Fine-tuning Robustness in Euphemism Detection Benchmarks

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec

Intermediate Language Fine-Tuning for Cross-Lingual Euphemism Detection in XLM-R

Euphemisms are culturally variable and often ambiguous, posing challenges for language models, espec