Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Evaluating Vision-Language Models as a Zero-Shot Learning Alternative to You Only Look Once and Optical Character Recognition for Nigerian License Plate Recognition

Domain:

digital infrastructure

Record type:

paper
Creator:
TijMusMuhAli
Publisher:
arXiv
Host:avatar
License Plate Recognition (LPR) systems are critical tools in traffic monitoring, security enforcement, and urban mobility management. Traditional LPR systems often rely on a multi-stage pipeline involving object detection using You Only Look Once (YOLO) and Optical Character Recognition (OCR), which suffer from limitations such as high resource demands, poor performance in unstructured environments, and the need for large annotated datasets. This study explores the potential of Vision-Language Models (VLMs) as a unified, zeroshot learning solution for Nigerian license plate recognition. Using a curated dataset of 88 challenging real-world images collected in Nigeria, we evaluate five selected VLMs: Gemini 2.0 Flash Exp (Google DeepMind), Qwen2.5-VL-7B-Instruct (Alibaba), GPT-4o (OpenAI), Claude 4 Sonnet (Anthropic), and Llama 3.2 Vision 90b (Meta). Results based on Character Error Rate (CER) reveal that Gemini and Qwen significantly outperform other models in both accuracy and robustness, on the challenging image scenarios. This work highlights the practical advantages of VLMs over YOLO+OCR, questions the claims by model providers, and compares the performances of the VLMs.

Visit

doi.orgarxiv.org

Tasks

computer visionoptical character recognition

Tags

Computer Vision and Pattern Recognition (cs.CV)FOS: Computer and information sciences

Licenses

arXiv.org perpetual, non-exclusive licensehttp://arxiv.org/licenses/nonexclusive-distrib/1.0/

Similar

CAMARINE: A FISH SPECIES RECOGNITION SYSTEM THROUGH YOU ONLY LOOK ONCEComputer Vision for License Plate Recognition ChallengeApplication of Machine Learning for Automatic Number Plate Recognition Using Optical Character Recognition EngineCocoa insect pest detection and counting using computer vision (YOLO: You Only Look Once)ETHIOPIAN CAR LICENSE PLATE RECOGNITION USING DEEP LEARNINGCAR MODEL & LICENSE PLATE RECOGNITION

CAMARINE: A FISH SPECIES RECOGNITION SYSTEM THROUGH YOU ONLY LOOK ONCE

Data collection for marine sciences has always been arduous, mainly because of cost. The higher the

Computer Vision for License Plate Recognition Challenge

Use computer vision to detect and recognise Tunisian vehicle license plates
The data provided in this challenge is divided into 3 sets:
A train set of 900 images, where each image contains only one car and one license plate. The annotations provided are

Application of Machine Learning for Automatic Number Plate Recognition Using Optical Character Recognition Engine

In Nigeria, traffic control management, vehicle real- time tracking and vehicle identification have

Cocoa insect pest detection and counting using computer vision (YOLO: You Only Look Once)

Cocoa insect pest detection and counting using computer vision (YOLO: You Only Look Once)

Poster presented at the Deep Learning Indaba 2023 by SONNA FEUCHIOFACK  BOREL JORDAN

ETHIOPIAN CAR LICENSE PLATE RECOGNITION USING DEEP LEARNING

The focus of this research is to develop a system that assist humans in reading car
license plat

CAR MODEL & LICENSE PLATE RECOGNITION

CAR MODEL & LICENSE PLATE RECOGNITION

Poster presented at the Deep Learning Indaba 2023 by Jihed Selmi