Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Early detection of food safety risks using BERT and large language models

Domaine:

natural language processinghealthcare

Type de record:

paper
Créateur:
MohSouMos
Éditeur:
Institute of Advanced Engineering and Science
Hôte:
Sentiment analysis can be a powerful tool in safeguarding public health. This allows authorities to investigate and take action before a foodborne illness outbreak spreads. This paper introduces a novel system that proactively empowers restaurants to identify potential food safety hazards and hygiene regulation violations. The system leverages the power of natural language processing (NLP) to analyze Arabic restaurant reviews left by customers. By fine-tuning a pre-trained BERT mini-Arabic model on three targeted datasets: Sentiment Twitter Corpus, an Algerian dialect dataset, and an Arabic restaurant dataset, the system achieves an impressive accuracy of 91%. Additionally, the system caters to spoken feedback by accepting audio reviews. We utilized Whisper AI for accurate text transcription, followed by classification using a fine-tuned Gemini model from Google on Algerian local comments and others generated using large language models (LLMs) through few-shot learning techniques, reaching an accuracy of 93%. Notably, both models operate independently and concurrently. Leveraging RESTful APIs, the system integrates the solved sub-solutions from each microservice into a fusion layer for a comprehensive restaurant evaluation. This multifaceted approach delivers remarkable results for both modern standard Arabic (MSA) and the Algerian dialect, demonstrating its effectiveness in addressing restaurant food safety concerns.

Visit

doi.org

Tasks

automatic speech recognitionsentiment analysisspeech processingtext classification

Languages

Arabic, Algerian Spoken

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

Context-Based Question Answering Using Large Language BERT Variant Models for Low Resourced Sesotho sa Leboa LanguageAbusive and Threatening Language Detection in Urdu using Boosting based and BERT based models: A Comparative ApproachMachine Translation Hallucination Detection for Low and High Resource Languages using Large Language ModelsIndicSafeEval: Safety Robustness of Large Language Models under Multilingual Persuasive Jailbreak AttacksSchema Generation for Large Knowledge Graphs Using Large Language ModelsThe Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text

Context-Based Question Answering Using Large Language BERT Variant Models for Low Resourced Sesotho sa Leboa Language

Abusive and Threatening Language Detection in Urdu using Boosting based and BERT based models: A Comparative Approach

Online hatred is a growing concern on many social media platforms. To address this issue, different

Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models

Recent advancements in massively multilingual machine translation systems have significantly enhance

IndicSafeEval: Safety Robustness of Large Language Models under Multilingual Persuasive Jailbreak Attacks

Large language models (LLMs) are increasingly used in multilingual settings, yet their safety is sti

Schema Generation for Large Knowledge Graphs Using Large Language Models

Schemas play a vital role in ensuring data quality and supporting usability in the Semantic Web and

The Arabic AI Fingerprint: Stylometric Analysis and Detection of Large Language Models Text

Large Language Models (LLMs) have achieved unprecedented capabilities in generating human-like text,