Logo Lanfrica
  • Home
  • Atlas
  • Insights
  • Docs
  • Sign in

© 2026 Lanfrica. All rights reserved. All copyrights of the resources shown on the Lanfrica website belong to the original copyright holders, unless explicitly stated otherwise.

Fine-Tuning and Evaluating Conversational AI for Agricultural Advisory

Domain:

agriculturenatural language processing

Record type:

papermodelsoftware
Creator:
SinGanSinPed
Host:avatar
Large Language Models show promise for agricultural advisory, yet vanilla models exhibit unsupported recommendations, generic advice lacking specific, actionable detail, and communication styles misaligned with smallholder farmer needs. In high stakes agricultural contexts, where recommendation accuracy has direct consequences for farmer outcomes, these limitations pose challenges for responsible deployment. We present a hybrid LLM architecture that decouples factual retrieval from conversational delivery: supervised fine-tuning with LoRA on expert-curated GOLDEN FACTS (atomic, verified units of agricultural knowledge) optimizes fact recall, while a separate stitching layer transforms retrieved facts into culturally appropriate, safety-aware responses. Our evaluation framework, DG-EVAL, performs atomic fact verification (measuring recall, precision, and contradiction detection) against expert-curated ground truth rather than Wikipedia or retrieved documents. Experiments across multiple model configurations on crops and queries from Bihar, India show that fine-tuning on curated data substantially improves fact recall and F1, while maintaining high relevance. Using a fine-tuned smaller model achieves comparable or better factual quality at a fraction of the cost of frontier models. A stitching layer further improves safety subscores while maintaining high conversational quality. We release the farmerchat-prompts library to enable reproducible development of domain-specific agricultural AI. 22 pages, 5 figures, 9 tables

Visit

arxiv.org

Tasks

question answering

Tags

Computation and LanguageArtificial IntelligenceMachine Learning

Similar

Where Does Retrieval Fail? Evaluating RAG Architectures for Agricultural AdvisoryAI for Agricultural Advisory and Financial Services for Smallholder Farmers in TanzaniaAgricultural Advisory Management System with AI Powered Chat-BotComparison of Intermediate-Task Fine-Tuning and Multilingual Fine-Tuning for Zero-Shot Low-Resource Language AccuracyBeyond Pretraining Bias: Evaluating Language-Adaptive Fine-Tuning Strategies for Sentiment and Topic Classification in Three Nigerian LanguagesFine Tuning Methods for Low-resource Languages

Where Does Retrieval Fail? Evaluating RAG Architectures for Agricultural Advisory

Retrieval quality in RAG systems is commonly reported as a single aggregate score, which can hide la

AI for Agricultural Advisory and Financial Services for Smallholder Farmers in Tanzania

A majority of smallholder farmers in Tanzania are only able to communicate through the Kiswahili spo

Agricultural Advisory Management System with AI Powered Chat-Bot

Agriculture remains the backbone of Malawi’s economy, supporting the livelihoods of a majority of it

Comparison of Intermediate-Task Fine-Tuning and Multilingual Fine-Tuning for Zero-Shot Low-Resource Language Accuracy

Accuracy of English-language Question Answering (QA) systems has improved significantly in recent ye

Beyond Pretraining Bias: Evaluating Language-Adaptive Fine-Tuning Strategies for Sentiment and Topic Classification in Three Nigerian Languages

This paper presents a controlled empirical study comparing four adaptation strategies — zero-shot pr

Fine Tuning Methods for Low-resource Languages

The rise of Large Language Models has not been inclusive of all cultures. The models are mostly trai