← all papers · overview

Evaluating Large Language Models On Historical Health Crisis Knowledge In Resource-limited Settings: A Hybrid Multi-metric Study

Abstract

Large Language Models (LLMs) offer significant potential for delivering health information. However, their reliability in low-resource contexts remains uncertain. This study evaluates GPT-4, Gemini Pro, Llama~3, and Mistral-7B on health crisis-related enquiries concerning COVID-19, dengue, the Nipah virus, and Chikungunya in the low-resource context of Bangladesh. We constructed a question--answer

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).