← all papers · overview

Assessing Large Language Models For Medical QA: Zero-shot And Llm-as-a-judge Evaluation

Abstract

Recently, Large Language Models (LLMs) have gained significant traction in medical domain, especially in developing a QA systems to Medical QA systems for enhancing access to healthcare in low-resourced settings. This paper compares five LLMs deployed between April 2024 and August 2025 for medical QA, using the iCliniq dataset, containing 38,000 medical questions and answers of diverse specialties

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).