← all papers · overview

D-NLP At Semeval-2024 Task 2: Evaluating Clinical Inference Capabilities Of Large Language Models

Abstract

Large language models (LLMs) have garnered significant attention and widespread usage due to their impressive performance in various tasks. However, they are not without their own set of challenges, including issues such as hallucinations, factual inconsistencies, and limitations in numerical-quantitative reasoning. Evaluating LLMs in miscellaneous reasoning tasks remains an active area of researc

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).