← all papers · overview

Psyeval: A Suite Of Mental Health Related Tasks For Evaluating Large Language Models

Abstract

Evaluating Large Language Models (LLMs) in the mental health domain poses distinct challenged from other domains, given the subtle and highly subjective nature of symptoms that exhibit significant variability among individuals. This paper presents PsyEval, the first comprehensive suite of mental health-related tasks for evaluating LLMs. PsyEval encompasses five sub-tasks that evaluate three critic

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).