← all papers · overview

Evaluating Language Models For Generating And Judging Programming Feedback

Abstract

The emergence of large language models (LLMs) has transformed research and practice across a wide range of domains. Within the computing education research (CER) domain, LLMs have garnered significant attention, particularly in the context of learning programming. Much of the work on LLMs in CER, however, has focused on applying and evaluating proprietary models. In this article, we evaluate the e

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).