← all papers · overview

Automatic Detection Of Llm-generated Code: A Comparative Case Study Of Contemporary Models Across Function And Class Granularities

Abstract

The adoption of Large Language Models (LLMs) for code generation risks incorporating vulnerable code into software systems. Existing detectors face two critical limitations: a lack of systematic cross-model validation and opaque "black box" operation. We address this through a comparative study of code generated by four distinct LLMs: GPT-3.5, Claude 3 Haiku, Claude Haiku 4.5, and GPT-OSS. Analy

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).