← all papers · overview

CAKE: Cloud Architecture Knowledge Evaluation Of Large Language Models

Abstract

In today's software architecture, large language models (LLMs) serve as software architecture co-pilots. However, no benchmark currently exists to evaluate large language models' actual understanding of cloud-native software architecture. For this reason we present a benchmark called CAKE, which consists of 188 expert-validated questions covering four cognitive levels of Bloom's revised taxonomy -

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).