Models
loadingβ¦
loadingβ¦
Models is one of the most active areas in Awesome AI for Code β 37 papers in this collection, evaluated on datasets like AIR-Bench, MTEB, RouterEval. A strong starting point is "Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception".