Safety
loadingβ¦
loadingβ¦
Safety is one of the most active areas in Awesome AI Agents β 672 papers in this collection, evaluated on datasets like HotpotQA, AgentDojo, HumanEval. A strong starting point is "AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security".