Skip to yearly menu bar Skip to main content


Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety

Tyler Crosse ⋅ Catherine Ge-Wang ⋅ Benjamin Hadad ⋅ Joachim Schaeffer ⋅ Ram Potham ⋅ Tyler Tracy

Abstract

Chat is not available.