Security Challenges in AI Agent Deployment: Insights from a Large Scale Public Competition
agent robustnessai agent vulnerabilitiesattack success ratesbenchmark constructioncritical vulnerabilitiesdefense improvementsdeployment policiesfrontier llmsinference compute budgetmodel capabilitypolicy violationsprompt injection attacksred teamingtransferabilityuniversality
AI agents are rapidly being deployed across diverse industries, but can they adhere to deployment policies under attacks? We organized a one-month red teaming challenge