NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
prompt injection attacks
3 papers
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
Security Challenges in AI Agent Deployment: Insights from a Large Scale Public Competition