NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Chuan Guo
4 papers
Meta FAIR
AdvPrefix: An Objective for Nuanced LLM Jailbreaks
AgentDAM: Privacy Leakage Evaluation for Autonomous Web Agents
Rethinking the Role of Verbatim Memorization in LLM Privacy
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks