NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Hanna Wallach
3 papers
Microsoft
Comparison requires valid measurement: Rethinking attack success rate comparisons in AI red teaming
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
Validating LLM-as-a-Judge Systems under Rating Indeterminacy