NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Yixiao Huang
3 papers
UC Berkeley
Generalization or Hallucination? Understanding Out-of-Context Reasoning in Transformers
OVERT: A Benchmark for Over-Refusal Evaluation on Text-to-Image Models
Understanding and Improving Fast Adversarial Training against $l_0$ Bounded Perturbations