NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Pin-Yu Chen
3 papers
IBM Research
Adaptive Distraction: Probing LLM Contextual Robustness with Automated Tree Search
CoP: Agentic Red-teaming for Large Language Models using Composition of Principles
Shape it Up! Restoring LLM Safety during Finetuning