NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Tsung-Yi Ho
3 papers
Department of Computer Science and Engineering, The Chinese University of Hong Kong
CARE: Decoding-Time Safety Alignment via Rollback and Introspection Intervention
CoP: Agentic Red-teaming for Large Language Models using Composition of Principles
PermLLM: Learnable Channel Permutation for N:M Sparse Large Language Models