NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Juntao Dai
3 papers
Peking University
InterMT: Multi-Turn Interleaved Preference Alignment with Human Feedback
Safe RLHF-V: Safe Reinforcement Learning from Multi-modal Human Feedback
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning