ZHAO-XIANG ZHANG
- DriveDPO: Policy Learning via Safety DPO For End-to-End Autonomous Driving
- KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation
- MVU-Eval: Towards Multi-Video Understanding Evaluation for Multimodal LLMs
- OmniBench: Towards The Future of Universal Omni-Language Models
- SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
- TC-Light: Temporally Coherent Generative Rendering for Realistic World Transfer