Juntao Li
- SCAN: Self-Denoising Monte Carlo Annotation for Robust Process Reward Learning
- Scaling Code-Assisted Chain-of-Thoughts and Instructions for Model Reasoning
- Thoughts Are All Over the Place: On the Underthinking of Long Reasoning Models
- XIFBench: Evaluating Large Language Models on Multilingual Instruction Following