Xiaolong Li
- SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
- SWE-SQL: Illuminating LLM Pathways to Solve User SQL Issues in Real-World Applications
- The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement
- Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training