Qingguo Chen
- Let the LLM Stick to Its Strengths: Learning to Route Economical LLM
- Multimodal Tabular Reasoning with Privileged Structured Information
- SPACE: Noise Contrastive Estimation Stabilizes Self-Play Fine-Tuning for Large Language Models
- Triplets Better Than Pairs: Towards Stable and Effective Self-Play Fine-Tuning for LLMs