Binhang Yuan
- AREAL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
- AtmosSci-Bench: Evaluating the Recent Advance of Large Language Model for Atmospheric Science
- Efficient Pre-Training of LLMs via Topology-Aware Communication Alignment on More Than 9600 GPUs
- Hallucination at a Glance: Controlled Visual Edits and Fine-Grained Multimodal Learning
- MeCeFO: Enhancing LLM Training Robustness via Fault-Tolerant Optimization
- Multi-step Visual Reasoning with Visual Tokens Scaling and Verification