Si Liu
- FACT: Mitigating Inconsistent Hallucinations in LLMs via Fact-Driven Alternating Code-Text Training
- RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation
- Towards Realistic Earth-Observation Constellation Scheduling: Benchmark and Methodology
- UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning