NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Kun Shao
3 papers
Huawei Noah's Ark Lab
Succeed or Learn Slowly: Sample Efficient Off-Policy Reinforcement Learning for Mobile App Control
ThinkBench: Dynamic Out-of-Distribution Evaluation for Robust LLM Reasoning
Uncertainty-quantified Rollout Policy Adaptation for Unlabelled Cross-domain Video Temporal Grounding