NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
policy update
3 papers
GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents
Multi-Objective Reinforcement Learning with Max-Min Criterion: A Game-Theoretic Approach
Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards