NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
dynamic adjustment
3 papers
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
LLM-Explorer: A Plug-in Reinforcement Learning Policy Exploration Enhancement Driven by Large Language Models
Sample-Efficient Multi-Round Generative Data Augmentation for Long-Tail Instance Segmentation