NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Qingru Zhang
3 papers
Microsoft CoreAI
Ask a Strong LLM Judge when Your Reward Model is Uncertain
Think-RM: Enabling Long-Horizon Reasoning in Generative Reward Models
Win Fast or Lose Slow: Balancing Speed and Accuracy in Latency-Sensitive Decisions of LLMs