NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Kevin Jamieson
3 papers
U Washington
Adapting to Stochastic and Adversarial Losses in Episodic MDPs with Aggregate Bandit Feedback
Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
On the Universal Near Optimality of Hedge in Combinatorial Settings