NeurIPS 2025 Explorer
Concepts
Authors
Glossary
johnsanterre.github.io
Taira Tsuchiya
3 papers
The University of Tokyo
Adapting to Stochastic and Adversarial Losses in Episodic MDPs with Aggregate Bandit Feedback
Bandit and Delayed Feedback in Online Structured Prediction
Online Inverse Linear Optimization: Efficient Logarithmic-Regret Algorithm, Robustness to Suboptimality, and Lower Bound