stochastic approximation
- A General-Purpose Theorem for High-Probability Bounds of Stochastic Approximation with Polyak Averaging
- Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features
- Finite-Sample Analysis of Policy Evaluation for Robust Average Reward Reinforcement Learning
- REINFORCE Converges to Optimal Policies with Any Learning Rate
- Statistical inference for Linear Stochastic Approximation with Markovian Noise