regret guarantees

In reinforcement learning and decision theory, regret guarantees provide bounds on the difference between the reward achieved by the learning algorithm and the optimal reward, guiding the evaluation of algorithm performance.

6 papers