offline reinforcement learning

A branch of reinforcement learning that deals with learning optimal policies from previously collected experience without interacting with the environment. It allows for the reuse of past data to improve model training.

20 papers