Policy Gradient
concepts · 2 notes linked
Related: Openai · Reinforcement Learning · Actor Critic · Deepmind · Q Learning · Model Based RL
Notes
- Intuitive RL: Intro to Advantage-Actor-Critic (A2C) | HackerNoon — Intuitive narrative introduction to the Advantage-Actor-Critic RL algorithm
- Reinforcement Learning algorithms — an intuitive overview — Survey of model-free and model-based RL algorithm families