Skip to content
∑
Praveen T N
Learning Library
Overview
Concepts
Flashcards
Writing
AI Feed
Graph
Portfolio ↗
Overview
Concepts
Flashcards
Writing
AI Feed
Graph
Search
Back to portfolio ↗
Library
/
Concepts
/
Reinforcement Learning
/
RL Foundations
RL Foundations
MDPs, value functions, TD learning, policy gradients, actor-critic, TRPO and PPO.
20
concepts
140
flashcards
146
minutes of reading
Start with Returns, Discounting, and Episodes
Drill the deck
All levels
beginner
intermediate
advanced
01
Returns, Discounting, and Episodes
The return is the quantity a reinforcement learning agent actually optimises; discounting controls how far into the future it looks, and whether interactions are episodic or continuing shapes which formulation applies.
beginner
7m
7 cards