☰
Notes
🔍
×
Statistics
Supervised Learning
Unsupervised Learning
Deep Learning
Reinforcement Learning
▼
Foundations & Concepts
29
Core Vocabulary
9
Markov Decision Processes
8
Value Functions
6
Returns & Rewards
6
▼
RL Paradigms
15
Learning Approaches
5
Problem Settings
5
Multi-Agent RL
5
Exploration vs Exploitation
9
Multi-Armed Bandits
7
▼
Tabular Methods
24
Dynamic Programming
5
Monte Carlo Methods
6
Temporal Difference Learning
7
Tabular Control Algorithms
6
Function Approximation
7
▼
Deep Reinforcement Learning
25
Value-Based Deep RL
7
Policy Gradient Methods
7
Actor-Critic
4
Trust Region & Modern Algorithms
7
Model-Based RL
8
▼
Reward Engineering & Imitation
9
Reward Design
4
Imitation Learning
5
Extensions & Variants
9
RL Frameworks & Environments
8
Applications
9
AI ML
>
Reinforcement Learni…
>
Deep Reinforcement L…
>
Trust Region & Moder…
Overview
◀
Map
Tree
Terms
Read
▶
Resources
Quiz
Play
Trust Region & Modern Algorithms
7
terms
Terms in this group
TRPO
PPO
DDPG
TD3
SAC (Soft Actor-Critic)
IMPALA
Distributional RL
Term
Open full page
↗
Loading...