Deep Q-Network (DQN)
Double DQN
Dueling DQN
Experience Replay
Prioritized Experience Replay
Rainbow DQN
Advantage Function
Baseline (Policy Gradient)
Deterministic Policy Gradient (DPG)
Policy Gradient Methods
Policy Gradient Theorem
Advantage Actor-Critic (A2C)
Asynchronous Advantage Actor-Critic (A3C)
Generalized Advantage Estimation (GAE)
DDPG
Distributional RL
IMPALA
PPO
SAC (Soft Actor-Critic)
TD3
-
S
OpenAI Spinning Up — Part 3: Intro to Policy Optimization (Baselines in Policy Gradients) Art… 3 terms
-
-
-
-
-
-
-
-
A
Deep Reinforcement Learning with Double Q-learning (van Hasselt et al. — the Double DQN paper) Pap… 1 term
-
-
-
-
-
A
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures Pap… 1 term
-
-
-
-
-
-
-
-
A
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor Pap… 1 term
-
-
-
No terms or links match your search.