Double Q-learning
Pairings in the atlas
- Decoupled max estimatorstandardOverestimation-free RLreinforcement-learning
Rivals: other methods for the same problems
- A2C
- A3C
- Actor-critic
- DDPG
- Deep Q-network
- Double DQN
- Dueling DQN
- Expected SARSA
- Generalized advantage estimation
- Hindsight experience replay
- IMPALA
- Monte Carlo control
- PPG
- Policy gradient
- Prioritized experience replay
- Proximal policy optimization
- Q-learning
- Rainbow
- Reward shaping
- SARSA
- Soft actor-critic
- TD3
- Temporal difference learning
- Trust region policy optimization