Six DQN improvements in one agent: double, dueling, prioritized replay, multi-step returns, distributional values and noisy exploration, with an ablation for each. The result is mostly the finding that they address different problems and therefore add up.
supersedescorrects · extends
classifiesspecializes · part-of
substitutes forapproximates · alternative-to
depends onrequires · validates
Colour is the family; a dashed line is the second member of it.
Drag to pan · scroll to zoom · click a node to open it
This node
extendsadds capability to Deep Q-Networkevery fix had been measured alone, against the same unimproved baseline