Dueling Network

AI and Machine Learning · Reinforcement Learning · 2016 · dueling-network.yaml

Splits the head into a state-value stream and a per-action advantage stream, then adds them back. How good a state is gets learned once instead of separately inside every action's estimate.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G deep-q-network Deep Q-Network dueling-network Dueling Network dueling-network->deep-q-network in most states the action barely matters, yet each is estimated alone

This node

References