Weight Decay

AI and Machine Learning · Training and Optimisation · weight-decay.yaml

Shrinks weights toward zero every step. The oldest regulariser, and the one whose interaction with adaptive optimisers turned out to be subtler than anyone assumed.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G dropout Dropout early-stopping Early Stopping weight-decay Weight Decay early-stopping->weight-decay stops before overfitting rather than penalising capacity weight-decay->dropout constrains weight magnitude rather than co- adaptation

This node

Referenced by

References