Weight Initialization

AI and Machine Learning · Training and Optimisation · 2010 · weight-initialization.yaml

Chooses the starting variance so activations neither shrink nor blow up as they propagate. A deep network started wrong does not train slowly; it does not train.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G he-initialization He Initialization weight-initialization Weight Initialization he-initialization->weight-initialization the usual variance assumes a symmetric activation

Referenced by

References