Mixed-Precision Training

AI and Machine Learning · Efficiency and Deployment · 2017 · mixed-precision-training.yaml

Stores and multiplies in sixteen bits while keeping a master copy of the weights and the accumulations in thirty-two, with a loss scale to keep small gradients representable.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G half-precision Half Precision mixed-precision-training Mixed-Precision Training mixed-precision-training->half-precision the halved bandwidth is the entire point stochastic-gradient-descent Stochastic Gradient Descent mixed-precision-training->stochastic-gradient-descent full-precision training is bandwidth bound

This node

References