JEPA

AI and Machine Learning · Representation Learning · 2023 · also: Joint-Embedding Predictive Architecture, I-JEPA · jepa.yaml

Predicts the representation of a masked region rather than its pixels. The argument is that pixel reconstruction forces the model to model detail that is genuinely unpredictable and carries no meaning.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G jepa JEPA masked-autoencoder Masked Autoencoder jepa->masked-autoencoder pixel targets spend capacity on unpredictable detail v-jepa V-JEPA v-jepa->jepa predicts masked spacetime regions across video frames

This node

Referenced by

References