V-JEPA

AI and Machine Learning · Representation Learning · 2024 · v-jepa.yaml

Extends the same idea to video, predicting the features of masked spacetime regions. Learns motion and object permanence from unlabelled footage alone.

Colour is the family; a dashed line is the second member of it.

Drag to pan · scroll to zoom · click a node to open it
G genie Genie v-jepa V-JEPA genie->v-jepa the dynamics are learned over video representations jepa JEPA v-jepa->jepa predicts masked spacetime regions across video frames

This node

Referenced by

References