MATH · IN · MODELS

JEPA (Joint Embedding Predictive Architecture)

Independent / academic

By model (4)

ViT-S/16 (I-JEPA objective, independently trained per view: smallNORB/nuScenes/ImageNet-1k)
V-JEPA (video joint-embedding predictive architecture)

Papers

Latent Video Prediction Learns Better World Models (2026), The Invisible Hand of Physics: When Video Diffusion Models Know More Than They Show (2026), Social-JEPA: Emergent Geometric Isomorphism in Independently Trained World Models (2026), Interpreting Physics in Video World Models (2026), Do Video Foundation Models Understand Intuitive Physics? A Layerwise Probing Analysis (2026)