— web clipping, 469 words — published 2026-03-29
Thread by @JulianSaks
**JulianSaks** @JulianSaks 2026-03-29
let’s break this list down so it’s actually useful
• JEPA / H-JEPA: avoids predicting every single pixel (too expensive) and rather predicts in latent space. H-JEPA adds hierarchy - short term details vs long term planning ie. how humans actually learn
• I-JEPA: built for very efficient vision models. Masks image patches and predicts the semantics and in doing so bypasses heavy compute of traditional autoencoders
• MC-JEPA & V-JEPA: both of these are built for videos. MC-JEPA separates content (what an object is) vs motion (how it moves). V-JEPA masks video features with no text labels making it perfect of action tracking at scale
• Audio-JEPA: filters out background noise by treating sounds like visuals
• Point-JEPA & 3D-JEPA: used primarily in AVs. Uses LiDAR point clouds & volumetric grids
• ACT-JEPA: filters out real world noise to learn manipulation tasks efficiently via imitation learning
• V-JEPA 2: predicts future physical states of the world caused by an action before it happens
• LeJEPA: replaces techniques like masking with an Energy-Based Model (EBM) which mathematically prevents "feature collapse" & ensures the model scales reliably as data increases
• Causal-JEPA: for learning true cause-and-effect physics by applying object level masking
• V-JEPA 2.1: great for spatial grounding since it combines a dense predictive loss across image & video
• LeWorldModel: built directly on LeJEPA's math but super compact - 15M params
• ThinkJEPA: uses dense physical prediction with VLM reasoning. Best used when long-term strategy is needed
> 2026-03-29
>
> 14 most important and influential types of JEPA
>
> ▪️ JEPA / H-JEPA
>
> ▪️ I-JEPA
>
> ▪️ MC-JEPA
>
> ▪️ V-JEPA
>
> ▪️ Audio-JEPA
>
> ▪️ Point-JEPA
>
> ▪️ 3D-JEPA
>
> ▪️ ACT-JEPA
>
> ▪️ V-JEPA 2
>
> ▪️ LeJEPA
>
> ▪️ Causal-JEPA
>
> ▪️ V-JEPA 2.1
>
> ▪️ LeWorldModel
>
> ▪️ ThinkJEPA
>
> Save the list and check this out to explore these
>
> [image]
---
**ScholarBuild** @scholaread1 [2026-03-30](https://x.com/scholaread1/status/2038547644143079667)
Excellent breakdown—do you see the JEPA family eventually replacing Transformers for long-horizon reasoning in physical AI?
---
**JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038568881422209216)
if we can generate enough dreams in parallel over long horizons possibly but time will tell :)
---
**Jakie PLA** @3DPrintAficio [2026-03-29](https://x.com/3DPrintAficio/status/2038323152682377417)
Point-JEPA using LiDAR for volumetric understanding - exactly what I want in manufacturing vision. Spatial reasoning without compute bloat? OH YES.
---
**JulianSaks** @JulianSaks [2026-03-29](https://x.com/JulianSaks/status/2038332603070091559)
😆💯
---
**Francesco** @0xfrankly [2026-03-30](https://x.com/0xfrankly/status/2038524921689554986)
What would you use for 3d medical imaging tasksv
---
**JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038539594849808779)
I just came across this actually!
---
**Nick** @nickmhc [2026-03-30](https://x.com/nickmhc/status/2038414627256889542)
I appreciate that LeWorldModel is probably named after LeCun but kinda sounds like it’s named after LeBron
---
**JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038498255902785629)
LeBron should pick up claude code and build a JEPA 😆😆