← Timeline

@JulianSaks

@JulianSaks on X

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@JulianSaks

— web clipping, 469 words — published 2026-03-29

Thread by @JulianSaks

**JulianSaks** @JulianSaks 2026-03-29 let’s break this list down so it’s actually useful • JEPA / H-JEPA: avoids predicting every single pixel (too expensive) and rather predicts in latent space. H-JEPA adds hierarchy - short term details vs long term planning ie. how humans actually learn • I-JEPA: built for very efficient vision models. Masks image patches and predicts the semantics and in doing so bypasses heavy compute of traditional autoencoders • MC-JEPA & V-JEPA: both of these are built for videos. MC-JEPA separates content (what an object is) vs motion (how it moves). V-JEPA masks video features with no text labels making it perfect of action tracking at scale • Audio-JEPA: filters out background noise by treating sounds like visuals • Point-JEPA & 3D-JEPA: used primarily in AVs. Uses LiDAR point clouds & volumetric grids • ACT-JEPA: filters out real world noise to learn manipulation tasks efficiently via imitation learning • V-JEPA 2: predicts future physical states of the world caused by an action before it happens • LeJEPA: replaces techniques like masking with an Energy-Based Model (EBM) which mathematically prevents "feature collapse" & ensures the model scales reliably as data increases • Causal-JEPA: for learning true cause-and-effect physics by applying object level masking • V-JEPA 2.1: great for spatial grounding since it combines a dense predictive loss across image & video • LeWorldModel: built directly on LeJEPA's math but super compact - 15M params • ThinkJEPA: uses dense physical prediction with VLM reasoning. Best used when long-term strategy is needed > 2026-03-29 > > 14 most important and influential types of JEPA > > ▪️ JEPA / H-JEPA > > ▪️ I-JEPA > > ▪️ MC-JEPA > > ▪️ V-JEPA > > ▪️ Audio-JEPA > > ▪️ Point-JEPA > > ▪️ 3D-JEPA > > ▪️ ACT-JEPA > > ▪️ V-JEPA 2 > > ▪️ LeJEPA > > ▪️ Causal-JEPA > > ▪️ V-JEPA 2.1 > > ▪️ LeWorldModel > > ▪️ ThinkJEPA > > Save the list and check this out to explore these > > [image] --- **ScholarBuild** @scholaread1 [2026-03-30](https://x.com/scholaread1/status/2038547644143079667) Excellent breakdown—do you see the JEPA family eventually replacing Transformers for long-horizon reasoning in physical AI? --- **JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038568881422209216) if we can generate enough dreams in parallel over long horizons possibly but time will tell :) --- **Jakie PLA** @3DPrintAficio [2026-03-29](https://x.com/3DPrintAficio/status/2038323152682377417) Point-JEPA using LiDAR for volumetric understanding - exactly what I want in manufacturing vision. Spatial reasoning without compute bloat? OH YES. --- **JulianSaks** @JulianSaks [2026-03-29](https://x.com/JulianSaks/status/2038332603070091559) 😆💯 --- **Francesco** @0xfrankly [2026-03-30](https://x.com/0xfrankly/status/2038524921689554986) What would you use for 3d medical imaging tasksv --- **JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038539594849808779) I just came across this actually! --- **Nick** @nickmhc [2026-03-30](https://x.com/nickmhc/status/2038414627256889542) I appreciate that LeWorldModel is probably named after LeCun but kinda sounds like it’s named after LeBron --- **JulianSaks** @JulianSaks [2026-03-30](https://x.com/JulianSaks/status/2038498255902785629) LeBron should pick up claude code and build a JEPA 😆😆