Drooid Logo
Back to today’s briefing

Story perspectives

Meta's V-JEPA: AI Learns Surprise from Video Insights

12/7/2025

26 5

1 of 1

Story summary
  • Meta researchers developed Video Joint Embedding Predictive Architecture (V-JEPA) that learns about the world through videos.
  • V-JEPA shows a notion of "surprise" when information contradicts its learned knowledge, akin to infant object permanence.
  • Unlike traditional pixel-space models, which misinterpret irrelevant details, V-JEPA aims to better interpret complex scenes.
  • Cognitive scientist Micha Heilbron says the claims are plausible and the results are intriguing.