Mido Assran
- V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
2025/06/11 by Mido Assran, Assran, Mido, Adrien Bardes +59 · 2 voices · 180 citations
Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Robotics (cs.RO) #Social Robot Interaction and HRI #cs.AI #cs.CV #cs.LG #cs.RO
- Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
2025/04/19 by Sergio Arnaud, Arnaud, Sergio, Paul McVay +41 · 1 voice · 13 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #I.2.10 #I.2.6 #I.2.9 #I.3.7 #I.4.6 #I.4.8 #Robotics (cs.RO) #cs.AI #cs.CV #cs.RO
- Hierarchical Planning with Latent World Models
2026/04/03 by Wancong Zhang, Basile Terver, Artem Zholus +8 · 2 voices · 1 citation
Computer Science · #cs.LG
- V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning
2026/03/15 by Lorenzo Mur-Labadia, Matthew J. Muckley, Matthew Muckley +7 · 2 voices · 2 citations
Computer Science · Engineering · #Action recognition #Anticipation (artificial intelligence) #Deep learning #Encoder #Feature (linguistics) #Human Pose and Action Recognition #Key (lock) #Multimodal Machine Learning Applications #Representation (politics) #Robot Manipulation and Learning #Training (meteorology) #cs.CV