Causal-JEPA: Learning World Models through Object-Level Latent Masking
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nam, Heejeong, Lidec, Quentin Le, Maes, Lucas, LeCun, Yann, Balestriero, Randall |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
stable-worldmodel-v1: Reproducible World Modeling Research and Evaluation
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures
von: Huang, Hai, et al.
Veröffentlicht: (2025)
von: Huang, Hai, et al.
Veröffentlicht: (2025)
Learning by Reconstruction Produces Uninformative Features For Perception
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
Fast and Exact Enumeration of Deep Networks Partitions Regions
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
Value-guided action planning with JEPA world models
von: Destrade, Matthieu, et al.
Veröffentlicht: (2025)
von: Destrade, Matthieu, et al.
Veröffentlicht: (2025)
Semantic Tube Prediction: Beating LLM Data Efficiency with JEPA
von: Huang, Hai, et al.
Veröffentlicht: (2026)
von: Huang, Hai, et al.
Veröffentlicht: (2026)
Gaussian Embeddings: How JEPAs Secretly Learn Your Data Density
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
Learning Latent Action World Models In The Wild
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
An Information-Theoretic Perspective on Variance-Invariance-Covariance Regularization
von: Shwartz-Ziv, Ravid, et al.
Veröffentlicht: (2023)
von: Shwartz-Ziv, Ravid, et al.
Veröffentlicht: (2023)
Learning and Leveraging World Models in Visual Representation Learning
von: Garrido, Quentin, et al.
Veröffentlicht: (2024)
von: Garrido, Quentin, et al.
Veröffentlicht: (2024)
Variance Covariance Regularization Enforces Pairwise Independence in Self-Supervised Representations
von: Mialon, Grégoire, et al.
Veröffentlicht: (2022)
von: Mialon, Grégoire, et al.
Veröffentlicht: (2022)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
von: Zhou, Gaoyue, et al.
Veröffentlicht: (2024)
von: Zhou, Gaoyue, et al.
Veröffentlicht: (2024)
Why AI systems don't learn and what to do about it: Lessons on autonomous learning from cognitive science
von: Dupoux, Emmanuel, et al.
Veröffentlicht: (2026)
von: Dupoux, Emmanuel, et al.
Veröffentlicht: (2026)
stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
Rectified LpJEPA: Joint-Embedding Predictive Architectures with Sparse and Maximum-Entropy Representations
von: Kuang, Yilun, et al.
Veröffentlicht: (2026)
von: Kuang, Yilun, et al.
Veröffentlicht: (2026)
JEPA as a Neural Tokenizer: Learning Robust Speech Representations with Density Adaptive Attention
von: Ioannides, Georgios, et al.
Veröffentlicht: (2025)
von: Ioannides, Georgios, et al.
Veröffentlicht: (2025)
Video Representation Learning with Joint-Embedding Predictive Architectures
von: Drozdov, Katrina, et al.
Veröffentlicht: (2024)
von: Drozdov, Katrina, et al.
Veröffentlicht: (2024)
Navigation World Models
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?
von: Terver, Basile, et al.
Veröffentlicht: (2025)
von: Terver, Basile, et al.
Veröffentlicht: (2025)
The Spike, the Sparse and the Sink: Anatomy of Massive Activations and Attention Sinks
von: Sun, Shangwen, et al.
Veröffentlicht: (2026)
von: Sun, Shangwen, et al.
Veröffentlicht: (2026)
AI Must Embrace Specialization via Superhuman Adaptable Intelligence
von: Goldfeder, Judah, et al.
Veröffentlicht: (2026)
von: Goldfeder, Judah, et al.
Veröffentlicht: (2026)
A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures
von: Terver, Basile, et al.
Veröffentlicht: (2026)
von: Terver, Basile, et al.
Veröffentlicht: (2026)
Blockwise Self-Supervised Learning at Scale
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
World Models for Learning Dexterous Hand-Object Interactions from Human Videos
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
Light-weight probing of unsupervised representations for Reinforcement Learning
von: Zhang, Wancong, et al.
Veröffentlicht: (2022)
von: Zhang, Wancong, et al.
Veröffentlicht: (2022)
Learning from Reward-Free Offline Data: A Case for Planning with Latent Dynamics Models
von: Sobal, Vlad, et al.
Veröffentlicht: (2025)
von: Sobal, Vlad, et al.
Veröffentlicht: (2025)
Introduction to Latent Variable Energy-Based Models: A Path Towards Autonomous Machine Intelligence
von: Dawid, Anna, et al.
Veröffentlicht: (2023)
von: Dawid, Anna, et al.
Veröffentlicht: (2023)
Towards Causal Representation Learning with Observable Sources as Auxiliaries
von: Kim, Kwonho, et al.
Veröffentlicht: (2025)
von: Kim, Kwonho, et al.
Veröffentlicht: (2025)
Variance-Covariance Regularization Improves Representation Learning
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
SAFE: A Novel Approach to AI Weather Evaluation through Stratified Assessments of Forecasts over Earth
von: Masi, Nick, et al.
Veröffentlicht: (2025)
von: Masi, Nick, et al.
Veröffentlicht: (2025)
Revisiting Feature Prediction for Learning Visual Representations from Video
von: Bardes, Adrien, et al.
Veröffentlicht: (2024)
von: Bardes, Adrien, et al.
Veröffentlicht: (2024)
Why and How Auxiliary Tasks Improve JEPA Representations
von: Yu, Jiacan, et al.
Veröffentlicht: (2025)
von: Yu, Jiacan, et al.
Veröffentlicht: (2025)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
von: Shani, Chen, et al.
Veröffentlicht: (2025)
von: Shani, Chen, et al.
Veröffentlicht: (2025)
ALLoRA: Adaptive Learning Rate Mitigates LoRA Fatal Flaws
von: Huang, Hai, et al.
Veröffentlicht: (2024)
von: Huang, Hai, et al.
Veröffentlicht: (2024)
Hierarchical Planning with Latent World Models
von: Zhang, Wancong, et al.
Veröffentlicht: (2026)
von: Zhang, Wancong, et al.
Veröffentlicht: (2026)
Task Priors: Enhancing Model Evaluation by Considering the Entire Space of Downstream Tasks
von: Patel, Niket, et al.
Veröffentlicht: (2025)
von: Patel, Niket, et al.
Veröffentlicht: (2025)
Transformers without Normalization
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
Intuitive physics understanding emerges from self-supervised pretraining on natural videos
von: Garrido, Quentin, et al.
Veröffentlicht: (2025)
von: Garrido, Quentin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
von: Maes, Lucas, et al.
Veröffentlicht: (2026) -
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
von: Balestriero, Randall, et al.
Veröffentlicht: (2025) -
stable-worldmodel-v1: Reproducible World Modeling Research and Evaluation
von: Maes, Lucas, et al.
Veröffentlicht: (2026) -
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures
von: Huang, Hai, et al.
Veröffentlicht: (2025) -
Learning by Reconstruction Produces Uninformative Features For Perception
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)