Navigation World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bar, Amir, Zhou, Gaoyue, Tran, Danny, Darrell, Trevor, LeCun, Yann |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Whole-Body Conditioned Egocentric Video Prediction
von: Bai, Yutong, et al.
Veröffentlicht: (2025)
von: Bai, Yutong, et al.
Veröffentlicht: (2025)
World Models for Learning Dexterous Hand-Object Interactions from Human Videos
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025)
EgoPet: Egomotion and Interaction Data from an Animal's Perspective
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
Stochastic positional embeddings improve masked image modeling
von: Bar, Amir, et al.
Veröffentlicht: (2023)
von: Bar, Amir, et al.
Veröffentlicht: (2023)
Learning by Reconstruction Produces Uninformative Features For Perception
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
Lifting Embodied World Models for Planning and Control
von: Wang, Alex N., et al.
Veröffentlicht: (2026)
von: Wang, Alex N., et al.
Veröffentlicht: (2026)
Hierarchical World Models as Visual Whole-Body Humanoid Controllers
von: Hansen, Nicklas, et al.
Veröffentlicht: (2024)
von: Hansen, Nicklas, et al.
Veröffentlicht: (2024)
Learning and Leveraging World Models in Visual Representation Learning
von: Garrido, Quentin, et al.
Veröffentlicht: (2024)
von: Garrido, Quentin, et al.
Veröffentlicht: (2024)
From Generated Human Videos to Physically Plausible Robot Trajectories
von: Ni, James, et al.
Veröffentlicht: (2025)
von: Ni, James, et al.
Veröffentlicht: (2025)
Gaussian Embeddings: How JEPAs Secretly Learn Your Data Density
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
von: Balestriero, Randall, et al.
Veröffentlicht: (2025)
Blockwise Self-Supervised Learning at Scale
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
von: Siddiqui, Shoaib Ahmed, et al.
Veröffentlicht: (2023)
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning
von: Zhou, Gaoyue, et al.
Veröffentlicht: (2024)
von: Zhou, Gaoyue, et al.
Veröffentlicht: (2024)
Variance-Covariance Regularization Improves Representation Learning
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2023)
Transformers without Normalization
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
von: Assran, Mido, et al.
Veröffentlicht: (2025)
von: Assran, Mido, et al.
Veröffentlicht: (2025)
Forgotten Polygons: Multimodal Large Language Models are Shape-Blind
von: Rudman, William, et al.
Veröffentlicht: (2025)
von: Rudman, William, et al.
Veröffentlicht: (2025)
Aether: Geometric-Aware Unified World Modeling
von: Aether Team, et al.
Veröffentlicht: (2025)
von: Aether Team, et al.
Veröffentlicht: (2025)
Revisiting Feature Prediction for Learning Visual Representations from Video
von: Bardes, Adrien, et al.
Veröffentlicht: (2024)
von: Bardes, Adrien, et al.
Veröffentlicht: (2024)
A hierarchical loss and its problems when classifying non-hierarchically
von: Wu, Cinna, et al.
Veröffentlicht: (2017)
von: Wu, Cinna, et al.
Veröffentlicht: (2017)
Learning Latent Action World Models In The Wild
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
von: Garrido, Quentin, et al.
Veröffentlicht: (2026)
Vision-Language Models Create Cross-Modal Task Representations
von: Luo, Grace, et al.
Veröffentlicht: (2024)
von: Luo, Grace, et al.
Veröffentlicht: (2024)
Video Representation Learning with Joint-Embedding Predictive Architectures
von: Drozdov, Katrina, et al.
Veröffentlicht: (2024)
von: Drozdov, Katrina, et al.
Veröffentlicht: (2024)
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
Learning to Navigate Socially Through Proactive Risk Perception
von: Xiao, Erjia, et al.
Veröffentlicht: (2025)
von: Xiao, Erjia, et al.
Veröffentlicht: (2025)
What Drives Success in Physical Planning with Joint-Embedding Predictive World Models?
von: Terver, Basile, et al.
Veröffentlicht: (2025)
von: Terver, Basile, et al.
Veröffentlicht: (2025)
AutoWorld: Scaling Multi-Agent Traffic Simulation with Self-Supervised World Models
von: Pourkeshavatz, Mozhgan, et al.
Veröffentlicht: (2026)
von: Pourkeshavatz, Mozhgan, et al.
Veröffentlicht: (2026)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
von: Lancaster, Patrick, et al.
Veröffentlicht: (2023)
von: Lancaster, Patrick, et al.
Veröffentlicht: (2023)
Enhancing 3D Point Cloud Classification with ModelNet-R and Point-SkipNet
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
Cosmos World Foundation Model Platform for Physical AI
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
World Simulation with Video Foundation Models for Physical AI
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
Real-World Robot Applications of Foundation Models: A Review
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
Learning Humanoid Navigation from Human Data
von: Wang, Weizhuo, et al.
Veröffentlicht: (2026)
von: Wang, Weizhuo, et al.
Veröffentlicht: (2026)
Point-GN: A Non-Parametric Network Using Gaussian Positional Encoding for Point Cloud Classification
von: Mohammadi, Marzieh, et al.
Veröffentlicht: (2024)
von: Mohammadi, Marzieh, et al.
Veröffentlicht: (2024)
Robustness Is a Function, Not a Number: A Factorized Comprehensive Study of OOD Robustness in Vision-Based Driving
von: Mallak, Amir, et al.
Veröffentlicht: (2026)
von: Mallak, Amir, et al.
Veröffentlicht: (2026)
ExoPredicator: Learning Abstract Models of Dynamic Worlds for Robot Planning
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
TD-MPC2: Scalable, Robust World Models for Continuous Control
von: Hansen, Nicklas, et al.
Veröffentlicht: (2023)
von: Hansen, Nicklas, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Whole-Body Conditioned Egocentric Video Prediction
von: Bai, Yutong, et al.
Veröffentlicht: (2025) -
World Models for Learning Dexterous Hand-Object Interactions from Human Videos
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2025) -
EgoPet: Egomotion and Interaction Data from an Animal's Perspective
von: Bar, Amir, et al.
Veröffentlicht: (2024) -
Stochastic positional embeddings improve masked image modeling
von: Bar, Amir, et al.
Veröffentlicht: (2023) -
Learning by Reconstruction Produces Uninformative Features For Perception
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)