MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Yejin, Pumacay, Wilbert, Rayyan, Omar, Argus, Max, Han, Winson, VanderBilt, Eli, Salvador, Jordi, Deshpande, Abhay, Hendrix, Rose, Jauhri, Snehal, Liu, Shuo, Shafiullah, Nur Muhammad Mahi, Guru, Maya, Eftekhar, Ainaz, Farley, Karen, Clay, Donovan, Duan, Jiafei, Guru, Arjun, Wolters, Piper, Herrasti, Alvaro, Lee, Ying-Chun, Chalvatzaki, Georgia, Cui, Yuchen, Farhadi, Ali, Fox, Dieter, Krishna, Ranjay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation
von: Deshpande, Abhay, et al.
Veröffentlicht: (2026)
von: Deshpande, Abhay, et al.
Veröffentlicht: (2026)
MolmoAct: Action Reasoning Models that can Reason in Space
von: Lee, Jason, et al.
Veröffentlicht: (2025)
von: Lee, Jason, et al.
Veröffentlicht: (2025)
MolmoAct2: Action Reasoning Models for Real-world Deployment
von: Fang, Haoquan, et al.
Veröffentlicht: (2026)
von: Fang, Haoquan, et al.
Veröffentlicht: (2026)
The One RING: a Robotic Indoor Navigation Generalist
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2024)
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2024)
Selective Visual Representations Improve Convergence and Generalization for Embodied AI
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2023)
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2023)
Learning Any-View 6DoF Robotic Grasping in Cluttered Scenes via Neural Surface Rendering
von: Jauhri, Snehal, et al.
Veröffentlicht: (2023)
von: Jauhri, Snehal, et al.
Veröffentlicht: (2023)
Whole-Body Mobile Manipulation using Offline Reinforcement Learning on Sub-optimal Controllers
von: Jauhri, Snehal, et al.
Veröffentlicht: (2026)
von: Jauhri, Snehal, et al.
Veröffentlicht: (2026)
Active-Perceptive Motion Generation for Mobile Manipulation
von: Jauhri, Snehal, et al.
Veröffentlicht: (2023)
von: Jauhri, Snehal, et al.
Veröffentlicht: (2023)
GraspMolmo: Generalizable Task-Oriented Grasping via Large-Scale Synthetic Data Generation
von: Deshpande, Abhay, et al.
Veröffentlicht: (2025)
von: Deshpande, Abhay, et al.
Veröffentlicht: (2025)
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
von: Pumacay, Wilbert, et al.
Veröffentlicht: (2024)
von: Pumacay, Wilbert, et al.
Veröffentlicht: (2024)
SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
von: Ehsani, Kiana, et al.
Veröffentlicht: (2023)
von: Ehsani, Kiana, et al.
Veröffentlicht: (2023)
UniFField: A Generalizable Unified Neural Feature Field for Visual, Semantic, and Spatial Uncertainties in Any Scene
von: Maurer, Christian, et al.
Veröffentlicht: (2025)
von: Maurer, Christian, et al.
Veröffentlicht: (2025)
2HandedAfforder: Learning Precise Actionable Bimanual Affordances from Human Videos
von: Heidinger, Marvin, et al.
Veröffentlicht: (2025)
von: Heidinger, Marvin, et al.
Veröffentlicht: (2025)
Convergent Functions, Divergent Forms
von: Jeon, Hyeonseong, et al.
Veröffentlicht: (2025)
von: Jeon, Hyeonseong, et al.
Veröffentlicht: (2025)
MolmoWeb: Open Visual Web Agent and Open Data for the Open Web
von: Gupta, Tanmay, et al.
Veröffentlicht: (2026)
von: Gupta, Tanmay, et al.
Veröffentlicht: (2026)
6DOPE-GS: Online 6D Object Pose Estimation using Gaussian Splatting
von: Jin, Yufeng, et al.
Veröffentlicht: (2024)
von: Jin, Yufeng, et al.
Veröffentlicht: (2024)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
SAM2Act: Integrating Visual Foundation Model with A Memory Architecture for Robotic Manipulation
von: Fang, Haoquan, et al.
Veröffentlicht: (2025)
von: Fang, Haoquan, et al.
Veröffentlicht: (2025)
Holodeck: Language Guided Generation of 3D Embodied AI Environments
von: Yang, Yue, et al.
Veröffentlicht: (2023)
von: Yang, Yue, et al.
Veröffentlicht: (2023)
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
von: Yuan, Wentao, et al.
Veröffentlicht: (2024)
von: Yuan, Wentao, et al.
Veröffentlicht: (2024)
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
von: Clark, Christopher, et al.
Veröffentlicht: (2026)
von: Clark, Christopher, et al.
Veröffentlicht: (2026)
Arguing The Unarguable
von: Guru Dev Teeluckdharry
Veröffentlicht: (2026)
von: Guru Dev Teeluckdharry
Veröffentlicht: (2026)
Le Judiciaire et la Franc-Maçonnerie
von: Teeluckdharry, Guru Dev
Veröffentlicht: (2025)
von: Teeluckdharry, Guru Dev
Veröffentlicht: (2025)
Rook decomposition of the Partition function
von: Sharan, N. Guru
Veröffentlicht: (2025)
von: Sharan, N. Guru
Veröffentlicht: (2025)
Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents
von: Singhi, Nishad, et al.
Veröffentlicht: (2026)
von: Singhi, Nishad, et al.
Veröffentlicht: (2026)
Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
von: Deitke, Matt, et al.
Veröffentlicht: (2024)
von: Deitke, Matt, et al.
Veröffentlicht: (2024)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
von: Duan, Jiafei, et al.
Veröffentlicht: (2024)
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding
von: Clark, Christopher, et al.
Veröffentlicht: (2026)
von: Clark, Christopher, et al.
Veröffentlicht: (2026)
Optical Control of Ferroaxial Order
von: He, Zhiren, et al.
Veröffentlicht: (2023)
von: He, Zhiren, et al.
Veröffentlicht: (2023)
Deep Generative Models in Robotics: A Survey on Learning from Multimodal Demonstrations
von: Urain, Julen, et al.
Veröffentlicht: (2024)
von: Urain, Julen, et al.
Veröffentlicht: (2024)
PointArena: Probing Multimodal Grounding Through Language-Guided Pointing
von: Cheng, Long, et al.
Veröffentlicht: (2025)
von: Cheng, Long, et al.
Veröffentlicht: (2025)
The Mordell-Tornheim zeta function: Kronecker limit type formula, Series Evaluations and Applications
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
Mordell--Tornheim zeta function: Kronecker limit type formulas and Special values
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
Partitions with Durfee triangles of fixed size
von: Sharan, N. Guru, et al.
Veröffentlicht: (2025)
von: Sharan, N. Guru, et al.
Veröffentlicht: (2025)
Advancement of constant and progressive load multi‐cycle indentation method on surface properties characterization of polymers
von: Soumya Ranjan Guru, et al.
Veröffentlicht: (2024)
von: Soumya Ranjan Guru, et al.
Veröffentlicht: (2024)
From Stimulus to Response: Understanding the Causes and Outcomes of Consumer Activism
von: Sajith Narayanan, et al.
Veröffentlicht: (2024)
von: Sajith Narayanan, et al.
Veröffentlicht: (2024)
Equivalence criteria for the two--term functional equations for Herglotz--Zagier functions
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
von: Sathyanarayana, Sumukha, et al.
Veröffentlicht: (2025)
Phonon-Mediated Third-Harmonic Generation in Diamond
von: Zheng, Jiaoyang, et al.
Veröffentlicht: (2023)
von: Zheng, Jiaoyang, et al.
Veröffentlicht: (2023)
Dynamic magnetic response in ABA type trilayered systems and compensation phenomenon
von: Guru, Enakshi, et al.
Veröffentlicht: (2024)
von: Guru, Enakshi, et al.
Veröffentlicht: (2024)
Iterations of Meromorphic Functions involving Sine
von: Kumar, Gaurav, et al.
Veröffentlicht: (2025)
von: Kumar, Gaurav, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation
von: Deshpande, Abhay, et al.
Veröffentlicht: (2026) -
MolmoAct: Action Reasoning Models that can Reason in Space
von: Lee, Jason, et al.
Veröffentlicht: (2025) -
MolmoAct2: Action Reasoning Models for Real-world Deployment
von: Fang, Haoquan, et al.
Veröffentlicht: (2026) -
The One RING: a Robotic Indoor Navigation Generalist
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2024) -
Selective Visual Representations Improve Convergence and Generalization for Embodied AI
von: Eftekhar, Ainaz, et al.
Veröffentlicht: (2023)