GraspMolmo: Generalizable Task-Oriented Grasping via Large-Scale Synthetic Data Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Deshpande, Abhay, Deng, Yuquan, Ray, Arijit, Salvador, Jordi, Han, Winson, Duan, Jiafei, Zeng, Kuo-Hao, Zhu, Yuke, Krishna, Ranjay, Hendrix, Rose |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MolmoAct: Action Reasoning Models that can Reason in Space
by: Lee, Jason, et al.
Published: (2025)
by: Lee, Jason, et al.
Published: (2025)
MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation
by: Deshpande, Abhay, et al.
Published: (2026)
by: Deshpande, Abhay, et al.
Published: (2026)
FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation Models
by: Tang, Chao, et al.
Published: (2024)
by: Tang, Chao, et al.
Published: (2024)
MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
by: Kim, Yejin, et al.
Published: (2026)
by: Kim, Yejin, et al.
Published: (2026)
The One RING: a Robotic Indoor Navigation Generalist
by: Eftekhar, Ainaz, et al.
Published: (2024)
by: Eftekhar, Ainaz, et al.
Published: (2024)
OVAL-Grasp: Open-Vocabulary Affordance Localization for Task Oriented Grasping
by: Tong, Edmond, et al.
Published: (2025)
by: Tong, Edmond, et al.
Published: (2025)
SAT: Dynamic Spatial Aptitude Training for Multimodal Language Models
by: Ray, Arijit, et al.
Published: (2024)
by: Ray, Arijit, et al.
Published: (2024)
MolmoPoint: Better Pointing for VLMs with Grounding Tokens
by: Clark, Christopher, et al.
Published: (2026)
by: Clark, Christopher, et al.
Published: (2026)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
MolmoAct2: Action Reasoning Models for Real-world Deployment
by: Fang, Haoquan, et al.
Published: (2026)
by: Fang, Haoquan, et al.
Published: (2026)
SegGrasp: Zero-Shot Task-Oriented Grasping via Semantic and Geometric Guided Segmentation
by: Li, Haosheng, et al.
Published: (2024)
by: Li, Haosheng, et al.
Published: (2024)
Selective Visual Representations Improve Convergence and Generalization for Embodied AI
by: Eftekhar, Ainaz, et al.
Published: (2023)
by: Eftekhar, Ainaz, et al.
Published: (2023)
AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter
by: Tang, Yingbo, et al.
Published: (2025)
by: Tang, Yingbo, et al.
Published: (2025)
HGDiffuser: Efficient Task-Oriented Grasp Generation via Human-Guided Grasp Diffusion Models
by: Huang, Dehao, et al.
Published: (2025)
by: Huang, Dehao, et al.
Published: (2025)
GRIM: Task-Oriented Grasping with Conditioning on Generative Examples
by: Shailesh, et al.
Published: (2025)
by: Shailesh, et al.
Published: (2025)
GraspLDP: Towards Generalizable Grasping Policy via Latent Diffusion
by: Xiang, Enda, et al.
Published: (2026)
by: Xiang, Enda, et al.
Published: (2026)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
by: Li, Samuel, et al.
Published: (2024)
by: Li, Samuel, et al.
Published: (2024)
MolmoWeb: Open Visual Web Agent and Open Data for the Open Web
by: Gupta, Tanmay, et al.
Published: (2026)
by: Gupta, Tanmay, et al.
Published: (2026)
AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance
by: Wei, Yi-Lin, et al.
Published: (2025)
by: Wei, Yi-Lin, et al.
Published: (2025)
UltraDexGrasp: Learning Universal Dexterous Grasping for Bimanual Robots with Synthetic Data
by: Yang, Sizhe, et al.
Published: (2026)
by: Yang, Sizhe, et al.
Published: (2026)
SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
by: Ehsani, Kiana, et al.
Published: (2023)
by: Ehsani, Kiana, et al.
Published: (2023)
Grasp Like Humans: Learning Generalizable Multi-Fingered Grasping from Human Proprioceptive Sensorimotor Integration
by: Guo, Ce, et al.
Published: (2025)
by: Guo, Ce, et al.
Published: (2025)
Visual Imitation Learning of Task-Oriented Object Grasping and Rearrangement
by: Cai, Yichen, et al.
Published: (2024)
by: Cai, Yichen, et al.
Published: (2024)
DexTOG: Learning Task-Oriented Dexterous Grasp with Language
by: Zhang, Jieyi, et al.
Published: (2025)
by: Zhang, Jieyi, et al.
Published: (2025)
Volumetric Reconstruction From Partial Views for Task-Oriented Grasping
by: Yan, Fujian, et al.
Published: (2025)
by: Yan, Fujian, et al.
Published: (2025)
OmniDexGrasp: Generalizable Dexterous Grasping via Foundation Model and Force Feedback
by: Wei, Yi-Lin, et al.
Published: (2025)
by: Wei, Yi-Lin, et al.
Published: (2025)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
by: Liu, An-Lun, et al.
Published: (2025)
by: Liu, An-Lun, et al.
Published: (2025)
Task-Oriented 6-DoF Grasp Pose Detection in Clutters
by: Wang, An-Lan, et al.
Published: (2025)
by: Wang, An-Lan, et al.
Published: (2025)
ZeroDexGrasp: Zero-Shot Task-Oriented Dexterous Grasp Synthesis with Prompt-Based Multi-Stage Semantic Reasoning
by: Jian, Juntao, et al.
Published: (2025)
by: Jian, Juntao, et al.
Published: (2025)
MPGNet: Learning Move-Push-Grasping Synergy for Target-Oriented Grasping in Occluded Scenes
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
by: Deng, Shengliang, et al.
Published: (2025)
by: Deng, Shengliang, et al.
Published: (2025)
NeuGrasp: Generalizable Neural Surface Reconstruction with Background Priors for Material-Agnostic Object Grasp Detection
by: Fan, Qingyu, et al.
Published: (2025)
by: Fan, Qingyu, et al.
Published: (2025)
Task-Oriented Grasping Using Reinforcement Learning with a Contextual Reward Machine
by: Li, Hui, et al.
Published: (2025)
by: Li, Hui, et al.
Published: (2025)
DexGANGrasp: Dexterous Generative Adversarial Grasping Synthesis for Task-Oriented Manipulation
by: Feng, Qian, et al.
Published: (2024)
by: Feng, Qian, et al.
Published: (2024)
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding
by: Clark, Christopher, et al.
Published: (2026)
by: Clark, Christopher, et al.
Published: (2026)
FuseGrasp: Radar-Camera Fusion for Robotic Grasping of Transparent Objects
by: Deng, Hongyu, et al.
Published: (2025)
by: Deng, Hongyu, et al.
Published: (2025)
EVE: Enabling Anyone to Train Robots using Augmented Reality
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
VLS: Steering Pretrained Robot Policies via Vision-Language Models
by: Liu, Shuo, et al.
Published: (2026)
by: Liu, Shuo, et al.
Published: (2026)
GAT-Grasp: Gesture-Driven Affordance Transfer for Task-Aware Robotic Grasping
by: Wang, Ruixiang, et al.
Published: (2025)
by: Wang, Ruixiang, et al.
Published: (2025)
DAGDiff: Guiding Dual-Arm Grasp Diffusion to Stable and Collision-Free Grasps
by: Karim, Md Faizal, et al.
Published: (2025)
by: Karim, Md Faizal, et al.
Published: (2025)
Similar Items
-
MolmoAct: Action Reasoning Models that can Reason in Space
by: Lee, Jason, et al.
Published: (2025) -
MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation
by: Deshpande, Abhay, et al.
Published: (2026) -
FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation Models
by: Tang, Chao, et al.
Published: (2024) -
MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
by: Kim, Yejin, et al.
Published: (2026) -
The One RING: a Robotic Indoor Navigation Generalist
by: Eftekhar, Ainaz, et al.
Published: (2024)