VLAD-Grasp: Zero-shot Grasp Detection via Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kulshrestha, Manav, Bukhari, S. Talha, Conover, Damon, Bera, Aniket |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Move or Not to Move: Constraint-based Planning Enables Zero-Shot Generalization for Interactive Navigation
by: Vashisth, Apoorva, et al.
Published: (2026)
by: Vashisth, Apoorva, et al.
Published: (2026)
Scalable Multi-Robot Informative Path Planning for Target Mapping via Deep Reinforcement Learning
by: Vashisth, Apoorva, et al.
Published: (2024)
by: Vashisth, Apoorva, et al.
Published: (2024)
Dynamic Obstacle Avoidance through Uncertainty-Based Adaptive Planning with Diffusion
by: Punyamoorty, Vineet, et al.
Published: (2024)
by: Punyamoorty, Vineet, et al.
Published: (2024)
Mollified Value Learning
by: Viswanath, Hrishikesh, et al.
Published: (2026)
by: Viswanath, Hrishikesh, et al.
Published: (2026)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)
by: Daiya, Divyanshu, et al.
Published: (2024)
Grasp-HGN: Grasping the Unexpected
by: Zandigohar, Mehrshad, et al.
Published: (2025)
by: Zandigohar, Mehrshad, et al.
Published: (2025)
Variational Shape Inference for Grasp Diffusion on SE(3)
by: Bukhari, S. Talha, et al.
Published: (2025)
by: Bukhari, S. Talha, et al.
Published: (2025)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
by: Yamada, Jun, et al.
Published: (2025)
by: Yamada, Jun, et al.
Published: (2025)
Differentiable Inverse Graphics for Zero-shot Scene Reconstruction and Robot Grasping
by: Arriaga, Octavio, et al.
Published: (2026)
by: Arriaga, Octavio, et al.
Published: (2026)
HessianForge: Scalable LiDAR reconstruction with Physics-Informed Neural Representation and Smoothness Energy Constraints
by: Viswanath, Hrishikesh, et al.
Published: (2025)
by: Viswanath, Hrishikesh, et al.
Published: (2025)
Show and Grasp: Few-shot Semantic Segmentation for Robot Grasping through Zero-shot Foundation Models
by: Barcellona, Leonardo, et al.
Published: (2024)
by: Barcellona, Leonardo, et al.
Published: (2024)
Grasp Anything: Combining Teacher-Augmented Policy Gradient Learning with Instance Segmentation to Grasp Arbitrary Objects
by: Mosbach, Malte, et al.
Published: (2024)
by: Mosbach, Malte, et al.
Published: (2024)
Go-SLAM: Grounded Object Segmentation and Localization with Gaussian Splatting SLAM
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Supervised Mixture-of-Experts for Surgical Grasping and Retraction
by: Mazza, Lorenzo, et al.
Published: (2026)
by: Mazza, Lorenzo, et al.
Published: (2026)
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
by: Lee, Sungjae, et al.
Published: (2025)
by: Lee, Sungjae, et al.
Published: (2025)
Learning Adaptive Dexterous Grasping from Single Demonstrations
by: Shi, Liangzhi, et al.
Published: (2025)
by: Shi, Liangzhi, et al.
Published: (2025)
Label-Efficient Grasp Joint Prediction with Point-JEPA
by: Guzelkabaagac, Jed, et al.
Published: (2025)
by: Guzelkabaagac, Jed, et al.
Published: (2025)
FFHFlow: Diverse and Uncertainty-Aware Dexterous Grasp Generation via Flow Variational Inference
by: Feng, Qian, et al.
Published: (2024)
by: Feng, Qian, et al.
Published: (2024)
Optimizing Crowd-Aware Multi-Agent Path Finding through Local Communication with Graph Neural Networks
by: Pham, Phu, et al.
Published: (2023)
by: Pham, Phu, et al.
Published: (2023)
Offline-to-online Reinforcement Learning for Image-based Grasping with Scarce Demonstrations
by: Chan, Bryan, et al.
Published: (2024)
by: Chan, Bryan, et al.
Published: (2024)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
by: Li, Samuel, et al.
Published: (2024)
by: Li, Samuel, et al.
Published: (2024)
QuadWBG: Generalizable Quadrupedal Whole-Body Grasping
by: Wang, Jilong, et al.
Published: (2024)
by: Wang, Jilong, et al.
Published: (2024)
Combining Shape Completion and Grasp Prediction for Fast and Versatile Grasping with a Multi-Fingered Hand
by: Humt, Matthias, et al.
Published: (2023)
by: Humt, Matthias, et al.
Published: (2023)
Reinforcement Learning-Based Bionic Reflex Control for Anthropomorphic Robotic Grasping exploiting Domain Randomization
by: Basumatary, Hirakjyoti, et al.
Published: (2023)
by: Basumatary, Hirakjyoti, et al.
Published: (2023)
VCoT-Grasp: Grasp Foundation Models with Visual Chain-of-Thought Reasoning for Language-driven Grasp Generation
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping
by: Zhong, Yifan, et al.
Published: (2025)
by: Zhong, Yifan, et al.
Published: (2025)
DexGrasp-Zero: A Morphology-Aligned Policy for Zero-Shot Cross-Embodiment Dexterous Grasping
by: Wu, Yuliang, et al.
Published: (2026)
by: Wu, Yuliang, et al.
Published: (2026)
HAVEN: Hierarchical Adversary-aware Visibility-Enabled Navigation with Cover Utilization using Deep Transformer Q-Networks
by: Chauhan, Mihir, et al.
Published: (2025)
by: Chauhan, Mihir, et al.
Published: (2025)
Corner-Grasp: Multi-Action Grasp Detection and Active Gripper Adaptation for Grasping in Cluttered Environments
by: Son, Yeong Gwang, et al.
Published: (2025)
by: Son, Yeong Gwang, et al.
Published: (2025)
Grasp2Grasp: Vision-Based Dexterous Grasp Translation via Schrödinger Bridges
by: Zhong, Tao, et al.
Published: (2025)
by: Zhong, Tao, et al.
Published: (2025)
GraspGF: Learning Score-based Grasping Primitive for Human-assisting Dexterous Grasping
by: Wu, Tianhao, et al.
Published: (2023)
by: Wu, Tianhao, et al.
Published: (2023)
IFG: Internet-Scale Guidance for Functional Grasping Generation
by: Liu, Ray Muxin, et al.
Published: (2025)
by: Liu, Ray Muxin, et al.
Published: (2025)
OmniDexVLG: Learning Dexterous Grasp Generation from Vision Language Model-Guided Grasp Semantics, Taxonomy and Functional Affordance
by: Zhang, Lei, et al.
Published: (2025)
by: Zhang, Lei, et al.
Published: (2025)
SECOND-Grasp: Semantic Contact-guided Dexterous Grasping
by: Shin, Han Yi, et al.
Published: (2026)
by: Shin, Han Yi, et al.
Published: (2026)
GraspFactory: A Large Object-Centric Grasping Dataset
by: Srinivas, Srinidhi Kalgundi, et al.
Published: (2025)
by: Srinivas, Srinidhi Kalgundi, et al.
Published: (2025)
QuickGrasp: Lightweight Antipodal Grasp Planning with Point Clouds
by: Ravie, Navin Sriram, et al.
Published: (2025)
by: Ravie, Navin Sriram, et al.
Published: (2025)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
IMPACT: Intelligent Motion Planning with Acceptable Contact Trajectories via Vision-Language Models
by: Ling, Yiyang, et al.
Published: (2025)
by: Ling, Yiyang, et al.
Published: (2025)
VLMgineer: Vision Language Models as Robotic Toolsmiths
by: Gao, George Jiayuan, et al.
Published: (2025)
by: Gao, George Jiayuan, et al.
Published: (2025)
Vision Language Models are In-Context Value Learners
by: Ma, Yecheng Jason, et al.
Published: (2024)
by: Ma, Yecheng Jason, et al.
Published: (2024)
Similar Items
-
To Move or Not to Move: Constraint-based Planning Enables Zero-Shot Generalization for Interactive Navigation
by: Vashisth, Apoorva, et al.
Published: (2026) -
Scalable Multi-Robot Informative Path Planning for Target Mapping via Deep Reinforcement Learning
by: Vashisth, Apoorva, et al.
Published: (2024) -
Dynamic Obstacle Avoidance through Uncertainty-Based Adaptive Planning with Diffusion
by: Punyamoorty, Vineet, et al.
Published: (2024) -
Mollified Value Learning
by: Viswanath, Hrishikesh, et al.
Published: (2026) -
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)