Language-driven Grasp Detection with Mask-guided Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Van Vo, Tuan, Vu, Minh Nhat, Huang, Baoru, Vuong, An, Le, Ngan, Vo, Thieu, Nguyen, Anh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lightweight Language-driven Grasp Detection using Conditional Consistency Model
by: Nguyen, Nghia, et al.
Published: (2024)
by: Nguyen, Nghia, et al.
Published: (2024)
Language-Driven 6-DoF Grasp Detection Using Negative Prompt Guidance
by: Nguyen, Toan, et al.
Published: (2024)
by: Nguyen, Toan, et al.
Published: (2024)
Language-driven Grasp Detection
by: Vuong, An Dinh, et al.
Published: (2024)
by: Vuong, An Dinh, et al.
Published: (2024)
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
by: Nguyen, Nghia, et al.
Published: (2024)
by: Nguyen, Nghia, et al.
Published: (2024)
GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning
by: Nguyen, Huy Hoang, et al.
Published: (2024)
by: Nguyen, Huy Hoang, et al.
Published: (2024)
EgoMusic-driven Human Dance Motion Estimation with Skeleton Mamba
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
Learning Human Motion with Temporally Conditional Mamba
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
Autonomous Catheterization with Open-source Simulator and Expert Trajectory
by: Jianu, Tudor, et al.
Published: (2024)
by: Jianu, Tudor, et al.
Published: (2024)
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning
by: Van Vo, Tuan, et al.
Published: (2026)
by: Van Vo, Tuan, et al.
Published: (2026)
HabiCrowd: A High Performance Simulator for Crowd-Aware Visual Navigation
by: Vuong, An Dinh, et al.
Published: (2023)
by: Vuong, An Dinh, et al.
Published: (2023)
Rethinking Progression of Memory State in Robotic Manipulation: An Object-Centric Perspective
by: Chung, Nhat, et al.
Published: (2025)
by: Chung, Nhat, et al.
Published: (2025)
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
by: Hanyu, Taisei, et al.
Published: (2025)
by: Hanyu, Taisei, et al.
Published: (2025)
More Reliable Pseudo-labels, Better Performance: A Generalized Approach to Single Positive Multi-label Learning
by: Tran, Luong, et al.
Published: (2025)
by: Tran, Luong, et al.
Published: (2025)
Amodal Instance Segmentation with Diffusion Shape Prior Estimation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
Your Vision-Language-Action Model Already Has Attention Heads For Path Deviation Detection
by: Jeong, Jaehwan, et al.
Published: (2026)
by: Jeong, Jaehwan, et al.
Published: (2026)
Lang2Lift: A Language-Guided Autonomous Forklift System for Outdoor Industrial Pallet Handling
by: Nguyen, Huy Hoang, et al.
Published: (2025)
by: Nguyen, Huy Hoang, et al.
Published: (2025)
SemLT3D: Semantic-Guided Expert Distillation for Camera-only Long-Tailed 3D Object Detection
by: Vo, Hao, et al.
Published: (2026)
by: Vo, Hao, et al.
Published: (2026)
Planning Robot Placement for Object Grasping
by: Saini, Manish, et al.
Published: (2024)
by: Saini, Manish, et al.
Published: (2024)
A Distributed Multi-Modal Sensing Approach for Human Activity Recognition in Real-Time Human-Robot Collaboration
by: Belcamino, Valerio, et al.
Published: (2026)
by: Belcamino, Valerio, et al.
Published: (2026)
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model
by: Vo, Khoa, et al.
Published: (2024)
by: Vo, Khoa, et al.
Published: (2024)
Shape2Animal: Creative Animal Generation from Natural Silhouettes
by: Tran, Quoc-Duy, et al.
Published: (2025)
by: Tran, Quoc-Duy, et al.
Published: (2025)
Counting to Four is still a Chore for VLMs
by: Anh, Duy Le Dinh, et al.
Published: (2026)
by: Anh, Duy Le Dinh, et al.
Published: (2026)
SIGMA: A Physics-Based Benchmark for Gas Chimney Understanding in Seismic Images
by: Truong, Bao, et al.
Published: (2026)
by: Truong, Bao, et al.
Published: (2026)
DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning
by: Lee, Junha, et al.
Published: (2026)
by: Lee, Junha, et al.
Published: (2026)
Graspness Discovery in Clutters for Fast and Accurate Grasp Detection
by: Wang, Chenxi, et al.
Published: (2024)
by: Wang, Chenxi, et al.
Published: (2024)
A comparison of extended object tracking with multi-modal sensors in indoor environment
by: Shuai, Jiangtao, et al.
Published: (2024)
by: Shuai, Jiangtao, et al.
Published: (2024)
Online Trajectory Replanner for Dynamically Grasping Irregular Objects
by: Vu, Minh Nhat, et al.
Published: (2025)
by: Vu, Minh Nhat, et al.
Published: (2025)
DenseMTL: Cross-task Attention Mechanism for Dense Multi-task Learning
by: Lopes, Ivan, et al.
Published: (2022)
by: Lopes, Ivan, et al.
Published: (2022)
ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning
by: Van Vo, Tuan, et al.
Published: (2025)
by: Van Vo, Tuan, et al.
Published: (2025)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
by: Jiang, Zebin, et al.
Published: (2025)
by: Jiang, Zebin, et al.
Published: (2025)
PANDORA: Pixel-wise Attention Dissolution and Latent Guidance for Zero-Shot Object Removal
by: Vo, Dinh-Khoi, et al.
Published: (2026)
by: Vo, Dinh-Khoi, et al.
Published: (2026)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
by: Nguyen, Minh Anh, et al.
Published: (2026)
by: Nguyen, Minh Anh, et al.
Published: (2026)
SpikeGrasp: A Benchmark for 6-DoF Grasp Pose Detection from Stereo Spike Streams
by: Gao, Zhuoheng, et al.
Published: (2025)
by: Gao, Zhuoheng, et al.
Published: (2025)
TARGO: Benchmarking Target-driven Object Grasping under Occlusions
by: Xia, Yan, et al.
Published: (2024)
by: Xia, Yan, et al.
Published: (2024)
FunGrasp: Functional Grasping for Diverse Dexterous Hands
by: Huang, Linyi, et al.
Published: (2024)
by: Huang, Linyi, et al.
Published: (2024)
Sim-to-Real Grasp Detection with Global-to-Local RGB-D Adaptation
by: Ma, Haoxiang, et al.
Published: (2024)
by: Ma, Haoxiang, et al.
Published: (2024)
Spatial RoboGrasp: Generalized Robotic Grasping Control Policy
by: Huang, Yiqi, et al.
Published: (2025)
by: Huang, Yiqi, et al.
Published: (2025)
GraspLDP: Towards Generalizable Grasping Policy via Latent Diffusion
by: Xiang, Enda, et al.
Published: (2026)
by: Xiang, Enda, et al.
Published: (2026)
Similar Items
-
Lightweight Language-driven Grasp Detection using Conditional Consistency Model
by: Nguyen, Nghia, et al.
Published: (2024) -
Language-Driven 6-DoF Grasp Detection Using Negative Prompt Guidance
by: Nguyen, Toan, et al.
Published: (2024) -
Language-driven Grasp Detection
by: Vuong, An Dinh, et al.
Published: (2024) -
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
by: Nguyen, Nghia, et al.
Published: (2024) -
GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning
by: Nguyen, Huy Hoang, et al.
Published: (2024)