MapleGrasp: Mask-guided Feature Pooling for Language-driven Efficient Robotic Grasping
Fuente:
arXiv
Saved in:
| Main Authors: | Bhat, Vineet, Patel, Naman, Krishnamurthy, Prashanth, Karri, Ramesh, Khorrami, Farshad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2024)
by: Bhat, Vineet, et al.
Published: (2024)
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
by: Vineet Bhat, et al.
Published: (2025)
by: Vineet Bhat, et al.
Published: (2025)
Grounding LLMs For Robot Task Planning Using Closed-loop State Feedback
by: Bhat, Vineet, et al.
Published: (2024)
by: Bhat, Vineet, et al.
Published: (2024)
3D CAVLA: Leveraging Depth and 3D Context to Generalize Vision Language Action Models for Unseen Tasks
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
by: Patel, Naman, et al.
Published: (2025)
by: Patel, Naman, et al.
Published: (2025)
SALSA: Swift Adaptive Lightweight Self-Attention for Enhanced LiDAR Place Recognition
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
BOP-ASK: Object-Interaction Reasoning for Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
Enabling Deep Visibility into VxWorks-Based Embedded Controllers in Cyber-Physical Systems for Anomaly Detection
by: Krishnamurthy, Prashanth, et al.
Published: (2025)
by: Krishnamurthy, Prashanth, et al.
Published: (2025)
SCAMPER -- Synchrophasor Covert chAnnel for Malicious and Protective ERrands
by: Krishnamurthy, Prashanth, et al.
Published: (2025)
by: Krishnamurthy, Prashanth, et al.
Published: (2025)
Real-Time Multi-Modal Subcomponent-Level Measurements for Trustworthy System Monitoring and Malware Detection
by: Khorrami, Farshad, et al.
Published: (2025)
by: Khorrami, Farshad, et al.
Published: (2025)
Efficient and Distributed Large-Scale 3D Map Registration using Tomographic Features
by: Unlu, Halil Utku, et al.
Published: (2024)
by: Unlu, Halil Utku, et al.
Published: (2024)
Proactive Hierarchical Control Barrier Function-Based Safety Prioritization in Close Human-Robot Interaction Scenarios
by: Maithani, Patanjali, et al.
Published: (2025)
by: Maithani, Patanjali, et al.
Published: (2025)
Open-Architecture End-to-End System for Real-World Autonomous Robot Navigation
by: Devarakonda, Venkata Naren, et al.
Published: (2024)
by: Devarakonda, Venkata Naren, et al.
Published: (2024)
Safe Multi-Robotic Arm Interaction via 3D Convex Shapes
by: Kaypak, Ali Umut, et al.
Published: (2025)
by: Kaypak, Ali Umut, et al.
Published: (2025)
Language-driven Grasp Detection with Mask-guided Attention
by: Van Vo, Tuan, et al.
Published: (2024)
by: Van Vo, Tuan, et al.
Published: (2024)
Sailing Through Point Clouds: Safe Navigation Using Point Cloud Based Control Barrier Functions
by: Dai, Bolun, et al.
Published: (2024)
by: Dai, Bolun, et al.
Published: (2024)
Differentiable Optimization Based Time-Varying Control Barrier Functions for Dynamic Obstacle Avoidance
by: Dai, Bolun, et al.
Published: (2023)
by: Dai, Bolun, et al.
Published: (2023)
RESCORE: LLM-Driven Simulation Recovery in Control Systems Research Papers
by: Bhat, Vineet, et al.
Published: (2026)
by: Bhat, Vineet, et al.
Published: (2026)
Tracking Real-time Anomalies in Cyber-Physical Systems Through Dynamic Behavioral Analysis
by: Krishnamurthy, Prashanth, et al.
Published: (2024)
by: Krishnamurthy, Prashanth, et al.
Published: (2024)
FlashMix: Fast Map-Free LiDAR Localization via Feature Mixing and Contrastive-Constrained Accelerated Training
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
Distributed Inverse Dynamics Control for Quadruped Robots using Geometric Optimization
by: Khandelwal, Nimesh, et al.
Published: (2024)
by: Khandelwal, Nimesh, et al.
Published: (2024)
RoboPEPP: Vision-Based Robot Pose and Joint Angle Estimation through Embedding Predictive Pre-Training
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
CLIPScope: Enhancing Zero-Shot OOD Detection with Bayesian Scoring
by: Fu, Hao, et al.
Published: (2024)
by: Fu, Hao, et al.
Published: (2024)
Compliant Control of Quadruped Robots for Assistive Load Carrying
by: Khandelwal, Nimesh, et al.
Published: (2025)
by: Khandelwal, Nimesh, et al.
Published: (2025)
OSVI-WM: One-Shot Visual Imitation for Unseen Tasks using World-Model-Guided Trajectory Generation
by: Goswami, Raktim Gautam, et al.
Published: (2025)
by: Goswami, Raktim Gautam, et al.
Published: (2025)
Grasp as You Say: Language-guided Dexterous Grasp Generation
by: Wei, Yi-Lin, et al.
Published: (2024)
by: Wei, Yi-Lin, et al.
Published: (2024)
A Differentiable Distance Metric for Robotics Through Generalized Alternating Projection
by: Gonçalves, Vinicius M., et al.
Published: (2025)
by: Gonçalves, Vinicius M., et al.
Published: (2025)
SENTAUR: Security EnhaNced Trojan Assessment Using LLMs Against Undesirable Revisions
by: Bhandari, Jitendra, et al.
Published: (2024)
by: Bhandari, Jitendra, et al.
Published: (2024)
REMaQE: Reverse Engineering Math Equations from Executables
by: Udeshi, Meet, et al.
Published: (2023)
by: Udeshi, Meet, et al.
Published: (2023)
SHIELD: A Host-Independent Framework for Ransomware Detection using Deep Filesystem Features
by: Raz, Md, et al.
Published: (2025)
by: Raz, Md, et al.
Published: (2025)
A Parameter-Efficient Tuning Framework for Language-guided Object Grounding and Robot Grasping
by: Yu, Houjian, et al.
Published: (2024)
by: Yu, Houjian, et al.
Published: (2024)
SECOND-Grasp: Semantic Contact-guided Dexterous Grasping
by: Shin, Han Yi, et al.
Published: (2026)
by: Shin, Han Yi, et al.
Published: (2026)
VCoT-Grasp: Grasp Foundation Models with Visual Chain-of-Thought Reasoning for Language-driven Grasp Generation
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance
by: Wei, Yi-Lin, et al.
Published: (2025)
by: Wei, Yi-Lin, et al.
Published: (2025)
GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System
by: Nguyen, Quang, et al.
Published: (2025)
by: Nguyen, Quang, et al.
Published: (2025)
GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning
by: Nguyen, Huy Hoang, et al.
Published: (2024)
by: Nguyen, Huy Hoang, et al.
Published: (2024)
MultiTalk: Introspective and Extrospective Dialogue for Human-Environment-LLM Alignment
by: Devarakonda, Venkata Naren, et al.
Published: (2024)
by: Devarakonda, Venkata Naren, et al.
Published: (2024)
Data-Efficient System Identification via Lipschitz Neural Networks
by: Wei, Shiqing, et al.
Published: (2024)
by: Wei, Shiqing, et al.
Published: (2024)
LangGrasp: Leveraging Fine-Tuned LLMs for Language Interactive Robot Grasping with Ambiguous Instructions
by: Lin, Yunhan, et al.
Published: (2025)
by: Lin, Yunhan, et al.
Published: (2025)
RT-Grasp: Reasoning Tuning Robotic Grasping via Multi-modal Large Language Model
by: Xu, Jinxuan, et al.
Published: (2024)
by: Xu, Jinxuan, et al.
Published: (2024)
Similar Items
-
HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2024) -
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
by: Vineet Bhat, et al.
Published: (2025) -
Grounding LLMs For Robot Task Planning Using Closed-loop State Feedback
by: Bhat, Vineet, et al.
Published: (2024) -
3D CAVLA: Leveraging Depth and 3D Context to Generalize Vision Language Action Models for Unseen Tasks
by: Bhat, Vineet, et al.
Published: (2025) -
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
by: Patel, Naman, et al.
Published: (2025)