CollaBot: Vision-Language Guided Simultaneous Collaborative Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Kun, Chen, Gaoming, Ma, Shentao, Jin, Ninglong, Zhao, Guangbao, Ding, Mingyu, Xiong, Zhenhua, Pan, Jia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
P2 Explore: Efficient Exploration in Unknown Cluttered Environment with Floor Plan Prediction
by: Song, Kun, et al.
Published: (2024)
by: Song, Kun, et al.
Published: (2024)
Distributed Motion Control of Multiple Mobile Manipulators for Reducing Interaction Wrench in Object Manipulation
by: Liu, Wenhang, et al.
Published: (2024)
by: Liu, Wenhang, et al.
Published: (2024)
Multi-Robot Rendezvous in Unknown Environment with Limited Communication
by: Song, Kun, et al.
Published: (2024)
by: Song, Kun, et al.
Published: (2024)
RHAML: Rendezvous-based Hierarchical Architecture for Mutual Localization
by: Chen, Gaoming, et al.
Published: (2024)
by: Chen, Gaoming, et al.
Published: (2024)
A Novel Planning Framework for Complex Flipping Manipulation of Multiple Mobile Manipulators
by: Liu, Wenhang, et al.
Published: (2023)
by: Liu, Wenhang, et al.
Published: (2023)
KiloBot: A Programming Language for Deploying Perception-Guided Industrial Manipulators at Scale
by: Gao, Wei, et al.
Published: (2024)
by: Gao, Wei, et al.
Published: (2024)
PinchBot: Long-Horizon Deformable Manipulation with Guided Diffusion Policy
by: Bartsch, Alison, et al.
Published: (2025)
by: Bartsch, Alison, et al.
Published: (2025)
Locomotion as Manipulation with ReachBot
by: Chen, Tony G., et al.
Published: (2024)
by: Chen, Tony G., et al.
Published: (2024)
DeRi-Bot: Learning to Collaboratively Manipulate Rigid Objects via Deformable Objects
by: Wang, Zixing, et al.
Published: (2023)
by: Wang, Zixing, et al.
Published: (2023)
Vision-Language Model Predictive Control for Manipulation Planning and Trajectory Generation
by: Chen, Jiaming, et al.
Published: (2025)
by: Chen, Jiaming, et al.
Published: (2025)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
by: Zhao, Han, et al.
Published: (2025)
by: Zhao, Han, et al.
Published: (2025)
ShakingBot: Dynamic Manipulation for Bagging
by: Gu, Ningquan, et al.
Published: (2023)
by: Gu, Ningquan, et al.
Published: (2023)
VIP: Vision Instructed Pre-training for Robotic Manipulation
by: Li, Zhuoling, et al.
Published: (2024)
by: Li, Zhuoling, et al.
Published: (2024)
Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action
by: Li, Pengteng, et al.
Published: (2026)
by: Li, Pengteng, et al.
Published: (2026)
ToddlerBot: Open-Source ML-Compatible Humanoid Platform for Loco-Manipulation
by: Shi, Haochen, et al.
Published: (2025)
by: Shi, Haochen, et al.
Published: (2025)
VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
by: Zhao, Wentao, et al.
Published: (2024)
by: Zhao, Wentao, et al.
Published: (2024)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
Does Peer Observation Help? Vision-Sharing Collaboration for Vision-Language Navigation
by: Jin, Qunchao, et al.
Published: (2026)
by: Jin, Qunchao, et al.
Published: (2026)
One Hand to Rule Them All: Canonical Representations for Unified Dexterous Manipulation
by: Wei, Zhenyu, et al.
Published: (2026)
by: Wei, Zhenyu, et al.
Published: (2026)
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
by: Tang, Weiliang, et al.
Published: (2025)
by: Tang, Weiliang, et al.
Published: (2025)
DoorBot: Closed-Loop Task Planning and Manipulation for Door Opening in the Wild with Haptic Feedback
by: Wang, Zhi, et al.
Published: (2025)
by: Wang, Zhi, et al.
Published: (2025)
VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis
by: Gu, Songen, et al.
Published: (2026)
by: Gu, Songen, et al.
Published: (2026)
A Novel Semi-Coupled Hierarchical Motion Planning Framework for Cooperative Transportation of Multiple Mobile Manipulators
by: Zhang, Heng, et al.
Published: (2022)
by: Zhang, Heng, et al.
Published: (2022)
ReMoBot: Retrieval-Based Few-Shot Imitation Learning for Mobile Manipulation with Vision Foundation Models
by: Zhang, Yuying, et al.
Published: (2024)
by: Zhang, Yuying, et al.
Published: (2024)
FrankenBot: Brain-Morphic Modular Orchestration for Robotic Manipulation with Vision-Language Models
by: Wang, Shiyi, et al.
Published: (2025)
by: Wang, Shiyi, et al.
Published: (2025)
iFlyBot-VLM Technical Report
by: Nie, Xin, et al.
Published: (2025)
by: Nie, Xin, et al.
Published: (2025)
LatBot: Distilling Universal Latent Actions for Vision-Language-Action Models
by: Li, Zuolei, et al.
Published: (2025)
by: Li, Zuolei, et al.
Published: (2025)
TSP-Bot: Robotic TSP Pen Art using High-DoF Manipulators
by: Song, Daeun, et al.
Published: (2022)
by: Song, Daeun, et al.
Published: (2022)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
by: Yuan, Puzhen, et al.
Published: (2025)
by: Yuan, Puzhen, et al.
Published: (2025)
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
by: Zhou, Jiaming, et al.
Published: (2025)
by: Zhou, Jiaming, et al.
Published: (2025)
Rethinking Intermediate Representation for VLM-based Robot Manipulation
by: Tang, Weiliang, et al.
Published: (2025)
by: Tang, Weiliang, et al.
Published: (2025)
VLATest: Testing and Evaluating Vision-Language-Action Models for Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
by: Mahmoud, Yara, et al.
Published: (2026)
by: Mahmoud, Yara, et al.
Published: (2026)
DepthCache: Depth-Guided Training-Free Visual Token Merging for Vision-Language-Action Model Inference
by: Li, Yuquan, et al.
Published: (2026)
by: Li, Yuquan, et al.
Published: (2026)
Vision-Guided Loco-Manipulation with a Snake Robot
by: Salagame, Adarsh, et al.
Published: (2025)
by: Salagame, Adarsh, et al.
Published: (2025)
RichMap: A Reachability Map Balancing Precision, Efficiency, and Flexibility for Rich Robot Manipulation Tasks
by: Lu, Yupu, et al.
Published: (2026)
by: Lu, Yupu, et al.
Published: (2026)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
by: Sun, Jianli, et al.
Published: (2026)
by: Sun, Jianli, et al.
Published: (2026)
AToM-Bot: Embodied Fulfillment of Unspoken Human Needs with Affective Theory of Mind
by: Ding, Wei, et al.
Published: (2024)
by: Ding, Wei, et al.
Published: (2024)
Forecast-aware Gaussian Splatting for Predictive 3D Representation in Language-Guided Pick-and-Place Manipulation
by: Jia, Kaixin, et al.
Published: (2026)
by: Jia, Kaixin, et al.
Published: (2026)
VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation
by: Zhou, Huayi, et al.
Published: (2025)
by: Zhou, Huayi, et al.
Published: (2025)
Similar Items
-
P2 Explore: Efficient Exploration in Unknown Cluttered Environment with Floor Plan Prediction
by: Song, Kun, et al.
Published: (2024) -
Distributed Motion Control of Multiple Mobile Manipulators for Reducing Interaction Wrench in Object Manipulation
by: Liu, Wenhang, et al.
Published: (2024) -
Multi-Robot Rendezvous in Unknown Environment with Limited Communication
by: Song, Kun, et al.
Published: (2024) -
RHAML: Rendezvous-based Hierarchical Architecture for Mutual Localization
by: Chen, Gaoming, et al.
Published: (2024) -
A Novel Planning Framework for Complex Flipping Manipulation of Multiple Mobile Manipulators
by: Liu, Wenhang, et al.
Published: (2023)