ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Enyu, Raval, Vedant, Zhang, Hejia, Mao, Jiageng, Shangguan, Zeyu, Nikolaidis, Stefanos, Wang, Yue, Seita, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GPT-Fabric: Smoothing and Folding Fabric by Leveraging Pre-Trained Foundation Models
von: Raval, Vedant, et al.
Veröffentlicht: (2024)
von: Raval, Vedant, et al.
Veröffentlicht: (2024)
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
von: Chow, Wei, et al.
Veröffentlicht: (2025)
von: Chow, Wei, et al.
Veröffentlicht: (2025)
The MOTIF Hand: A Robotic Hand for Multimodal Observations with Thermal, Inertial, and Force Sensors
von: Zhou, Hanyang, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyang, et al.
Veröffentlicht: (2025)
HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction
von: Shi, Zhonghao, et al.
Veröffentlicht: (2025)
von: Shi, Zhonghao, et al.
Veröffentlicht: (2025)
OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation
von: Wang, Kuanning, et al.
Veröffentlicht: (2026)
von: Wang, Kuanning, et al.
Veröffentlicht: (2026)
iManip: Skill-Incremental Learning for Robotic Manipulation
von: Zheng, Zexin, et al.
Veröffentlicht: (2025)
von: Zheng, Zexin, et al.
Veröffentlicht: (2025)
SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation
von: Huang, Chengyue, et al.
Veröffentlicht: (2026)
von: Huang, Chengyue, et al.
Veröffentlicht: (2026)
GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
Red-Teaming Vision-Language-Action Models via Quality Diversity Prompt Generation for Robust Robot Policies
von: Srikanth, Siddharth, et al.
Veröffentlicht: (2026)
von: Srikanth, Siddharth, et al.
Veröffentlicht: (2026)
Designing Robot Identity: The Role of Voice, Clothing, and Task on Robot Gender Perception
von: Dennler, Nathaniel S., et al.
Veröffentlicht: (2024)
von: Dennler, Nathaniel S., et al.
Veröffentlicht: (2024)
Robot Learning from Any Images
von: Zhao, Siheng, et al.
Veröffentlicht: (2025)
von: Zhao, Siheng, et al.
Veröffentlicht: (2025)
ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models
von: Song, Zirui, et al.
Veröffentlicht: (2025)
von: Song, Zirui, et al.
Veröffentlicht: (2025)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
Manip4Care: Robotic Manipulation of Human Limbs for Solving Assistive Tasks
von: Koh, Yubin, et al.
Veröffentlicht: (2025)
von: Koh, Yubin, et al.
Veröffentlicht: (2025)
Singing the Body Electric: The Impact of Robot Embodiment on User Expectations
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
Granular Loco-Manipulation: Repositioning Rocks Through Strategic Sand Avalanche
von: Hu, Haodi, et al.
Veröffentlicht: (2025)
von: Hu, Haodi, et al.
Veröffentlicht: (2025)
OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
von: Li, Ying, et al.
Veröffentlicht: (2025)
von: Li, Ying, et al.
Veröffentlicht: (2025)
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
Quality Diversity for Robot Learning: Limitations and Future Directions
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
von: Batra, Sumeet, et al.
Veröffentlicht: (2024)
Enabling Adaptive Agent Training in Open-Ended Simulators by Targeting Diversity
von: Costales, Robby, et al.
Veröffentlicht: (2024)
von: Costales, Robby, et al.
Veröffentlicht: (2024)
Concurrent Prehensile and Nonprehensile Manipulation: A Practical Approach to Multi-Stage Dexterous Tasks
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
von: Jiang, Hao, et al.
Veröffentlicht: (2026)
HeteroGenManip: Generalizable Manipulation For Heterogeneous Object Interactions
von: Shen, Zhenhao, et al.
Veröffentlicht: (2026)
von: Shen, Zhenhao, et al.
Veröffentlicht: (2026)
VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
von: Zhao, Wentao, et al.
Veröffentlicht: (2024)
von: Zhao, Wentao, et al.
Veröffentlicht: (2024)
The RoSiD Tool: Empowering Users to Design Multimodal Signals for Human-Robot Collaboration
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
Learning Granular Media Avalanche Behavior for Indirectly Manipulating Obstacles on a Granular Slope
von: Hu, Haodi, et al.
Veröffentlicht: (2024)
von: Hu, Haodi, et al.
Veröffentlicht: (2024)
Sequential Multi-Object Grasping with One Dexterous Hand
von: He, Sicheng, et al.
Veröffentlicht: (2025)
von: He, Sicheng, et al.
Veröffentlicht: (2025)
ResponsibleRobotBench: Benchmarking Responsible Robot Manipulation using Multi-modal Large Language Models
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
DreamPlan: Efficient Reinforcement Fine-Tuning of Vision-Language Planners via Video World Models
von: Jia, Emily Yue-Ting, et al.
Veröffentlicht: (2026)
von: Jia, Emily Yue-Ting, et al.
Veröffentlicht: (2026)
Soft and Compliant Contact-Rich Hair Manipulation and Care
von: Yoo, Uksang, et al.
Veröffentlicht: (2025)
von: Yoo, Uksang, et al.
Veröffentlicht: (2025)
RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments
von: Murooka, Masaki, et al.
Veröffentlicht: (2025)
von: Murooka, Masaki, et al.
Veröffentlicht: (2025)
Humanoid Everyday: A Comprehensive Robotic Dataset for Open-World Humanoid Manipulation
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2025)
ManipArena: Comprehensive Real-world Evaluation of Reasoning-Oriented Generalist Robot Manipulation
von: Sun, Yu, et al.
Veröffentlicht: (2026)
von: Sun, Yu, et al.
Veröffentlicht: (2026)
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
ArtiBench and ArtiBrain: Benchmarking Generalizable Vision-Language Articulated Object Manipulation
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
von: Wu, Yuhan, et al.
Veröffentlicht: (2025)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
von: Dennler, Nathaniel, et al.
Veröffentlicht: (2024)
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
Robots of the Lost Arc: Self-Supervised Learning to Dynamically Manipulate Fixed-Endpoint Cables
von: Zhang, Harry, et al.
Veröffentlicht: (2020)
von: Zhang, Harry, et al.
Veröffentlicht: (2020)
Bench-Push: Benchmarking Pushing-based Navigation and Manipulation Tasks for Mobile Robots
von: Zhong, Ninghan, et al.
Veröffentlicht: (2025)
von: Zhong, Ninghan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GPT-Fabric: Smoothing and Folding Fabric by Leveraging Pre-Trained Foundation Models
von: Raval, Vedant, et al.
Veröffentlicht: (2024) -
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
von: Chow, Wei, et al.
Veröffentlicht: (2025) -
The MOTIF Hand: A Robotic Hand for Multimodal Observations with Thermal, Inertial, and Force Sensors
von: Zhou, Hanyang, et al.
Veröffentlicht: (2025) -
HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction
von: Shi, Zhonghao, et al.
Veröffentlicht: (2025) -
OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation
von: Wang, Kuanning, et al.
Veröffentlicht: (2026)