Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Kaiyuan, Xie, Shuangyu, Ma, Zehan, Sanketi, Pannag R, Goldberg, Ken |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robo-DM: Data Management For Large Robot Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025)
Blox-Net: Generative Design-for-Robot-Assembly Using VLM Supervision, Physics Simulation, and a Robot with Reset
von: Goldberg, Andrew, et al.
Veröffentlicht: (2024)
von: Goldberg, Andrew, et al.
Veröffentlicht: (2024)
LAVQA: A Latency-Aware Visual Question Answering Framework for Shared Autonomy in Self-Driving Vehicles
von: Xie, Shuangyu, et al.
Veröffentlicht: (2025)
von: Xie, Shuangyu, et al.
Veröffentlicht: (2025)
OXE-AugE: A Large-Scale Robot Augmentation of OXE for Scaling Cross-Embodiment Policy Learning
von: Ji, Guanhua, et al.
Veröffentlicht: (2025)
von: Ji, Guanhua, et al.
Veröffentlicht: (2025)
DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
von: Khazatsky, Alexander, et al.
Veröffentlicht: (2024)
von: Khazatsky, Alexander, et al.
Veröffentlicht: (2024)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins
von: Mu, Yao, et al.
Veröffentlicht: (2025)
von: Mu, Yao, et al.
Veröffentlicht: (2025)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)
von: Mu, Yao, et al.
Veröffentlicht: (2024)
von: Mu, Yao, et al.
Veröffentlicht: (2024)
Energy Efficient Planning for Repetitive Heterogeneous Tasks in Precision Agriculture
von: Xie, Shuangyu, et al.
Veröffentlicht: (2025)
von: Xie, Shuangyu, et al.
Veröffentlicht: (2025)
RoboUniView: Visual-Language Model with Unified View Representation for Robotic Manipulation
von: Liu, Fanfan, et al.
Veröffentlicht: (2024)
von: Liu, Fanfan, et al.
Veröffentlicht: (2024)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
von: Jiang, Feng, et al.
Veröffentlicht: (2026)
von: Jiang, Feng, et al.
Veröffentlicht: (2026)
RoboDexVLM: Visual Language Model-Enabled Task Planning and Motion Control for Dexterous Robot Manipulation
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
AutoRT: Embodied Foundation Models for Large Scale Orchestration of Robotic Agents
von: Ahn, Michael, et al.
Veröffentlicht: (2024)
von: Ahn, Michael, et al.
Veröffentlicht: (2024)
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
Simulation to Rules: A Dual-VLM Framework for Formal Visual Planning
von: Hao, Yilun, et al.
Veröffentlicht: (2025)
von: Hao, Yilun, et al.
Veröffentlicht: (2025)
RoboPilot: Generalizable Dynamic Robotic Manipulation with Dual-thinking Modes
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
von: Liu, Xinyi, et al.
Veröffentlicht: (2025)
RoboPlayground: Democratizing Robotic Evaluation through Structured Physical Domains
von: Wang, Yi Ru, et al.
Veröffentlicht: (2026)
von: Wang, Yi Ru, et al.
Veröffentlicht: (2026)
RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
von: Wang, Siyin, et al.
Veröffentlicht: (2025)
von: Wang, Siyin, et al.
Veröffentlicht: (2025)
RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation
von: Ying, Chenduo, et al.
Veröffentlicht: (2025)
von: Ying, Chenduo, et al.
Veröffentlicht: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
von: Li, Huiqiong, et al.
Veröffentlicht: (2026)
Robots of the Lost Arc: Self-Supervised Learning to Dynamically Manipulate Fixed-Endpoint Cables
von: Zhang, Harry, et al.
Veröffentlicht: (2020)
von: Zhang, Harry, et al.
Veröffentlicht: (2020)
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs
von: Hu, Zichao, et al.
Veröffentlicht: (2024)
von: Hu, Zichao, et al.
Veröffentlicht: (2024)
Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers
von: Jain, Vidhi, et al.
Veröffentlicht: (2024)
von: Jain, Vidhi, et al.
Veröffentlicht: (2024)
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
von: Zhang, Shiduo, et al.
Veröffentlicht: (2024)
von: Zhang, Shiduo, et al.
Veröffentlicht: (2024)
Robot Collapse: Supply Chain Backdoor Attacks Against VLM-based Robotic Manipulation
von: Wang, Xianlong, et al.
Veröffentlicht: (2024)
von: Wang, Xianlong, et al.
Veröffentlicht: (2024)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
von: Wang, Boyang, et al.
Veröffentlicht: (2026)
Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks
von: Parekh, Amit, et al.
Veröffentlicht: (2024)
von: Parekh, Amit, et al.
Veröffentlicht: (2024)
Autonomous Frontier-Based Exploration with VLM Guidance
von: Aitha, Aarush, et al.
Veröffentlicht: (2026)
von: Aitha, Aarush, et al.
Veröffentlicht: (2026)
MonoDuo: Using One Robot Arm to Learn Bimanual Policies
von: Bajamahal, Sandeep, et al.
Veröffentlicht: (2026)
von: Bajamahal, Sandeep, et al.
Veröffentlicht: (2026)
Score the Steps, Not Just the Goal: VLM-Based Subgoal Evaluation for Robotic Manipulation
von: ElMallah, Ramy, et al.
Veröffentlicht: (2025)
von: ElMallah, Ramy, et al.
Veröffentlicht: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
von: Jiang, Guangqi, et al.
Veröffentlicht: (2024)
Leveraging Adaptive Group Negotiation for Heterogeneous Multi-Robot Collaboration with Large Language Models
von: Song, Siqi, et al.
Veröffentlicht: (2025)
von: Song, Siqi, et al.
Veröffentlicht: (2025)
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
von: Jia, Qingxuan, et al.
Veröffentlicht: (2025)
von: Jia, Qingxuan, et al.
Veröffentlicht: (2025)
RoboCasa365: A Large-Scale Simulation Framework for Training and Benchmarking Generalist Robots
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2026)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2026)
RoboLight: A Dataset with Linearly Composable Illumination for Robotic Manipulation
von: Jin, Shutong, et al.
Veröffentlicht: (2026)
von: Jin, Shutong, et al.
Veröffentlicht: (2026)
Robot Data Curation with Mutual Information Estimators
von: Hejna, Joey, et al.
Veröffentlicht: (2025)
von: Hejna, Joey, et al.
Veröffentlicht: (2025)
MCD: Diverse Large-Scale Multi-Campus Dataset for Robot Perception
von: Nguyen, Thien-Minh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thien-Minh, et al.
Veröffentlicht: (2024)
KiloBot: A Programming Language for Deploying Perception-Guided Industrial Manipulators at Scale
von: Gao, Wei, et al.
Veröffentlicht: (2024)
von: Gao, Wei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Robo-DM: Data Management For Large Robot Datasets
von: Chen, Kaiyuan, et al.
Veröffentlicht: (2025) -
Blox-Net: Generative Design-for-Robot-Assembly Using VLM Supervision, Physics Simulation, and a Robot with Reset
von: Goldberg, Andrew, et al.
Veröffentlicht: (2024) -
LAVQA: A Latency-Aware Visual Question Answering Framework for Shared Autonomy in Self-Driving Vehicles
von: Xie, Shuangyu, et al.
Veröffentlicht: (2025) -
OXE-AugE: A Large-Scale Robot Augmentation of OXE for Scaling Cross-Embodiment Policy Learning
von: Ji, Guanhua, et al.
Veröffentlicht: (2025) -
DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset
von: Khazatsky, Alexander, et al.
Veröffentlicht: (2024)