Robo-MUTUAL: Robotic Multimodal Task Specification via Unimodal Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Jianxiong, Wang, Zhihao, Zheng, Jinliang, Zhou, Xiaoai, Wang, Guanming, Song, Guanglu, Liu, Yu, Liu, Jingjing, Zhang, Ya-Qin, Yu, Junzhi, Zhan, Xianyuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PhysiAgent: An Embodied Agent Framework in Physical World
di: Wang, Zhihao, et al.
Pubblicazione: (2025)
di: Wang, Zhihao, et al.
Pubblicazione: (2025)
Universal Actions for Enhanced Embodied Foundation Models
di: Zheng, Jinliang, et al.
Pubblicazione: (2025)
di: Zheng, Jinliang, et al.
Pubblicazione: (2025)
Demystifying Action Space Design for Robotic Manipulation Policies
di: Feng, Yuchun, et al.
Pubblicazione: (2026)
di: Feng, Yuchun, et al.
Pubblicazione: (2026)
Efficient Robotic Policy Learning via Latent Space Backward Planning
di: Liu, Dongxiu, et al.
Pubblicazione: (2025)
di: Liu, Dongxiu, et al.
Pubblicazione: (2025)
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning
di: Li, Jianxiong, et al.
Pubblicazione: (2024)
di: Li, Jianxiong, et al.
Pubblicazione: (2024)
Flow Matching-Based Autonomous Driving Planning with Advanced Interactive Behavior Modeling
di: Tan, Tianyi, et al.
Pubblicazione: (2025)
di: Tan, Tianyi, et al.
Pubblicazione: (2025)
Instruction-Guided Visual Masking
di: Zheng, Jinliang, et al.
Pubblicazione: (2024)
di: Zheng, Jinliang, et al.
Pubblicazione: (2024)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
di: Zheng, Yinan, et al.
Pubblicazione: (2024)
di: Zheng, Yinan, et al.
Pubblicazione: (2024)
X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
di: Zheng, Jinliang, et al.
Pubblicazione: (2025)
di: Zheng, Jinliang, et al.
Pubblicazione: (2025)
Dichotomous Diffusion Policy Optimization
di: Liang, Ruiming, et al.
Pubblicazione: (2025)
di: Liang, Ruiming, et al.
Pubblicazione: (2025)
Diffusion-Based Planning for Autonomous Driving with Flexible Guidance
di: Zheng, Yinan, et al.
Pubblicazione: (2025)
di: Zheng, Yinan, et al.
Pubblicazione: (2025)
RoboClaw: An Agentic Framework for Scalable Long-Horizon Robotic Tasks
di: Li, Ruiying, et al.
Pubblicazione: (2026)
di: Li, Ruiying, et al.
Pubblicazione: (2026)
OpenRoboCare: A Multimodal Multi-Task Expert Demonstration Dataset for Robot Caregiving
di: Liang, Xiaoyu, et al.
Pubblicazione: (2025)
di: Liang, Xiaoyu, et al.
Pubblicazione: (2025)
RoboBERT: An End-to-end Multimodal Robotic Manipulation Model
di: Wang, Sicheng, et al.
Pubblicazione: (2025)
di: Wang, Sicheng, et al.
Pubblicazione: (2025)
RoboMP$^2$: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language Models
di: Lv, Qi, et al.
Pubblicazione: (2024)
di: Lv, Qi, et al.
Pubblicazione: (2024)
Skill Expansion and Composition in Parameter Space
di: Liu, Tenglong, et al.
Pubblicazione: (2025)
di: Liu, Tenglong, et al.
Pubblicazione: (2025)
RoboTron-Mani: All-in-One Multimodal Large Model for Robotic Manipulation
di: Yan, Feng, et al.
Pubblicazione: (2024)
di: Yan, Feng, et al.
Pubblicazione: (2024)
RoboKube: Establishing a New Foundation for the Cloud Native Evolution in Robotics
di: Liu, Yu, et al.
Pubblicazione: (2024)
di: Liu, Yu, et al.
Pubblicazione: (2024)
RoboLLM: Robotic Vision Tasks Grounded on Multimodal Large Language Models
di: Long, Zijun, et al.
Pubblicazione: (2023)
di: Long, Zijun, et al.
Pubblicazione: (2023)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
di: Li, Pengxiang, et al.
Pubblicazione: (2025)
di: Li, Pengxiang, et al.
Pubblicazione: (2025)
RoboCodeX: Multimodal Code Generation for Robotic Behavior Synthesis
di: Mu, Yao, et al.
Pubblicazione: (2024)
di: Mu, Yao, et al.
Pubblicazione: (2024)
xTED: Cross-Domain Adaptation via Diffusion-Based Trajectory Editing
di: Niu, Haoyi, et al.
Pubblicazione: (2024)
di: Niu, Haoyi, et al.
Pubblicazione: (2024)
Towards Robust Zero-Shot Reinforcement Learning
di: Zheng, Kexin, et al.
Pubblicazione: (2025)
di: Zheng, Kexin, et al.
Pubblicazione: (2025)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
di: Wang, Siyin, et al.
Pubblicazione: (2025)
di: Wang, Siyin, et al.
Pubblicazione: (2025)
RoboDexVLM: Visual Language Model-Enabled Task Planning and Motion Control for Dexterous Robot Manipulation
di: Liu, Haichao, et al.
Pubblicazione: (2025)
di: Liu, Haichao, et al.
Pubblicazione: (2025)
RoboPanoptes: The All-seeing Robot with Whole-body Dexterity
di: Xu, Xiaomeng, et al.
Pubblicazione: (2025)
di: Xu, Xiaomeng, et al.
Pubblicazione: (2025)
RoboCOIN: An Open-Sourced Bimanual Robotic Data Collection for Integrated Manipulation
di: Wu, Shihan, et al.
Pubblicazione: (2025)
di: Wu, Shihan, et al.
Pubblicazione: (2025)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
di: Wu, Hao, et al.
Pubblicazione: (2026)
di: Wu, Hao, et al.
Pubblicazione: (2026)
RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
di: Lei, Huashuo, et al.
Pubblicazione: (2026)
di: Lei, Huashuo, et al.
Pubblicazione: (2026)
RoboCoder: Robotic Learning from Basic Skills to General Tasks with Large Language Models
di: Li, Jingyao, et al.
Pubblicazione: (2024)
di: Li, Jingyao, et al.
Pubblicazione: (2024)
RoboMatch: A Unified Mobile-Manipulation Teleoperation Platform with Auto-Matching Network Architecture for Long-Horizon Tasks
di: Liu, Hanyu, et al.
Pubblicazione: (2025)
di: Liu, Hanyu, et al.
Pubblicazione: (2025)
RoboRetriever: Single-Camera Robot Object Retrieval via Active and Interactive Perception with Dynamic Scene Graph
di: Wang, Hecheng, et al.
Pubblicazione: (2025)
di: Wang, Hecheng, et al.
Pubblicazione: (2025)
RoboInter: A Holistic Intermediate Representation Suite Towards Robotic Manipulation
di: Li, Hao, et al.
Pubblicazione: (2026)
di: Li, Hao, et al.
Pubblicazione: (2026)
RoboBPP: Benchmarking Robotic Online Bin Packing with Physics-based Simulation
di: Wang, Zhoufeng, et al.
Pubblicazione: (2025)
di: Wang, Zhoufeng, et al.
Pubblicazione: (2025)
ReSPIRe: Informative and Reusable Belief Tree Search for Robot Probabilistic Search and Tracking in Unknown Environments
di: Zhou, Kangjie, et al.
Pubblicazione: (2025)
di: Zhou, Kangjie, et al.
Pubblicazione: (2025)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
di: Li, Huiqiong, et al.
Pubblicazione: (2026)
di: Li, Huiqiong, et al.
Pubblicazione: (2026)
RoboTAG: End-to-end Robot Configuration Estimation via Topological Alignment Graph
di: Liu, Yifan, et al.
Pubblicazione: (2025)
di: Liu, Yifan, et al.
Pubblicazione: (2025)
RoboGPT-R1: Enhancing Robot Planning with Reinforcement Learning
di: Liu, Jinrui, et al.
Pubblicazione: (2025)
di: Liu, Jinrui, et al.
Pubblicazione: (2025)
RoboRouter: Training-Free Policy Routing for Robotic Manipulation
di: Chen, Yiteng, et al.
Pubblicazione: (2026)
di: Chen, Yiteng, et al.
Pubblicazione: (2026)
RoboCAS: A Benchmark for Robotic Manipulation in Complex Object Arrangement Scenarios
di: Zheng, Liming, et al.
Pubblicazione: (2024)
di: Zheng, Liming, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PhysiAgent: An Embodied Agent Framework in Physical World
di: Wang, Zhihao, et al.
Pubblicazione: (2025) -
Universal Actions for Enhanced Embodied Foundation Models
di: Zheng, Jinliang, et al.
Pubblicazione: (2025) -
Demystifying Action Space Design for Robotic Manipulation Policies
di: Feng, Yuchun, et al.
Pubblicazione: (2026) -
Efficient Robotic Policy Learning via Latent Space Backward Planning
di: Liu, Dongxiu, et al.
Pubblicazione: (2025) -
DecisionNCE: Embodied Multimodal Representations via Implicit Preference Learning
di: Li, Jianxiong, et al.
Pubblicazione: (2024)