Object-Centric Instruction Augmentation for Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Junjie, Zhu, Yichen, Zhu, Minjie, Li, Jinming, Xu, Zhiyuan, Che, Zhengping, Shen, Chaomin, Peng, Yaxin, Liu, Dong, Feng, Feifei, Tang, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
Visual Robotic Manipulation with Depth-Aware Pretraining
von: Wang, Wanying, et al.
Veröffentlicht: (2024)
von: Wang, Wanying, et al.
Veröffentlicht: (2024)
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
von: Zhu, Minjie, et al.
Veröffentlicht: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
MMRo: Are Multimodal LLMs Eligible as the Brain for In-Home Robotics?
von: Li, Jinming, et al.
Veröffentlicht: (2024)
von: Li, Jinming, et al.
Veröffentlicht: (2024)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
CoA-VLA: Improving Vision-Language-Action Models via Visual-Textual Chain-of-Affordance
von: Li, Jinming, et al.
Veröffentlicht: (2024)
von: Li, Jinming, et al.
Veröffentlicht: (2024)
Let Me Show You: Learning by Retrieving from Egocentric Video for Robotic Manipulation
von: Zhu, Yichen, et al.
Veröffentlicht: (2025)
von: Zhu, Yichen, et al.
Veröffentlicht: (2025)
A Survey on Robotics with Foundation Models: toward Embodied AI
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
WorldEval: World Model as Real-World Robot Policies Evaluator
von: Li, Yaxuan, et al.
Veröffentlicht: (2025)
von: Li, Yaxuan, et al.
Veröffentlicht: (2025)
Learning from Imperfect Demonstrations with Self-Supervision for Robotic Manipulation
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
von: Li, Chengmeng, et al.
Veröffentlicht: (2025)
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
von: Wang, Xinhua, et al.
Veröffentlicht: (2026)
von: Wang, Xinhua, et al.
Veröffentlicht: (2026)
Mipha: A Comprehensive Overhaul of Multimodal Assistant with Small Language Models
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
von: Zhu, Minjie, et al.
Veröffentlicht: (2024)
CRAFT: Adapting VLA Models to Contact-rich Manipulation via Force-aware Curriculum Fine-tuning
von: Zhang, Yike, et al.
Veröffentlicht: (2026)
von: Zhang, Yike, et al.
Veröffentlicht: (2026)
HACTS: a Human-As-Copilot Teleoperation System for Robot Learning
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Embodied Agents
von: Zhu, Yichen, et al.
Veröffentlicht: (2024)
von: Zhu, Yichen, et al.
Veröffentlicht: (2024)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
von: Zeng, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zeng, Qiyuan, et al.
Veröffentlicht: (2025)
Language-Guided Object-Centric Diffusion Policy for Generalizable and Collision-Aware Robotic Manipulation
von: Li, Hang, et al.
Veröffentlicht: (2024)
von: Li, Hang, et al.
Veröffentlicht: (2024)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Li, Xiaoqi, et al.
Veröffentlicht: (2025)
SPOT: SE(3) Pose Trajectory Diffusion for Object-Centric Manipulation
von: Hsu, Cheng-Chun, et al.
Veröffentlicht: (2024)
von: Hsu, Cheng-Chun, et al.
Veröffentlicht: (2024)
Object-Centric Kinodynamic Planning for Nonprehensile Robot Rearrangement Manipulation
von: Ren, Kejia, et al.
Veröffentlicht: (2024)
von: Ren, Kejia, et al.
Veröffentlicht: (2024)
HumanoidExo: Scalable Whole-Body Humanoid Manipulation via Wearable Exoskeleton
von: Zhong, Rui, et al.
Veröffentlicht: (2025)
von: Zhong, Rui, et al.
Veröffentlicht: (2025)
Diffusion Trajectory-guided Policy for Long-horizon Robot Manipulation
von: Fan, Shichao, et al.
Veröffentlicht: (2025)
von: Fan, Shichao, et al.
Veröffentlicht: (2025)
dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
von: Wen, Junjie, et al.
Veröffentlicht: (2025)
Disentangled Object-Centric Image Representation for Robotic Manipulation
von: Emukpere, David, et al.
Veröffentlicht: (2025)
von: Emukpere, David, et al.
Veröffentlicht: (2025)
OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
von: Pan, Mingjie, et al.
Veröffentlicht: (2025)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
von: Zhao, Yinuo, et al.
Veröffentlicht: (2024)
von: Zhao, Yinuo, et al.
Veröffentlicht: (2024)
User-Centric Object Navigation: A Benchmark with Integrated User Habits for Personalized Embodied Object Search
von: Wang, Hongcheng, et al.
Veröffentlicht: (2026)
von: Wang, Hongcheng, et al.
Veröffentlicht: (2026)
Spotlighting Task-Relevant Features: Object-Centric Representations for Better Generalization in Robotic Manipulation
von: Chapin, Alexandre, et al.
Veröffentlicht: (2026)
von: Chapin, Alexandre, et al.
Veröffentlicht: (2026)
MLA: A Multisensory Language-Action Model for Multimodal Understanding and Forecasting in Robotic Manipulation
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Liu, Zhuoyang, et al.
Veröffentlicht: (2025)
A Deep Reinforcement Learning Environment for Particle Robot Navigation and Object Manipulation
von: Shen, Jeremy, et al.
Veröffentlicht: (2022)
von: Shen, Jeremy, et al.
Veröffentlicht: (2022)
Object-Centric Representations Improve Policy Generalization in Robot Manipulation
von: Chapin, Alexandre, et al.
Veröffentlicht: (2025)
von: Chapin, Alexandre, et al.
Veröffentlicht: (2025)
Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation
von: Heng, Liang, et al.
Veröffentlicht: (2025)
von: Heng, Liang, et al.
Veröffentlicht: (2025)
Fresh-CL: Feature Realignment through Experts on Hypersphere in Continual Learning
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhongyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
von: Zhu, Minjie, et al.
Veröffentlicht: (2024) -
Visual Robotic Manipulation with Depth-Aware Pretraining
von: Wang, Wanying, et al.
Veröffentlicht: (2024) -
Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation
von: Zhu, Minjie, et al.
Veröffentlicht: (2024) -
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
von: Zhu, Minjie, et al.
Veröffentlicht: (2025) -
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)