From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Yifu, Cui, Haiqin, Chen, Yibin, Dong, Zibin, Ni, Fei, Kou, Longxin, Liu, Jinyi, Li, Pengyi, Zheng, Yan, Hao, Jianye |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025)
by: Cui, Haiqin, et al.
Published: (2025)
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
by: Dong, Zibin, et al.
Published: (2025)
by: Dong, Zibin, et al.
Published: (2025)
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024)
by: Dong, Zibin, et al.
Published: (2024)
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
by: Liu, Jinyi, et al.
Published: (2024)
by: Liu, Jinyi, et al.
Published: (2024)
DiffuserLite: Towards Real-time Diffusion Planning
by: Dong, Zibin, et al.
Published: (2024)
by: Dong, Zibin, et al.
Published: (2024)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
by: Chen, Yibin, et al.
Published: (2024)
by: Chen, Yibin, et al.
Published: (2024)
ActionCodec: What Makes for Good Action Tokenizers
by: Dong, Zibin, et al.
Published: (2026)
by: Dong, Zibin, et al.
Published: (2026)
ForceFlow: Learning to Feel and Act via Contact-Driven Flow Matching
by: Zhang, Shuoheng, et al.
Published: (2026)
by: Zhang, Shuoheng, et al.
Published: (2026)
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Conditioning Matters: Training Diffusion Policies is Faster Than You Think
by: Dong, Zibin, et al.
Published: (2025)
by: Dong, Zibin, et al.
Published: (2025)
Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction
by: Kerr, Justin, et al.
Published: (2024)
by: Kerr, Justin, et al.
Published: (2024)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
by: Zhao, Kai, et al.
Published: (2023)
by: Zhao, Kai, et al.
Published: (2023)
From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation
by: Liu, Yibin, et al.
Published: (2026)
by: Liu, Yibin, et al.
Published: (2026)
Spatial-Temporal Graph Diffusion Policy with Kinematic Modeling for Bimanual Robotic Manipulation
by: Lv, Qi, et al.
Published: (2025)
by: Lv, Qi, et al.
Published: (2025)
Optimizing Robotic Manipulation with Decision-RWKV: A Recurrent Sequence Modeling Approach for Lifelong Learning
by: Dong, Yujian, et al.
Published: (2024)
by: Dong, Yujian, et al.
Published: (2024)
Bridging Language and Action: A Survey of Language-Conditioned Robot Manipulation
by: Yao, Xiangtong, et al.
Published: (2023)
by: Yao, Xiangtong, et al.
Published: (2023)
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making
by: Liu, Jun, et al.
Published: (2026)
by: Liu, Jun, et al.
Published: (2026)
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
ResponsibleRobotBench: Benchmarking Responsible Robot Manipulation using Multi-modal Large Language Models
by: Zhang, Lei, et al.
Published: (2025)
by: Zhang, Lei, et al.
Published: (2025)
From Chaos to Order: The Atomic Reasoner Framework for Fine-grained Reasoning in Large Language Models
by: Liu, Jinyi, et al.
Published: (2025)
by: Liu, Jinyi, et al.
Published: (2025)
DexCanvas: Bridging Human Demonstrations and Robot Learning for Dexterous Manipulation
by: Xu, Xinyue, et al.
Published: (2025)
by: Xu, Xinyue, et al.
Published: (2025)
Action-Sketcher: From Reasoning to Action via Visual Sketches for Long-Horizon Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2026)
by: Tan, Huajie, et al.
Published: (2026)
Robot See, Robot Do: Imitation Reward for Noisy Financial Environments
by: Goluža, Sven, et al.
Published: (2024)
by: Goluža, Sven, et al.
Published: (2024)
Embodied Arena: A Comprehensive, Unified, and Evolving Evaluation Platform for Embodied AI
by: Ni, Fei, et al.
Published: (2025)
by: Ni, Fei, et al.
Published: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
by: Huang, Wenlong, et al.
Published: (2024)
by: Huang, Wenlong, et al.
Published: (2024)
EEG-Driven AR-Robot System for Zero-Touch Grasping Manipulation
by: Wang, Junzhe, et al.
Published: (2025)
by: Wang, Junzhe, et al.
Published: (2025)
BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
by: Yu, Chenhao, et al.
Published: (2026)
by: Yu, Chenhao, et al.
Published: (2026)
AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
by: Tian, Yang, et al.
Published: (2024)
by: Tian, Yang, et al.
Published: (2024)
Hyperbolic Multiview Pretraining for Robotic Manipulation
by: Yang, Jin, et al.
Published: (2026)
by: Yang, Jin, et al.
Published: (2026)
Seeing, Saying, Solving: An LLM-to-TL Framework for Cooperative Robots
by: Choe, Dan BW, et al.
Published: (2025)
by: Choe, Dan BW, et al.
Published: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
by: Bai, Yongjie, et al.
Published: (2025)
by: Bai, Yongjie, et al.
Published: (2025)
Seeing the Bigger Picture: 3D Latent Mapping for Mobile Manipulation Policy Learning
by: Kim, Sunghwan, et al.
Published: (2025)
by: Kim, Sunghwan, et al.
Published: (2025)
Learning When to See and When to Feel: Adaptive Vision-Torque Fusion for Contact-Aware Manipulation
by: Lei, Jiuzhou, et al.
Published: (2026)
by: Lei, Jiuzhou, et al.
Published: (2026)
Haptic-ACT: Bridging Human Intuition with Compliant Robotic Manipulation via Immersive VR
by: Li, Kelin, et al.
Published: (2024)
by: Li, Kelin, et al.
Published: (2024)
RGBManip: Monocular Image-based Robotic Manipulation through Active Object Pose Estimation
by: An, Boshi, et al.
Published: (2023)
by: An, Boshi, et al.
Published: (2023)
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation
by: Liu, Yuanzhe, et al.
Published: (2026)
by: Liu, Yuanzhe, et al.
Published: (2026)
PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation
by: Guo, Pengyuan, et al.
Published: (2026)
by: Guo, Pengyuan, et al.
Published: (2026)
Similar Items
-
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025) -
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025) -
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
by: Dong, Zibin, et al.
Published: (2025) -
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024) -
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
by: Liu, Jinyi, et al.
Published: (2024)