Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jinyi, Yuan, Yifu, Hao, Jianye, Ni, Fei, Fu, Lingzhi, Chen, Yibin, Zheng, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025)
by: Cui, Haiqin, et al.
Published: (2025)
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
by: Dong, Zibin, et al.
Published: (2025)
by: Dong, Zibin, et al.
Published: (2025)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
by: Chen, Yibin, et al.
Published: (2024)
by: Chen, Yibin, et al.
Published: (2024)
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024)
by: Dong, Zibin, et al.
Published: (2024)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
by: Zhao, Kai, et al.
Published: (2023)
by: Zhao, Kai, et al.
Published: (2023)
ActionCodec: What Makes for Good Action Tokenizers
by: Dong, Zibin, et al.
Published: (2026)
by: Dong, Zibin, et al.
Published: (2026)
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
by: Li, Zaijing, et al.
Published: (2026)
by: Li, Zaijing, et al.
Published: (2026)
ForceFlow: Learning to Feel and Act via Contact-Driven Flow Matching
by: Zhang, Shuoheng, et al.
Published: (2026)
by: Zhang, Shuoheng, et al.
Published: (2026)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
by: Zhang, Yihao, et al.
Published: (2025)
by: Zhang, Yihao, et al.
Published: (2025)
ARCap: Collecting High-quality Human Demonstrations for Robot Learning with Augmented Reality Feedback
by: Chen, Sirui, et al.
Published: (2024)
by: Chen, Sirui, et al.
Published: (2024)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
by: Singh, Harsh, et al.
Published: (2024)
by: Singh, Harsh, et al.
Published: (2024)
Hybrid Framework for Robotic Manipulation: Integrating Reinforcement Learning and Large Language Models
by: Saad, Md, et al.
Published: (2026)
by: Saad, Md, et al.
Published: (2026)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Lifelong Language-Conditioned Robotic Manipulation Learning
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
Hierarchical Language Models for Semantic Navigation and Manipulation in an Aerial-Ground Robotic System
by: Liu, Haokun, et al.
Published: (2025)
by: Liu, Haokun, et al.
Published: (2025)
Large Language Models for Robotics: A Survey
by: Zeng, Fanlong, et al.
Published: (2023)
by: Zeng, Fanlong, et al.
Published: (2023)
HybridFlow: A Two-Step Generative Policy for Robotic Manipulation
by: Dong, Zhenchen, et al.
Published: (2026)
by: Dong, Zhenchen, et al.
Published: (2026)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
by: Liu, Jinkun, et al.
Published: (2026)
by: Liu, Jinkun, et al.
Published: (2026)
Physically Grounded Vision-Language Models for Robotic Manipulation
by: Gao, Jensen, et al.
Published: (2023)
by: Gao, Jensen, et al.
Published: (2023)
ManiFoundation Model for General-Purpose Robotic Manipulation of Contact Synthesis with Arbitrary Objects and Robots
by: Xu, Zhixuan, et al.
Published: (2024)
by: Xu, Zhixuan, et al.
Published: (2024)
RoboTron-Mani: All-in-One Multimodal Large Model for Robotic Manipulation
by: Yan, Feng, et al.
Published: (2024)
by: Yan, Feng, et al.
Published: (2024)
OpenNav: Open-World Navigation with Multimodal Large Language Models
by: Yuan, Mingfeng, et al.
Published: (2025)
by: Yuan, Mingfeng, et al.
Published: (2025)
SEVO: Semantic-Enhanced Virtual Observation for Robust VLA Manipulation via Active Illumination and Data-Centric Collection
by: Fang, Tianchonghui, et al.
Published: (2026)
by: Fang, Tianchonghui, et al.
Published: (2026)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
by: Yan, Hongyu, et al.
Published: (2026)
by: Yan, Hongyu, et al.
Published: (2026)
Embodied AI in Mobile Robots: Coverage Path Planning with Large Language Models
by: Kong, Xiangrui, et al.
Published: (2024)
by: Kong, Xiangrui, et al.
Published: (2024)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
MLLM-Fabric: Multimodal Large Language Model-Driven Robotic Framework for Fabric Sorting and Selection
by: Wang, Liman, et al.
Published: (2025)
by: Wang, Liman, et al.
Published: (2025)
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models
by: Turcato, Niccolò, et al.
Published: (2025)
by: Turcato, Niccolò, et al.
Published: (2025)
Large Language Models for Robotics: Opportunities, Challenges, and Perspectives
by: Wang, Jiaqi, et al.
Published: (2024)
by: Wang, Jiaqi, et al.
Published: (2024)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Toward Generalist Neural Motion Planners for Robotic Manipulators: Challenges and Opportunities
by: Soleymanzadeh, Davood, et al.
Published: (2026)
by: Soleymanzadeh, Davood, et al.
Published: (2026)
Exploring Embodied Multimodal Large Models: Development, Datasets, and Future Directions
by: Chen, Shoubin, et al.
Published: (2025)
by: Chen, Shoubin, et al.
Published: (2025)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
by: Zhang, Rongyu, et al.
Published: (2025)
by: Zhang, Rongyu, et al.
Published: (2025)
RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design
by: Chen, Tianxing, et al.
Published: (2026)
by: Chen, Tianxing, et al.
Published: (2026)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
by: Zhao, Enyu, et al.
Published: (2025)
by: Zhao, Enyu, et al.
Published: (2025)
Similar Items
-
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025) -
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025) -
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
by: Dong, Zibin, et al.
Published: (2025) -
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025) -
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
by: Chen, Yibin, et al.
Published: (2024)