Robix: A Unified Model for Robot Interaction, Reasoning and Planning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Huang, Zhang, Mengxi, Dong, Heng, Li, Wei, Wang, Zixuan, Zhang, Qifeng, Tian, Xueyun, Hu, Yucheng, Li, Hang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PPAD: Iterative Interactions of Prediction and Planning for End-to-end Autonomous Driving
von: Chen, Zhili, et al.
Veröffentlicht: (2023)
von: Chen, Zhili, et al.
Veröffentlicht: (2023)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
RoboScape-R: Unified Reward-Observation World Models for Generalizable Robotics Training via RL
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
von: Wang, Beichen, et al.
Veröffentlicht: (2024)
Driving Animatronic Robot Facial Expression From Speech
von: Li, Boren, et al.
Veröffentlicht: (2024)
von: Li, Boren, et al.
Veröffentlicht: (2024)
Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models
von: Ling, Yiran, et al.
Veröffentlicht: (2026)
von: Ling, Yiran, et al.
Veröffentlicht: (2026)
RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
von: Mao, Weixin, et al.
Veröffentlicht: (2024)
Towards Generalizable Robotic Manipulation in Dynamic Environments
von: Fang, Heng, et al.
Veröffentlicht: (2026)
von: Fang, Heng, et al.
Veröffentlicht: (2026)
Unified Vision-Language-Action Model
von: Wang, Yuqi, et al.
Veröffentlicht: (2025)
von: Wang, Yuqi, et al.
Veröffentlicht: (2025)
Closed Loop Interactive Embodied Reasoning for Robot Manipulation
von: Nazarczuk, Michal, et al.
Veröffentlicht: (2024)
von: Nazarczuk, Michal, et al.
Veröffentlicht: (2024)
Exploring 3D Reasoning-Driven Planning: From Implicit Human Intentions to Route-Aware Activity Planning
von: Jiang, Xueying, et al.
Veröffentlicht: (2025)
von: Jiang, Xueying, et al.
Veröffentlicht: (2025)
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
von: Huang, Wenlong, et al.
Veröffentlicht: (2024)
von: Huang, Wenlong, et al.
Veröffentlicht: (2024)
A Step Toward World Models: A Survey on Robotic Manipulation
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2025)
Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation
von: Ding, Hongyu, et al.
Veröffentlicht: (2026)
von: Ding, Hongyu, et al.
Veröffentlicht: (2026)
Solving Motion Planning Tasks with a Scalable Generative Model
von: Hu, Yihan, et al.
Veröffentlicht: (2024)
von: Hu, Yihan, et al.
Veröffentlicht: (2024)
EARL: Towards a Unified Analysis-Guided Reinforcement Learning Framework for Egocentric Interaction Reasoning and Pixel Grounding
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
von: Dai, Tingjun, et al.
Veröffentlicht: (2026)
DiffAD: A Unified Diffusion Modeling Approach for Autonomous Driving
von: Wang, Tao, et al.
Veröffentlicht: (2025)
von: Wang, Tao, et al.
Veröffentlicht: (2025)
RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
von: Ji, Yuheng, et al.
Veröffentlicht: (2025)
von: Ji, Yuheng, et al.
Veröffentlicht: (2025)
rt-RISeg: Real-Time Model-Free Robot Interactive Segmentation for Active Instance-Level Object Understanding
von: Qian, Howard H., et al.
Veröffentlicht: (2025)
von: Qian, Howard H., et al.
Veröffentlicht: (2025)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
von: Zhang, Jianke, et al.
Veröffentlicht: (2024)
von: Zhang, Jianke, et al.
Veröffentlicht: (2024)
Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement Learning
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
LLM-Grounded Dynamic Task Planning with Hierarchical Temporal Logic for Human-Aware Multi-Robot Collaboration
von: Hu, Shuyuan, et al.
Veröffentlicht: (2026)
von: Hu, Shuyuan, et al.
Veröffentlicht: (2026)
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
von: Shao, Rui, et al.
Veröffentlicht: (2025)
von: Shao, Rui, et al.
Veröffentlicht: (2025)
UniDWM: Towards a Unified Driving World Model via Multifaceted Representation Learning
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning
von: Chen, Hao, et al.
Veröffentlicht: (2026)
von: Chen, Hao, et al.
Veröffentlicht: (2026)
VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models
von: Gao, Chongkai, et al.
Veröffentlicht: (2025)
von: Gao, Chongkai, et al.
Veröffentlicht: (2025)
Data-Centric Evolution in Autonomous Driving: A Comprehensive Survey of Big Data System, Data Mining, and Closed-Loop Technologies
von: Li, Lincan, et al.
Veröffentlicht: (2024)
von: Li, Lincan, et al.
Veröffentlicht: (2024)
Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation
von: Hu, Zixuan, et al.
Veröffentlicht: (2026)
von: Hu, Zixuan, et al.
Veröffentlicht: (2026)
SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning
von: Liu, Yuecheng, et al.
Veröffentlicht: (2025)
von: Liu, Yuecheng, et al.
Veröffentlicht: (2025)
VR-Robo: A Real-to-Sim-to-Real Framework for Visual Robot Navigation and Locomotion
von: Zhu, Shaoting, et al.
Veröffentlicht: (2025)
von: Zhu, Shaoting, et al.
Veröffentlicht: (2025)
UPTor: Unified 3D Human Pose Dynamics and Trajectory Prediction for Human-Robot Interaction
von: Nilavadi, Nisarga, et al.
Veröffentlicht: (2025)
von: Nilavadi, Nisarga, et al.
Veröffentlicht: (2025)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
von: Hu, Yucheng, et al.
Veröffentlicht: (2024)
von: Hu, Yucheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PPAD: Iterative Interactions of Prediction and Planning for End-to-end Autonomous Driving
von: Chen, Zhili, et al.
Veröffentlicht: (2023) -
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
von: Chen, Yandu, et al.
Veröffentlicht: (2025) -
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023) -
SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation
von: Li, Wei, et al.
Veröffentlicht: (2025) -
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
von: Wang, Guokang, et al.
Veröffentlicht: (2024)