Vidarc: Embodied Video Diffusion Model for Closed-loop Control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Feng, Yao, Xiang, Chendong, Mao, Xinyi, Tan, Hengkai, Zhang, Zuyue, Huang, Shuhe, Zheng, Kaiwen, Liu, Haitian, Su, Hang, Zhu, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
von: Feng, Yao, et al.
Veröffentlicht: (2025)
von: Feng, Yao, et al.
Veröffentlicht: (2025)
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
von: Tan, Hengkai, et al.
Veröffentlicht: (2025)
von: Tan, Hengkai, et al.
Veröffentlicht: (2025)
Fourier Controller Networks for Real-Time Decision-Making in Embodied Learning
von: Tan, Hengkai, et al.
Veröffentlicht: (2024)
von: Tan, Hengkai, et al.
Veröffentlicht: (2024)
ManiBox: Enhancing Embodied Spatial Generalization via Scalable Simulation Data Generations
von: Tan, Hengkai, et al.
Veröffentlicht: (2024)
von: Tan, Hengkai, et al.
Veröffentlicht: (2024)
Motus: A Unified Latent Action World Model
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
MotuBrain: An Advanced World Action Model for Robot Control
von: MotuBrain Team, et al.
Veröffentlicht: (2026)
von: MotuBrain Team, et al.
Veröffentlicht: (2026)
H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
von: Liu, Songming, et al.
Veröffentlicht: (2024)
von: Liu, Songming, et al.
Veröffentlicht: (2024)
Embodied Long Horizon Manipulation with Closed-loop Code Generation and Incremental Few-shot Adaptation
von: Meng, Yuan, et al.
Veröffentlicht: (2025)
von: Meng, Yuan, et al.
Veröffentlicht: (2025)
RDT2: Exploring the Scaling Limit of UMI Data Towards Zero-Shot Cross-Embodiment Generalization
von: Liu, Songming, et al.
Veröffentlicht: (2026)
von: Liu, Songming, et al.
Veröffentlicht: (2026)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
von: Zhao, Min, et al.
Veröffentlicht: (2024)
von: Zhao, Min, et al.
Veröffentlicht: (2024)
Rethinking Closed-loop Planning Framework for Imitation-based Model Integrating Prediction and Planning
von: Guo, Jiayu, et al.
Veröffentlicht: (2024)
von: Guo, Jiayu, et al.
Veröffentlicht: (2024)
GE-Sim 2.0: A Roadmap Towards Comprehensive Closed-loop Video World Simulators for Robotic Manipulation
von: Qiu, Boxiang, et al.
Veröffentlicht: (2026)
von: Qiu, Boxiang, et al.
Veröffentlicht: (2026)
FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Evaluating Uncertainty-based Failure Detection for Closed-Loop LLM Planners
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
Causal Composition Diffusion Model for Closed-loop Traffic Generation
von: Lin, Haohong, et al.
Veröffentlicht: (2024)
von: Lin, Haohong, et al.
Veröffentlicht: (2024)
Closed-loop Control of Steerable Balloon Endoscopes for Robot-assisted Transcatheter Intracardiac Procedures
von: McCandless, Max, et al.
Veröffentlicht: (2025)
von: McCandless, Max, et al.
Veröffentlicht: (2025)
Theoretical Closed-loop Stability Bounds for Dynamical System Coupled with Diffusion Policies
von: Lauzier, Gabriel, et al.
Veröffentlicht: (2025)
von: Lauzier, Gabriel, et al.
Veröffentlicht: (2025)
Learning Free Terminal Time Optimal Closed-loop Control of Manipulators
von: Hu, Wei, et al.
Veröffentlicht: (2023)
von: Hu, Wei, et al.
Veröffentlicht: (2023)
Promptable Closed-loop Traffic Simulation
von: Tan, Shuhan, et al.
Veröffentlicht: (2024)
von: Tan, Shuhan, et al.
Veröffentlicht: (2024)
Analyzing Key Objectives in Human-to-Robot Retargeting for Dexterous Manipulation
von: Xin, Chendong, et al.
Veröffentlicht: (2025)
von: Xin, Chendong, et al.
Veröffentlicht: (2025)
Towards High-Consistency Embodied World Model with Multi-View Trajectory Videos
von: Su, Taiyi, et al.
Veröffentlicht: (2025)
von: Su, Taiyi, et al.
Veröffentlicht: (2025)
Hindsight Planner: A Closed-Loop Few-Shot Planner for Embodied Instruction Following
von: Yang, Yuxiao, et al.
Veröffentlicht: (2024)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2024)
Domain Randomization for Robust, Affordable and Effective Closed-loop Control of Soft Robots
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2023)
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2023)
Evaluation as Evolution: Transforming Adversarial Diffusion into Closed-Loop Curricula for Autonomous Vehicles
von: Guo, Yicheng, et al.
Veröffentlicht: (2026)
von: Guo, Yicheng, et al.
Veröffentlicht: (2026)
EHC-MM: Embodied Holistic Control for Mobile Manipulation
von: Wang, Jiawen, et al.
Veröffentlicht: (2024)
von: Wang, Jiawen, et al.
Veröffentlicht: (2024)
EmbodiedCity: A Benchmark Platform for Embodied Agent in Real-world City Environment
von: Gao, Chen, et al.
Veröffentlicht: (2024)
von: Gao, Chen, et al.
Veröffentlicht: (2024)
CLEA: Closed-Loop Embodied Agent for Enhancing Task Execution in Dynamic Environments
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
von: Lei, Mingcong, et al.
Veröffentlicht: (2025)
Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
PlanAgent: A Multi-modal Large Language Agent for Closed-loop Vehicle Motion Planning
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Language-Guided Object-Centric Diffusion Policy for Generalizable and Collision-Aware Robotic Manipulation
von: Li, Hang, et al.
Veröffentlicht: (2024)
von: Li, Hang, et al.
Veröffentlicht: (2024)
EO-1: An Open Unified Embodied Foundation Model for General Robot Control
von: Qu, Delin, et al.
Veröffentlicht: (2025)
von: Qu, Delin, et al.
Veröffentlicht: (2025)
RealMirror: A Comprehensive, Open-Source Vision-Language-Action Platform for Embodied AI
von: Tai, Cong, et al.
Veröffentlicht: (2025)
von: Tai, Cong, et al.
Veröffentlicht: (2025)
PCASim: Promptable Closed-loop Adversarial Simulation for Urban Traffic Environment
von: Zhang, Chuancheng, et al.
Veröffentlicht: (2026)
von: Zhang, Chuancheng, et al.
Veröffentlicht: (2026)
$\mathcal{P}^3$: Toward Versatile Embodied Agents
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
Coherence-Driven Multimodal Safety Dialogue with Active Learning for Embodied Agents
von: Hassan, Sabit, et al.
Veröffentlicht: (2024)
von: Hassan, Sabit, et al.
Veröffentlicht: (2024)
Closed-loop Multi-step Planning
von: Lafratta, Giulia, et al.
Veröffentlicht: (2024)
von: Lafratta, Giulia, et al.
Veröffentlicht: (2024)
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models
von: Long, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: Long, Xiaoxiao, et al.
Veröffentlicht: (2025)
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis
von: Lang, Xiaolei, et al.
Veröffentlicht: (2026)
von: Lang, Xiaolei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Vidar: Embodied Video Diffusion Model for Generalist Manipulation
von: Feng, Yao, et al.
Veröffentlicht: (2025) -
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation
von: Tan, Hengkai, et al.
Veröffentlicht: (2025) -
Fourier Controller Networks for Real-Time Decision-Making in Embodied Learning
von: Tan, Hengkai, et al.
Veröffentlicht: (2024) -
ManiBox: Enhancing Embodied Spatial Generalization via Scalable Simulation Data Generations
von: Tan, Hengkai, et al.
Veröffentlicht: (2024) -
Motus: A Unified Latent Action World Model
von: Bi, Hongzhe, et al.
Veröffentlicht: (2025)