TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Yuteng, Wang, Haoran, Bai, Ruofei, Li, Zhengguo, Li, Jun, Yee, Meng, Chuah, Yau, Wei Yun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph-based SLAM-Aware Exploration with Prior Topo-Metric Information
by: Bai, Ruofei, et al.
Published: (2023)
by: Bai, Ruofei, et al.
Published: (2023)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
by: Hu, Yucheng, et al.
Published: (2026)
by: Hu, Yucheng, et al.
Published: (2026)
Realm: Real-Time Line-of-Sight Maintenance in Multi-Robot Navigation with Unknown Obstacles
by: Bai, Ruofei, et al.
Published: (2025)
by: Bai, Ruofei, et al.
Published: (2025)
Interleave-VLA: Enhancing Robot Manipulation with Interleaved Image-Text Instructions
by: Fan, Cunxin, et al.
Published: (2025)
by: Fan, Cunxin, et al.
Published: (2025)
Plane Detection and Ranking via Model Information Optimization
by: Zhong, Daoxin, et al.
Published: (2025)
by: Zhong, Daoxin, et al.
Published: (2025)
BlockVLA: Accelerating Autoregressive VLA via Block Diffusion Finetuning
by: Wang, Ruiheng, et al.
Published: (2026)
by: Wang, Ruiheng, et al.
Published: (2026)
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
by: Li, Boyu, et al.
Published: (2026)
by: Li, Boyu, et al.
Published: (2026)
Traversability analysis with vision and terrain probing for safe legged robot navigation
by: Haddeler, Garen, et al.
Published: (2022)
by: Haddeler, Garen, et al.
Published: (2022)
Beyond Imitation: Reinforcement Learning Fine-Tuning for Adaptive Diffusion Navigation Policies
by: Sheng, Junhe, et al.
Published: (2026)
by: Sheng, Junhe, et al.
Published: (2026)
ImagiNav: Scalable Embodied Navigation via Generative Visual Prediction and Inverse Dynamics
by: Chen, Jie, et al.
Published: (2026)
by: Chen, Jie, et al.
Published: (2026)
Collaborative Graph Exploration with Reduced Pose-SLAM Uncertainty via Submodular Optimization
by: Bai, Ruofei, et al.
Published: (2024)
by: Bai, Ruofei, et al.
Published: (2024)
Line-of-Sight-Constrained Multi-Robot Mapless Navigation via Polygonal Visible Regions
by: Bai, Ruofei, et al.
Published: (2026)
by: Bai, Ruofei, et al.
Published: (2026)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
by: Liu, Qiqi, et al.
Published: (2026)
by: Liu, Qiqi, et al.
Published: (2026)
World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy
by: Liu, Xiaokang, et al.
Published: (2026)
by: Liu, Xiaokang, et al.
Published: (2026)
NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models
by: Zhu, Ziyue, et al.
Published: (2026)
by: Zhu, Ziyue, et al.
Published: (2026)
LLaDA-VLA: Vision Language Diffusion Action Models
by: Wen, Yuqing, et al.
Published: (2025)
by: Wen, Yuqing, et al.
Published: (2025)
Open-Loop Planning, Closed-Loop Verification: Speculative Verification for VLA
by: Wang, Zihua, et al.
Published: (2026)
by: Wang, Zihua, et al.
Published: (2026)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
by: Sun, Nan, et al.
Published: (2025)
by: Sun, Nan, et al.
Published: (2025)
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models
by: Li, Yixuan, et al.
Published: (2025)
by: Li, Yixuan, et al.
Published: (2025)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
by: Shen, Boyang, et al.
Published: (2026)
by: Shen, Boyang, et al.
Published: (2026)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
by: Sun, Jianli, et al.
Published: (2026)
by: Sun, Jianli, et al.
Published: (2026)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
by: Huang, Zhiyu, et al.
Published: (2026)
by: Huang, Zhiyu, et al.
Published: (2026)
D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
by: Guo, Yucheng, et al.
Published: (2026)
by: Guo, Yucheng, et al.
Published: (2026)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
by: Zhu, Juan, et al.
Published: (2026)
by: Zhu, Juan, et al.
Published: (2026)
Make Your VLA More Robust Without More Data By Interleaving Motion Planning
by: Choe, Dan BW, et al.
Published: (2026)
by: Choe, Dan BW, et al.
Published: (2026)
PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance
by: Zheng, Yupeng, et al.
Published: (2026)
by: Zheng, Yupeng, et al.
Published: (2026)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
by: Zhong, Linqing, et al.
Published: (2026)
by: Zhong, Linqing, et al.
Published: (2026)
VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
by: Zhang, Chushan, et al.
Published: (2026)
by: Zhang, Chushan, et al.
Published: (2026)
ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
by: Li, Puhao, et al.
Published: (2025)
by: Li, Puhao, et al.
Published: (2025)
Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
by: Bai, Shuanghao, et al.
Published: (2026)
by: Bai, Shuanghao, et al.
Published: (2026)
ElegantVLA: Learning When to Think for Efficient Vision-Language-Action Models
by: Li, Ye, et al.
Published: (2026)
by: Li, Ye, et al.
Published: (2026)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
by: Liang, Zhixuan, et al.
Published: (2025)
by: Liang, Zhixuan, et al.
Published: (2025)
RynnVLA-002: A Unified Vision-Language-Action and World Model
by: Cen, Jun, et al.
Published: (2025)
by: Cen, Jun, et al.
Published: (2025)
WorldVLA: Towards Autoregressive Action World Model
by: Cen, Jun, et al.
Published: (2025)
by: Cen, Jun, et al.
Published: (2025)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
by: Luo, Jingzhou, et al.
Published: (2026)
by: Luo, Jingzhou, et al.
Published: (2026)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
by: Li, Qiwei, et al.
Published: (2026)
by: Li, Qiwei, et al.
Published: (2026)
SmoothVLA: Aligning Vision-Language-Action Models with Physical Constraints via Intrinsic Smoothness Optimization
by: Li, Jiashun, et al.
Published: (2026)
by: Li, Jiashun, et al.
Published: (2026)
Similar Items
-
Graph-based SLAM-Aware Exploration with Prior Topo-Metric Information
by: Bai, Ruofei, et al.
Published: (2023) -
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
by: Hu, Yucheng, et al.
Published: (2026) -
Realm: Real-Time Line-of-Sight Maintenance in Multi-Robot Navigation with Unknown Obstacles
by: Bai, Ruofei, et al.
Published: (2025) -
Interleave-VLA: Enhancing Robot Manipulation with Interleaved Image-Text Instructions
by: Fan, Cunxin, et al.
Published: (2025) -
Plane Detection and Ranking via Model Information Optimization
by: Zhong, Daoxin, et al.
Published: (2025)