Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Chenyv, Tan, Wentao, Zhu, Lei, Li, Fengling, Li, Jingjing, Yang, Guoli, Shen, Heng Tao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MOTIF: Learning Action Motifs for Few-shot Cross-Embodiment Transfer
di: Zhi, Heng, et al.
Pubblicazione: (2026)
di: Zhi, Heng, et al.
Pubblicazione: (2026)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
di: Ye, Wencheng, et al.
Pubblicazione: (2025)
di: Ye, Wencheng, et al.
Pubblicazione: (2025)
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
di: Chen, Yipeng, et al.
Pubblicazione: (2026)
di: Chen, Yipeng, et al.
Pubblicazione: (2026)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
di: Yu, Wenda, et al.
Pubblicazione: (2026)
di: Yu, Wenda, et al.
Pubblicazione: (2026)
A Step Toward World Models: A Survey on Robotic Manipulation
di: Zhang, Peng-Fei, et al.
Pubblicazione: (2025)
di: Zhang, Peng-Fei, et al.
Pubblicazione: (2025)
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching
di: Chen, Jiayi, et al.
Pubblicazione: (2026)
di: Chen, Jiayi, et al.
Pubblicazione: (2026)
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
di: Shen, Boyang, et al.
Pubblicazione: (2026)
di: Shen, Boyang, et al.
Pubblicazione: (2026)
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models
di: Xu, Siyuan, et al.
Pubblicazione: (2026)
di: Xu, Siyuan, et al.
Pubblicazione: (2026)
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
di: Jiang, Anqing, et al.
Pubblicazione: (2025)
di: Jiang, Anqing, et al.
Pubblicazione: (2025)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
di: Shang, Shuyao, et al.
Pubblicazione: (2026)
di: Shang, Shuyao, et al.
Pubblicazione: (2026)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
di: Li, Chengmeng, et al.
Pubblicazione: (2025)
di: Li, Chengmeng, et al.
Pubblicazione: (2025)
Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning
di: Wen, Junjie, et al.
Pubblicazione: (2024)
di: Wen, Junjie, et al.
Pubblicazione: (2024)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
di: Li, Yajie, et al.
Pubblicazione: (2026)
di: Li, Yajie, et al.
Pubblicazione: (2026)
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention
di: Xiao, Lei, et al.
Pubblicazione: (2025)
di: Xiao, Lei, et al.
Pubblicazione: (2025)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
di: Yin, Zhenhan, et al.
Pubblicazione: (2025)
di: Yin, Zhenhan, et al.
Pubblicazione: (2025)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
di: Sun, Jingwen, et al.
Pubblicazione: (2026)
di: Sun, Jingwen, et al.
Pubblicazione: (2026)
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
di: Wang, Yating, et al.
Pubblicazione: (2025)
di: Wang, Yating, et al.
Pubblicazione: (2025)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
di: Li, Qiwei, et al.
Pubblicazione: (2026)
di: Li, Qiwei, et al.
Pubblicazione: (2026)
Uni-World VLA: Interleaved World Modeling and Planning for Autonomous Driving
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
di: Liu, Qiqi, et al.
Pubblicazione: (2026)
SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning
di: Wang, Hanzhen, et al.
Pubblicazione: (2025)
di: Wang, Hanzhen, et al.
Pubblicazione: (2025)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
di: Zhu, Minjie, et al.
Pubblicazione: (2025)
di: Zhu, Minjie, et al.
Pubblicazione: (2025)
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
di: Zhou, Zhongyi, et al.
Pubblicazione: (2025)
Sparse Imagination for Efficient Visual World Model Planning
di: Chun, Junha, et al.
Pubblicazione: (2025)
di: Chun, Junha, et al.
Pubblicazione: (2025)
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models
di: Zhang, Borong, et al.
Pubblicazione: (2025)
di: Zhang, Borong, et al.
Pubblicazione: (2025)
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
di: Zhang, Wenyao, et al.
Pubblicazione: (2025)
di: Zhang, Wenyao, et al.
Pubblicazione: (2025)
VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
di: Xu, Siyu, et al.
Pubblicazione: (2025)
di: Xu, Siyu, et al.
Pubblicazione: (2025)
PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding
di: Song, Wenxuan, et al.
Pubblicazione: (2025)
di: Song, Wenxuan, et al.
Pubblicazione: (2025)
GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
di: Qian, Jingjing, et al.
Pubblicazione: (2025)
di: Qian, Jingjing, et al.
Pubblicazione: (2025)
VLA-RFT: Vision-Language-Action Reinforcement Fine-tuning with Verified Rewards in World Simulators
di: Li, Hengtao, et al.
Pubblicazione: (2025)
di: Li, Hengtao, et al.
Pubblicazione: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
di: Wen, Junjie, et al.
Pubblicazione: (2024)
di: Wen, Junjie, et al.
Pubblicazione: (2024)
MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots
di: Huang, Ting, et al.
Pubblicazione: (2025)
di: Huang, Ting, et al.
Pubblicazione: (2025)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration
di: Zhang, Ninghao, et al.
Pubblicazione: (2026)
di: Zhang, Ninghao, et al.
Pubblicazione: (2026)
DriveWorld-VLA: Unified Latent-Space World Modeling with Vision-Language-Action for Autonomous Driving
di: jia, Feiyang, et al.
Pubblicazione: (2026)
di: jia, Feiyang, et al.
Pubblicazione: (2026)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
di: Li, Runhao, et al.
Pubblicazione: (2025)
di: Li, Runhao, et al.
Pubblicazione: (2025)
dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought
di: Wen, Junjie, et al.
Pubblicazione: (2025)
di: Wen, Junjie, et al.
Pubblicazione: (2025)
SABER: A Scalable Action-Based Embodied Dataset for Real-World VLA Adaptation
di: Menga, Narsimha, et al.
Pubblicazione: (2026)
di: Menga, Narsimha, et al.
Pubblicazione: (2026)
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
di: Zhang, Qiyao, et al.
Pubblicazione: (2026)
di: Zhang, Qiyao, et al.
Pubblicazione: (2026)
Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation
di: Heng, Liang, et al.
Pubblicazione: (2025)
di: Heng, Liang, et al.
Pubblicazione: (2025)
LLaDA-VLA: Vision Language Diffusion Action Models
di: Wen, Yuqing, et al.
Pubblicazione: (2025)
di: Wen, Yuqing, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MOTIF: Learning Action Motifs for Few-shot Cross-Embodiment Transfer
di: Zhi, Heng, et al.
Pubblicazione: (2026) -
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
di: Ye, Wencheng, et al.
Pubblicazione: (2025) -
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
di: Chen, Yipeng, et al.
Pubblicazione: (2026) -
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
di: Yu, Wenda, et al.
Pubblicazione: (2026) -
A Step Toward World Models: A Survey on Robotic Manipulation
di: Zhang, Peng-Fei, et al.
Pubblicazione: (2025)