EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Chushan, Lu, Ruihan, Tong, Jinguang, Li, Xuesong, Wang, Yikai, Li, Hongdong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3D-IDE: 3D Implicit Depth Emergent
by: Zhang, Chushan, et al.
Published: (2026)
by: Zhang, Chushan, et al.
Published: (2026)
Implicit Action Chunking for Smooth Continuous Control
by: Liang, Bosun, et al.
Published: (2026)
by: Liang, Bosun, et al.
Published: (2026)
Reinforcement Learning with Action Chunking
by: Li, Qiyang, et al.
Published: (2025)
by: Li, Qiyang, et al.
Published: (2025)
Learning Native Continuation for Action Chunking Flow Policies
by: Liu, Yufeng, et al.
Published: (2026)
by: Liu, Yufeng, et al.
Published: (2026)
Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control
by: Posadas-Nava, Alejandro, et al.
Published: (2025)
by: Posadas-Nava, Alejandro, et al.
Published: (2025)
Causal Scene Narration with Runtime Safety Supervision for Vision-Language-Action Driving
by: Li, Yun, et al.
Published: (2026)
by: Li, Yun, et al.
Published: (2026)
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
by: Sun, Yuteng, et al.
Published: (2026)
by: Sun, Yuteng, et al.
Published: (2026)
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
by: Chen, Lingling, et al.
Published: (2026)
by: Chen, Lingling, et al.
Published: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
by: Miao, Cui, et al.
Published: (2025)
by: Miao, Cui, et al.
Published: (2025)
Improving Robotic Manipulation Robustness via NICE Scene Surgery
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
Leave No Observation Behind: Real-time Correction for VLA Action Chunks
by: Sendai, Kohei, et al.
Published: (2025)
by: Sendai, Kohei, et al.
Published: (2025)
Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling
by: Liu, Yuejiang, et al.
Published: (2024)
by: Liu, Yuejiang, et al.
Published: (2024)
Open-Vocabulary Spatio-Temporal Scene Graph for Robot Perception and Teleoperation Planning
by: Wang, Yi, et al.
Published: (2025)
by: Wang, Yi, et al.
Published: (2025)
CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors
by: Li, Dachong, et al.
Published: (2026)
by: Li, Dachong, et al.
Published: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
by: Zhang, Rongyu, et al.
Published: (2025)
by: Zhang, Rongyu, et al.
Published: (2025)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
by: Wang, Qiuyue, et al.
Published: (2026)
by: Wang, Qiuyue, et al.
Published: (2026)
Personalized Robotic Object Rearrangement from Scene Context
by: Ramachandruni, Kartik, et al.
Published: (2025)
by: Ramachandruni, Kartik, et al.
Published: (2025)
VacuumVLA: Boosting VLA Capabilities via a Unified Suction and Gripping Tool for Complex Robotic Manipulation
by: Zhou, Hui, et al.
Published: (2025)
by: Zhou, Hui, et al.
Published: (2025)
Emerging Extrinsic Dexterity in Cluttered Scenes via Dynamics-aware Policy Learning
by: Zheng, Yixin, et al.
Published: (2026)
by: Zheng, Yixin, et al.
Published: (2026)
Mixture of Horizons in Action Chunking
by: Jing, Dong, et al.
Published: (2025)
by: Jing, Dong, et al.
Published: (2025)
SmoothVLA: Aligning Vision-Language-Action Models with Physical Constraints via Intrinsic Smoothness Optimization
by: Li, Jiashun, et al.
Published: (2026)
by: Li, Jiashun, et al.
Published: (2026)
Training-Time Action Conditioning for Efficient Real-Time Chunking
by: Black, Kevin, et al.
Published: (2025)
by: Black, Kevin, et al.
Published: (2025)
WorldVLA: Towards Autoregressive Action World Model
by: Cen, Jun, et al.
Published: (2025)
by: Cen, Jun, et al.
Published: (2025)
VeriGraph: Scene Graphs for Execution Verifiable Robot Planning
by: Ekpo, Daniel, et al.
Published: (2024)
by: Ekpo, Daniel, et al.
Published: (2024)
Bi-LAT: Bilateral Control-Based Imitation Learning via Natural Language and Action Chunking with Transformers
by: Kobayashi, Takumi, et al.
Published: (2025)
by: Kobayashi, Takumi, et al.
Published: (2025)
Actor-Critic for Continuous Action Chunks: A Reinforcement Learning Framework for Long-Horizon Robotic Manipulation with Sparse Reward
by: Yang, Jiarui, et al.
Published: (2025)
by: Yang, Jiarui, et al.
Published: (2025)
LiteVLA-Edge: Quantized On-Device Multimodal Control for Embedded Robotics
by: Williams, Justin, et al.
Published: (2026)
by: Williams, Justin, et al.
Published: (2026)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
by: Jiang, Hanxiao, et al.
Published: (2024)
by: Jiang, Hanxiao, et al.
Published: (2024)
Toward Accurate Long-Horizon Robotic Manipulation: Language-to-Action with Foundation Models via Scene Graphs
by: Dinesh, Sushil Samuel, et al.
Published: (2025)
by: Dinesh, Sushil Samuel, et al.
Published: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
by: Zeng, Qixin, et al.
Published: (2026)
by: Zeng, Qixin, et al.
Published: (2026)
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
by: Han, Gaoge, et al.
Published: (2026)
by: Han, Gaoge, et al.
Published: (2026)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)
by: Zhao, Zhikai, et al.
Published: (2026)
VLA-AN: An Efficient and Onboard Vision-Language-Action Framework for Aerial Navigation in Complex Environments
by: Wu, Yuze, et al.
Published: (2025)
by: Wu, Yuze, et al.
Published: (2025)
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
by: Liu, Ruixun, et al.
Published: (2025)
by: Liu, Ruixun, et al.
Published: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
Learning Bimanual Manipulation via Action Chunking and Inter-Arm Coordination with Transformers
by: Motoda, Tomohiro, et al.
Published: (2025)
by: Motoda, Tomohiro, et al.
Published: (2025)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
by: Li, Haoyun, et al.
Published: (2025)
by: Li, Haoyun, et al.
Published: (2025)
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
by: Hu, Yutong, et al.
Published: (2026)
by: Hu, Yutong, et al.
Published: (2026)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
by: Yan, Hongyu, et al.
Published: (2026)
by: Yan, Hongyu, et al.
Published: (2026)
Similar Items
-
3D-IDE: 3D Implicit Depth Emergent
by: Zhang, Chushan, et al.
Published: (2026) -
Implicit Action Chunking for Smooth Continuous Control
by: Liang, Bosun, et al.
Published: (2026) -
Reinforcement Learning with Action Chunking
by: Li, Qiyang, et al.
Published: (2025) -
Learning Native Continuation for Action Chunking Flow Policies
by: Liu, Yufeng, et al.
Published: (2026) -
Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control
by: Posadas-Nava, Alejandro, et al.
Published: (2025)