CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Dachong, Chen, ZhuangZhuang, Zhang, Jin, Li, Jianqiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
por: Qu, Delin, et al.
Publicado: (2025)
por: Qu, Delin, et al.
Publicado: (2025)
SmoothVLA: Aligning Vision-Language-Action Models with Physical Constraints via Intrinsic Smoothness Optimization
por: Li, Jiashun, et al.
Publicado: (2026)
por: Li, Jiashun, et al.
Publicado: (2026)
SEVO: Semantic-Enhanced Virtual Observation for Robust VLA Manipulation via Active Illumination and Data-Centric Collection
por: Fang, Tianchonghui, et al.
Publicado: (2026)
por: Fang, Tianchonghui, et al.
Publicado: (2026)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
por: Liu, Chenyv, et al.
Publicado: (2026)
por: Liu, Chenyv, et al.
Publicado: (2026)
DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping
por: Zhong, Yifan, et al.
Publicado: (2025)
por: Zhong, Yifan, et al.
Publicado: (2025)
WorldVLA: Towards Autoregressive Action World Model
por: Cen, Jun, et al.
Publicado: (2025)
por: Cen, Jun, et al.
Publicado: (2025)
SA-VLA: Spatially-Aware Flow-Matching for Vision-Language-Action Reinforcement Learning
por: Pan, Xu, et al.
Publicado: (2026)
por: Pan, Xu, et al.
Publicado: (2026)
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning
por: Zhang, Borong, et al.
Publicado: (2025)
por: Zhang, Borong, et al.
Publicado: (2025)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
por: Jin, Ruofan, et al.
Publicado: (2026)
por: Jin, Ruofan, et al.
Publicado: (2026)
EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
por: Zhang, Chushan, et al.
Publicado: (2026)
por: Zhang, Chushan, et al.
Publicado: (2026)
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
por: Sun, Yuteng, et al.
Publicado: (2026)
por: Sun, Yuteng, et al.
Publicado: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
por: Zhang, Rongyu, et al.
Publicado: (2025)
por: Zhang, Rongyu, et al.
Publicado: (2025)
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
por: Hu, Xintong, et al.
Publicado: (2026)
por: Hu, Xintong, et al.
Publicado: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
por: Zeng, Qixin, et al.
Publicado: (2026)
por: Zeng, Qixin, et al.
Publicado: (2026)
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
por: Han, Gaoge, et al.
Publicado: (2026)
por: Han, Gaoge, et al.
Publicado: (2026)
VLA-AN: An Efficient and Onboard Vision-Language-Action Framework for Aerial Navigation in Complex Environments
por: Wu, Yuze, et al.
Publicado: (2025)
por: Wu, Yuze, et al.
Publicado: (2025)
VacuumVLA: Boosting VLA Capabilities via a Unified Suction and Gripping Tool for Complex Robotic Manipulation
por: Zhou, Hui, et al.
Publicado: (2025)
por: Zhou, Hui, et al.
Publicado: (2025)
Can Explicit Physical Feasibility Benefit VLA Learning? An Empirical Study
por: Wei, Yubai, et al.
Publicado: (2026)
por: Wei, Yubai, et al.
Publicado: (2026)
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
por: Liu, Ruixun, et al.
Publicado: (2025)
por: Liu, Ruixun, et al.
Publicado: (2025)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
por: Xiong, Zheng, et al.
Publicado: (2025)
por: Xiong, Zheng, et al.
Publicado: (2025)
AnchorRefine: Synergy-Manipulation Based on Trajectory Anchor and Residual Refinement for Vision-Language-Action Models
por: Jia, Tingzheng, et al.
Publicado: (2026)
por: Jia, Tingzheng, et al.
Publicado: (2026)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
por: Wang, Qiuyue, et al.
Publicado: (2026)
por: Wang, Qiuyue, et al.
Publicado: (2026)
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models
por: Hu, Yutong, et al.
Publicado: (2026)
por: Hu, Yutong, et al.
Publicado: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
por: Miao, Cui, et al.
Publicado: (2025)
por: Miao, Cui, et al.
Publicado: (2025)
Pure Vision Language Action (VLA) Models: A Comprehensive Survey
por: Zhang, Dapeng, et al.
Publicado: (2025)
por: Zhang, Dapeng, et al.
Publicado: (2025)
End-to-End Dexterous Arm-Hand VLA Policies via Shared Autonomy: VR Teleoperation Augmented by Autonomous Hand VLA Policy for Efficient Data Collection
por: Cui, Yu, et al.
Publicado: (2025)
por: Cui, Yu, et al.
Publicado: (2025)
DropVLA: An Action-Level Backdoor Attack on Vision-Language-Action Models
por: Xu, Zonghuan, et al.
Publicado: (2025)
por: Xu, Zonghuan, et al.
Publicado: (2025)
EndoVLA: Dual-Phase Vision-Language-Action Model for Autonomous Tracking in Endoscopy
por: Ng, Chi Kit, et al.
Publicado: (2025)
por: Ng, Chi Kit, et al.
Publicado: (2025)
UrbanVLA: A Vision-Language-Action Model for Urban Micromobility
por: Li, Anqi, et al.
Publicado: (2025)
por: Li, Anqi, et al.
Publicado: (2025)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
por: Zhao, Ziyan, et al.
Publicado: (2025)
por: Zhao, Ziyan, et al.
Publicado: (2025)
Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion
por: Li, Zhuo, et al.
Publicado: (2025)
por: Li, Zhuo, et al.
Publicado: (2025)
RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models
por: Liufu, Weijia, et al.
Publicado: (2026)
por: Liufu, Weijia, et al.
Publicado: (2026)
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
por: Chen, Xinyi, et al.
Publicado: (2025)
por: Chen, Xinyi, et al.
Publicado: (2025)
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
por: Chen, Lingling, et al.
Publicado: (2026)
por: Chen, Lingling, et al.
Publicado: (2026)
Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment
por: Zhou, Kaijun, et al.
Publicado: (2026)
por: Zhou, Kaijun, et al.
Publicado: (2026)
Efficient Multi-Task Reinforcement Learning via Task-Specific Action Correction
por: Feng, Jinyuan, et al.
Publicado: (2024)
por: Feng, Jinyuan, et al.
Publicado: (2024)
Event-Grounded Sparse Autoencoders for Vision-Language-Action Policies
por: Jin, Xinchen, et al.
Publicado: (2026)
por: Jin, Xinchen, et al.
Publicado: (2026)
Sci-VLA: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
por: Pang, Yiwen, et al.
Publicado: (2026)
por: Pang, Yiwen, et al.
Publicado: (2026)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
por: Su, Taiyi, et al.
Publicado: (2026)
por: Su, Taiyi, et al.
Publicado: (2026)
VLA Models Are More Generalizable Than You Think: Revisiting Physical and Spatial Modeling
por: Li, Weiqi, et al.
Publicado: (2025)
por: Li, Weiqi, et al.
Publicado: (2025)
Ejemplares similares
-
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
por: Qu, Delin, et al.
Publicado: (2025) -
SmoothVLA: Aligning Vision-Language-Action Models with Physical Constraints via Intrinsic Smoothness Optimization
por: Li, Jiashun, et al.
Publicado: (2026) -
SEVO: Semantic-Enhanced Virtual Observation for Robust VLA Manipulation via Active Illumination and Data-Centric Collection
por: Fang, Tianchonghui, et al.
Publicado: (2026) -
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
por: Liu, Chenyv, et al.
Publicado: (2026) -
DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping
por: Zhong, Yifan, et al.
Publicado: (2025)