DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Sixu, Qing, Yunpeng, Liu, Litao, Zhou, Ming, Jin, Ruixing, Fan, Xiaoyi, Liu, Guiliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Grounding Sim-to-Real Generalization in Dexterous Manipulation: An Empirical Study with Vision-Language-Action Models
von: Jin, Ruixing, et al.
Veröffentlicht: (2026)
von: Jin, Ruixing, et al.
Veröffentlicht: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
von: Yang, Zebin, et al.
Veröffentlicht: (2026)
von: Yang, Zebin, et al.
Veröffentlicht: (2026)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
von: Li, Meng, et al.
Veröffentlicht: (2025)
von: Li, Meng, et al.
Veröffentlicht: (2025)
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
von: Zhai, Shaopeng, et al.
Veröffentlicht: (2025)
von: Zhai, Shaopeng, et al.
Veröffentlicht: (2025)
From Reaction to Anticipation: Proactive Failure Recovery through Agentic Task Graph for Robotic Manipulation
von: Xu, Sheng, et al.
Veröffentlicht: (2026)
von: Xu, Sheng, et al.
Veröffentlicht: (2026)
UnderwaterVLA: Dual-brain Vision-Language-Action architecture for Autonomous Underwater Navigation
von: Wang, Zhangyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhangyuan, et al.
Veröffentlicht: (2025)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
HWC-Loco: A Hierarchical Whole-Body Control Approach to Robust Humanoid Locomotion
von: Lin, Sixu, et al.
Veröffentlicht: (2025)
von: Lin, Sixu, et al.
Veröffentlicht: (2025)
VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
von: Si, Shengyu, et al.
Veröffentlicht: (2026)
von: Si, Shengyu, et al.
Veröffentlicht: (2026)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
ElegantVLA: Learning When to Think for Efficient Vision-Language-Action Models
von: Li, Ye, et al.
Veröffentlicht: (2026)
von: Li, Ye, et al.
Veröffentlicht: (2026)
SignBot: Learning Human-to-Humanoid Sign Language Interaction
von: Qiao, Guanren, et al.
Veröffentlicht: (2025)
von: Qiao, Guanren, et al.
Veröffentlicht: (2025)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
von: Huang, Zhiyu, et al.
Veröffentlicht: (2026)
von: Huang, Zhiyu, et al.
Veröffentlicht: (2026)
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
von: Wang, Yating, et al.
Veröffentlicht: (2025)
von: Wang, Yating, et al.
Veröffentlicht: (2025)
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
RedVLA: Physical Red Teaming for Vision-Language-Action Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2026)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
UrbanVLA: A Vision-Language-Action Model for Urban Micromobility
von: Li, Anqi, et al.
Veröffentlicht: (2025)
von: Li, Anqi, et al.
Veröffentlicht: (2025)
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
von: Sun, Lin, et al.
Veröffentlicht: (2025)
von: Sun, Lin, et al.
Veröffentlicht: (2025)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
von: Zhu, Xiang, et al.
Veröffentlicht: (2026)
DyPho-SLAM : Real-time Photorealistic SLAM in Dynamic Environments
von: Liu, Yi, et al.
Veröffentlicht: (2025)
von: Liu, Yi, et al.
Veröffentlicht: (2025)
RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models
von: Liufu, Weijia, et al.
Veröffentlicht: (2026)
von: Liufu, Weijia, et al.
Veröffentlicht: (2026)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
FPC-VLA: A Vision-Language-Action Framework with a Supervisor for Failure Prediction and Correction
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
von: Yang, Yifan, et al.
Veröffentlicht: (2025)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
von: Sun, Nan, et al.
Veröffentlicht: (2025)
von: Sun, Nan, et al.
Veröffentlicht: (2025)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
von: Jin, Ruofan, et al.
Veröffentlicht: (2026)
von: Jin, Ruofan, et al.
Veröffentlicht: (2026)
MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent
von: Fu, Yuxia, et al.
Veröffentlicht: (2025)
von: Fu, Yuxia, et al.
Veröffentlicht: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
ST4VLA: Spatially Guided Training for Vision-Language-Action Models
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
von: Ye, Jinhui, et al.
Veröffentlicht: (2026)
GAT-Grasp: Gesture-Driven Affordance Transfer for Task-Aware Robotic Grasping
von: Wang, Ruixiang, et al.
Veröffentlicht: (2025)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2025)
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
von: Li, Xiaoqi, et al.
Veröffentlicht: (2026)
von: Li, Xiaoqi, et al.
Veröffentlicht: (2026)
RynnVLA-002: A Unified Vision-Language-Action and World Model
von: Cen, Jun, et al.
Veröffentlicht: (2025)
von: Cen, Jun, et al.
Veröffentlicht: (2025)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
von: Sun, Jianli, et al.
Veröffentlicht: (2026)
von: Sun, Jianli, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Grounding Sim-to-Real Generalization in Dexterous Manipulation: An Empirical Study with Vision-Language-Action Models
von: Jin, Ruixing, et al.
Veröffentlicht: (2026) -
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
von: Zheng, Zihao, et al.
Veröffentlicht: (2026) -
DySL-VLA: Efficient Vision-Language-Action Model Inference via Dynamic-Static Layer-Skipping for Robot Manipulation
von: Yang, Zebin, et al.
Veröffentlicht: (2026) -
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
von: Li, Meng, et al.
Veröffentlicht: (2025) -
A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning
von: Zhai, Shaopeng, et al.
Veröffentlicht: (2025)