RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Haoran, Guo, Yongjian, Guan, Zhong, Di, Shuai, Bai, Xiaodong, Long, Jing, Zhao, Tianyun, Luo, Mingxi, Zhao, Hongke, Wu, Likang, Deng, Xiaotie, Chu, Xu, Xiao, Xi, Wen, Sheng, Gong, Yicheng, Xiong, Junwu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction
von: Guan, Zhong, et al.
Veröffentlicht: (2026)
von: Guan, Zhong, et al.
Veröffentlicht: (2026)
D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
Sword: Style-Robust World Models as Simulators via Dynamic Latent Bootstrapping for VLA Policy Post-Training
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
VLA-RAIL: A Real-Time Asynchronous Inference Linker for VLA Models and Robots
von: Zhao, Yongsheng, et al.
Veröffentlicht: (2025)
von: Zhao, Yongsheng, et al.
Veröffentlicht: (2025)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
von: Li, Haozhan, et al.
Veröffentlicht: (2025)
Attention Mechanisms Perspective: Exploring LLM Processing of Graph-Structured Data
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
Enhancing Collaborative Semantics of Language Model-Driven Recommendations via Graph-Aware Learning
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
LangTopo: Aligning Language Descriptions of Graphs with Tokenized Topological Modeling
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
Recall-Extend Dynamics: Enhancing Small Language Models through Controlled Exploration and Refined Offline Integration
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
AsyncVLA: An Asynchronous VLA for Fast and Robust Navigation on the Edge
von: Hirose, Noriaki, et al.
Veröffentlicht: (2026)
von: Hirose, Noriaki, et al.
Veröffentlicht: (2026)
BlockVLA: Accelerating Autoregressive VLA via Block Diffusion Finetuning
von: Wang, Ruiheng, et al.
Veröffentlicht: (2026)
von: Wang, Ruiheng, et al.
Veröffentlicht: (2026)
ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
von: Chen, Yuhui, et al.
Veröffentlicht: (2025)
von: Chen, Yuhui, et al.
Veröffentlicht: (2025)
Multi-View Empowered Structural Graph Wordification for Language Models
von: Liu, Zipeng, et al.
Veröffentlicht: (2024)
von: Liu, Zipeng, et al.
Veröffentlicht: (2024)
How Social is It? A Benchmark for LLMs' Capabilities in Multi-user Multi-turn Social Agent Tasks
von: Wu, Yusen, et al.
Veröffentlicht: (2025)
von: Wu, Yusen, et al.
Veröffentlicht: (2025)
Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning
von: Liu, Yuan, et al.
Veröffentlicht: (2026)
von: Liu, Yuan, et al.
Veröffentlicht: (2026)
Hierarchical Semantic RL: Tackling the Problem of Dynamic Action Space for RL-based Recommendations
von: Wang, Minmao, et al.
Veröffentlicht: (2025)
von: Wang, Minmao, et al.
Veröffentlicht: (2025)
NS-VLA: Towards Neuro-Symbolic Vision-Language-Action Models
von: Zhu, Ziyue, et al.
Veröffentlicht: (2026)
von: Zhu, Ziyue, et al.
Veröffentlicht: (2026)
Reinventing Clinical Dialogue: Agentic Paradigms for LLM Enabled Healthcare Communication
von: Zhi, Xiaoquan, et al.
Veröffentlicht: (2025)
von: Zhi, Xiaoquan, et al.
Veröffentlicht: (2025)
GANPrompt: Enhancing Robustness in LLM-Based Recommendations with GAN-Enhanced Diversity Prompts
von: Li, Xinyu, et al.
Veröffentlicht: (2024)
von: Li, Xinyu, et al.
Veröffentlicht: (2024)
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL
von: Jiang, Zhennan, et al.
Veröffentlicht: (2026)
von: Jiang, Zhennan, et al.
Veröffentlicht: (2026)
AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
von: Guo, Yucheng, et al.
Veröffentlicht: (2026)
SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
Thousand-GPU Large-Scale Training and Optimization Recipe for AI-Native Cloud Embodied Intelligence Infrastructure
von: Guo, Yongjian, et al.
Veröffentlicht: (2026)
von: Guo, Yongjian, et al.
Veröffentlicht: (2026)
Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
von: Wei, Yifei, et al.
Veröffentlicht: (2026)
LANE: Logic Alignment of Non-tuning Large Language Models and Online Recommendation Systems for Explainable Reason Generation
von: Zhao, Hongke, et al.
Veröffentlicht: (2024)
von: Zhao, Hongke, et al.
Veröffentlicht: (2024)
AsyncShield: A Plug-and-Play Edge Adapter for Asynchronous Cloud-based VLA Navigation
von: Yang, Kai, et al.
Veröffentlicht: (2026)
von: Yang, Kai, et al.
Veröffentlicht: (2026)
A Pragmatic VLA Foundation Model
von: Wu, Wei, et al.
Veröffentlicht: (2026)
von: Wu, Wei, et al.
Veröffentlicht: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
DepthVLA: Enhancing Vision-Language-Action Models with Depth-Aware Spatial Reasoning
von: Yuan, Tianyuan, et al.
Veröffentlicht: (2025)
von: Yuan, Tianyuan, et al.
Veröffentlicht: (2025)
EchoVLA: Synergistic Declarative Memory for VLA-Driven Mobile Manipulation
von: Lin, Min, et al.
Veröffentlicht: (2025)
von: Lin, Min, et al.
Veröffentlicht: (2025)
NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models
von: Huang, Wen, et al.
Veröffentlicht: (2026)
von: Huang, Wen, et al.
Veröffentlicht: (2026)
World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy
von: Liu, Xiaokang, et al.
Veröffentlicht: (2026)
von: Liu, Xiaokang, et al.
Veröffentlicht: (2026)
VLA-GSE: Boosting Parameter-Efficient Fine-Tuning in VLA with Generalized and Specialized Experts
von: Jiang, Yuhua, et al.
Veröffentlicht: (2026)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2026)
DroneVLA: VLA based Aerial Manipulation
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
What Can RL Bring to VLA Generalization? An Empirical Study
von: Liu, Jijia, et al.
Veröffentlicht: (2025)
von: Liu, Jijia, et al.
Veröffentlicht: (2025)
Refined Policy Distillation: From VLA Generalists to RL Experts
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
von: Jülg, Tobias, et al.
Veröffentlicht: (2025)
What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models
von: Peng, Yuanfang, et al.
Veröffentlicht: (2026)
von: Peng, Yuanfang, et al.
Veröffentlicht: (2026)
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
von: Ye, Guo, et al.
Veröffentlicht: (2025)
von: Ye, Guo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction
von: Guan, Zhong, et al.
Veröffentlicht: (2026) -
D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models
von: Guo, Yucheng, et al.
Veröffentlicht: (2026) -
Sword: Style-Robust World Models as Simulators via Dynamic Latent Bootstrapping for VLA Policy Post-Training
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026) -
VLA-RAIL: A Real-Time Asynchronous Inference Linker for VLA Models and Robots
von: Zhao, Yongsheng, et al.
Veröffentlicht: (2025) -
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
von: Li, Haozhan, et al.
Veröffentlicht: (2025)