ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Rushuai, Wang, Hecheng, Liu, Chiming, Yan, Xiaohan, Wang, Yunlong, Du, Xuan, Yue, Shuoyu, Liu, Yongcheng, Zhang, Chuheng, Qi, Lizhe, Chen, Yi, Shan, Wei, Yao, Maoqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
von: Zhong, Linqing, et al.
Veröffentlicht: (2026)
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
EnerVerse-AC: Envisioning Embodied Environments with Action Condition
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025)
Continually Evolving Skill Knowledge in Vision Language Action Model
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)
Interactive Post-Training for Vision-Language-Action Models
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
von: Tan, Shuhan, et al.
Veröffentlicht: (2025)
RoboRetriever: Single-Camera Robot Object Retrieval via Active and Interactive Perception with Dynamic Scene Graph
von: Wang, Hecheng, et al.
Veröffentlicht: (2025)
von: Wang, Hecheng, et al.
Veröffentlicht: (2025)
CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models
von: Liu, Zhi
Veröffentlicht: (2026)
von: Liu, Zhi
Veröffentlicht: (2026)
CF-VLA: Efficient Coarse-to-Fine Action Generation for Vision-Language-Action Policies
von: Du, Fan, et al.
Veröffentlicht: (2026)
von: Du, Fan, et al.
Veröffentlicht: (2026)
METIS: Multi-Source Egocentric Training for Integrated Dexterous Vision-Language-Action Model
von: Fu, Yankai, et al.
Veröffentlicht: (2025)
von: Fu, Yankai, et al.
Veröffentlicht: (2025)
Context-Action Embedding Learning for Off-Policy Evaluation in Contextual Bandits
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
von: Chandak, Kushagra, et al.
Veröffentlicht: (2025)
Toward Embodiment Equivariant Vision-Language-Action Policy
von: Chen, Anzhe, et al.
Veröffentlicht: (2025)
von: Chen, Anzhe, et al.
Veröffentlicht: (2025)
SOP: A Scalable Online Post-Training System for Vision-Language-Action Models
von: Pan, Mingjie, et al.
Veröffentlicht: (2026)
von: Pan, Mingjie, et al.
Veröffentlicht: (2026)
Learning Action Embeddings for Off-Policy Evaluation
von: Cief, Matej, et al.
Veröffentlicht: (2023)
von: Cief, Matej, et al.
Veröffentlicht: (2023)
HBVLA: Pushing 1-Bit Post-Training Quantization for Vision-Language-Action Models
von: Yan, Xin, et al.
Veröffentlicht: (2026)
von: Yan, Xin, et al.
Veröffentlicht: (2026)
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
von: Zhang, Jingxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Jingxuan, et al.
Veröffentlicht: (2026)
POTEC: Off-Policy Learning for Large Action Spaces via Two-Stage Policy Decomposition
von: Saito, Yuta, et al.
Veröffentlicht: (2024)
von: Saito, Yuta, et al.
Veröffentlicht: (2024)
SkeletonAgent: An Agentic Interaction Framework for Skeleton-based Action Recognition
von: Liu, Hongda, et al.
Veröffentlicht: (2025)
von: Liu, Hongda, et al.
Veröffentlicht: (2025)
Zero-shot Action Localization via the Confidence of Large Vision-Language Models
von: Aklilu, Josiah, et al.
Veröffentlicht: (2024)
von: Aklilu, Josiah, et al.
Veröffentlicht: (2024)
Off-OAB: Off-Policy Policy Gradient Method with Optimal Action-Dependent Baseline
von: Meng, Wenjia, et al.
Veröffentlicht: (2024)
von: Meng, Wenjia, et al.
Veröffentlicht: (2024)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
How Do VLAs Effectively Inherit from VLMs?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models
von: Xu, Siyuan, et al.
Veröffentlicht: (2026)
von: Xu, Siyuan, et al.
Veröffentlicht: (2026)
VITA: Vision-to-Action Flow Matching Policy
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
von: Gao, Dechen, et al.
Veröffentlicht: (2025)
VAM: Verbalized Action Masking for Controllable Exploration in RL Post-Training -- A Chess Case Study
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning
von: Zhao, Shiwan, et al.
Veröffentlicht: (2026)
von: Zhao, Shiwan, et al.
Veröffentlicht: (2026)
Efficient Off-Policy Learning for High-Dimensional Action Spaces
von: Otto, Fabian, et al.
Veröffentlicht: (2024)
von: Otto, Fabian, et al.
Veröffentlicht: (2024)
Bayesian Off-Policy Evaluation and Learning for Large Action Spaces
von: Aouali, Imad, et al.
Veröffentlicht: (2024)
von: Aouali, Imad, et al.
Veröffentlicht: (2024)
ActionStudio: A Lightweight Framework for Data and Training of Large Action Models
von: Zhang, Jianguo, et al.
Veröffentlicht: (2025)
von: Zhang, Jianguo, et al.
Veröffentlicht: (2025)
Training One Model to Master Cross-Level Agentic Actions via Reinforcement Learning
von: He, Kaichen, et al.
Veröffentlicht: (2025)
von: He, Kaichen, et al.
Veröffentlicht: (2025)
SRPO: Self-Referential Policy Optimization for Vision-Language-Action Models
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
von: Fei, Senyu, et al.
Veröffentlicht: (2025)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
von: Liang, Zhixuan, et al.
Veröffentlicht: (2025)
Frame Order Matters: A Temporal Sequence-Aware Model for Few-Shot Action Recognition
von: Li, Bozheng, et al.
Veröffentlicht: (2024)
von: Li, Bozheng, et al.
Veröffentlicht: (2024)
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
von: Neary, Cyrus, et al.
Veröffentlicht: (2025)
von: Neary, Cyrus, et al.
Veröffentlicht: (2025)
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
von: Han, Gaoge, et al.
Veröffentlicht: (2026)
von: Han, Gaoge, et al.
Veröffentlicht: (2026)
QuoVLA: Quotient Space for Vision-Language-Action Models
von: Wang, Xuan, et al.
Veröffentlicht: (2026)
von: Wang, Xuan, et al.
Veröffentlicht: (2026)
What Do Latent Action Models Actually Learn?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
FreezeVLA: Action-Freezing Attacks against Vision-Language-Action Models
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
DropVLA: An Action-Level Backdoor Attack on Vision-Language-Action Models
von: Xu, Zonghuan, et al.
Veröffentlicht: (2025)
von: Xu, Zonghuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025) -
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
von: Zhong, Linqing, et al.
Veröffentlicht: (2026) -
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
von: Yang, Rushuai, et al.
Veröffentlicht: (2025) -
EnerVerse-AC: Envisioning Embodied Environments with Action Condition
von: Jiang, Yuxin, et al.
Veröffentlicht: (2025) -
Continually Evolving Skill Knowledge in Vision Language Action Model
von: Wu, Yuxuan, et al.
Veröffentlicht: (2025)