FPC-VLA: A Vision-Language-Action Framework with a Supervisor for Failure Prediction and Correction
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Yifan, Duan, Zhixiang, Xie, Tianshi, Cao, Fuyu, Shen, Pinxi, Song, Peili, Jin, Piaopiao, Sun, Guokang, Xu, Shaoqing, You, Yangwei, Liu, Jingtai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Dual-Actor Fine-Tuning of VLA Models: A Talk-and-Tweak Human-in-the-Loop Approach
di: Jin, Piaopiao, et al.
Pubblicazione: (2025)
di: Jin, Piaopiao, et al.
Pubblicazione: (2025)
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
di: Yang, Yifan, et al.
Pubblicazione: (2025)
di: Yang, Yifan, et al.
Pubblicazione: (2025)
Simulating Automotive Radar with Lidar and Camera Inputs
di: Song, Peili, et al.
Pubblicazione: (2025)
di: Song, Peili, et al.
Pubblicazione: (2025)
AcL: Action Learner for Fault-Tolerant Quadruped Locomotion Control
di: Xu, Tianyu, et al.
Pubblicazione: (2025)
di: Xu, Tianyu, et al.
Pubblicazione: (2025)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
di: Yu, Wenda, et al.
Pubblicazione: (2026)
di: Yu, Wenda, et al.
Pubblicazione: (2026)
QuoVLA: Quotient Space for Vision-Language-Action Models
di: Wang, Xuan, et al.
Pubblicazione: (2026)
di: Wang, Xuan, et al.
Pubblicazione: (2026)
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
di: Sun, Lin, et al.
Pubblicazione: (2025)
di: Sun, Lin, et al.
Pubblicazione: (2025)
PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation
di: Cao, Haofan, et al.
Pubblicazione: (2026)
di: Cao, Haofan, et al.
Pubblicazione: (2026)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
di: Wang, Guokang, et al.
Pubblicazione: (2024)
di: Wang, Guokang, et al.
Pubblicazione: (2024)
LaST-VLA: Thinking in Latent Spatio-Temporal Space for Vision-Language-Action in Autonomous Driving
di: Luo, Yuechen, et al.
Pubblicazione: (2026)
di: Luo, Yuechen, et al.
Pubblicazione: (2026)
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
di: Shen, Boyang, et al.
Pubblicazione: (2026)
di: Shen, Boyang, et al.
Pubblicazione: (2026)
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
di: Chen, Lingling, et al.
Pubblicazione: (2026)
di: Chen, Lingling, et al.
Pubblicazione: (2026)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
di: Jin, Ruofan, et al.
Pubblicazione: (2026)
di: Jin, Ruofan, et al.
Pubblicazione: (2026)
EdgeVLA: Efficient Vision-Language-Action Models
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
STRONG-VLA: Decoupled Robustness Learning for Vision-Language-Action Models under Multimodal Perturbations
di: Xie, Yuhan, et al.
Pubblicazione: (2026)
di: Xie, Yuhan, et al.
Pubblicazione: (2026)
FutureVLA: Joint Visuomotor Prediction for Vision-Language-Action Model
di: Xu, Xiaoxu, et al.
Pubblicazione: (2026)
di: Xu, Xiaoxu, et al.
Pubblicazione: (2026)
Unleashing VLA Potentials in Autonomous Driving via Explicit Learning from Failures
di: Luo, Yuechen, et al.
Pubblicazione: (2026)
di: Luo, Yuechen, et al.
Pubblicazione: (2026)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
di: Luo, Jingzhou, et al.
Pubblicazione: (2026)
di: Luo, Jingzhou, et al.
Pubblicazione: (2026)
Vision-Language-Action (VLA) Models: Concepts, Progress, Applications and Challenges
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
di: Sapkota, Ranjan, et al.
Pubblicazione: (2025)
Counterfactual VLA: Self-Reflective Vision-Language-Action Model with Adaptive Reasoning
di: Peng, Zhenghao "Mark", et al.
Pubblicazione: (2025)
di: Peng, Zhenghao "Mark", et al.
Pubblicazione: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
di: Wei, Xiangyi, et al.
Pubblicazione: (2025)
di: Wei, Xiangyi, et al.
Pubblicazione: (2025)
Certifying optimality in nonconvex robust PCA
di: Gong, Pinxi, et al.
Pubblicazione: (2026)
di: Gong, Pinxi, et al.
Pubblicazione: (2026)
VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model
di: Li, Wenhao, et al.
Pubblicazione: (2026)
di: Li, Wenhao, et al.
Pubblicazione: (2026)
Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
di: Bai, Shuanghao, et al.
Pubblicazione: (2026)
di: Bai, Shuanghao, et al.
Pubblicazione: (2026)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
di: Ye, Wencheng, et al.
Pubblicazione: (2025)
di: Ye, Wencheng, et al.
Pubblicazione: (2025)
DyDexHandover: Human-like Bimanual Dynamic Dexterous Handover using RGB-only Perception
di: Zhou, Haoran, et al.
Pubblicazione: (2025)
di: Zhou, Haoran, et al.
Pubblicazione: (2025)
RedVLA: Physical Red Teaming for Vision-Language-Action Models
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
di: Guo, Xinyu, et al.
Pubblicazione: (2026)
di: Guo, Xinyu, et al.
Pubblicazione: (2026)
OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning
di: Lin, Fanqi, et al.
Pubblicazione: (2025)
di: Lin, Fanqi, et al.
Pubblicazione: (2025)
Pure Vision Language Action (VLA) Models: A Comprehensive Survey
di: Zhang, Dapeng, et al.
Pubblicazione: (2025)
di: Zhang, Dapeng, et al.
Pubblicazione: (2025)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
di: Liu, Chenyv, et al.
Pubblicazione: (2026)
di: Liu, Chenyv, et al.
Pubblicazione: (2026)
OpenVLA: An Open-Source Vision-Language-Action Model
di: Kim, Moo Jin, et al.
Pubblicazione: (2024)
di: Kim, Moo Jin, et al.
Pubblicazione: (2024)
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
di: Du, Zhiying, et al.
Pubblicazione: (2025)
di: Du, Zhiying, et al.
Pubblicazione: (2025)
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models
di: Li, Yixuan, et al.
Pubblicazione: (2025)
di: Li, Yixuan, et al.
Pubblicazione: (2025)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
di: Li, Qiwei, et al.
Pubblicazione: (2026)
di: Li, Qiwei, et al.
Pubblicazione: (2026)
ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models
di: Zhong, Linqing, et al.
Pubblicazione: (2026)
di: Zhong, Linqing, et al.
Pubblicazione: (2026)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
di: Xie, Haozhe, et al.
Pubblicazione: (2026)
di: Xie, Haozhe, et al.
Pubblicazione: (2026)
CRL-VLA: Continual Vision-Language-Action Learning
di: Zeng, Qixin, et al.
Pubblicazione: (2026)
di: Zeng, Qixin, et al.
Pubblicazione: (2026)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
di: Deng, Shengliang, et al.
Pubblicazione: (2025)
di: Deng, Shengliang, et al.
Pubblicazione: (2025)
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models
di: Xu, Siyuan, et al.
Pubblicazione: (2026)
di: Xu, Siyuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Dual-Actor Fine-Tuning of VLA Models: A Talk-and-Tweak Human-in-the-Loop Approach
di: Jin, Piaopiao, et al.
Pubblicazione: (2025) -
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
di: Yang, Yifan, et al.
Pubblicazione: (2025) -
Simulating Automotive Radar with Lidar and Camera Inputs
di: Song, Peili, et al.
Pubblicazione: (2025) -
AcL: Action Learner for Fault-Tolerant Quadruped Locomotion Control
di: Xu, Tianyu, et al.
Pubblicazione: (2025) -
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
di: Yu, Wenda, et al.
Pubblicazione: (2026)