Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zang, Zehua, Wang, Xi, Sun, Fuchun, Xu, Xiao, Lium, Lixiang, Zhou, Jiahuan, Li, Jiangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BayesTTA: Continual-Temporal Test-Time Adaptation for Vision-Language Models via Gaussian Discriminant Analysis
by: Cui, Shuang, et al.
Published: (2025)
by: Cui, Shuang, et al.
Published: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
by: Song, Fei, et al.
Published: (2025)
by: Song, Fei, et al.
Published: (2025)
Vision-Language Attribute Disentanglement and Reinforcement for Lifelong Person Re-Identification
by: Xu, Kunlun, et al.
Published: (2026)
by: Xu, Kunlun, et al.
Published: (2026)
RS-SSM: Refining Forgotten Specifics in State Space Model for Video Semantic Segmentation
by: Zhu, Kai, et al.
Published: (2026)
by: Zhu, Kai, et al.
Published: (2026)
Rethinking Generalizability and Discriminability of Self-Supervised Learning from Evolutionary Game Theory Perspective
by: Li, Jiangmeng, et al.
Published: (2024)
by: Li, Jiangmeng, et al.
Published: (2024)
Rethinking Misalignment in Vision-Language Model Adaptation from a Causal Perspective
by: Zhang, Yanan, et al.
Published: (2024)
by: Zhang, Yanan, et al.
Published: (2024)
Self-Reinforcing Prototype Evolution with Dual-Knowledge Cooperation for Semi-Supervised Lifelong Person Re-Identification
by: Xu, Kunlun, et al.
Published: (2025)
by: Xu, Kunlun, et al.
Published: (2025)
C^2Prompt: Class-aware Client Knowledge Interaction for Federated Continual Learning
by: Xu, Kunlun, et al.
Published: (2025)
by: Xu, Kunlun, et al.
Published: (2025)
Intention Action Anticipation Model with Guide-Feedback Loop Mechanism
by: Ma, Zongnan, et al.
Published: (2024)
by: Ma, Zongnan, et al.
Published: (2024)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
by: Li, Qiwei, et al.
Published: (2026)
by: Li, Qiwei, et al.
Published: (2026)
All-in-One Image Restoration via Causal-Deconfounding Wavelet-Disentangled Prompt Network
by: Wang, Bingnan, et al.
Published: (2026)
by: Wang, Bingnan, et al.
Published: (2026)
EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models
by: Bai, Zechen, et al.
Published: (2025)
by: Bai, Zechen, et al.
Published: (2025)
Learning Invariant Causal Mechanism from Vision-Language Models
by: Song, Zeen, et al.
Published: (2024)
by: Song, Zeen, et al.
Published: (2024)
On the Generalization and Causal Explanation in Self-Supervised Learning
by: Qiang, Wenwen, et al.
Published: (2024)
by: Qiang, Wenwen, et al.
Published: (2024)
On the Transferability and Discriminability of Repersentation Learning in Unsupervised Domain Adaptation
by: Qiang, Wenwen, et al.
Published: (2025)
by: Qiang, Wenwen, et al.
Published: (2025)
Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration
by: Xu, Jinglin, et al.
Published: (2026)
by: Xu, Jinglin, et al.
Published: (2026)
SCAP: Transductive Test-Time Adaptation via Supportive Clique-based Attribute Prompting
by: Zhang, Chenyu, et al.
Published: (2025)
by: Zhang, Chenyu, et al.
Published: (2025)
Flatness Guided Test-Time Adaptation for Vision-Language Models
by: Li, Aodi, et al.
Published: (2025)
by: Li, Aodi, et al.
Published: (2025)
AmPLe: Supporting Vision-Language Models via Adaptive-Debiased Ensemble Multi-Prompt Learning
by: Song, Fei, et al.
Published: (2025)
by: Song, Fei, et al.
Published: (2025)
Bayesian Test-Time Adaptation for Vision-Language Models
by: Zhou, Lihua, et al.
Published: (2025)
by: Zhou, Lihua, et al.
Published: (2025)
TTRV: Test-Time Reinforcement Learning for Vision Language Models
by: Singh, Akshit, et al.
Published: (2025)
by: Singh, Akshit, et al.
Published: (2025)
CAPrompt: Cyclic Prompt Aggregation for Pre-Trained Model Based Class Incremental Learning
by: Li, Qiwei, et al.
Published: (2024)
by: Li, Qiwei, et al.
Published: (2024)
Class-aware Domain Knowledge Fusion and Fission for Continual Test-Time Adaptation
by: Zhou, Jiahuan, et al.
Published: (2025)
by: Zhou, Jiahuan, et al.
Published: (2025)
Test-Time Training for Visual Foresight Vision-Language-Action Models
by: Park, Sangwu, et al.
Published: (2026)
by: Park, Sangwu, et al.
Published: (2026)
Supporting Vision-Language Model Inference with Confounder-pruning Knowledge Prompt
by: Li, Jiangmeng, et al.
Published: (2022)
by: Li, Jiangmeng, et al.
Published: (2022)
Efficient Test-Time Prompt Tuning for Vision-Language Models
by: Zhu, Yuhan, et al.
Published: (2024)
by: Zhu, Yuhan, et al.
Published: (2024)
Test-Time Consistency in Vision Language Models
by: Chou, Shih-Han, et al.
Published: (2025)
by: Chou, Shih-Han, et al.
Published: (2025)
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
by: Qiao, Xiaozhen, et al.
Published: (2025)
by: Qiao, Xiaozhen, et al.
Published: (2025)
On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations
by: Guo, Jianing, et al.
Published: (2025)
by: Guo, Jianing, et al.
Published: (2025)
Vision Graph Prompting via Semantic Low-Rank Decomposition
by: Ai, Zixiang, et al.
Published: (2025)
by: Ai, Zixiang, et al.
Published: (2025)
MAIN-VLA: Modeling Abstraction of Intention and eNvironment for Vision-Language-Action Models
by: Zhou, Zheyuan, et al.
Published: (2026)
by: Zhou, Zheyuan, et al.
Published: (2026)
Realistic Test-Time Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
Efficient Test-Time Adaptation of Vision-Language Models
by: Karmanov, Adilbek, et al.
Published: (2024)
by: Karmanov, Adilbek, et al.
Published: (2024)
Attention-Guided Patch-Wise Sparse Adversarial Attacks on Vision-Language-Action Models
by: Zhang, Naifu, et al.
Published: (2025)
by: Zhang, Naifu, et al.
Published: (2025)
DriveMA: Driving Vision-Language-Action Models with verifiable Meta-Actions
by: Zheng, Weicheng, et al.
Published: (2026)
by: Zheng, Weicheng, et al.
Published: (2026)
CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models
by: Song, Wenxuan, et al.
Published: (2026)
by: Song, Wenxuan, et al.
Published: (2026)
GAPrompt: Geometry-Aware Point Cloud Prompt for 3D Vision Model
by: Ai, Zixiang, et al.
Published: (2025)
by: Ai, Zixiang, et al.
Published: (2025)
LoRA-TTT: Low-Rank Test-Time Training for Vision-Language Models
by: Kojima, Yuto, et al.
Published: (2025)
by: Kojima, Yuto, et al.
Published: (2025)
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
by: Zhou, Xueyang, et al.
Published: (2025)
by: Zhou, Xueyang, et al.
Published: (2025)
Test-Time Hinting for Black-Box Vision-Language Models
by: Hou, Kaihua, et al.
Published: (2026)
by: Hou, Kaihua, et al.
Published: (2026)
Similar Items
-
BayesTTA: Continual-Temporal Test-Time Adaptation for Vision-Language Models via Gaussian Discriminant Analysis
by: Cui, Shuang, et al.
Published: (2025) -
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
by: Song, Fei, et al.
Published: (2025) -
Vision-Language Attribute Disentanglement and Reinforcement for Lifelong Person Re-Identification
by: Xu, Kunlun, et al.
Published: (2026) -
RS-SSM: Refining Forgotten Specifics in State Space Model for Video Semantic Segmentation
by: Zhu, Kai, et al.
Published: (2026) -
Rethinking Generalizability and Discriminability of Self-Supervised Learning from Evolutionary Game Theory Perspective
by: Li, Jiangmeng, et al.
Published: (2024)