Bidirectional Prototype-Reward co-Evolution for Test-Time Adaptation of Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Qiao, Xiaozhen, Huang, Peng, Yuan, Jiakang, Guo, Xianda, Ye, Bowen, Xue, Chaocan, Zheng, Ye, Sun, Zhe, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
by: Qiao, Xiaozhen, et al.
Published: (2025)
by: Qiao, Xiaozhen, et al.
Published: (2025)
ComKD-CLIP: Comprehensive Knowledge Distillation for Contrastive Language-Image Pre-traning Model
by: Chen, Yifan, et al.
Published: (2024)
by: Chen, Yifan, et al.
Published: (2024)
Mitigating Long-Tail Bias in HOI Detection via Adaptive Diversity Cache
by: Jiang, Yuqiu, et al.
Published: (2025)
by: Jiang, Yuqiu, et al.
Published: (2025)
Prototype-Based Test-Time Adaptation of Vision-Language Models
by: Huang, Zhaohong, et al.
Published: (2026)
by: Huang, Zhaohong, et al.
Published: (2026)
Test-Time Adaptation for Tactile-Vision-Language Models
by: Ye, Chuyang, et al.
Published: (2026)
by: Ye, Chuyang, et al.
Published: (2026)
Bayesian Test-Time Adaptation for Vision-Language Models
by: Zhou, Lihua, et al.
Published: (2025)
by: Zhou, Lihua, et al.
Published: (2025)
Bridging Perception and Planning: Towards End-to-End Planning for Signal Temporal Logic Tasks
by: Ye, Bowen, et al.
Published: (2025)
by: Ye, Bowen, et al.
Published: (2025)
VLRMBench: A Comprehensive and Challenging Benchmark for Vision-Language Reward Models
by: Ruan, Jiacheng, et al.
Published: (2025)
by: Ruan, Jiacheng, et al.
Published: (2025)
Test-Time Adaptation with CLIP Reward for Zero-Shot Generalization in Vision-Language Models
by: Zhao, Shuai, et al.
Published: (2023)
by: Zhao, Shuai, et al.
Published: (2023)
Advancing Visual Reliability: Color-Accurate Underwater Image Enhancement for Real-Time Underwater Missions
by: Zhou, Yiqiang, et al.
Published: (2026)
by: Zhou, Yiqiang, et al.
Published: (2026)
Similarity-Guided Layer-Adaptive Vision Transformer for UAV Tracking
by: Xue, Chaocan, et al.
Published: (2025)
by: Xue, Chaocan, et al.
Published: (2025)
Discover Your Neighbors: Advanced Stable Test-Time Adaptation in Dynamic World
by: Jiang, Qinting, et al.
Published: (2024)
by: Jiang, Qinting, et al.
Published: (2024)
DOTA: Distributional Test-Time Adaptation of Vision-Language Models
by: Han, Zongbo, et al.
Published: (2024)
by: Han, Zongbo, et al.
Published: (2024)
From Captions to Rewards (CAREVL): Leveraging Large Language Model Experts for Enhanced Reward Modeling in Large Vision-Language Models
by: Dai, Muzhi, et al.
Published: (2025)
by: Dai, Muzhi, et al.
Published: (2025)
ProtoTTA: Prototype-Guided Test-Time Adaptation
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
Decoupled Prototype Learning for Reliable Test-Time Adaptation
by: Wang, Guowei, et al.
Published: (2024)
by: Wang, Guowei, et al.
Published: (2024)
Advancing Test-Time Adaptation in Wild Acoustic Test Settings
by: Liu, Hongfu, et al.
Published: (2023)
by: Liu, Hongfu, et al.
Published: (2023)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence
by: An, Hongjun, et al.
Published: (2026)
by: An, Hongjun, et al.
Published: (2026)
CREST: Cross-modal Resonance through Evidential Deep Learning for Enhanced Zero-Shot Learning
by: Huang, Haojian, et al.
Published: (2024)
by: Huang, Haojian, et al.
Published: (2024)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
by: Zhang, Ye, et al.
Published: (2026)
by: Zhang, Ye, et al.
Published: (2026)
Adaptive Cascading Network for Continual Test-Time Adaptation
by: Nguyen, Kien X., et al.
Published: (2024)
by: Nguyen, Kien X., et al.
Published: (2024)
Vision-Language Model Selection and Reuse for Downstream Adaptation
by: Tan, Hao-Zhe, et al.
Published: (2025)
by: Tan, Hao-Zhe, et al.
Published: (2025)
Realistic Test-Time Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
Noisy Test-Time Adaptation in Vision-Language Models
by: Cao, Chentao, et al.
Published: (2025)
by: Cao, Chentao, et al.
Published: (2025)
Efficient Test-Time Adaptation of Vision-Language Models
by: Karmanov, Adilbek, et al.
Published: (2024)
by: Karmanov, Adilbek, et al.
Published: (2024)
ThanoRA: Task Heterogeneity-Aware Multi-Task Low-Rank Adaptation
by: Liang, Jian, et al.
Published: (2025)
by: Liang, Jian, et al.
Published: (2025)
Feature-Based Instance Neighbor Discovery: Advanced Stable Test-Time Adaptation in Dynamic World
by: Jiang, Qinting, et al.
Published: (2025)
by: Jiang, Qinting, et al.
Published: (2025)
LSTM-MAS: A Long Short-Term Memory Inspired Multi-Agent System for Long-Context Understanding
by: Jiang, Yichen, et al.
Published: (2026)
by: Jiang, Yichen, et al.
Published: (2026)
Ultra-Light Test-Time Adaptation for Vision--Language Models
by: Kim, Byunghyun
Published: (2025)
by: Kim, Byunghyun
Published: (2025)
Online Gaussian Test-Time Adaptation of Vision-Language Models
by: Fuchs, Clément, et al.
Published: (2025)
by: Fuchs, Clément, et al.
Published: (2025)
CLIPTTA: Robust Contrastive Vision-Language Test-Time Adaptation
by: Lafon, Marc, et al.
Published: (2025)
by: Lafon, Marc, et al.
Published: (2025)
Negation-Aware Test-Time Adaptation for Vision-Language Models
by: Han, Haochen, et al.
Published: (2025)
by: Han, Haochen, et al.
Published: (2025)
Frustratingly Easy Test-Time Adaptation of Vision-Language Models
by: Farina, Matteo, et al.
Published: (2024)
by: Farina, Matteo, et al.
Published: (2024)
Flatness Guided Test-Time Adaptation for Vision-Language Models
by: Li, Aodi, et al.
Published: (2025)
by: Li, Aodi, et al.
Published: (2025)
TTP: Test-Time Padding for Adversarial Detection and Robust Adaptation on Vision-Language Models
by: Li, Zhiwei, et al.
Published: (2025)
by: Li, Zhiwei, et al.
Published: (2025)
Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models
by: Chen, Xinyu, et al.
Published: (2025)
by: Chen, Xinyu, et al.
Published: (2025)
All-in-One: Transferring Vision Foundation Models into Stereo Matching
by: Zhou, Jingyi, et al.
Published: (2024)
by: Zhou, Jingyi, et al.
Published: (2024)
BoostAdapter: Improving Vision-Language Test-Time Adaptation via Regional Bootstrapping
by: Zhang, Taolin, et al.
Published: (2024)
by: Zhang, Taolin, et al.
Published: (2024)
When Normality Shifts: Risk-Aware Test-Time Adaptation for Unsupervised Tabular Anomaly Detection
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
Similar Items
-
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
by: Qiao, Xiaozhen, et al.
Published: (2025) -
ComKD-CLIP: Comprehensive Knowledge Distillation for Contrastive Language-Image Pre-traning Model
by: Chen, Yifan, et al.
Published: (2024) -
Mitigating Long-Tail Bias in HOI Detection via Adaptive Diversity Cache
by: Jiang, Yuqiu, et al.
Published: (2025) -
Prototype-Based Test-Time Adaptation of Vision-Language Models
by: Huang, Zhaohong, et al.
Published: (2026) -
Test-Time Adaptation for Tactile-Vision-Language Models
by: Ye, Chuyang, et al.
Published: (2026)