Efficient Test-Time Adaptation of Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Karmanov, Adilbek, Guan, Dayan, Lu, Shijian, Saddik, Abdulmotaleb El, Xing, Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dimensional Coactivation for Representational Consistency in Frozen Vision Foundation Models
von: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Veröffentlicht: (2026)
von: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Veröffentlicht: (2026)
From ChatGPT to DeepSeek AI: A Comprehensive Analysis of Evolution, Deviation, and Future Implications in AI-Language Models
von: Singh, Simrandeep, et al.
Veröffentlicht: (2025)
von: Singh, Simrandeep, et al.
Veröffentlicht: (2025)
Weakly Supervised 3D Open-vocabulary Segmentation
von: Liu, Kunhao, et al.
Veröffentlicht: (2023)
von: Liu, Kunhao, et al.
Veröffentlicht: (2023)
Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026)
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026)
TextureMeDefect: LLM-based Defect Texture Generation for Railway Components on Mobile Devices
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2024)
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2024)
Real-Time Oriented Object Detection Transformer in Remote Sensing Images
von: Ding, Zeyu, et al.
Veröffentlicht: (2026)
von: Ding, Zeyu, et al.
Veröffentlicht: (2026)
Content-Adaptive Image Retouching Guided by Attribute-Based Text Representation
von: Zhu, Hancheng, et al.
Veröffentlicht: (2025)
von: Zhu, Hancheng, et al.
Veröffentlicht: (2025)
Realistic Test-Time Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
Bayesian Test-Time Adaptation for Vision-Language Models
von: Zhou, Lihua, et al.
Veröffentlicht: (2025)
von: Zhou, Lihua, et al.
Veröffentlicht: (2025)
LOD-Net: Locality-Aware 3D Object Detection Using Multi-Scale Transformer Network
von: Khan, Mustaqeem, et al.
Veröffentlicht: (2026)
von: Khan, Mustaqeem, et al.
Veröffentlicht: (2026)
VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors
von: Belal, Atif, et al.
Veröffentlicht: (2025)
von: Belal, Atif, et al.
Veröffentlicht: (2025)
Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
von: Sreenivas, Manogna, et al.
Veröffentlicht: (2024)
von: Sreenivas, Manogna, et al.
Veröffentlicht: (2024)
EVOKE: Emotion Enabled Virtual Avatar Mapping Using Optimized Knowledge Distillation
von: Nadeem, Maryam, et al.
Veröffentlicht: (2024)
von: Nadeem, Maryam, et al.
Veröffentlicht: (2024)
Vision-Language Models for Vision Tasks: A Survey
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
Ultra-Light Test-Time Adaptation for Vision--Language Models
von: Kim, Byunghyun
Veröffentlicht: (2025)
von: Kim, Byunghyun
Veröffentlicht: (2025)
Prototype-Based Test-Time Adaptation of Vision-Language Models
von: Huang, Zhaohong, et al.
Veröffentlicht: (2026)
von: Huang, Zhaohong, et al.
Veröffentlicht: (2026)
Online Gaussian Test-Time Adaptation of Vision-Language Models
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
Negation-Aware Test-Time Adaptation for Vision-Language Models
von: Han, Haochen, et al.
Veröffentlicht: (2025)
von: Han, Haochen, et al.
Veröffentlicht: (2025)
Flatness Guided Test-Time Adaptation for Vision-Language Models
von: Li, Aodi, et al.
Veröffentlicht: (2025)
von: Li, Aodi, et al.
Veröffentlicht: (2025)
ETTA: Efficient Test-Time Adaptation for Vision-Language Models through Dynamic Embedding Updates
von: Dastmalchi, Hamidreza, et al.
Veröffentlicht: (2025)
von: Dastmalchi, Hamidreza, et al.
Veröffentlicht: (2025)
SOAP: Style-Omniscient Animatable Portraits
von: Liao, Tingting, et al.
Veröffentlicht: (2025)
von: Liao, Tingting, et al.
Veröffentlicht: (2025)
DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis
von: Gu, Yuming, et al.
Veröffentlicht: (2025)
von: Gu, Yuming, et al.
Veröffentlicht: (2025)
RQFormer: Rotated Query Transformer for End-to-End Oriented Object Detection
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2023)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2023)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
von: Nguyen, Khanh-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Khanh-Binh, et al.
Veröffentlicht: (2025)
OrientedFormer: An End-to-End Transformer-Based Oriented Object Detector in Remote Sensing Images
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2024)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2024)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
Test-Time Adaptation of Vision-Language Models for Open-Vocabulary Semantic Segmentation
von: Noori, Mehrdad, et al.
Veröffentlicht: (2025)
von: Noori, Mehrdad, et al.
Veröffentlicht: (2025)
Mitigating Cache Noise in Test-Time Adaptation for Large Vision-Language Models
von: Zhai, Haotian, et al.
Veröffentlicht: (2025)
von: Zhai, Haotian, et al.
Veröffentlicht: (2025)
CLIPTTA: Robust Contrastive Vision-Language Test-Time Adaptation
von: Lafon, Marc, et al.
Veröffentlicht: (2025)
von: Lafon, Marc, et al.
Veröffentlicht: (2025)
Efficient Test-Time Prompt Tuning for Vision-Language Models
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
Frustratingly Easy Test-Time Adaptation of Vision-Language Models
von: Farina, Matteo, et al.
Veröffentlicht: (2024)
von: Farina, Matteo, et al.
Veröffentlicht: (2024)
Advancing Reliable Test-Time Adaptation of Vision-Language Models under Visual Variations
von: Liang, Yiwen, et al.
Veröffentlicht: (2025)
von: Liang, Yiwen, et al.
Veröffentlicht: (2025)
Bidirectional Prototype-Reward co-Evolution for Test-Time Adaptation of Vision-Language Models
von: Qiao, Xiaozhen, et al.
Veröffentlicht: (2025)
von: Qiao, Xiaozhen, et al.
Veröffentlicht: (2025)
A Lost Opportunity for Vision-Language Models: A Comparative Study of Online Test-Time Adaptation for Vision-Language Models
von: Döbler, Mario, et al.
Veröffentlicht: (2024)
von: Döbler, Mario, et al.
Veröffentlicht: (2024)
Characterizing Continual Learning Scenarios and Strategies for Audio Analysis
von: Bhatt, Ruchi, et al.
Veröffentlicht: (2024)
von: Bhatt, Ruchi, et al.
Veröffentlicht: (2024)
Empirical Studies of Large Scale Environment Scanning by Consumer Electronics
von: Wang, Mengyuan, et al.
Veröffentlicht: (2025)
von: Wang, Mengyuan, et al.
Veröffentlicht: (2025)
Fast-Slow Test-Time Adaptation for Online Vision-and-Language Navigation
von: Gao, Junyu, et al.
Veröffentlicht: (2023)
von: Gao, Junyu, et al.
Veröffentlicht: (2023)
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models
von: Hu, Xuefeng, et al.
Veröffentlicht: (2024)
von: Hu, Xuefeng, et al.
Veröffentlicht: (2024)
Class-Aware Prototype Learning with Negative Contrast for Test-Time Adaptation of Vision-Language Models
von: Qiao, Xiaozhen, et al.
Veröffentlicht: (2025)
von: Qiao, Xiaozhen, et al.
Veröffentlicht: (2025)
Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift
von: Chen, Lixian, et al.
Veröffentlicht: (2026)
von: Chen, Lixian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Dimensional Coactivation for Representational Consistency in Frozen Vision Foundation Models
von: Saddik, Izaldein Al-Zyoud Abdulmotaleb El
Veröffentlicht: (2026) -
From ChatGPT to DeepSeek AI: A Comprehensive Analysis of Evolution, Deviation, and Future Implications in AI-Language Models
von: Singh, Simrandeep, et al.
Veröffentlicht: (2025) -
Weakly Supervised 3D Open-vocabulary Segmentation
von: Liu, Kunhao, et al.
Veröffentlicht: (2023) -
Segmentation-Guided Spatial Indexing for Generalizable and Explainable Deepfake Detection
von: Al-Zyoud, Izaldein, et al.
Veröffentlicht: (2026) -
TextureMeDefect: LLM-based Defect Texture Generation for Railway Components on Mobile Devices
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2024)