Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Khanh-Binh, Bui, Phuoc-Nguyen, Choo, Hyunseung, Nguyen, Duc Thanh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accelerating Conditional Prompt Learning via Masked Image Modeling for Vision-Language Models
by: Bui, Phuoc-Nguyen, et al.
Published: (2025)
by: Bui, Phuoc-Nguyen, et al.
Published: (2025)
Attn-Adapter: Attention Is All You Need for Online Few-shot Learner of Vision-Language Model
by: Bui, Phuoc-Nguyen, et al.
Published: (2025)
by: Bui, Phuoc-Nguyen, et al.
Published: (2025)
Unsupervised Domain Adaptation with SAM-RefiSeR for Enhanced Brain Tumor Segmentation
by: Imans, Dillan, et al.
Published: (2026)
by: Imans, Dillan, et al.
Published: (2026)
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis
by: Bui, Phuoc-Nguyen, et al.
Published: (2024)
by: Bui, Phuoc-Nguyen, et al.
Published: (2024)
Representation Learning with Semantic-aware Instance and Sparse Token Alignments
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
Clinical Graph-Mediated Distillation for Unpaired MRI-to-CFI Hypertension Prediction
by: Imans, Dillan, et al.
Published: (2026)
by: Imans, Dillan, et al.
Published: (2026)
Frequency Adapter with SAM for Generalized Medical Image Segmentation
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
Response-Aware Multimodal Learning for Post-Treatment Visual Acuity Forecasting
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)
Dual Strategies for Test-Time Adaptation
by: Phuong, Nam Nguyen, et al.
Published: (2026)
by: Phuong, Nam Nguyen, et al.
Published: (2026)
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
by: Quang, Ngoc Bui Lam, et al.
Published: (2025)
by: Quang, Ngoc Bui Lam, et al.
Published: (2025)
SOUPLE: Enhancing Audio-Visual Localization and Segmentation with Learnable Prompt Contexts
by: Nguyen, Khanh Binh, et al.
Published: (2026)
by: Nguyen, Khanh Binh, et al.
Published: (2026)
Symmetric masking strategy enhances the performance of Masked Image Modeling
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
Retro: Reusing teacher projection head for efficient embedding distillation on Lightweight Models via Self-supervised Learning
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
by: Nguyen, Huu Tien, et al.
Published: (2025)
by: Nguyen, Huu Tien, et al.
Published: (2025)
IQBench: How "Smart'' Are Vision-Language Models? A Study with Human IQ Tests
by: Pham, Tan-Hanh, et al.
Published: (2025)
by: Pham, Tan-Hanh, et al.
Published: (2025)
A model-agnostic active learning approach for animal detection from camera traps
by: Nguyen, Thi Thu Thuy, et al.
Published: (2025)
by: Nguyen, Thi Thu Thuy, et al.
Published: (2025)
Efficient and Concise Explanations for Object Detection with Gaussian-Class Activation Mapping Explainer
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
by: Nguyen, Quoc Khanh, et al.
Published: (2024)
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
by: Pham, Duc-Hai, et al.
Published: (2024)
by: Pham, Duc-Hai, et al.
Published: (2024)
Learning to Stop Overthinking at Test Time
by: Bao, Hieu Tran, et al.
Published: (2025)
by: Bao, Hieu Tran, et al.
Published: (2025)
Mitigating Cache Noise in Test-Time Adaptation for Large Vision-Language Models
by: Zhai, Haotian, et al.
Published: (2025)
by: Zhai, Haotian, et al.
Published: (2025)
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Bellman Optimal Stepsize Straightening of Flow-Matching Models
by: Nguyen, Bao, et al.
Published: (2023)
by: Nguyen, Bao, et al.
Published: (2023)
CLIPping the Deception: Adapting Vision-Language Models for Universal Deepfake Detection
by: Khan, Sohail Ahmed, et al.
Published: (2024)
by: Khan, Sohail Ahmed, et al.
Published: (2024)
SAVE: Segment Audio-Visual Easy way using Segment Anything Model
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
by: Nguyen, Khanh-Binh, et al.
Published: (2024)
Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates
by: Shum, Ka Chun, et al.
Published: (2023)
by: Shum, Ka Chun, et al.
Published: (2023)
SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation
by: Pham, Duc-Hai, et al.
Published: (2024)
by: Pham, Duc-Hai, et al.
Published: (2024)
Color Alignment in Diffusion
by: Shum, Ka Chun, et al.
Published: (2025)
by: Shum, Ka Chun, et al.
Published: (2025)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
by: Truong, Quang-Trung, et al.
Published: (2024)
by: Truong, Quang-Trung, et al.
Published: (2024)
Collaborative Perceiver: Elevating Vision-based 3D Object Detection via Local Density-Aware Spatial Occupancy
by: Yuan, Jicheng, et al.
Published: (2025)
by: Yuan, Jicheng, et al.
Published: (2025)
N-EIoU-YOLOv9: A Signal-Aware Bounding Box Regression Loss for Lightweight Mobile Detection of Rice Leaf Diseases
by: Duc, Dung Ta Nguyen, et al.
Published: (2026)
by: Duc, Dung Ta Nguyen, et al.
Published: (2026)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Non-Learning Low-Light Stereo Vision
by: Wang, Jason, et al.
Published: (2026)
by: Wang, Jason, et al.
Published: (2026)
Domain Generalization through Spatial Relation Induction over Visual Primitives
by: Nguyen, Dat, et al.
Published: (2026)
by: Nguyen, Dat, et al.
Published: (2026)
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
by: To-Thanh, Dat, et al.
Published: (2026)
by: To-Thanh, Dat, et al.
Published: (2026)
The Art of Camouflage: Few-Shot Learning for Animal Detection and Segmentation
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
by: Nguyen, Thanh-Danh, et al.
Published: (2023)
Deep-Wide Learning Assistance for Insect Pest Classification
by: Nguyen, Toan, et al.
Published: (2024)
by: Nguyen, Toan, et al.
Published: (2024)
ITSELF: Attention Guided Fine-Grained Alignment for Vision-Language Retrieval
by: Nguyen, Tien-Huy, et al.
Published: (2026)
by: Nguyen, Tien-Huy, et al.
Published: (2026)
PADM: A Physics-aware Diffusion Model for Attenuation Correction
by: Pham, Trung Kien, et al.
Published: (2025)
by: Pham, Trung Kien, et al.
Published: (2025)
LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models
by: Debnath, Soumyaratna, et al.
Published: (2026)
by: Debnath, Soumyaratna, et al.
Published: (2026)
Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
Similar Items
-
Accelerating Conditional Prompt Learning via Masked Image Modeling for Vision-Language Models
by: Bui, Phuoc-Nguyen, et al.
Published: (2025) -
Attn-Adapter: Attention Is All You Need for Online Few-shot Learner of Vision-Language Model
by: Bui, Phuoc-Nguyen, et al.
Published: (2025) -
Unsupervised Domain Adaptation with SAM-RefiSeR for Enhanced Brain Tumor Segmentation
by: Imans, Dillan, et al.
Published: (2026) -
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis
by: Bui, Phuoc-Nguyen, et al.
Published: (2024) -
Representation Learning with Semantic-aware Instance and Sparse Token Alignments
by: Bui, Phuoc-Nguyen, et al.
Published: (2026)