Advancing Reliable Test-Time Adaptation of Vision-Language Models under Visual Variations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Yiwen, Chen, Hui, Xiong, Yizhe, Zhou, Zihan, Lyu, Mengyao, Lin, Zijia, Niu, Shuaicheng, Zhao, Sicheng, Han, Jungong, Ding, Guiguang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neutralizing Token Aggregation via Information Augmentation for Efficient Test-Time Adaptation
von: Xiong, Yizhe, et al.
Veröffentlicht: (2025)
von: Xiong, Yizhe, et al.
Veröffentlicht: (2025)
Learn from the Learnt: Source-Free Active Domain Adaptation via Contrastive Sampling and Visual Persistence
von: Lyu, Mengyao, et al.
Veröffentlicht: (2024)
von: Lyu, Mengyao, et al.
Veröffentlicht: (2024)
Towards Efficient Vision-Language Tuning: More Information Density, More Generalizability
von: Hao, Tianxiang, et al.
Veröffentlicht: (2023)
von: Hao, Tianxiang, et al.
Veröffentlicht: (2023)
CAIT: Triple-Win Compression towards High Accuracy, Fast Inference, and Favorable Transferability For ViTs
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
RepViT-SAM: Towards Real-Time Segmenting Anything
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
PYRA: Parallel Yielding Re-Activation for Training-Inference Efficient Task Adaptation
von: Xiong, Yizhe, et al.
Veröffentlicht: (2024)
von: Xiong, Yizhe, et al.
Veröffentlicht: (2024)
YOLOE: Real-Time Seeing Anything
von: Wang, Ao, et al.
Veröffentlicht: (2025)
von: Wang, Ao, et al.
Veröffentlicht: (2025)
LSNet: See Large, Focus Small
von: Wang, Ao, et al.
Veröffentlicht: (2025)
von: Wang, Ao, et al.
Veröffentlicht: (2025)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Scaffold-BPE: Enhancing Byte Pair Encoding for Large Language Models with Simple and Effective Scaffold Token Removal
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
UniAttn: Reducing Inference Costs via Softmax Unification for Post-Training LLMs
von: Xiong, Yizhe, et al.
Veröffentlicht: (2025)
von: Xiong, Yizhe, et al.
Veröffentlicht: (2025)
YOLOv10: Real-Time End-to-End Object Detection
von: Wang, Ao, et al.
Veröffentlicht: (2024)
von: Wang, Ao, et al.
Veröffentlicht: (2024)
[CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
von: Wang, Ao, et al.
Veröffentlicht: (2024)
von: Wang, Ao, et al.
Veröffentlicht: (2024)
Temporal Scaling Law for Large Language Models
von: Xiong, Yizhe, et al.
Veröffentlicht: (2024)
von: Xiong, Yizhe, et al.
Veröffentlicht: (2024)
POEM: Explore Unexplored Reliable Samples to Enhance Test-Time Adaptation
von: Yi, Chang'an, et al.
Veröffentlicht: (2025)
von: Yi, Chang'an, et al.
Veröffentlicht: (2025)
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models
von: Sun, Fengyuan, et al.
Veröffentlicht: (2025)
von: Sun, Fengyuan, et al.
Veröffentlicht: (2025)
LBPE: Long-token-first Tokenization to Improve Large Language Models
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
Cream of the Crop: Harvesting Rich, Scalable and Transferable Multi-Modal Data for Instruction Fine-Tuning
von: Lyu, Mengyao, et al.
Veröffentlicht: (2025)
von: Lyu, Mengyao, et al.
Veröffentlicht: (2025)
Fast Quiet-STaR: Thinking Without Thought Tokens
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
PrefixKV: Adaptive Prefix KV Cache is What Vision Instruction-Following Models Need for Efficient Generation
von: Wang, Ao, et al.
Veröffentlicht: (2024)
von: Wang, Ao, et al.
Veröffentlicht: (2024)
LLMI3D: MLLM-based 3D Perception from a Single 2D Image
von: Yang, Fan, et al.
Veröffentlicht: (2024)
von: Yang, Fan, et al.
Veröffentlicht: (2024)
One-Dimensional Adapter to Rule Them All: Concepts, Diffusion Models and Erasing Applications
von: Lyu, Mengyao, et al.
Veröffentlicht: (2023)
von: Lyu, Mengyao, et al.
Veröffentlicht: (2023)
Context Enhancement with Reconstruction as Sequence for Unified Unsupervised Anomaly Detection
von: Yang, Hui-Yue, et al.
Veröffentlicht: (2024)
von: Yang, Hui-Yue, et al.
Veröffentlicht: (2024)
More is Better: Deep Domain Adaptation with Multiple Sources
von: Zhao, Sicheng, et al.
Veröffentlicht: (2024)
von: Zhao, Sicheng, et al.
Veröffentlicht: (2024)
Finedeep: Mitigating Sparse Activation in Dense LLMs via Multi-Layer Fine-Grained Experts
von: Pan, Leiyu, et al.
Veröffentlicht: (2025)
von: Pan, Leiyu, et al.
Veröffentlicht: (2025)
Tracking and Segmenting Anything in Any Modality
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
DPCore: Dynamic Prompt Coreset for Continual Test-Time Adaptation
von: Zhang, Yunbei, et al.
Veröffentlicht: (2024)
von: Zhang, Yunbei, et al.
Veröffentlicht: (2024)
Modality Reliability Guided Multimodal Recommendation
von: Dong, Xue, et al.
Veröffentlicht: (2025)
von: Dong, Xue, et al.
Veröffentlicht: (2025)
DSMoE: Matrix-Partitioned Experts with Dynamic Routing for Computation-Efficient Dense LLMs
von: Lv, Minxuan, et al.
Veröffentlicht: (2025)
von: Lv, Minxuan, et al.
Veröffentlicht: (2025)
DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval
von: Shen, Leqi, et al.
Veröffentlicht: (2025)
von: Shen, Leqi, et al.
Veröffentlicht: (2025)
Test-Time Model Adaptation with Only Forward Passes
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2024)
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2024)
Self-Bootstrapping for Versatile Test-Time Adaptation
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2025)
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2025)
CartesianMoE: Boosting Knowledge Sharing among Experts via Cartesian Product Routing in Mixture-of-Experts
von: Su, Zhenpeng, et al.
Veröffentlicht: (2024)
von: Su, Zhenpeng, et al.
Veröffentlicht: (2024)
Fully Test-Time Adaptation for Monocular 3D Object Detection
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
Promptable Anomaly Segmentation with SAM Through Self-Perception Tuning
von: Yang, Hui-Yue, et al.
Veröffentlicht: (2024)
von: Yang, Hui-Yue, et al.
Veröffentlicht: (2024)
Breaking the Stage Barrier: A Novel Single-Stage Approach to Long Context Extension for Large Language Models
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
von: Lian, Haoran, et al.
Veröffentlicht: (2024)
Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
Variational Continual Test-Time Adaptation
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
von: Lyu, Fan, et al.
Veröffentlicht: (2024)
Learning to Generate Gradients for Test-Time Adaptation via Test-Time Training Layers
von: Deng, Qi, et al.
Veröffentlicht: (2024)
von: Deng, Qi, et al.
Veröffentlicht: (2024)
Test-Time Model Adaptation for Quantized Neural Networks
von: Deng, Zeshuai, et al.
Veröffentlicht: (2025)
von: Deng, Zeshuai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neutralizing Token Aggregation via Information Augmentation for Efficient Test-Time Adaptation
von: Xiong, Yizhe, et al.
Veröffentlicht: (2025) -
Learn from the Learnt: Source-Free Active Domain Adaptation via Contrastive Sampling and Visual Persistence
von: Lyu, Mengyao, et al.
Veröffentlicht: (2024) -
Towards Efficient Vision-Language Tuning: More Information Density, More Generalizability
von: Hao, Tianxiang, et al.
Veröffentlicht: (2023) -
CAIT: Triple-Win Compression towards High Accuracy, Fast Inference, and Favorable Transferability For ViTs
von: Wang, Ao, et al.
Veröffentlicht: (2023) -
RepViT-SAM: Towards Real-Time Segmenting Anything
von: Wang, Ao, et al.
Veröffentlicht: (2023)