Towards Difficulty-Agnostic Efficient Transfer Learning for Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Yongjin, Ko, Jongwoo, Yun, Se-Young |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Bayesian Principles Improve Prompt Learning In Vision-Language Models
por: Kim, Mingyu, et al.
Publicado: (2025)
por: Kim, Mingyu, et al.
Publicado: (2025)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
por: Ha, Jiwoo, et al.
Publicado: (2026)
por: Ha, Jiwoo, et al.
Publicado: (2026)
Towards Cross-modal Backward-compatible Representation Learning for Vision-Language Models
por: Jang, Young Kyun, et al.
Publicado: (2024)
por: Jang, Young Kyun, et al.
Publicado: (2024)
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
por: Kim, Hayeon, et al.
Publicado: (2026)
por: Kim, Hayeon, et al.
Publicado: (2026)
Efficient Personalization of Quantized Diffusion Model without Backpropagation
por: Seo, Hoigi, et al.
Publicado: (2025)
por: Seo, Hoigi, et al.
Publicado: (2025)
QAVA: Query-Agnostic Visual Attack to Large Vision-Language Models
por: Zhang, Yudong, et al.
Publicado: (2025)
por: Zhang, Yudong, et al.
Publicado: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
por: Seo, Hoigi, et al.
Publicado: (2025)
por: Seo, Hoigi, et al.
Publicado: (2025)
UniSAFE: A Comprehensive Benchmark for Safety Evaluation of Unified Multimodal Models
por: Lee, Segyu, et al.
Publicado: (2026)
por: Lee, Segyu, et al.
Publicado: (2026)
Reasoning under Vision: Understanding Visual-Spatial Cognition in Vision-Language Models for CAPTCHA
por: Song, Python, et al.
Publicado: (2025)
por: Song, Python, et al.
Publicado: (2025)
Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
por: Kwon, JuneHyoung, et al.
Publicado: (2026)
por: Kwon, JuneHyoung, et al.
Publicado: (2026)
Efficient Few-Shot Learning in Remote Sensing: Fusing Vision and Vision-Language Models
por: Chua, Jia Yun, et al.
Publicado: (2025)
por: Chua, Jia Yun, et al.
Publicado: (2025)
Towards Calibrated Robust Fine-Tuning of Vision-Language Models
por: Oh, Changdae, et al.
Publicado: (2023)
por: Oh, Changdae, et al.
Publicado: (2023)
AVA: Towards Agentic Video Analytics with Vision Language Models
por: Yan, Yuxuan, et al.
Publicado: (2025)
por: Yan, Yuxuan, et al.
Publicado: (2025)
HAWAII: Hierarchical Visual Knowledge Transfer for Efficient Vision-Language Models
por: Wang, Yimu, et al.
Publicado: (2025)
por: Wang, Yimu, et al.
Publicado: (2025)
Efficient Few-Shot Continual Learning in Vision-Language Models
por: Panos, Aristeidis, et al.
Publicado: (2025)
por: Panos, Aristeidis, et al.
Publicado: (2025)
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
por: Cui, Kaiyuan, et al.
Publicado: (2026)
por: Cui, Kaiyuan, et al.
Publicado: (2026)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
por: Seo, Hoigi, et al.
Publicado: (2026)
por: Seo, Hoigi, et al.
Publicado: (2026)
Dropout Prompt Learning: Towards Robust and Adaptive Vision-Language Models
por: Chen, Biao, et al.
Publicado: (2025)
por: Chen, Biao, et al.
Publicado: (2025)
Versatile Incremental Learning: Towards Class and Domain-Agnostic Incremental Learning
por: Park, Min-Yeong, et al.
Publicado: (2024)
por: Park, Min-Yeong, et al.
Publicado: (2024)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
por: Zeng, Fanhu, et al.
Publicado: (2025)
por: Zeng, Fanhu, et al.
Publicado: (2025)
Patch Rebirth: Toward Fast and Transferable Model Inversion of Vision Transformers
por: Heo, Seongsoo, et al.
Publicado: (2025)
por: Heo, Seongsoo, et al.
Publicado: (2025)
Contribution-based Low-Rank Adaptation with Pre-training Model for Real Image Restoration
por: Park, Donwon, et al.
Publicado: (2024)
por: Park, Donwon, et al.
Publicado: (2024)
Perturb a Model, Not an Image: Towards Robust Privacy Protection via Anti-Personalized Diffusion Models
por: Lee, Tae-Young, et al.
Publicado: (2025)
por: Lee, Tae-Young, et al.
Publicado: (2025)
SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction
por: Zhou, Yang, et al.
Publicado: (2024)
por: Zhou, Yang, et al.
Publicado: (2024)
TinyLVLM-eHub: Towards Comprehensive and Efficient Evaluation for Large Vision-Language Models
por: Shao, Wenqi, et al.
Publicado: (2023)
por: Shao, Wenqi, et al.
Publicado: (2023)
Index-Preserving Lightweight Token Pruning for Efficient Document Understanding in Vision-Language Models
por: Son, Jaemin, et al.
Publicado: (2025)
por: Son, Jaemin, et al.
Publicado: (2025)
Synergistic Integration of Coordinate Network and Tensorial Feature for Improving Neural Radiance Fields from Sparse Inputs
por: Kim, Mingyu, et al.
Publicado: (2024)
por: Kim, Mingyu, et al.
Publicado: (2024)
DistiLLM: Towards Streamlined Distillation for Large Language Models
por: Ko, Jongwoo, et al.
Publicado: (2024)
por: Ko, Jongwoo, et al.
Publicado: (2024)
ProtoDCS: Towards Robust and Efficient Open-Set Test-Time Adaptation for Vision-Language Models
por: Luo, Wei, et al.
Publicado: (2026)
por: Luo, Wei, et al.
Publicado: (2026)
FedSOL: Stabilized Orthogonal Learning with Proximal Restrictions in Federated Learning
por: Lee, Gihun, et al.
Publicado: (2023)
por: Lee, Gihun, et al.
Publicado: (2023)
Localized Concept Erasure for Text-to-Image Diffusion Models Using Training-Free Gated Low-Rank Adaptation
por: Lee, Byung Hyun, et al.
Publicado: (2025)
por: Lee, Byung Hyun, et al.
Publicado: (2025)
4D Gaussian Splatting in the Wild with Uncertainty-Aware Regularization
por: Kim, Mijeong, et al.
Publicado: (2024)
por: Kim, Mijeong, et al.
Publicado: (2024)
When Robots Obey the Patch: Universal Transferable Patch Attacks on Vision-Language-Action Models
por: Lu, Hui, et al.
Publicado: (2025)
por: Lu, Hui, et al.
Publicado: (2025)
WAVER: Writing-style Agnostic Text-Video Retrieval via Distilling Vision-Language Models Through Open-Vocabulary Knowledge
por: Le, Huy, et al.
Publicado: (2023)
por: Le, Huy, et al.
Publicado: (2023)
Learning to Look: Cognitive Attention Alignment with Vision-Language Models
por: Yang, Ryan L., et al.
Publicado: (2025)
por: Yang, Ryan L., et al.
Publicado: (2025)
Advancing Efficient Brain Tumor Multi-Class Classification -- New Insights from the Vision Mamba Model in Transfer Learning
por: Lai, Yinyi, et al.
Publicado: (2024)
por: Lai, Yinyi, et al.
Publicado: (2024)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
por: Park, Jihwan, et al.
Publicado: (2025)
por: Park, Jihwan, et al.
Publicado: (2025)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
por: Kwak, Jaehyun, et al.
Publicado: (2025)
por: Kwak, Jaehyun, et al.
Publicado: (2025)
Empirical Recipes for Efficient and Compact Vision-Language Models
por: Huang, Jiabo, et al.
Publicado: (2026)
por: Huang, Jiabo, et al.
Publicado: (2026)
Learning Domain Agnostic Latent Embeddings of 3D Faces for Zero-shot Animal Expression Transfer
por: Wang, Yue, et al.
Publicado: (2026)
por: Wang, Yue, et al.
Publicado: (2026)
Ejemplares similares
-
Bayesian Principles Improve Prompt Learning In Vision-Language Models
por: Kim, Mingyu, et al.
Publicado: (2025) -
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
por: Ha, Jiwoo, et al.
Publicado: (2026) -
Towards Cross-modal Backward-compatible Representation Learning for Vision-Language Models
por: Jang, Young Kyun, et al.
Publicado: (2024) -
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
por: Kim, Hayeon, et al.
Publicado: (2026) -
Efficient Personalization of Quantized Diffusion Model without Backpropagation
por: Seo, Hoigi, et al.
Publicado: (2025)