Mind the Interference: Retaining Pre-trained Knowledge in Parameter Efficient Continual Learning of Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Longxiang, Tian, Zhuotao, Li, Kai, He, Chunming, Zhou, Hantao, Zhao, Hengshuang, Li, Xiu, Jia, Jiaya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniQA: Unified Vision-Language Pre-training for Image Quality and Aesthetic Assessment
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
Unified Language-driven Zero-shot Domain Adaptation
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
von: Yang, Senqiao, et al.
Veröffentlicht: (2025)
Scalable Language Model with Generalized Continual Learning
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
VisionZip: Longer is Better but Not Necessary in Vision Language Models
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
von: Peng, Bohao, et al.
Veröffentlicht: (2024)
GroupContrast: Semantic-aware Self-supervised Representation Learning for 3D Understanding
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
von: Wang, Chengyao, et al.
Veröffentlicht: (2024)
Retaining and Enhancing Pre-trained Knowledge in Vision-Language Models with Prompt Ensembling
von: Kim, Donggeun, et al.
Veröffentlicht: (2024)
von: Kim, Donggeun, et al.
Veröffentlicht: (2024)
Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?
von: Qu, Tianyuan, et al.
Veröffentlicht: (2025)
von: Qu, Tianyuan, et al.
Veröffentlicht: (2025)
Gamma: Toward Generic Image Assessment with Mixture of Assessment Experts
von: Zhou, Hantao, et al.
Veröffentlicht: (2025)
von: Zhou, Hantao, et al.
Veröffentlicht: (2025)
An Empirical Study of Parameter Efficient Fine-tuning on Vision-Language Pre-train Model
von: Tian, Yuxin, et al.
Veröffentlicht: (2024)
von: Tian, Yuxin, et al.
Veröffentlicht: (2024)
Reti-Diff: Illumination Degradation Image Restoration with Retinex-based Latent Diffusion Model
von: He, Chunming, et al.
Veröffentlicht: (2023)
von: He, Chunming, et al.
Veröffentlicht: (2023)
LISA: Reasoning Segmentation via Large Language Model
von: Lai, Xin, et al.
Veröffentlicht: (2023)
von: Lai, Xin, et al.
Veröffentlicht: (2023)
Real-world Image Dehazing with Coherence-based Pseudo Labeling and Cooperative Unfolding Network
von: Fang, Chengyu, et al.
Veröffentlicht: (2024)
von: Fang, Chengyu, et al.
Veröffentlicht: (2024)
Diffusion Models in Low-Level Vision: A Survey
von: He, Chunming, et al.
Veröffentlicht: (2024)
von: He, Chunming, et al.
Veröffentlicht: (2024)
Exploiting Discriminative Codebook Prior for Autoregressive Image Generation
von: Tang, Longxiang, et al.
Veröffentlicht: (2025)
von: Tang, Longxiang, et al.
Veröffentlicht: (2025)
Enhancing LLM Knowledge Learning through Generalization
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
PRISM: Rethinking Scattered Atmosphere Reconstruction as a Unified Understanding and Generation Model for Real-world Dehazing
von: Fang, Chengyu, et al.
Veröffentlicht: (2026)
von: Fang, Chengyu, et al.
Veröffentlicht: (2026)
Exploiting the Semantic Knowledge of Pre-trained Text-Encoders for Continual Learning
von: Yu, Lu, et al.
Veröffentlicht: (2024)
von: Yu, Lu, et al.
Veröffentlicht: (2024)
VIP: Vision Instructed Pre-training for Robotic Manipulation
von: Li, Zhuoling, et al.
Veröffentlicht: (2024)
von: Li, Zhuoling, et al.
Veröffentlicht: (2024)
LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model
von: Yang, Senqiao, et al.
Veröffentlicht: (2023)
von: Yang, Senqiao, et al.
Veröffentlicht: (2023)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
von: Lai, Xin, et al.
Veröffentlicht: (2024)
von: Lai, Xin, et al.
Veröffentlicht: (2024)
Revisiting Continual Semantic Segmentation with Pre-trained Vision Models
von: Zhang, Duzhen, et al.
Veröffentlicht: (2025)
von: Zhang, Duzhen, et al.
Veröffentlicht: (2025)
Video Object Segmentation with Dynamic Query Modulation
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
von: Zhou, Hantao, et al.
Veröffentlicht: (2024)
Continual Retinal Vision-Language Pre-training upon Incremental Imaging Modalities
von: Yao, Yuang, et al.
Veröffentlicht: (2025)
von: Yao, Yuang, et al.
Veröffentlicht: (2025)
GridMask Data Augmentation
von: Chen, Pengguang, et al.
Veröffentlicht: (2020)
von: Chen, Pengguang, et al.
Veröffentlicht: (2020)
LogoSticker: Inserting Logos into Diffusion Models for Customized Generation
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2024)
From Noisy Traces to Stable Gradients: Bias-Variance Optimized Preference Optimization for Aligning Large Reasoning Models
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
Stratified GRPO: Handling Structural Heterogeneity in Reinforcement Learning of LLM Search Agents
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
von: Zhu, Mingkang, et al.
Veröffentlicht: (2025)
Photon: Speedup Volume Understanding with Efficient Multimodal Large Language Models
von: Fang, Chengyu, et al.
Veröffentlicht: (2026)
von: Fang, Chengyu, et al.
Veröffentlicht: (2026)
MultiBooth: Towards Generating All Your Concepts in an Image from Text
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhu, Chenyang, et al.
Veröffentlicht: (2024)
Efficient and Long-Tailed Generalization for Pre-trained Vision-Language Model
von: Shi, Jiang-Xin, et al.
Veröffentlicht: (2024)
von: Shi, Jiang-Xin, et al.
Veröffentlicht: (2024)
Personalized Continual EEG Decoding: Retaining and Transferring Knowledge
von: Li, Dan, et al.
Veröffentlicht: (2024)
von: Li, Dan, et al.
Veröffentlicht: (2024)
SCALER: SAM-Enhanced Collaborative Learning for Label-Deficient Concealed Object Segmentation
von: He, Chunming, et al.
Veröffentlicht: (2025)
von: He, Chunming, et al.
Veröffentlicht: (2025)
VLPose: Bridging the Domain Gap in Pose Estimation with Language-Vision Tuning
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
von: Li, Jingyao, et al.
Veröffentlicht: (2024)
Efficient Vision-Language Pre-training by Cluster Masking
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
Segment Concealed Objects with Incomplete Supervision
von: He, Chunming, et al.
Veröffentlicht: (2025)
von: He, Chunming, et al.
Veröffentlicht: (2025)
A Survey of Camouflaged Object Detection and Beyond
von: Xiao, Fengyang, et al.
Veröffentlicht: (2024)
von: Xiao, Fengyang, et al.
Veröffentlicht: (2024)
Integrating Extra Modality Helps Segmentor Find Camouflaged Objects Well
von: Fang, Chengyu, et al.
Veröffentlicht: (2025)
von: Fang, Chengyu, et al.
Veröffentlicht: (2025)
Freeze the backbones: A Parameter-Efficient Contrastive Approach to Robust Medical Vision-Language Pre-training
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
von: Qin, Jiuming, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
UniQA: Unified Vision-Language Pre-training for Image Quality and Aesthetic Assessment
von: Zhou, Hantao, et al.
Veröffentlicht: (2024) -
Unified Language-driven Zero-shot Domain Adaptation
von: Yang, Senqiao, et al.
Veröffentlicht: (2024) -
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning
von: Yang, Senqiao, et al.
Veröffentlicht: (2025) -
Scalable Language Model with Generalized Continual Learning
von: Peng, Bohao, et al.
Veröffentlicht: (2024) -
VisionZip: Longer is Better but Not Necessary in Vision Language Models
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)