DesCLIP: Robust Continual Learning via General Attribute Descriptions for VLM-Based Visual Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | He, Chiyuan, Qiu, Zihuan, Meng, Fanman, Xu, Linfeng, Wu, Qingbo, Li, Hongliang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continual Learning with Vision-Language Models via Semantic-Geometry Preservation
by: He, Chiyuan, et al.
Published: (2026)
by: He, Chiyuan, et al.
Published: (2026)
MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model Merging
by: Qiu, Zihuan, et al.
Published: (2025)
by: Qiu, Zihuan, et al.
Published: (2025)
Distribution-Level Memory Recall for Continual Learning: Preserving Knowledge and Avoiding Confusion
by: Cheng, Shaoxu, et al.
Published: (2024)
by: Cheng, Shaoxu, et al.
Published: (2024)
Leveraging Pre-Trained Models for Multimodal Class-Incremental Learning under Adaptive Fusion
by: Chen, Yukun, et al.
Published: (2025)
by: Chen, Yukun, et al.
Published: (2025)
Towards Continual Egocentric Activity Recognition: A Multi-modal Egocentric Activity Dataset for Continual Learning
by: Xu, Linfeng, et al.
Published: (2023)
by: Xu, Linfeng, et al.
Published: (2023)
Closing the Oracle Gap: Increment Vector Transformation for Class Incremental Learning
by: Qiu, Zihuan, et al.
Published: (2025)
by: Qiu, Zihuan, et al.
Published: (2025)
Cross-modal Cognitive Consensus guided Audio-Visual Segmentation
by: Shi, Zhaofeng, et al.
Published: (2023)
by: Shi, Zhaofeng, et al.
Published: (2023)
Few-Shot Continual Learning for Activity Recognition in Classroom Surveillance Images
by: Qian, Yilei, et al.
Published: (2024)
by: Qian, Yilei, et al.
Published: (2024)
Null-Space Filtering for Data-Free Continual Model Merging: Preserving Stability, Promoting Plasticity
by: Qiu, Zihuan, et al.
Published: (2025)
by: Qiu, Zihuan, et al.
Published: (2025)
FireRescue: A UAV-Based Dataset and Enhanced YOLO Model for Object Detection in Fire Rescue Scenes
by: Xu, Qingyu, et al.
Published: (2025)
by: Xu, Qingyu, et al.
Published: (2025)
Unsupervised Ego- and Exo-centric Dense Procedural Activity Captioning via Gaze Consensus Adaptation
by: Shi, Zhaofeng, et al.
Published: (2025)
by: Shi, Zhaofeng, et al.
Published: (2025)
Cognition Transferring and Decoupling for Text-supervised Egocentric Semantic Segmentation
by: Shi, Zhaofeng, et al.
Published: (2024)
by: Shi, Zhaofeng, et al.
Published: (2024)
No Re-Train, More Gain: Upgrading Backbones with Diffusion model for Pixel-Wise and Weakly-Supervised Few-Shot Segmentation
by: Chen, Shuai, et al.
Published: (2024)
by: Chen, Shuai, et al.
Published: (2024)
CMaP-SAM: Contraction Mapping Prior for SAM-driven Few-shot Segmentation
by: Chen, Shuai, et al.
Published: (2025)
by: Chen, Shuai, et al.
Published: (2025)
Test-time Ego-Exo-centric Adaptation for Action Anticipation via Multi-Label Prototype Growing and Dual-Clue Consistency
by: Shi, Zhaofeng, et al.
Published: (2026)
by: Shi, Zhaofeng, et al.
Published: (2026)
SAVA-X: Ego-to-Exo Imitation Error Detection via Scene-Adaptive View Alignment and Bidirectional Cross View Fusion
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
On the Adversarial Robustness of Learning-based Image Compression Against Rate-Distortion Attacks
by: Wu, Chenhao, et al.
Published: (2024)
by: Wu, Chenhao, et al.
Published: (2024)
Attention-disentangled Uniform Orthogonal Feature Space Optimization for Few-shot Object Detection
by: Zhao, Taijin, et al.
Published: (2025)
by: Zhao, Taijin, et al.
Published: (2025)
CMP: A Composable Meta Prompt for SAM-Based Cross-Domain Few-Shot Segmentation
by: Chen, Shuai, et al.
Published: (2025)
by: Chen, Shuai, et al.
Published: (2025)
DFR: A Decompose-Fuse-Reconstruct Framework for Multi-Modal Few-Shot Segmentation
by: Chen, Shuai, et al.
Published: (2025)
by: Chen, Shuai, et al.
Published: (2025)
Your Demands Deserve More Bits: Referring Semantic Image Compression at Ultra-low Bitrate
by: Wu, Chenhao, et al.
Published: (2025)
by: Wu, Chenhao, et al.
Published: (2025)
ARIC: An Activity Recognition Dataset in Classroom Surveillance Images
by: Xu, Linfeng, et al.
Published: (2024)
by: Xu, Linfeng, et al.
Published: (2024)
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition
by: Baron, Ethan, et al.
Published: (2024)
by: Baron, Ethan, et al.
Published: (2024)
GRSDet: Learning to Generate Local Reverse Samples for Few-shot Object Detection
by: Mei, Hefei, et al.
Published: (2023)
by: Mei, Hefei, et al.
Published: (2023)
WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition
by: Ning, Shan, et al.
Published: (2026)
by: Ning, Shan, et al.
Published: (2026)
ArtVLM: Attribute Recognition Through Vision-Based Prefix Language Modeling
by: Zhu, William Yicheng, et al.
Published: (2024)
by: Zhu, William Yicheng, et al.
Published: (2024)
WD-FQDet: Multispectral Detection Transformer via Wavelet Decomposition and Frequency-aware Query Learning
by: Yang, Chunjin, et al.
Published: (2026)
by: Yang, Chunjin, et al.
Published: (2026)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool
by: Ma, Yue, et al.
Published: (2024)
by: Ma, Yue, et al.
Published: (2024)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
Quasi-stationary Slice Detection-Based Robust Respiration Rate Estimation under Large-scale Random Body Movement
by: Xu, Chendong, et al.
Published: (2026)
by: Xu, Chendong, et al.
Published: (2026)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
by: Lai, Zhengfeng, et al.
Published: (2023)
by: Lai, Zhengfeng, et al.
Published: (2023)
Not Just What's There: Enabling CLIP to Comprehend Negated Visual Descriptions Without Fine-tuning
by: Xiao, Junhao, et al.
Published: (2026)
by: Xiao, Junhao, et al.
Published: (2026)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024)
by: Zhang, Zaiwei, et al.
Published: (2024)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
by: Alyami, Sarah, et al.
Published: (2025)
by: Alyami, Sarah, et al.
Published: (2025)
TangZihuan/Cox-MK-file: Cox-MK-file
by: TangZihuan
Published: (2026)
by: TangZihuan
Published: (2026)
AREA: Attribute Extraction and Aggregation for CLIP-Based Class-Incremental Learning
by: Xie, Zhen-Hao, et al.
Published: (2026)
by: Xie, Zhen-Hao, et al.
Published: (2026)
One VLM to Keep it Learning: Generation and Balancing for Data-free Continual Visual Question Answering
by: Das, Deepayan, et al.
Published: (2024)
by: Das, Deepayan, et al.
Published: (2024)
EvdCLIP: Improving Vision-Language Retrieval with Entity Visual Descriptions from Large Language Models
by: Meng, GuangHao, et al.
Published: (2025)
by: Meng, GuangHao, et al.
Published: (2025)
Embedding Shift Dissection on CLIP: Effects of Augmentations on VLM's Representation Learning
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
Similar Items
-
Continual Learning with Vision-Language Models via Semantic-Geometry Preservation
by: He, Chiyuan, et al.
Published: (2026) -
MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model Merging
by: Qiu, Zihuan, et al.
Published: (2025) -
Distribution-Level Memory Recall for Continual Learning: Preserving Knowledge and Avoiding Confusion
by: Cheng, Shaoxu, et al.
Published: (2024) -
Leveraging Pre-Trained Models for Multimodal Class-Incremental Learning under Adaptive Fusion
by: Chen, Yukun, et al.
Published: (2025) -
Towards Continual Egocentric Activity Recognition: A Multi-modal Egocentric Activity Dataset for Continual Learning
by: Xu, Linfeng, et al.
Published: (2023)