Enhancing Visual Continual Learning with Language-Guided Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ni, Bolin, Zhao, Hongbo, Zhang, Chenghao, Hu, Ke, Meng, Gaofeng, Zhang, Zhaoxiang, Xiang, Shiming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reusable Architecture Growth for Continual Stereo Matching
von: Zhang, Chenghao, et al.
Veröffentlicht: (2024)
von: Zhang, Chenghao, et al.
Veröffentlicht: (2024)
Practical Continual Forgetting for Pre-trained Vision Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
Continual Forgetting for Pre-trained Vision Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
Defying Imbalanced Forgetting in Class Incremental Learning
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
MLLM-CL: Continual Learning for Multimodal Large Language Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
von: Xu, Shixiong, et al.
Veröffentlicht: (2024)
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
A Survey of Low-shot Vision-Language Model Adaptation via Representer Theorem
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
AddressVLM: Cross-view Alignment Tuning for Image Address Localization using Large Vision-Language Models
von: Xu, Shixiong, et al.
Veröffentlicht: (2025)
von: Xu, Shixiong, et al.
Veröffentlicht: (2025)
Free Lunch for Generating Effective Outlier Supervision
von: Pei, Sen, et al.
Veröffentlicht: (2023)
von: Pei, Sen, et al.
Veröffentlicht: (2023)
OpenSatMap: A Fine-grained High-resolution Satellite Dataset for Large-scale Map Construction
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024)
Rethinking Comprehensive Benchmark for Chart Understanding: A Perspective from Scientific Literature
von: Shen, Lingdong, et al.
Veröffentlicht: (2024)
von: Shen, Lingdong, et al.
Veröffentlicht: (2024)
R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning
von: Yang, Qi, et al.
Veröffentlicht: (2025)
von: Yang, Qi, et al.
Veröffentlicht: (2025)
Semi-parametric Memory Consolidation: Towards Brain-like Deep Continual Learning
von: Liu, Geng, et al.
Veröffentlicht: (2025)
von: Liu, Geng, et al.
Veröffentlicht: (2025)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
Beyond Sequential Distance: Inter-Modal Distance Invariant Position Encoding
von: Chen, Lin, et al.
Veröffentlicht: (2026)
von: Chen, Lin, et al.
Veröffentlicht: (2026)
VisuCraft: Enhancing Large Vision-Language Models for Complex Visual-Guided Creative Content Generation via Structured Information Extraction
von: Jiang, Rongxin, et al.
Veröffentlicht: (2025)
von: Jiang, Rongxin, et al.
Veröffentlicht: (2025)
Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
FCL-COD: Weakly Supervised Camouflaged Object Detection with Frequency-aware and Contrastive Learning
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
von: Ni, Jingchen, et al.
Veröffentlicht: (2026)
Ambiguity-Guided Learnable Distribution Calibration for Semi-Supervised Few-Shot Class-Incremental Learning
von: Lyu, Fan, et al.
Veröffentlicht: (2025)
von: Lyu, Fan, et al.
Veröffentlicht: (2025)
Re-ranking Reasoning Context with Tree Search Makes Large Vision-Language Models Stronger
von: Yang, Qi, et al.
Veröffentlicht: (2025)
von: Yang, Qi, et al.
Veröffentlicht: (2025)
Weakly Supervised 3D Object Detection with Multi-Stage Generalization
von: He, Jiawei, et al.
Veröffentlicht: (2023)
von: He, Jiawei, et al.
Veröffentlicht: (2023)
Continuous Speculative Decoding for Autoregressive Image Generation
von: Wang, Zili, et al.
Veröffentlicht: (2024)
von: Wang, Zili, et al.
Veröffentlicht: (2024)
EvoVLMA: Evolutionary Vision-Language Model Adaptation
von: Ding, Kun, et al.
Veröffentlicht: (2025)
von: Ding, Kun, et al.
Veröffentlicht: (2025)
SEA: Supervised Embedding Alignment for Token-Level Visual-Textual Integration in MLLMs
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
Concept-Guided Prompt Learning for Generalization in Vision-Language Models
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
General Geometry-aware Weakly Supervised 3D Object Detection
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
PADReg: Physics-Aware Deformable Registration Guided by Contact Force for Ultrasound Sequences
von: Geng, Yimeng, et al.
Veröffentlicht: (2025)
von: Geng, Yimeng, et al.
Veröffentlicht: (2025)
Continual Vision-Language Learning for Remote Sensing: Benchmarking and Analysis
von: Weng, Xingxing, et al.
Veröffentlicht: (2026)
von: Weng, Xingxing, et al.
Veröffentlicht: (2026)
MixSup: Mixed-grained Supervision for Label-efficient LiDAR-based 3D Object Detection
von: Yang, Yuxue, et al.
Veröffentlicht: (2024)
von: Yang, Yuxue, et al.
Veröffentlicht: (2024)
VisionNVS: Self-Supervised Inpainting for Novel View Synthesis under the Virtual-Shift Paradigm
von: Lu, Hongbo, et al.
Veröffentlicht: (2026)
von: Lu, Hongbo, et al.
Veröffentlicht: (2026)
CL-VISTA: Benchmarking Continual Learning in Video Large Language Models
von: Guo, Haiyang, et al.
Veröffentlicht: (2026)
von: Guo, Haiyang, et al.
Veröffentlicht: (2026)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
von: Liu, Xiwei, et al.
Veröffentlicht: (2025)
von: Liu, Xiwei, et al.
Veröffentlicht: (2025)
Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction
von: Zhang, Quan, et al.
Veröffentlicht: (2025)
von: Zhang, Quan, et al.
Veröffentlicht: (2025)
Taming Modality Entanglement in Continual Audio-Visual Segmentation
von: Hong, Yuyang, et al.
Veröffentlicht: (2025)
von: Hong, Yuyang, et al.
Veröffentlicht: (2025)
Enhanced Continual Learning of Vision-Language Models with Model Fusion
von: Gao, Haoyuan, et al.
Veröffentlicht: (2025)
von: Gao, Haoyuan, et al.
Veröffentlicht: (2025)
SeaVIS: Sound-Enhanced Association for Online Audio-Visual Instance Segmentation
von: Zhu, Yingjian, et al.
Veröffentlicht: (2026)
von: Zhu, Yingjian, et al.
Veröffentlicht: (2026)
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
Compositional Kronecker Context Optimization for Vision-Language Models
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reusable Architecture Growth for Continual Stereo Matching
von: Zhang, Chenghao, et al.
Veröffentlicht: (2024) -
Practical Continual Forgetting for Pre-trained Vision Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025) -
Continual Forgetting for Pre-trained Vision Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2024) -
Defying Imbalanced Forgetting in Class Incremental Learning
von: Xu, Shixiong, et al.
Veröffentlicht: (2024) -
VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)