Generalizable Knowledge Distillation from Vision Foundation Models for Semantic Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Lv, Chonghua, Zhao, Dong, Wang, Shuang, Quan, Dou, Huyan, Ning, Sebe, Nicu, Zhong, Zhun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FisherTune: Fisher-Guided Robust Tuning of Vision Foundation Models for Domain Generalized Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2025)
di: Zhao, Dong, et al.
Pubblicazione: (2025)
Multi-Expert Learning Framework with the State Space Model for Optical and SAR Image Registration
di: Wang, Wei, et al.
Pubblicazione: (2025)
di: Wang, Wei, et al.
Pubblicazione: (2025)
Stable Neighbor Denoising for Source-free Domain Adaptive Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2024)
di: Zhao, Dong, et al.
Pubblicazione: (2024)
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2026)
di: Zhao, Dong, et al.
Pubblicazione: (2026)
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
di: Wang, Wei, et al.
Pubblicazione: (2026)
di: Wang, Wei, et al.
Pubblicazione: (2026)
Textual Knowledge Matters: Cross-Modality Co-Teaching for Generalized Visual Class Discovery
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
TAR: Text Semantic Assisted Cross-modal Image Registration Framework for Optical and SAR Images
di: Cai, Zhuoyu, et al.
Pubblicazione: (2026)
di: Cai, Zhuoyu, et al.
Pubblicazione: (2026)
Prototypical Hash Encoding for On-the-Fly Fine-Grained Category Discovery
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
Generalized Fine-Grained Category Discovery with Multi-Granularity Conceptual Experts
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
Large-scale Pre-trained Models are Surprisingly Strong in Incremental Novel Class Discovery
di: Liu, Mingxuan, et al.
Pubblicazione: (2023)
di: Liu, Mingxuan, et al.
Pubblicazione: (2023)
Democratizing Fine-grained Visual Recognition with Large Language Models
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
Vision-Language Semantic Aggregation Leveraging Foundation Model for Generalizable Medical Image Segmentation
di: Yu, Wenjun, et al.
Pubblicazione: (2025)
di: Yu, Wenjun, et al.
Pubblicazione: (2025)
Lightweight Adapter Learning for More Generalized Remote Sensing Change Detection
di: Quan, Dou, et al.
Pubblicazione: (2025)
di: Quan, Dou, et al.
Pubblicazione: (2025)
Generate, Refine, and Encode: Leveraging Synthesized Novel Samples for On-the-Fly Fine-Grained Category Discovery
di: Liu, Xiao, et al.
Pubblicazione: (2025)
di: Liu, Xiao, et al.
Pubblicazione: (2025)
Open-World Deepfake Attribution via Confidence-Aware Asymmetric Learning
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
Enhancing Robustness of Vision-Language Models through Orthogonality Learning and Self-Regularization
di: Li, Jinlong, et al.
Pubblicazione: (2024)
di: Li, Jinlong, et al.
Pubblicazione: (2024)
In defense of the two-stage framework for open-set domain adaptive semantic segmentation
di: Ren, Wenqi, et al.
Pubblicazione: (2026)
di: Ren, Wenqi, et al.
Pubblicazione: (2026)
3D Weakly Supervised Semantic Segmentation with 2D Vision-Language Guidance
di: Xu, Xiaoxu, et al.
Pubblicazione: (2024)
di: Xu, Xiaoxu, et al.
Pubblicazione: (2024)
CrossEarth: Geospatial Vision Foundation Model for Domain Generalizable Remote Sensing Semantic Segmentation
di: Gong, Ziyang, et al.
Pubblicazione: (2024)
di: Gong, Ziyang, et al.
Pubblicazione: (2024)
CLIP is Strong Enough to Fight Back: Test-time Counterattacks towards Zero-shot Adversarial Robustness of CLIP
di: Xing, Songlong, et al.
Pubblicazione: (2025)
di: Xing, Songlong, et al.
Pubblicazione: (2025)
Cues3D: Unleashing the Power of Sole NeRF for Consistent and Unique Instances in Open-Vocabulary 3D Panoptic Segmentation
di: Xue, Feng, et al.
Pubblicazione: (2025)
di: Xue, Feng, et al.
Pubblicazione: (2025)
Finetune Like You Pretrain: Boosting Zero-shot Adversarial Robustness in Vision-language Models
di: Xing, Songlong, et al.
Pubblicazione: (2026)
di: Xing, Songlong, et al.
Pubblicazione: (2026)
CLNet: Cross-View Correspondence Makes a Stronger Geo-Localizationer
di: Cao, Xianwei, et al.
Pubblicazione: (2025)
di: Cao, Xianwei, et al.
Pubblicazione: (2025)
Vision+X: A Survey on Multimodal Learning in the Light of Data
di: Zhu, Ye, et al.
Pubblicazione: (2022)
di: Zhu, Ye, et al.
Pubblicazione: (2022)
Diffusion-Guided Knowledge Distillation for Weakly-Supervised Low-Light Semantic Segmentation
di: Wang, Chunyan, et al.
Pubblicazione: (2025)
di: Wang, Chunyan, et al.
Pubblicazione: (2025)
Large Language Models for Multimodal Deformable Image Registration
di: Ma, Mingrui, et al.
Pubblicazione: (2024)
di: Ma, Mingrui, et al.
Pubblicazione: (2024)
Task-Specific Knowledge Distillation from the Vision Foundation Model for Enhanced Medical Image Segmentation
di: Liang, Pengchen, et al.
Pubblicazione: (2025)
di: Liang, Pengchen, et al.
Pubblicazione: (2025)
CrossEarth-SAR: A SAR-Centric and Billion-Scale Geospatial Foundation Model for Domain Generalizable Semantic Segmentation
di: Ye, Ziqi, et al.
Pubblicazione: (2026)
di: Ye, Ziqi, et al.
Pubblicazione: (2026)
CE-SDWV: Effective and Efficient Concept Erasure for Text-to-Image Diffusion Models via a Semantic-Driven Word Vocabulary
di: Tu, Jiahang, et al.
Pubblicazione: (2025)
di: Tu, Jiahang, et al.
Pubblicazione: (2025)
Annotation Free Semantic Segmentation with Vision Foundation Models
di: Seifi, Soroush, et al.
Pubblicazione: (2024)
di: Seifi, Soroush, et al.
Pubblicazione: (2024)
RankFeat&RankWeight: Rank-1 Feature/Weight Removal for Out-of-distribution Detection
di: Song, Yue, et al.
Pubblicazione: (2023)
di: Song, Yue, et al.
Pubblicazione: (2023)
A Closer Look at Conditional Prompt Tuning for Vision-Language Models
di: Zhang, Ji, et al.
Pubblicazione: (2025)
di: Zhang, Ji, et al.
Pubblicazione: (2025)
ZeroReg: Zero-Shot Point Cloud Registration with Foundation Models
di: Wang, Weijie, et al.
Pubblicazione: (2023)
di: Wang, Weijie, et al.
Pubblicazione: (2023)
Structural and Statistical Texture Knowledge Distillation for Semantic Segmentation
di: Ji, Deyi, et al.
Pubblicazione: (2023)
di: Ji, Deyi, et al.
Pubblicazione: (2023)
Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models
di: Li, Jinlong, et al.
Pubblicazione: (2026)
di: Li, Jinlong, et al.
Pubblicazione: (2026)
Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation
di: Dong, Jiahua, et al.
Pubblicazione: (2025)
di: Dong, Jiahua, et al.
Pubblicazione: (2025)
Generalizable Semantic Vision Query Generation for Zero-shot Panoptic and Semantic Segmentation
di: Chen, Jialei, et al.
Pubblicazione: (2024)
di: Chen, Jialei, et al.
Pubblicazione: (2024)
Rethinking the Learning Paradigm for Facial Expression Recognition
di: Wang, Weijie, et al.
Pubblicazione: (2022)
di: Wang, Weijie, et al.
Pubblicazione: (2022)
Unbiased Semantic Decoding with Vision Foundation Models for Few-shot Segmentation
di: Wang, Jin, et al.
Pubblicazione: (2025)
di: Wang, Jin, et al.
Pubblicazione: (2025)
Reverse Personalization
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
FisherTune: Fisher-Guided Robust Tuning of Vision Foundation Models for Domain Generalized Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2025) -
Multi-Expert Learning Framework with the State Space Model for Optical and SAR Image Registration
di: Wang, Wei, et al.
Pubblicazione: (2025) -
Stable Neighbor Denoising for Source-free Domain Adaptive Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2024) -
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2026) -
BGG: Bridging the Geometric Gap between Cross-View images by Vision Foundation Model Adaptation for Geo-Localization
di: Wang, Wei, et al.
Pubblicazione: (2026)