Seeing Further on the Shoulders of Giants: Knowledge Inheritance for Vision Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Jiabo, Chen, Chen, Lyu, Lingjuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UNIFORM: Unifying Knowledge from Large-scale and Diverse Pre-trained Models
von: Wang, Yimu, et al.
Veröffentlicht: (2025)
von: Wang, Yimu, et al.
Veröffentlicht: (2025)
Empirical Recipes for Efficient and Compact Vision-Language Models
von: Huang, Jiabo, et al.
Veröffentlicht: (2026)
von: Huang, Jiabo, et al.
Veröffentlicht: (2026)
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
von: Wang, Ziyao, et al.
Veröffentlicht: (2026)
See Further for Parameter Efficient Fine-tuning by Standing on the Shoulders of Decomposition
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
Closer to Reality: Practical Semi-Supervised Federated Learning for Foundation Model Adaptation
von: Sun, Guangyu, et al.
Veröffentlicht: (2025)
von: Sun, Guangyu, et al.
Veröffentlicht: (2025)
Detecting, Explaining, and Mitigating Memorization in Diffusion Models
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
RT-DETRv4: Painlessly Furthering Real-Time Object Detection with Vision Foundation Models
von: Liao, Zijun, et al.
Veröffentlicht: (2025)
von: Liao, Zijun, et al.
Veröffentlicht: (2025)
See Further When Clear: Curriculum Consistency Model
von: Liu, Yunpeng, et al.
Veröffentlicht: (2024)
von: Liu, Yunpeng, et al.
Veröffentlicht: (2024)
COALA: A Practical and Vision-Centric Federated Learning Platform
von: Zhuang, Weiming, et al.
Veröffentlicht: (2024)
von: Zhuang, Weiming, et al.
Veröffentlicht: (2024)
Standing on the Shoulders of Giants: Reprogramming Visual-Language Model for General Deepfake Detection
von: Lin, Kaiqing, et al.
Veröffentlicht: (2024)
von: Lin, Kaiqing, et al.
Veröffentlicht: (2024)
Evaluating and Mitigating IP Infringement in Visual Generative AI
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
Task-Specific Knowledge Distillation from the Vision Foundation Model for Enhanced Medical Image Segmentation
von: Liang, Pengchen, et al.
Veröffentlicht: (2025)
von: Liang, Pengchen, et al.
Veröffentlicht: (2025)
A Simple Background Augmentation Method for Object Detection with Diffusion Model
von: Li, Yuhang, et al.
Veröffentlicht: (2024)
von: Li, Yuhang, et al.
Veröffentlicht: (2024)
Towards Fundamentally Scalable Model Selection: Asymptotically Fast Update and Selection
von: Wang, Wenxiao, et al.
Veröffentlicht: (2024)
von: Wang, Wenxiao, et al.
Veröffentlicht: (2024)
Training-Free Layout-to-Image Generation with Marginal Attention Constraints
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
CoCAViT: Compact Vision Transformer with Robust Global Coordination
von: Wang, Xuyang, et al.
Veröffentlicht: (2025)
von: Wang, Xuyang, et al.
Veröffentlicht: (2025)
Replay-Free Continual Low-Rank Adaptation with Dynamic Memory
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
von: Chen, Huancheng, et al.
Veröffentlicht: (2024)
When Alignment Fails: Multimodal Adversarial Attacks on Vision-Language-Action Models
von: Yan, Yuping, et al.
Veröffentlicht: (2025)
von: Yan, Yuping, et al.
Veröffentlicht: (2025)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Vision-Language Models Can't See the Obvious
von: Dahou, Yasser, et al.
Veröffentlicht: (2025)
von: Dahou, Yasser, et al.
Veröffentlicht: (2025)
On the Limits of Token Reduction for Efficient Unified Vision Language Training
von: Chen, Siyi, et al.
Veröffentlicht: (2026)
von: Chen, Siyi, et al.
Veröffentlicht: (2026)
NocPlace: Nocturnal Visual Place Recognition via Generative and Inherited Knowledge Transfer
von: Liu, Bingxi, et al.
Veröffentlicht: (2024)
von: Liu, Bingxi, et al.
Veröffentlicht: (2024)
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
von: Guo, Jianyuan, et al.
Veröffentlicht: (2024)
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
von: Cai, Zhejia, et al.
Veröffentlicht: (2025)
von: Cai, Zhejia, et al.
Veröffentlicht: (2025)
DINOReg: Strong Point Cloud Registration with Vision Foundation Model
von: Chen, Congjia, et al.
Veröffentlicht: (2025)
von: Chen, Congjia, et al.
Veröffentlicht: (2025)
Generalizable Knowledge Distillation from Vision Foundation Models for Semantic Segmentation
von: Lv, Chonghua, et al.
Veröffentlicht: (2026)
von: Lv, Chonghua, et al.
Veröffentlicht: (2026)
WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model
von: Zhang, Songyan, et al.
Veröffentlicht: (2024)
von: Zhang, Songyan, et al.
Veröffentlicht: (2024)
Rendering-Refined Stable Diffusion for Privacy Compliant Synthetic Data
von: Patwari, Kartik, et al.
Veröffentlicht: (2024)
von: Patwari, Kartik, et al.
Veröffentlicht: (2024)
A Unified Low-level Foundation Model for Enhancing Pathology Image Quality
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
DIAGNOSIS: Detecting Unauthorized Data Usages in Text-to-image Diffusion Models
von: Wang, Zhenting, et al.
Veröffentlicht: (2023)
von: Wang, Zhenting, et al.
Veröffentlicht: (2023)
When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models
von: Ortu, Francesco, et al.
Veröffentlicht: (2025)
von: Ortu, Francesco, et al.
Veröffentlicht: (2025)
Implicit Modeling for Transferability Estimation of Vision Foundation Models
von: Zheng, Yaoyan, et al.
Veröffentlicht: (2025)
von: Zheng, Yaoyan, et al.
Veröffentlicht: (2025)
All-in-One: Transferring Vision Foundation Models into Stereo Matching
von: Zhou, Jingyi, et al.
Veröffentlicht: (2024)
von: Zhou, Jingyi, et al.
Veröffentlicht: (2024)
Adapting Vision Foundation Models for Real-time Ultrasound Image Segmentation
von: Zhang, Xiaoran, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoran, et al.
Veröffentlicht: (2025)
A Breast Vision Pathology Foundation Model for Real-world Clinical Utility
von: Xu, Yingxue, et al.
Veröffentlicht: (2026)
von: Xu, Yingxue, et al.
Veröffentlicht: (2026)
Efficient Transfer Learning for Video-language Foundation Models
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
Bootstrapping SparseFormers from Vision Foundation Models
von: Gao, Ziteng, et al.
Veröffentlicht: (2023)
von: Gao, Ziteng, et al.
Veröffentlicht: (2023)
Seeing the Unseen: Towards Zero-Shot Inspection for Wind Turbine Blades using Knowledge-Augmented Vision Language Models
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
von: Zhang, Yang, et al.
Veröffentlicht: (2025)
CO-SPY: Combining Semantic and Pixel Features to Detect Synthetic Images by AI
von: Cheng, Siyuan, et al.
Veröffentlicht: (2025)
von: Cheng, Siyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
UNIFORM: Unifying Knowledge from Large-scale and Diverse Pre-trained Models
von: Wang, Yimu, et al.
Veröffentlicht: (2025) -
Empirical Recipes for Efficient and Compact Vision-Language Models
von: Huang, Jiabo, et al.
Veröffentlicht: (2026) -
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
von: Wang, Ziyao, et al.
Veröffentlicht: (2026) -
See Further for Parameter Efficient Fine-tuning by Standing on the Shoulders of Decomposition
von: Si, Chongjie, et al.
Veröffentlicht: (2024) -
Closer to Reality: Practical Semi-Supervised Federated Learning for Foundation Model Adaptation
von: Sun, Guangyu, et al.
Veröffentlicht: (2025)