Implicit Modeling for Transferability Estimation of Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Yaoyan, Wang, Huiqun, Zhou, Nan, Huang, Di |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoVFT: Context-aware Visual Fine-tuning for Multimodal Large Language Models
by: Zhou, Nan, et al.
Published: (2026)
by: Zhou, Nan, et al.
Published: (2026)
Crowd-SAM: SAM as a Smart Annotator for Object Detection in Crowded Scenes
by: Cai, Zhi, et al.
Published: (2024)
by: Cai, Zhi, et al.
Published: (2024)
Deep Common Feature Mining for Efficient Video Semantic Segmentation
by: Zheng, Yaoyan, et al.
Published: (2024)
by: Zheng, Yaoyan, et al.
Published: (2024)
EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs
by: Chen, Zhenghao, et al.
Published: (2026)
by: Chen, Zhenghao, et al.
Published: (2026)
All-in-One: Transferring Vision Foundation Models into Stereo Matching
by: Zhou, Jingyi, et al.
Published: (2024)
by: Zhou, Jingyi, et al.
Published: (2024)
Towards Training-free Anomaly Detection with Vision and Language Foundation Models
by: Zhang, Jinjin, et al.
Published: (2025)
by: Zhang, Jinjin, et al.
Published: (2025)
Understanding the Transfer Limits of Vision Foundation Models
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation
by: Tang, Jiaqi, et al.
Published: (2026)
by: Tang, Jiaqi, et al.
Published: (2026)
TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration
by: Guo, Yiwei, et al.
Published: (2024)
by: Guo, Yiwei, et al.
Published: (2024)
Towards Unbiased Source-Free Object Detection via Vision Foundation Models
by: Cai, Zhi, et al.
Published: (2026)
by: Cai, Zhi, et al.
Published: (2026)
Efficient Transfer Learning for Video-language Foundation Models
by: Chen, Haoxing, et al.
Published: (2024)
by: Chen, Haoxing, et al.
Published: (2024)
A multimodal vision foundation model for generalizable knee pathology
by: Yu, Kang, et al.
Published: (2026)
by: Yu, Kang, et al.
Published: (2026)
ImFace++: A Sophisticated Nonlinear 3D Morphable Face Model with Implicit Neural Representations
by: Zheng, Mingwu, et al.
Published: (2023)
by: Zheng, Mingwu, et al.
Published: (2023)
On the Use of Hierarchical Vision Foundation Models for Low-Cost Human Mesh Recovery and Pose Estimation
by: Tarashima, Shuhei, et al.
Published: (2025)
by: Tarashima, Shuhei, et al.
Published: (2025)
SimMAT: Exploring Transferability from Vision Foundation Models to Any Image Modality
by: Lei, Chenyang, et al.
Published: (2024)
by: Lei, Chenyang, et al.
Published: (2024)
Enhancing Representation in Medical Vision-Language Foundation Models via Multi-Scale Information Extraction Techniques
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
by: Guo, Jianyuan, et al.
Published: (2024)
by: Guo, Jianyuan, et al.
Published: (2024)
Bootstrapping SparseFormers from Vision Foundation Models
by: Gao, Ziteng, et al.
Published: (2023)
by: Gao, Ziteng, et al.
Published: (2023)
iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models
by: Cui, Xuezhi, et al.
Published: (2026)
by: Cui, Xuezhi, et al.
Published: (2026)
Proxy Robustness in Vision Language Models is Effortlessly Transferable
by: Fu, Xiaowei, et al.
Published: (2026)
by: Fu, Xiaowei, et al.
Published: (2026)
Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following
by: Wang, Shijing, et al.
Published: (2026)
by: Wang, Shijing, et al.
Published: (2026)
Towards Vision-Language Geo-Foundation Model: A Survey
by: Zhou, Yue, et al.
Published: (2024)
by: Zhou, Yue, et al.
Published: (2024)
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)
by: Zheng, Anlin, et al.
Published: (2026)
Temporal-Guided Visual Foundation Models for Event-Based Vision
by: Xia, Ruihao, et al.
Published: (2025)
by: Xia, Ruihao, et al.
Published: (2025)
VisBias: Measuring Explicit and Implicit Social Biases in Vision Language Models
by: Huang, Jen-tse, et al.
Published: (2025)
by: Huang, Jen-tse, et al.
Published: (2025)
TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction
by: Zhang, Fengyi, et al.
Published: (2025)
by: Zhang, Fengyi, et al.
Published: (2025)
Unleashing Foundation Vision Models: Adaptive Transfer for Diverse Data-Limited Scientific Domains
by: Li, Qiankun, et al.
Published: (2025)
by: Li, Qiankun, et al.
Published: (2025)
Seeing Further on the Shoulders of Giants: Knowledge Inheritance for Vision Foundation Models
by: Huang, Jiabo, et al.
Published: (2025)
by: Huang, Jiabo, et al.
Published: (2025)
Sapiens: Foundation for Human Vision Models
by: Khirodkar, Rawal, et al.
Published: (2024)
by: Khirodkar, Rawal, et al.
Published: (2024)
M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision
by: Zhou, Kailai, et al.
Published: (2025)
by: Zhou, Kailai, et al.
Published: (2025)
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
by: Liu, Fan, et al.
Published: (2023)
by: Liu, Fan, et al.
Published: (2023)
GLOV: Guided Large Language Models as Implicit Optimizers for Vision Language Models
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
by: Mirza, M. Jehanzeb, et al.
Published: (2024)
Cross-Domain Transfer of Hyperspectral Foundation Models
by: Theisen, Nick, et al.
Published: (2026)
by: Theisen, Nick, et al.
Published: (2026)
Are Vision Foundation Models Foundational for Electron Microscopy Image Segmentation?
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
by: Fuster-Barceló, Caterina, et al.
Published: (2026)
Test-Time Adaptive Object Detection with Foundation Model
by: Gao, Yingjie, et al.
Published: (2025)
by: Gao, Yingjie, et al.
Published: (2025)
CellVTA: Enhancing Vision Foundation Models for Accurate Cell Segmentation and Classification
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
A Survey on Remote Sensing Foundation Models: From Vision to Multimodality
by: Huang, Ziyue, et al.
Published: (2025)
by: Huang, Ziyue, et al.
Published: (2025)
Visual Implicit Autoregressive Modeling
by: Jiang, Pengfei, et al.
Published: (2026)
by: Jiang, Pengfei, et al.
Published: (2026)
Universal Pansharpening Foundation Model
by: Wang, Hebaixu, et al.
Published: (2026)
by: Wang, Hebaixu, et al.
Published: (2026)
Similar Items
-
CoVFT: Context-aware Visual Fine-tuning for Multimodal Large Language Models
by: Zhou, Nan, et al.
Published: (2026) -
Crowd-SAM: SAM as a Smart Annotator for Object Detection in Crowded Scenes
by: Cai, Zhi, et al.
Published: (2024) -
Deep Common Feature Mining for Efficient Video Semantic Segmentation
by: Zheng, Yaoyan, et al.
Published: (2024) -
EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs
by: Chen, Zhenghao, et al.
Published: (2026) -
All-in-One: Transferring Vision Foundation Models into Stereo Matching
by: Zhou, Jingyi, et al.
Published: (2024)