Few-shot Adaptation of Multi-modal Foundation Models: A Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Fan, Zhang, Tianshu, Dai, Wenwen, Cai, Wenwen, Zhou, Xiaocong, Chen, Delong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Making Large Vision Language Models to be Good Few-shot Learners
by: Liu, Fan, et al.
Published: (2024)
by: Liu, Fan, et al.
Published: (2024)
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
by: Liu, Fan, et al.
Published: (2023)
by: Liu, Fan, et al.
Published: (2023)
Unbiased Semantic Decoding with Vision Foundation Models for Few-shot Segmentation
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
by: Chen, Tingxuan, et al.
Published: (2025)
by: Chen, Tingxuan, et al.
Published: (2025)
Retrieval-augmented Few-shot Medical Image Segmentation with Foundation Models
by: Zhao, Lin, et al.
Published: (2024)
by: Zhao, Lin, et al.
Published: (2024)
The Devil is in the Few Shots: Iterative Visual Knowledge Completion for Few-shot Learning
by: Li, Yaohui, et al.
Published: (2024)
by: Li, Yaohui, et al.
Published: (2024)
Data Adaptive Few-shot Multi Label Segmentation with Foundation Model
by: Reddy, Gurunath, et al.
Published: (2024)
by: Reddy, Gurunath, et al.
Published: (2024)
FishAI 2.0: Marine Fish Image Classification with Multi-modal Few-shot Learning
by: Yang, Chenghan, et al.
Published: (2025)
by: Yang, Chenghan, et al.
Published: (2025)
Few-shot Adaptation of Medical Vision-Language Models
by: Shakeri, Fereshteh, et al.
Published: (2024)
by: Shakeri, Fereshteh, et al.
Published: (2024)
Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation
by: Liu, Kuanghong, et al.
Published: (2025)
by: Liu, Kuanghong, et al.
Published: (2025)
Enhancing Test Time Adaptation with Few-shot Guidance
by: Luo, Siqi, et al.
Published: (2024)
by: Luo, Siqi, et al.
Published: (2024)
Rethinking Misalignment in Vision-Language Model Adaptation from a Causal Perspective
by: Zhang, Yanan, et al.
Published: (2024)
by: Zhang, Yanan, et al.
Published: (2024)
Cross-domain Multi-modal Few-shot Object Detection via Rich Text
by: Shangguan, Zeyu, et al.
Published: (2024)
by: Shangguan, Zeyu, et al.
Published: (2024)
Modeling Multi-modal Cross-interaction for Multi-label Few-shot Image Classification Based on Local Feature Selection
by: Yan, Kun, et al.
Published: (2024)
by: Yan, Kun, et al.
Published: (2024)
Improving Interpretability of Deep Active Learning for Flood Inundation Mapping Through Class Ambiguity Indices Using Multi-spectral Satellite Imagery
by: Lee, Hyunho, et al.
Published: (2024)
by: Lee, Hyunho, et al.
Published: (2024)
PM2: A New Prompting Multi-modal Model Paradigm for Few-shot Medical Image Classification
by: Wang, Zhenwei, et al.
Published: (2024)
by: Wang, Zhenwei, et al.
Published: (2024)
Task-Adapter: Task-specific Adaptation of Image Models for Few-shot Action Recognition
by: Cao, Congqi, et al.
Published: (2024)
by: Cao, Congqi, et al.
Published: (2024)
Boosting Pathology Foundation Models via Few-shot Prompt-tuning for Rare Cancer Subtyping
by: He, Dexuan, et al.
Published: (2025)
by: He, Dexuan, et al.
Published: (2025)
Causal Prompt Calibration Guided Segment Anything Model for Open-Vocabulary Multi-Entity Segmentation
by: Wang, Jingyao, et al.
Published: (2025)
by: Wang, Jingyao, et al.
Published: (2025)
Few-Shot Adaptation of Training-Free Foundation Model for 3D Medical Image Segmentation
by: He, Xingxin, et al.
Published: (2025)
by: He, Xingxin, et al.
Published: (2025)
Dynamic Prototype Adaptation with Distillation for Few-shot Point Cloud Segmentation
by: Liu, Jie, et al.
Published: (2024)
by: Liu, Jie, et al.
Published: (2024)
Cross-domain Few-shot Object Detection with Multi-modal Textual Enrichment
by: Shangguan, Zeyu, et al.
Published: (2025)
by: Shangguan, Zeyu, et al.
Published: (2025)
Dual-Adapter: Training-free Dual Adaptation for Few-shot Out-of-Distribution Detection
by: Chen, Xinyi, et al.
Published: (2024)
by: Chen, Xinyi, et al.
Published: (2024)
A Spatially Masked Adaptive Gated Network for multimodal post-flood water extent mapping using SAR and incomplete multispectral data
by: Lee, Hyunho, et al.
Published: (2025)
by: Lee, Hyunho, et al.
Published: (2025)
HAAF: Hierarchical Adaptation and Alignment of Foundation Models for Few-Shot Pathology Anomaly Detection
by: Yang, Chunze, et al.
Published: (2026)
by: Yang, Chunze, et al.
Published: (2026)
Multi-view Distillation based on Multi-modal Fusion for Few-shot Action Recognition(CLIP-$\mathrm{M^2}$DF)
by: Guo, Fei, et al.
Published: (2024)
by: Guo, Fei, et al.
Published: (2024)
Adapting a Pre-trained Single-Cell Foundation Model to Spatial Gene Expression Generation from Histology Images
by: Fang, Donghai, et al.
Published: (2026)
by: Fang, Donghai, et al.
Published: (2026)
Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation
by: Lai, Bolin, et al.
Published: (2024)
by: Lai, Bolin, et al.
Published: (2024)
Composed Multi-modal Retrieval: A Survey of Approaches and Applications
by: Zhang, Kun, et al.
Published: (2025)
by: Zhang, Kun, et al.
Published: (2025)
SeaS: Few-shot Industrial Anomaly Image Generation with Separation and Sharing Fine-tuning
by: Dai, Zhewei, et al.
Published: (2024)
by: Dai, Zhewei, et al.
Published: (2024)
FLIER: Few-shot Language Image Models Embedded with Latent Representations
by: Zhou, Zhinuo, et al.
Published: (2024)
by: Zhou, Zhinuo, et al.
Published: (2024)
SAMPO-Path: Segmentation Intent-Aligned Preference Optimization for Pathology Foundation Model Segmentation
by: Wu, Yonghuang, et al.
Published: (2025)
by: Wu, Yonghuang, et al.
Published: (2025)
Few-shot Object Localization
by: Ren, Yunhan, et al.
Published: (2024)
by: Ren, Yunhan, et al.
Published: (2024)
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
by: Yu, Zhixuan, et al.
Published: (2024)
by: Yu, Zhixuan, et al.
Published: (2024)
Few-shot Structure-Informed Machinery Part Segmentation with Foundation Models and Graph Neural Networks
by: Schwingshackl, Michael, et al.
Published: (2025)
by: Schwingshackl, Michael, et al.
Published: (2025)
Learning A Zero-shot Occupancy Network from Vision Foundation Models via Self-supervised Adaptation
by: Lin, Sihao, et al.
Published: (2025)
by: Lin, Sihao, et al.
Published: (2025)
Clustered-patch Element Connection for Few-shot Learning
by: Lai, Jinxiang, et al.
Published: (2023)
by: Lai, Jinxiang, et al.
Published: (2023)
Continual Few-shot Adaptation for Synthetic Fingerprint Detection
by: Benjamin, Joseph Geo, et al.
Published: (2026)
by: Benjamin, Joseph Geo, et al.
Published: (2026)
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
by: Zhu, Muzhi, et al.
Published: (2024)
by: Zhu, Muzhi, et al.
Published: (2024)
A Survey of Low-shot Vision-Language Model Adaptation via Representer Theorem
by: Ding, Kun, et al.
Published: (2024)
by: Ding, Kun, et al.
Published: (2024)
Similar Items
-
Making Large Vision Language Models to be Good Few-shot Learners
by: Liu, Fan, et al.
Published: (2024) -
RemoteCLIP: A Vision Language Foundation Model for Remote Sensing
by: Liu, Fan, et al.
Published: (2023) -
Unbiased Semantic Decoding with Vision Foundation Models for Few-shot Segmentation
by: Wang, Jin, et al.
Published: (2025) -
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
by: Chen, Tingxuan, et al.
Published: (2025) -
Retrieval-augmented Few-shot Medical Image Segmentation with Foundation Models
by: Zhao, Lin, et al.
Published: (2024)