FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Chen-Bin, Sha, Youyang, Liu, Longfei, Yu, Yongjun, Vong, Chi Man, Yu, Xuanlong, Shen, Xi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps
by: Yu, Xuanlong, et al.
Published: (2026)
by: Yu, Xuanlong, et al.
Published: (2026)
Real-Time Object Detection Meets DINOv3
by: Huang, Shihua, et al.
Published: (2025)
by: Huang, Shihua, et al.
Published: (2025)
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
by: Liu, Longfei, et al.
Published: (2026)
by: Liu, Longfei, et al.
Published: (2026)
Boosting Few-Shot Semantic Segmentation Via Segment Anything Model
by: Feng, Chen-Bin, et al.
Published: (2024)
by: Feng, Chen-Bin, et al.
Published: (2024)
OV-DEIM: Real-time DETR-Style Open-Vocabulary Object Detection with GridSynthetic Augmentation
by: Wang, Leilei, et al.
Published: (2026)
by: Wang, Leilei, et al.
Published: (2026)
AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors
by: Fučka, Matic, et al.
Published: (2026)
by: Fučka, Matic, et al.
Published: (2026)
Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection
by: Lundqvist, Lars, et al.
Published: (2026)
by: Lundqvist, Lars, et al.
Published: (2026)
From Misclassifications to Outliers: Joint Reliability Assessment in Classification
by: Li, Yang, et al.
Published: (2026)
by: Li, Yang, et al.
Published: (2026)
PLOOD: Partial Label Learning with Out-of-distribution Objects
by: Huang, Jintao, et al.
Published: (2024)
by: Huang, Jintao, et al.
Published: (2024)
Decoupled Prototype Matching with Vision Foundation Models for Few-Shot Industrial Object Detection
by: M., Hari Prasanth S., et al.
Published: (2026)
by: M., Hari Prasanth S., et al.
Published: (2026)
Weakly-supervised Semantic Segmentation via Dual-stream Contrastive Learning of Cross-image Contextual Information
by: Lai, Qi, et al.
Published: (2024)
by: Lai, Qi, et al.
Published: (2024)
VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models
by: Bi, Tianci, et al.
Published: (2025)
by: Bi, Tianci, et al.
Published: (2025)
Revisiting Few-Shot Object Detection with Vision-Language Models
by: Madan, Anish, et al.
Published: (2023)
by: Madan, Anish, et al.
Published: (2023)
The Solution for CVPR2024 Foundational Few-Shot Object Detection Challenge
by: Pan, Hongpeng, et al.
Published: (2024)
by: Pan, Hongpeng, et al.
Published: (2024)
3D Hand Mesh-Guided AI-Generated Malformed Hand Refinement with Hand Pose Transformation via Diffusion Model
by: Feng, Chen-Bin, et al.
Published: (2025)
by: Feng, Chen-Bin, et al.
Published: (2025)
SemanticStitch: Enhancing Image Coherence through Foreground-Aware Seam Carving
by: Jin, Ji-Ping, et al.
Published: (2025)
by: Jin, Ji-Ping, et al.
Published: (2025)
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
Object-level Correlation for Few-Shot Segmentation
by: Wen, Chunlin, et al.
Published: (2025)
by: Wen, Chunlin, et al.
Published: (2025)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
by: Kumar, Yogesh, et al.
Published: (2025)
by: Kumar, Yogesh, et al.
Published: (2025)
The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results
by: Qiu, Xingyu, et al.
Published: (2026)
by: Qiu, Xingyu, et al.
Published: (2026)
Towards Fine-Grained Vision-Language Alignment for Few-Shot Anomaly Detection
by: Fan, Yuanting, et al.
Published: (2025)
by: Fan, Yuanting, et al.
Published: (2025)
SURE: SUrvey REcipes for building reliable and robust deep networks
by: Li, Yuting, et al.
Published: (2024)
by: Li, Yuting, et al.
Published: (2024)
BioVFM-21M: Benchmarking and Scaling Self-Supervised Vision Foundation Models for Biomedical Image Analysis
by: Liu, Jiarun, et al.
Published: (2025)
by: Liu, Jiarun, et al.
Published: (2025)
CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion
by: Meng, Boyuan, et al.
Published: (2025)
by: Meng, Boyuan, et al.
Published: (2025)
Few-Shot Object Detection: Research Advances and Challenges
by: Xin, Zhimeng, et al.
Published: (2024)
by: Xin, Zhimeng, et al.
Published: (2024)
Few-Shot Object Detection with Sparse Context Transformers
by: Mei, Jie, et al.
Published: (2024)
by: Mei, Jie, et al.
Published: (2024)
Prototype-Driven Adaptation for Few-Shot Object Detection
by: Huang, Yushen, et al.
Published: (2025)
by: Huang, Yushen, et al.
Published: (2025)
VFM-ISRefiner: Towards Better Adapting Vision Foundation Models for Interactive Segmentation of Remote Sensing Images
by: Wang, Deliang, et al.
Published: (2025)
by: Wang, Deliang, et al.
Published: (2025)
Few-Shot Image Generation by Conditional Relaxing Diffusion Inversion
by: Cao, Yu, et al.
Published: (2024)
by: Cao, Yu, et al.
Published: (2024)
ForgeryTTT: Zero-Shot Image Manipulation Localization with Test-Time Training
by: Liu, Weihuang, et al.
Published: (2024)
by: Liu, Weihuang, et al.
Published: (2024)
VFM-Recon: Unlocking Cross-Domain Scene-Level Neural Reconstruction with Scale-Aligned Foundation Priors
by: Ming, Yuhang, et al.
Published: (2026)
by: Ming, Yuhang, et al.
Published: (2026)
AdaVFM: Adaptive Vision Foundation Models for Edge Intelligence via LLM-Guided Execution
by: Zhao, Yiwei, et al.
Published: (2026)
by: Zhao, Yiwei, et al.
Published: (2026)
VTFusion: A Vision-Text Multimodal Fusion Network for Few-Shot Anomaly Detection
by: Jiang, Yuxin, et al.
Published: (2026)
by: Jiang, Yuxin, et al.
Published: (2026)
Towards Generalized Few-Shot Open-Set Object Detection
by: Su, Binyi, et al.
Published: (2022)
by: Su, Binyi, et al.
Published: (2022)
Fine-Grained Prototypes Distillation for Few-Shot Object Detection
by: Wang, Zichen, et al.
Published: (2024)
by: Wang, Zichen, et al.
Published: (2024)
Generalization-Enhanced Few-Shot Object Detection in Remote Sensing
by: Lin, Hui, et al.
Published: (2025)
by: Lin, Hui, et al.
Published: (2025)
VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection
by: Han, Jianhong, et al.
Published: (2025)
by: Han, Jianhong, et al.
Published: (2025)
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
by: Li, Zhenhao, et al.
Published: (2025)
by: Li, Zhenhao, et al.
Published: (2025)
Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Decoupling Classifier for Boosting Few-shot Object Detection and Instance Segmentation
by: Gao, Bin-Bin, et al.
Published: (2025)
by: Gao, Bin-Bin, et al.
Published: (2025)
Similar Items
-
A Closer Look at Cross-Domain Few-Shot Object Detection: Fine-Tuning Matters and Parallel Decoder Helps
by: Yu, Xuanlong, et al.
Published: (2026) -
Real-Time Object Detection Meets DINOv3
by: Huang, Shihua, et al.
Published: (2025) -
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
by: Liu, Longfei, et al.
Published: (2026) -
Boosting Few-Shot Semantic Segmentation Via Segment Anything Model
by: Feng, Chen-Bin, et al.
Published: (2024) -
OV-DEIM: Real-time DETR-Style Open-Vocabulary Object Detection with GridSynthetic Augmentation
by: Wang, Leilei, et al.
Published: (2026)