VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yupeng, Han, Ruize, Guo, Ningnan, Feng, Wei, Wang, Song, Wan, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2025)
by: Zhang, Yupeng, et al.
Published: (2025)
LV-OSD: Language-Vision-Complementary Open-Set Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
Bridge: Basis-Driven Causal Inference Marries VFMs for Domain Generalization
by: Hong, Mingbo, et al.
Published: (2026)
by: Hong, Mingbo, et al.
Published: (2026)
Unveiling the Power of Self-supervision for Multi-view Multi-human Association and Tracking
by: Feng, Wei, et al.
Published: (2024)
by: Feng, Wei, et al.
Published: (2024)
VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection
by: Han, Jianhong, et al.
Published: (2025)
by: Han, Jianhong, et al.
Published: (2025)
From Indoor To Outdoor: Unsupervised Domain Adaptive Gait Recognition
by: Wang, Likai, et al.
Published: (2022)
by: Wang, Likai, et al.
Published: (2022)
Style-Adaptive Detection Transformer for Single-Source Domain Generalized Object Detection
by: Han, Jianhong, et al.
Published: (2025)
by: Han, Jianhong, et al.
Published: (2025)
Multi-Granularity Feature Calibration via VFM for Domain Generalized Semantic Segmentation
by: Li, Xinhui, et al.
Published: (2025)
by: Li, Xinhui, et al.
Published: (2025)
OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking
by: Liang, Haiji, et al.
Published: (2024)
by: Liang, Haiji, et al.
Published: (2024)
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
by: Feng, Chen-Bin, et al.
Published: (2026)
by: Feng, Chen-Bin, et al.
Published: (2026)
Stronger, Steadier & Superior: Geometric Consistency in Depth VFM Forges Domain Generalized Semantic Segmentation
by: Chen, Siyu, et al.
Published: (2025)
by: Chen, Siyu, et al.
Published: (2025)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
OCTrack: Benchmarking the Open-Corpus Multi-Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
Simulating Distribution Dynamics: Liquid Temporal Feature Evolution for Single-Domain Generalized Object Detection
by: Zhang, Zihao, et al.
Published: (2025)
by: Zhang, Zihao, et al.
Published: (2025)
Does Your VFM Speak Plant? The Botanical Grammar of Vision Foundation Models for Object Detection
by: Lundqvist, Lars, et al.
Published: (2026)
by: Lundqvist, Lars, et al.
Published: (2026)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
by: Ma, Yanbiao, et al.
Published: (2025)
by: Ma, Yanbiao, et al.
Published: (2025)
Single-Domain Generalized Object Detection by Balancing Domain Diversity and Invariance
by: He, Zhenwei, et al.
Published: (2025)
by: He, Zhenwei, et al.
Published: (2025)
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
GOOD: Towards Domain Generalized Orientated Object Detection
by: Bi, Qi, et al.
Published: (2024)
by: Bi, Qi, et al.
Published: (2024)
Phrase Grounding-based Style Transfer for Single-Domain Generalized Object Detection
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
RaffeSDG: Random Frequency Filtering enabled Single-source Domain Generalization for Medical Image Segmentation
by: Li, Heng, et al.
Published: (2024)
by: Li, Heng, et al.
Published: (2024)
Domain Similarity-Perceived Label Assignment for Domain Generalized Underwater Object Detection
by: Li, Xisheng, et al.
Published: (2023)
by: Li, Xisheng, et al.
Published: (2023)
CLIPVehicle: A Unified Framework for Vision-based Vehicle Search
by: Wang, Likai, et al.
Published: (2025)
by: Wang, Likai, et al.
Published: (2025)
FOUND: Fourier-based von Mises Distribution for Robust Single Domain Generalization in Object Detection
by: Wang, Mengzhu, et al.
Published: (2025)
by: Wang, Mengzhu, et al.
Published: (2025)
DragEntity: Trajectory Guided Video Generation using Entity and Positional Relationships
by: Wan, Zhang, et al.
Published: (2024)
by: Wan, Zhang, et al.
Published: (2024)
From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration
by: Qian, Zekun, et al.
Published: (2022)
by: Qian, Zekun, et al.
Published: (2022)
Online Reasoning Video Object Segmentation
by: Liu, Jinyuan, et al.
Published: (2026)
by: Liu, Jinyuan, et al.
Published: (2026)
Prompt-Driven Dynamic Object-Centric Learning for Single Domain Generalization
by: Li, Deng, et al.
Published: (2024)
by: Li, Deng, et al.
Published: (2024)
Improving Single Domain-Generalized Object Detection: A Focus on Diversification and Alignment
by: Danish, Muhammad Sohail, et al.
Published: (2024)
by: Danish, Muhammad Sohail, et al.
Published: (2024)
Unbiased Faster R-CNN for Single-source Domain Generalized Object Detection
by: Liu, Yajing, et al.
Published: (2024)
by: Liu, Yajing, et al.
Published: (2024)
Synthetic-To-Real Video Person Re-ID
by: Zhang, Xiangqun, et al.
Published: (2024)
by: Zhang, Xiangqun, et al.
Published: (2024)
What is the Added Value of UDA in the VFM Era?
by: Englert, Brunó B., et al.
Published: (2025)
by: Englert, Brunó B., et al.
Published: (2025)
Towards Single-Source Domain Generalized Object Detection via Causal Visual Prompts
by: Li, Chen, et al.
Published: (2025)
by: Li, Chen, et al.
Published: (2025)
G-NAS: Generalizable Neural Architecture Search for Single Domain Generalization Object Detection
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Robust Object Detection of Underwater Robot based on Domain Generalization
by: Song, Pinhao
Published: (2025)
by: Song, Pinhao
Published: (2025)
VFM-Recon: Unlocking Cross-Domain Scene-Level Neural Reconstruction with Scale-Aligned Foundation Priors
by: Ming, Yuhang, et al.
Published: (2026)
by: Ming, Yuhang, et al.
Published: (2026)
Incremental Object Detection with CLIP
by: Huang, Ziyue, et al.
Published: (2023)
by: Huang, Ziyue, et al.
Published: (2023)
Similar Items
-
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2025) -
LV-OSD: Language-Vision-Complementary Open-Set Object Detection
by: Zhang, Yupeng, et al.
Published: (2026) -
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2026) -
COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection
by: Zhang, Yupeng, et al.
Published: (2026) -
Bridge: Basis-Driven Causal Inference Marries VFMs for Domain Generalization
by: Hong, Mingbo, et al.
Published: (2026)