Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Xinhao, Li, Liulei, Pei, Gensheng, Chen, Tao, Pan, Jinshan, Yao, Yazhou, Wang, Wenguan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
by: Li, Liulei, et al.
Published: (2024)
by: Li, Liulei, et al.
Published: (2024)
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
by: Pei, Gensheng, et al.
Published: (2025)
by: Pei, Gensheng, et al.
Published: (2025)
Beyond Quadratic: Linear-Time Change Detection with RWKV
by: Yang, Zhenyu, et al.
Published: (2026)
by: Yang, Zhenyu, et al.
Published: (2026)
Poly Kernel Inception Network for Remote Sensing Detection
by: Cai, Xinhao, et al.
Published: (2024)
by: Cai, Xinhao, et al.
Published: (2024)
Clustering Propagation for Universal Medical Image Segmentation
by: Ding, Yuhang, et al.
Published: (2024)
by: Ding, Yuhang, et al.
Published: (2024)
DIFFVSGG: Diffusion-Driven Online Video Scene Graph Generation
by: Chen, Mu, et al.
Published: (2025)
by: Chen, Mu, et al.
Published: (2025)
Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
Relating CNN-Transformer Fusion Network for Change Detection
by: Gao, Yuhao, et al.
Published: (2024)
by: Gao, Yuhao, et al.
Published: (2024)
VideoMAC: Video Masked Autoencoders Meet ConvNets
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Towards Remote Sensing Change Detection with Neural Memory
by: Yang, Zhenyu, et al.
Published: (2026)
by: Yang, Zhenyu, et al.
Published: (2026)
A Light-weight Transformer-based Self-supervised Matching Network for Heterogeneous Images
by: Zhang, Wang, et al.
Published: (2024)
by: Zhang, Wang, et al.
Published: (2024)
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
by: Yin, Jianjian, et al.
Published: (2025)
by: Yin, Jianjian, et al.
Published: (2025)
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
by: Yin, Jianjian, et al.
Published: (2026)
by: Yin, Jianjian, et al.
Published: (2026)
General and Task-Oriented Video Segmentation
by: Chen, Mu, et al.
Published: (2024)
by: Chen, Mu, et al.
Published: (2024)
Efficiency Follows Global-Local Decoupling
by: Yang, Zhenyu, et al.
Published: (2026)
by: Yang, Zhenyu, et al.
Published: (2026)
Seeing the Unseen: A Frequency Prompt Guided Transformer for Image Restoration
by: Zhou, Shihao, et al.
Published: (2024)
by: Zhou, Shihao, et al.
Published: (2024)
PGP-DiffSR: Phase-Guided Progressive Pruning for Efficient Diffusion-based Image Super-Resolution
by: Yang, Zhongbao, et al.
Published: (2025)
by: Yang, Zhongbao, et al.
Published: (2025)
Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
by: Zhou, Bo, et al.
Published: (2026)
by: Zhou, Bo, et al.
Published: (2026)
Look-Around Before You Leap: High-Frequency Injected Transformer for Image Restoration
by: Zhou, Shihao, et al.
Published: (2024)
by: Zhou, Shihao, et al.
Published: (2024)
Visual Textualization for Image Prompted Object Detection
by: Wu, Yongjian, et al.
Published: (2025)
by: Wu, Yongjian, et al.
Published: (2025)
Frequency-based Matcher for Long-tailed Semantic Segmentation
by: Li, Shan, et al.
Published: (2024)
by: Li, Shan, et al.
Published: (2024)
Intra and Inter Parser-Prompted Transformers for Effective Image Restoration
by: Wang, Cong, et al.
Published: (2025)
by: Wang, Cong, et al.
Published: (2025)
FreeInpaint: Tuning-free Prompt Alignment and Visual Rationality Enhancement in Image Inpainting
by: Gong, Chao, et al.
Published: (2025)
by: Gong, Chao, et al.
Published: (2025)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
by: Yang, Shuai, et al.
Published: (2026)
by: Yang, Shuai, et al.
Published: (2026)
Efficient Concertormer for Image Deblurring and Beyond
by: Kuo, Pin-Hung, et al.
Published: (2024)
by: Kuo, Pin-Hung, et al.
Published: (2024)
Frequency Domain-Based Diffusion Model for Unpaired Image Dehazing
by: Liu, Chengxu, et al.
Published: (2025)
by: Liu, Chengxu, et al.
Published: (2025)
Neural Discrimination-Prompted Transformers for Efficient UHD Image Restoration and Enhancement
by: Wang, Cong, et al.
Published: (2026)
by: Wang, Cong, et al.
Published: (2026)
Towards Unbiased Source-Free Object Detection via Vision Foundation Models
by: Cai, Zhi, et al.
Published: (2026)
by: Cai, Zhi, et al.
Published: (2026)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Bidirectional Multi-Scale Implicit Neural Representations for Image Deraining
by: Chen, Xiang, et al.
Published: (2024)
by: Chen, Xiang, et al.
Published: (2024)
FaithDiff: Unleashing Diffusion Priors for Faithful Image Super-resolution
by: Chen, Junyang, et al.
Published: (2024)
by: Chen, Junyang, et al.
Published: (2024)
Neural Clustering based Visual Representation Learning
by: Chen, Guikun, et al.
Published: (2024)
by: Chen, Guikun, et al.
Published: (2024)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
by: Lee, Sanghoon, et al.
Published: (2026)
by: Lee, Sanghoon, et al.
Published: (2026)
Learning Human-Object Interaction as Groups
by: Hong, Jiajun, et al.
Published: (2025)
by: Hong, Jiajun, et al.
Published: (2025)
Efficient Visual State Space Model for Image Deblurring
by: Kong, Lingshun, et al.
Published: (2024)
by: Kong, Lingshun, et al.
Published: (2024)
Similar Items
-
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026) -
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
by: Cai, Xinhao, et al.
Published: (2026) -
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
by: Pei, Gensheng, et al.
Published: (2024) -
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
by: Li, Liulei, et al.
Published: (2024) -
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
by: Pei, Gensheng, et al.
Published: (2026)