Saved in:
| Main Authors: | Gao, Liying, Jiao, Bingliang, Wang, Peng, Zhang, Shizhou, Zhang, Hanwang, Zhang, Yanning |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2404.18695 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Prompt Selection for In-Context Learning Segmentation
by: Suo, Wei, et al.
Published: (2024)
by: Suo, Wei, et al.
Published: (2024)
Visual Prompt Tuning in Null Space for Continual Learning
by: Lu, Yue, et al.
Published: (2024)
by: Lu, Yue, et al.
Published: (2024)
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
by: Xing, Yinghui, et al.
Published: (2023)
by: Xing, Yinghui, et al.
Published: (2023)
Dynamic Textual Prompt For Rehearsal-free Lifelong Person Re-identification
by: Chen, Hongyu, et al.
Published: (2024)
by: Chen, Hongyu, et al.
Published: (2024)
Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning
by: Wu, Ruiqi, et al.
Published: (2024)
by: Wu, Ruiqi, et al.
Published: (2024)
AdaSemiCD: An Adaptive Semi-Supervised Change Detection Method Based on Pseudo-Label Evaluation
by: Lingyan, Ran, et al.
Published: (2024)
by: Lingyan, Ran, et al.
Published: (2024)
Adaptive Spatial Augmentation for Semi-supervised Semantic Segmentation
by: Ran, Lingyan, et al.
Published: (2025)
by: Ran, Lingyan, et al.
Published: (2025)
EMMA: Your Text-to-Image Diffusion Model Can Secretly Accept Multi-Modal Prompts
by: Han, Yucheng, et al.
Published: (2024)
by: Han, Yucheng, et al.
Published: (2024)
Knowing the Unknown: Interpretable Open-World Object Detection via Concept Decomposition Model
by: Lv, Xueqiang, et al.
Published: (2026)
by: Lv, Xueqiang, et al.
Published: (2026)
Demystifying Catastrophic Forgetting in Two-Stage Incremental Object Detector
by: Wu, Qirui, et al.
Published: (2025)
by: Wu, Qirui, et al.
Published: (2025)
CrossDiff: Exploring Self-Supervised Representation of Pansharpening via Cross-Predictive Diffusion Model
by: Xing, Yinghui, et al.
Published: (2024)
by: Xing, Yinghui, et al.
Published: (2024)
SeCap: Self-Calibrating and Adaptive Prompts for Cross-view Person Re-Identification in Aerial-Ground Networks
by: Wang, Shining, et al.
Published: (2025)
by: Wang, Shining, et al.
Published: (2025)
Prompt-Free Conditional Diffusion for Multi-object Image Augmentation
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
by: Yang, Jie, et al.
Published: (2024)
by: Yang, Jie, et al.
Published: (2024)
Cross-Platform Video Person ReID: A New Benchmark Dataset and Adaptation Approach
by: Zhang, Shizhou, et al.
Published: (2024)
by: Zhang, Shizhou, et al.
Published: (2024)
YOLO-IOD: Towards Real Time Incremental Object Detection
by: Zhang, Shizhou, et al.
Published: (2025)
by: Zhang, Shizhou, et al.
Published: (2025)
How to Handle Sketch-Abstraction in Sketch-Based Image Retrieval?
by: Koley, Subhadeep, et al.
Published: (2024)
by: Koley, Subhadeep, et al.
Published: (2024)
Frequency-Guided Spatial Adaptation for Camouflaged Object Detection
by: Zhang, Shizhou, et al.
Published: (2024)
by: Zhang, Shizhou, et al.
Published: (2024)
Prompt-aligned Gradient for Prompt Tuning
by: Zhu, Beier, et al.
Published: (2022)
by: Zhu, Beier, et al.
Published: (2022)
Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReID
by: Cheng, De, et al.
Published: (2023)
by: Cheng, De, et al.
Published: (2023)
Elevating All Zero-Shot Sketch-Based Image Retrieval Through Multimodal Prompt Learning
by: Singha, Mainak, et al.
Published: (2024)
by: Singha, Mainak, et al.
Published: (2024)
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Freeview Sketching: View-Aware Fine-Grained Sketch-Based Image Retrieval
by: Sain, Aneeshan, et al.
Published: (2024)
by: Sain, Aneeshan, et al.
Published: (2024)
Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval
by: Lyou, Eunyi, et al.
Published: (2024)
by: Lyou, Eunyi, et al.
Published: (2024)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
by: Zhang, Shizhou, et al.
Published: (2021)
by: Zhang, Shizhou, et al.
Published: (2021)
Multi-Modal Prompt Learning on Blind Image Quality Assessment
by: Pan, Wensheng, et al.
Published: (2024)
by: Pan, Wensheng, et al.
Published: (2024)
DuGI-MAE: Improving Infrared Mask Autoencoders via Dual-Domain Guidance
by: Xing, Yinghui, et al.
Published: (2025)
by: Xing, Yinghui, et al.
Published: (2025)
ModalPrompt: Towards Efficient Multimodal Continual Instruction Tuning with Dual-Modality Guided Prompt
by: Zeng, Fanhu, et al.
Published: (2024)
by: Zeng, Fanhu, et al.
Published: (2024)
Debiased Fine-Tuning for Vision-language Models by Prompt Regularization
by: Zhu, Beier, et al.
Published: (2023)
by: Zhu, Beier, et al.
Published: (2023)
DualTSR: Unified Dual-Diffusion Transformer for Scene Text Image Super-Resolution
by: Niu, Axi, et al.
Published: (2026)
by: Niu, Axi, et al.
Published: (2026)
A Plug-and-Play Method for Rare Human-Object Interactions Detection by Bridging Domain Gap
by: Zhang, Lijun, et al.
Published: (2024)
by: Zhang, Lijun, et al.
Published: (2024)
Enhancing CLIP Robustness via Cross-Modality Alignment
by: Zhu, Xingyu, et al.
Published: (2025)
by: Zhu, Xingyu, et al.
Published: (2025)
Towards Video Anomaly Retrieval from Video Anomaly Detection: New Benchmarks and Model
by: Wu, Peng, et al.
Published: (2023)
by: Wu, Peng, et al.
Published: (2023)
Biomed-DPT: Dual Modality Prompt Tuning for Biomedical Vision-Language Models
by: Peng, Wei, et al.
Published: (2025)
by: Peng, Wei, et al.
Published: (2025)
Relation-Aware Meta-Learning for Zero-shot Sketch-Based Image Retrieval
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
DMAT: A Dynamic Mask-Aware Transformer for Human De-occlusion
by: Liang, Guoqiang, et al.
Published: (2024)
by: Liang, Guoqiang, et al.
Published: (2024)
Towards Interactive Image Inpainting via Sketch Refinement
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
Enhancing Zero-Shot Vision Models by Label-Free Prompt Distribution Learning and Bias Correcting
by: Zhu, Xingyu, et al.
Published: (2024)
by: Zhu, Xingyu, et al.
Published: (2024)
Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval
by: Wang, Siyuan, et al.
Published: (2026)
by: Wang, Siyuan, et al.
Published: (2026)
Weakly Supervised Video Anomaly Detection and Localization with Spatio-Temporal Prompts
by: Wu, Peng, et al.
Published: (2024)
by: Wu, Peng, et al.
Published: (2024)
Similar Items
-
Visual Prompt Selection for In-Context Learning Segmentation
by: Suo, Wei, et al.
Published: (2024) -
Visual Prompt Tuning in Null Space for Continual Learning
by: Lu, Yue, et al.
Published: (2024) -
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
by: Xing, Yinghui, et al.
Published: (2023) -
Dynamic Textual Prompt For Rehearsal-free Lifelong Person Re-identification
by: Chen, Hongyu, et al.
Published: (2024) -
Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning
by: Wu, Ruiqi, et al.
Published: (2024)