Visual Textualization for Image Prompted Object Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Yongjian, Zhou, Yang, Saiyin, Jiya, Wei, Bingzheng, Xu, Yan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
por: Wu, Yongjian, et al.
Publicado: (2024)
por: Wu, Yongjian, et al.
Publicado: (2024)
SDPT: Synchronous Dual Prompt Tuning for Fusion-based Visual-Language Pre-trained Models
por: Zhou, Yang, et al.
Publicado: (2024)
por: Zhou, Yang, et al.
Publicado: (2024)
TAT: Task-Adaptive Transformer for All-in-One Medical Image Restoration
por: Yang, Zhiwen, et al.
Publicado: (2025)
por: Yang, Zhiwen, et al.
Publicado: (2025)
Textualize Visual Prompt for Image Editing via Diffusion Bridge
por: Xu, Pengcheng, et al.
Publicado: (2025)
por: Xu, Pengcheng, et al.
Publicado: (2025)
All-in-One Medical Image Restoration with Latent Diffusion-Enhanced Vector-Quantized Codebook Prior
por: Chen, Haowei, et al.
Publicado: (2025)
por: Chen, Haowei, et al.
Publicado: (2025)
Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
por: Cai, Xinhao, et al.
Publicado: (2025)
por: Cai, Xinhao, et al.
Publicado: (2025)
Textual and Visual Prompt Fusion for Image Editing via Step-Wise Alignment
por: Feng, Zhanbo, et al.
Publicado: (2023)
por: Feng, Zhanbo, et al.
Publicado: (2023)
Adaptive Prompt Learning with Negative Textual Semantics and Uncertainty Modeling for Universal Multi-Source Domain Adaptation
por: Yang, Yuxiang, et al.
Publicado: (2024)
por: Yang, Yuxiang, et al.
Publicado: (2024)
CTIS-QA: Clinical Template-Informed Slide-level Question Answering for Pathology
por: Lu, Hao, et al.
Publicado: (2026)
por: Lu, Hao, et al.
Publicado: (2026)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
por: Wu, Daiqing, et al.
Publicado: (2025)
por: Wu, Daiqing, et al.
Publicado: (2025)
TCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
por: Yao, Hantao, et al.
Publicado: (2023)
por: Yao, Hantao, et al.
Publicado: (2023)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
por: Wang, Zhifeng, et al.
Publicado: (2025)
por: Wang, Zhifeng, et al.
Publicado: (2025)
All-In-One Medical Image Restoration via Task-Adaptive Routing
por: Yang, Zhiwen, et al.
Publicado: (2024)
por: Yang, Zhiwen, et al.
Publicado: (2024)
T-Rex-Omni: Integrating Negative Visual Prompt in Generic Object Detection
por: Zhou, Jiazhou, et al.
Publicado: (2025)
por: Zhou, Jiazhou, et al.
Publicado: (2025)
Bringing Textual Prompt to AI-Generated Image Quality Assessment
por: Qu, Bowen, et al.
Publicado: (2024)
por: Qu, Bowen, et al.
Publicado: (2024)
Visual Consensus Prompting for Co-Salient Object Detection
por: Wang, Jie, et al.
Publicado: (2025)
por: Wang, Jie, et al.
Publicado: (2025)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
por: Guo, Pinxue, et al.
Publicado: (2024)
por: Guo, Pinxue, et al.
Publicado: (2024)
Parameterized Prompt for Incremental Object Detection
por: An, Zijia, et al.
Publicado: (2025)
por: An, Zijia, et al.
Publicado: (2025)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
Towards Generalizable AI-Generated Image Detection via Image-Adaptive Prompt Learning
por: Li, Yiheng, et al.
Publicado: (2025)
por: Li, Yiheng, et al.
Publicado: (2025)
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning
por: Wang, Yifan, et al.
Publicado: (2026)
por: Wang, Yifan, et al.
Publicado: (2026)
FEAT: Full-Dimensional Efficient Attention Transformer for Medical Video Generation
por: Wang, Huihan, et al.
Publicado: (2025)
por: Wang, Huihan, et al.
Publicado: (2025)
Towards Single-Source Domain Generalized Object Detection via Causal Visual Prompts
por: Li, Chen, et al.
Publicado: (2025)
por: Li, Chen, et al.
Publicado: (2025)
Restore-RWKV: Efficient and Effective Medical Image Restoration with RWKV
por: Yang, Zhiwen, et al.
Publicado: (2024)
por: Yang, Zhiwen, et al.
Publicado: (2024)
iDPA: Instance Decoupled Prompt Attention for Incremental Medical Object Detection
por: Yi, Huahui, et al.
Publicado: (2025)
por: Yi, Huahui, et al.
Publicado: (2025)
Region Attention Transformer for Medical Image Restoration
por: Yang, Zhiwen, et al.
Publicado: (2024)
por: Yang, Zhiwen, et al.
Publicado: (2024)
Explicit Visual Prompts for Visual Object Tracking
por: Shi, Liangtao, et al.
Publicado: (2024)
por: Shi, Liangtao, et al.
Publicado: (2024)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
por: Yang, Shuai, et al.
Publicado: (2026)
por: Yang, Shuai, et al.
Publicado: (2026)
Show or Tell? A Benchmark To Evaluate Visual and Textual Prompts in Semantic Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2025)
por: Rosi, Gabriele, et al.
Publicado: (2025)
Advancing Textual Prompt Learning with Anchored Attributes
por: Li, Zheng, et al.
Publicado: (2024)
por: Li, Zheng, et al.
Publicado: (2024)
VSCode: General Visual Salient and Camouflaged Object Detection with 2D Prompt Learning
por: Luo, Ziyang, et al.
Publicado: (2023)
por: Luo, Ziyang, et al.
Publicado: (2023)
Generalized Small Object Detection:A Point-Prompted Paradigm and Benchmark
por: Zhu, Haoran, et al.
Publicado: (2026)
por: Zhu, Haoran, et al.
Publicado: (2026)
Object-level Visual Prompts for Compositional Image Generation
por: Parmar, Gaurav, et al.
Publicado: (2025)
por: Parmar, Gaurav, et al.
Publicado: (2025)
Modality Prompts for Arbitrary Modality Salient Object Detection
por: Huang, Nianchang, et al.
Publicado: (2024)
por: Huang, Nianchang, et al.
Publicado: (2024)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
por: Li, Jiachen, et al.
Publicado: (2026)
por: Li, Jiachen, et al.
Publicado: (2026)
SPLF-SAM: Self-Prompting Segment Anything Model for Light Field Salient Object Detection
por: Xu, Qiyao, et al.
Publicado: (2025)
por: Xu, Qiyao, et al.
Publicado: (2025)
Unlocking Textual and Visual Wisdom: Open-Vocabulary 3D Object Detection Enhanced by Comprehensive Guidance from Text and Image
por: Jiao, Pengkun, et al.
Publicado: (2024)
por: Jiao, Pengkun, et al.
Publicado: (2024)
TP-GMOT: Tracking Generic Multiple Object by Textual Prompt with Motion-Appearance Cost (MAC) SORT
por: Anh, Duy Le Dinh, et al.
Publicado: (2024)
por: Anh, Duy Le Dinh, et al.
Publicado: (2024)
Token Coordinated Prompt Attention is Needed for Visual Prompting
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
Robust Tiny Object Detection in Aerial Images amidst Label Noise
por: Zhu, Haoran, et al.
Publicado: (2024)
por: Zhu, Haoran, et al.
Publicado: (2024)
Ejemplares similares
-
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
por: Wu, Yongjian, et al.
Publicado: (2024) -
SDPT: Synchronous Dual Prompt Tuning for Fusion-based Visual-Language Pre-trained Models
por: Zhou, Yang, et al.
Publicado: (2024) -
TAT: Task-Adaptive Transformer for All-in-One Medical Image Restoration
por: Yang, Zhiwen, et al.
Publicado: (2025) -
Textualize Visual Prompt for Image Editing via Diffusion Bridge
por: Xu, Pengcheng, et al.
Publicado: (2025) -
All-in-One Medical Image Restoration with Latent Diffusion-Enhanced Vector-Quantized Codebook Prior
por: Chen, Haowei, et al.
Publicado: (2025)