Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
Fuente:
arXiv
Guardado en:
| Autores principales: | Nozawa, Yuji, Lin, Yu-Chieh, Nakamura, Kazumoto, Ng, Youyang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
por: Nakamura, Kazumoto, et al.
Publicado: (2024)
por: Nakamura, Kazumoto, et al.
Publicado: (2024)
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
por: Nakata, Kengo, et al.
Publicado: (2024)
por: Nakata, Kengo, et al.
Publicado: (2024)
Revisiting Relevance Feedback for CLIP-based Interactive Image Retrieval
por: Nara, Ryoya, et al.
Publicado: (2024)
por: Nara, Ryoya, et al.
Publicado: (2024)
Human-Oriented Image Retrieval System (HORSE): A Neuro-Symbolic Approach to Optimizing Retrieval of Previewed Images
por: Weinberg, Abraham Itzhak
Publicado: (2025)
por: Weinberg, Abraham Itzhak
Publicado: (2025)
Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
por: Wang, Yifan, et al.
Publicado: (2025)
por: Wang, Yifan, et al.
Publicado: (2025)
AlbumFill: Album-Guided Reasoning and Retrieval for Personalized Image Completion
por: Tsai, Yu-Ju, et al.
Publicado: (2026)
por: Tsai, Yu-Ju, et al.
Publicado: (2026)
RAGAR: Retrieval Augmented Personalized Image Generation Guided by Recommendation
por: Ling, Run, et al.
Publicado: (2025)
por: Ling, Run, et al.
Publicado: (2025)
Large Language Model Informed Patent Image Retrieval
por: Lo, Hao-Cheng, et al.
Publicado: (2024)
por: Lo, Hao-Cheng, et al.
Publicado: (2024)
FIGROTD: A Friendly-to-Handle Dataset for Image Guided Retrieval with Optional Text
por: Le, Hoang-Bao, et al.
Publicado: (2025)
por: Le, Hoang-Bao, et al.
Publicado: (2025)
Active Learning via Classifier Impact and Greedy Selection for Interactive Image Retrieval
por: Bar, Leah, et al.
Publicado: (2024)
por: Bar, Leah, et al.
Publicado: (2024)
MLLM-Guided VLM Fine-Tuning with Joint Inference for Zero-Shot Composed Image Retrieval
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
UNION: A Lightweight Target Representation for Efficient Zero-Shot Image-Guided Retrieval with Optional Textual Queries
por: Le, Hoang-Bao, et al.
Publicado: (2025)
por: Le, Hoang-Bao, et al.
Publicado: (2025)
CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains
por: Takeda, Tomohisa, et al.
Publicado: (2026)
por: Takeda, Tomohisa, et al.
Publicado: (2026)
ITSELF: Attention Guided Fine-Grained Alignment for Vision-Language Retrieval
por: Nguyen, Tien-Huy, et al.
Publicado: (2026)
por: Nguyen, Tien-Huy, et al.
Publicado: (2026)
Entity Image and Mixed-Modal Image Retrieval Datasets
por: Blaga, Cristian-Ioan, et al.
Publicado: (2025)
por: Blaga, Cristian-Ioan, et al.
Publicado: (2025)
Retrieval-Guided Generation for Safer Histopathology Image Captioning
por: Hoq, Md. Enamul, et al.
Publicado: (2026)
por: Hoq, Md. Enamul, et al.
Publicado: (2026)
Chain-of-Thought Re-ranking for Image Retrieval Tasks
por: Wu, Shangrong, et al.
Publicado: (2025)
por: Wu, Shangrong, et al.
Publicado: (2025)
Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction
por: Zhang, Yao, et al.
Publicado: (2026)
por: Zhang, Yao, et al.
Publicado: (2026)
Attention Grounded Enhancement for Visual Document Retrieval
por: Cui, Wanqing, et al.
Publicado: (2025)
por: Cui, Wanqing, et al.
Publicado: (2025)
TIGER-FG: Text-Guided Implicit Fine-Grained Grounding for E-commerce Retrieval
por: Sun, Xinyu, et al.
Publicado: (2026)
por: Sun, Xinyu, et al.
Publicado: (2026)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
Leveraging Foundation Models for Content-Based Image Retrieval in Radiology
por: Denner, Stefan, et al.
Publicado: (2024)
por: Denner, Stefan, et al.
Publicado: (2024)
CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion
por: Gu, Geonmo, et al.
Publicado: (2023)
por: Gu, Geonmo, et al.
Publicado: (2023)
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories
por: Deng, Chenlong, et al.
Publicado: (2026)
por: Deng, Chenlong, et al.
Publicado: (2026)
Towards Text-Image Interleaved Retrieval
por: Zhang, Xin, et al.
Publicado: (2025)
por: Zhang, Xin, et al.
Publicado: (2025)
CoLLM: A Large Language Model for Composed Image Retrieval
por: Huynh, Chuong, et al.
Publicado: (2025)
por: Huynh, Chuong, et al.
Publicado: (2025)
Geometric Analysis of Self-Supervised Vision Representations for Semantic Image Retrieval
por: Rodríguez-Betancourt, Esteban, et al.
Publicado: (2026)
por: Rodríguez-Betancourt, Esteban, et al.
Publicado: (2026)
Language-only Efficient Training of Zero-shot Composed Image Retrieval
por: Gu, Geonmo, et al.
Publicado: (2023)
por: Gu, Geonmo, et al.
Publicado: (2023)
Validation of Whole-Slide Foundation Models for Image Retrieval in TCGA Data
por: Lei, Tianhao, et al.
Publicado: (2026)
por: Lei, Tianhao, et al.
Publicado: (2026)
Modality and Task Adaptation for Enhanced Zero-shot Composed Image Retrieval
por: Li, Haiwen, et al.
Publicado: (2024)
por: Li, Haiwen, et al.
Publicado: (2024)
EndoFinder: Online Image Retrieval for Explainable Colorectal Polyp Diagnosis
por: Yang, Ruijie, et al.
Publicado: (2024)
por: Yang, Ruijie, et al.
Publicado: (2024)
ADaFuSE: Adaptive Diffusion-generated Image and Text Fusion for Interactive Text-to-Image Retrieval
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
FGNet: Leveraging Feature-Guided Attention to Refine SAM2 for 3D EM Neuron Segmentation
por: Li, Zhenghua, et al.
Publicado: (2025)
por: Li, Zhenghua, et al.
Publicado: (2025)
Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation
por: Luo, Weiqing, et al.
Publicado: (2026)
por: Luo, Weiqing, et al.
Publicado: (2026)
Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs
por: Zhang, Huaying, et al.
Publicado: (2024)
por: Zhang, Huaying, et al.
Publicado: (2024)
A Resource-Efficient Training Framework for Remote Sensing Text--Image Retrieval
por: Zhang, Weihang, et al.
Publicado: (2025)
por: Zhang, Weihang, et al.
Publicado: (2025)
Iterative Optimal Attention and Local Model for Single Image Rain Streak Removal
por: Li, Xiangyu, et al.
Publicado: (2025)
por: Li, Xiangyu, et al.
Publicado: (2025)
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval
por: Agnolucci, Lorenzo, et al.
Publicado: (2024)
por: Agnolucci, Lorenzo, et al.
Publicado: (2024)
Leveraging High-Resolution Features for Improved Deep Hashing-based Image Retrieval
por: Berriche, Aymene, et al.
Publicado: (2024)
por: Berriche, Aymene, et al.
Publicado: (2024)
ProGEO: Generating Prompts through Image-Text Contrastive Learning for Visual Geo-localization
por: Mao, Chen, et al.
Publicado: (2024)
por: Mao, Chen, et al.
Publicado: (2024)
Ejemplares similares
-
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
por: Nakamura, Kazumoto, et al.
Publicado: (2024) -
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
por: Nakata, Kengo, et al.
Publicado: (2024) -
Revisiting Relevance Feedback for CLIP-based Interactive Image Retrieval
por: Nara, Ryoya, et al.
Publicado: (2024) -
Human-Oriented Image Retrieval System (HORSE): A Neuro-Symbolic Approach to Optimizing Retrieval of Previewed Images
por: Weinberg, Abraham Itzhak
Publicado: (2025) -
Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
por: Wang, Yifan, et al.
Publicado: (2025)