Revisiting Relevance Feedback for CLIP-based Interactive Image Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Nara, Ryoya, Lin, Yu-Chieh, Nozawa, Yuji, Ng, Youyang, Itoh, Goh, Torii, Osamu, Matsui, Yusuke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains
by: Takeda, Tomohisa, et al.
Published: (2026)
by: Takeda, Tomohisa, et al.
Published: (2026)
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
by: Nozawa, Yuji, et al.
Published: (2025)
by: Nozawa, Yuji, et al.
Published: (2025)
Adversarial Doodles: Interpretable and Human-drawable Attacks Provide Describable Insights
by: Nara, Ryoya, et al.
Published: (2023)
by: Nara, Ryoya, et al.
Published: (2023)
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
by: Nakamura, Kazumoto, et al.
Published: (2024)
by: Nakamura, Kazumoto, et al.
Published: (2024)
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
by: Nakata, Kengo, et al.
Published: (2024)
by: Nakata, Kengo, et al.
Published: (2024)
ProPy: Building Interactive Prompt Pyramids upon CLIP for Partially Relevant Video Retrieval
by: Pan, Yi, et al.
Published: (2025)
by: Pan, Yi, et al.
Published: (2025)
High-Frequency Anti-DreamBooth: Robust Defense against Personalized Image Synthesis
by: Onikubo, Takuto, et al.
Published: (2024)
by: Onikubo, Takuto, et al.
Published: (2024)
Noisy Label Refinement with Semantically Reliable Synthetic Images
by: Li, Yingxuan, et al.
Published: (2025)
by: Li, Yingxuan, et al.
Published: (2025)
ZoDi: Zero-Shot Domain Adaptation with Diffusion-Based Image Transfer
by: Azuma, Hiroki, et al.
Published: (2024)
by: Azuma, Hiroki, et al.
Published: (2024)
LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table
by: Matsui, Yusuke
Published: (2025)
by: Matsui, Yusuke
Published: (2025)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
by: Karimian, Banafsheh, et al.
Published: (2025)
by: Karimian, Banafsheh, et al.
Published: (2025)
SVGEditBench: A Benchmark Dataset for Quantitative Assessment of LLM's SVG Editing Capabilities
by: Nishina, Kunato, et al.
Published: (2024)
by: Nishina, Kunato, et al.
Published: (2024)
RouteExtract: A Modular Pipeline for Extracting Routes from Paper Maps
by: Kremser, Bjoern, et al.
Published: (2025)
by: Kremser, Bjoern, et al.
Published: (2025)
MaskedCLIP: Bridging the Masked and CLIP Space for Semi-Supervised Medical Vision-Language Pre-training
by: Zhu, Lei, et al.
Published: (2025)
by: Zhu, Lei, et al.
Published: (2025)
A Little More Like This: Text-to-Image Retrieval with Vision-Language Models Using Relevance Feedback
by: Khaertdinov, Bulat, et al.
Published: (2025)
by: Khaertdinov, Bulat, et al.
Published: (2025)
Region-Wise Correspondence Prediction between Manga Line Art Images
by: Li, Yingxuan, et al.
Published: (2025)
by: Li, Yingxuan, et al.
Published: (2025)
Leadership Assessment in Pediatric Intensive Care Unit Team Training
by: Ouyang, Liangyang, et al.
Published: (2025)
by: Ouyang, Liangyang, et al.
Published: (2025)
Revisiting Image Captioning Training Paradigm via Direct CLIP-based Optimization
by: Moratelli, Nicholas, et al.
Published: (2024)
by: Moratelli, Nicholas, et al.
Published: (2024)
Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval
by: Li, Jun, et al.
Published: (2026)
by: Li, Jun, et al.
Published: (2026)
Diffusion Feedback Helps CLIP See Better
by: Wang, Wenxuan, et al.
Published: (2024)
by: Wang, Wenxuan, et al.
Published: (2024)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
by: Kang, Bin, et al.
Published: (2025)
by: Kang, Bin, et al.
Published: (2025)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
by: Asgarov, Ali, et al.
Published: (2024)
by: Asgarov, Ali, et al.
Published: (2024)
Manga109Dialog: A Large-scale Dialogue Dataset for Comics Speaker Detection
by: Li, Yingxuan, et al.
Published: (2023)
by: Li, Yingxuan, et al.
Published: (2023)
DualFete: Revisiting Teacher-Student Interactions from a Feedback Perspective for Semi-supervised Medical Image Segmentation
by: Yi, Le, et al.
Published: (2025)
by: Yi, Le, et al.
Published: (2025)
CLIP-based Synergistic Knowledge Transfer for Text-based Person Retrieval
by: Liu, Yating, et al.
Published: (2023)
by: Liu, Yating, et al.
Published: (2023)
CLAIR: CLIP-Aided Weakly Supervised Zero-Shot Cross-Domain Image Retrieval
by: Tan, Chor Boon, et al.
Published: (2025)
by: Tan, Chor Boon, et al.
Published: (2025)
Multi-Perspective Subimage CLIP with Keyword Guidance for Remote Sensing Image-Text Retrieval
by: Li, Yifan, et al.
Published: (2026)
by: Li, Yifan, et al.
Published: (2026)
CLIP-driven rain perception: Adaptive deraining with pattern-aware network routing and mask-guided cross-attention
by: Guan, Cong, et al.
Published: (2025)
by: Guan, Cong, et al.
Published: (2025)
Optimizing CLIP Models for Image Retrieval with Maintained Joint-Embedding Alignment
by: Schall, Konstantin, et al.
Published: (2024)
by: Schall, Konstantin, et al.
Published: (2024)
CLIP Multi-modal Hashing for Multimedia Retrieval
by: Zhu, Jian, et al.
Published: (2024)
by: Zhu, Jian, et al.
Published: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
by: Xie, Jingyou, et al.
Published: (2024)
by: Xie, Jingyou, et al.
Published: (2024)
CLIP in Medical Imaging: A Survey
by: Zhao, Zihao, et al.
Published: (2023)
by: Zhao, Zihao, et al.
Published: (2023)
Neural Image Space Tessellation efect
by: Du, Youyang, et al.
Published: (2026)
by: Du, Youyang, et al.
Published: (2026)
Missing Target-Relevant Information Prediction with World Model for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2025)
by: Tang, Yuanmin, et al.
Published: (2025)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
by: Zhang, Zilun, et al.
Published: (2022)
by: Zhang, Zilun, et al.
Published: (2022)
From Global to Local: Social Bias Transfer in CLIP
by: Ramos, Ryan, et al.
Published: (2025)
by: Ramos, Ryan, et al.
Published: (2025)
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
by: Che, Chang, et al.
Published: (2024)
by: Che, Chang, et al.
Published: (2024)
Convergence Analysis of Blurring Mean Shift
by: Yamasaki, Ryoya, et al.
Published: (2024)
by: Yamasaki, Ryoya, et al.
Published: (2024)
Grounding Emotion Recognition with Visual Prototypes: VEGA -- Revisiting CLIP in MERC
by: Hu, Guanyu, et al.
Published: (2025)
by: Hu, Guanyu, et al.
Published: (2025)
Benchmarking PathCLIP for Pathology Image Analysis
by: Zheng, Sunyi, et al.
Published: (2024)
by: Zheng, Sunyi, et al.
Published: (2024)
Similar Items
-
CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains
by: Takeda, Tomohisa, et al.
Published: (2026) -
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
by: Nozawa, Yuji, et al.
Published: (2025) -
Adversarial Doodles: Interpretable and Human-drawable Attacks Provide Describable Insights
by: Nara, Ryoya, et al.
Published: (2023) -
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
by: Nakamura, Kazumoto, et al.
Published: (2024) -
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
by: Nakata, Kengo, et al.
Published: (2024)