Enhancing Supervised Composed Image Retrieval via Reasoning-Augmented Representation Engineering
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Jun, Dou, Hongjian, Zhang, Zhenyu, Li, Kai, Liu, Shaoguo, Gao, Tingting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning
di: Lv, Guannan, et al.
Pubblicazione: (2026)
di: Lv, Guannan, et al.
Pubblicazione: (2026)
CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval
di: Jiang, Xintong, et al.
Pubblicazione: (2024)
di: Jiang, Xintong, et al.
Pubblicazione: (2024)
CSMCIR: CoT-Enhanced Symmetric Alignment with Memory Bank for Composed Image Retrieval
di: Qian, Zhipeng, et al.
Pubblicazione: (2026)
di: Qian, Zhipeng, et al.
Pubblicazione: (2026)
TMCIR: Token Merge Benefits Composed Image Retrieval
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
di: Wang, Chaoyang, et al.
Pubblicazione: (2025)
MELT: Improve Composed Image Retrieval via the Modification Frequentation-Rarity Balance Network
di: Qiu, Guozhi, et al.
Pubblicazione: (2026)
di: Qiu, Guozhi, et al.
Pubblicazione: (2026)
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
di: Fu, Kang, et al.
Pubblicazione: (2026)
di: Fu, Kang, et al.
Pubblicazione: (2026)
Dual Relation Alignment for Composed Image Retrieval
di: Jiang, Xintong, et al.
Pubblicazione: (2023)
di: Jiang, Xintong, et al.
Pubblicazione: (2023)
Spherical Linear Interpolation and Text-Anchoring for Zero-shot Composed Image Retrieval
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval
di: Yang, Yuxin, et al.
Pubblicazione: (2025)
di: Yang, Yuxin, et al.
Pubblicazione: (2025)
Improving Composed Image Retrieval via Contrastive Learning with Scaling Positives and Negatives
di: Feng, Zhangchi, et al.
Pubblicazione: (2024)
di: Feng, Zhangchi, et al.
Pubblicazione: (2024)
Scale Up Composed Image Retrieval Learning via Modification Text Generation
di: Zhou, Yinan, et al.
Pubblicazione: (2025)
di: Zhou, Yinan, et al.
Pubblicazione: (2025)
FineCIR: Explicit Parsing of Fine-Grained Modification Semantics for Composed Image Retrieval
di: Li, Zixu, et al.
Pubblicazione: (2025)
di: Li, Zixu, et al.
Pubblicazione: (2025)
FBCIR: Balancing Cross-Modal Focuses in Composed Image Retrieval
di: Zhao, Chenchen, et al.
Pubblicazione: (2026)
di: Zhao, Chenchen, et al.
Pubblicazione: (2026)
Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval
di: Liu, Mingyu, et al.
Pubblicazione: (2026)
di: Liu, Mingyu, et al.
Pubblicazione: (2026)
Composed Vision-Language Retrieval for Skin Cancer Case Search via Joint Alignment of Global and Local Representations
di: Wang, Yuheng, et al.
Pubblicazione: (2026)
di: Wang, Yuheng, et al.
Pubblicazione: (2026)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
di: Kwak, Jaehyun, et al.
Pubblicazione: (2025)
di: Kwak, Jaehyun, et al.
Pubblicazione: (2025)
Reasoning-Augmented Representations for Multimodal Retrieval
di: Zhang, Jianrui, et al.
Pubblicazione: (2026)
di: Zhang, Jianrui, et al.
Pubblicazione: (2026)
ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing
di: Li, Lingen, et al.
Pubblicazione: (2025)
di: Li, Lingen, et al.
Pubblicazione: (2025)
good4cir: Generating Detailed Synthetic Captions for Composed Image Retrieval
di: Kolouju, Pranavi, et al.
Pubblicazione: (2025)
di: Kolouju, Pranavi, et al.
Pubblicazione: (2025)
Hybrid Global-Local Representation with Augmented Spatial Guidance for Zero-Shot Referring Image Segmentation
di: Liu, Ting, et al.
Pubblicazione: (2025)
di: Liu, Ting, et al.
Pubblicazione: (2025)
Decoupling the Image Perception and Multimodal Reasoning for Reasoning Segmentation with Digital Twin Representations
di: Li, Yizhen, et al.
Pubblicazione: (2025)
di: Li, Yizhen, et al.
Pubblicazione: (2025)
SeriesBench: A Benchmark for Narrative-Driven Drama Series Understanding
di: Zhang, Chenkai, et al.
Pubblicazione: (2025)
di: Zhang, Chenkai, et al.
Pubblicazione: (2025)
Knowledge Completes the Vision: A Multimodal Entity-aware Retrieval-Augmented Generation Framework for News Image Captioning
di: You, Xiaoxing, et al.
Pubblicazione: (2025)
di: You, Xiaoxing, et al.
Pubblicazione: (2025)
Retrieval-based Disentangled Representation Learning with Natural Language Supervision
di: Zhou, Jiawei, et al.
Pubblicazione: (2022)
di: Zhou, Jiawei, et al.
Pubblicazione: (2022)
Improving Medical Visual Reinforcement Fine-Tuning via Perception and Reasoning Augmentation
di: Yang, Guangjing, et al.
Pubblicazione: (2026)
di: Yang, Guangjing, et al.
Pubblicazione: (2026)
FashionMV: Product-Level Composed Image Retrieval with Multi-View Fashion Data
di: Yuan, Peng, et al.
Pubblicazione: (2026)
di: Yuan, Peng, et al.
Pubblicazione: (2026)
Enhancing Contrastive Learning for Retinal Imaging via Adjusted Augmentation Scales
di: Cheng, Zijie, et al.
Pubblicazione: (2025)
di: Cheng, Zijie, et al.
Pubblicazione: (2025)
A Comprehensive Survey on Composed Image Retrieval
di: Song, Xuemeng, et al.
Pubblicazione: (2025)
di: Song, Xuemeng, et al.
Pubblicazione: (2025)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
Enhanced Object Tracking by Self-Supervised Auxiliary Depth Estimation Learning
di: Wei, Zhenyu, et al.
Pubblicazione: (2024)
di: Wei, Zhenyu, et al.
Pubblicazione: (2024)
CubeComposer: Spatio-Temporal Autoregressive 4K 360° Video Generation from Perspective Video
di: Li, Lingen, et al.
Pubblicazione: (2026)
di: Li, Lingen, et al.
Pubblicazione: (2026)
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning
di: Shen, Yucheng, et al.
Pubblicazione: (2026)
di: Shen, Yucheng, et al.
Pubblicazione: (2026)
VLADriver-RAG: Retrieval-Augmented Vision-Language-Action Models for Autonomous Driving
di: Zhao, Rui, et al.
Pubblicazione: (2026)
di: Zhao, Rui, et al.
Pubblicazione: (2026)
RACap: Relation-Aware Prompting for Lightweight Retrieval-Augmented Image Captioning
di: Long, Xiaosheng, et al.
Pubblicazione: (2025)
di: Long, Xiaosheng, et al.
Pubblicazione: (2025)
Introducing 3D Representation for Medical Image Volume-to-Volume Translation via Score Fusion
di: Zhu, Xiyue, et al.
Pubblicazione: (2025)
di: Zhu, Xiyue, et al.
Pubblicazione: (2025)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
DualCap: Enhancing Lightweight Image Captioning via Dual Retrieval with Similar Scenes Visual Prompts
di: Li, Binbin, et al.
Pubblicazione: (2025)
di: Li, Binbin, et al.
Pubblicazione: (2025)
Multilingual Text-to-Image Person Retrieval via Bidirectional Relation Reasoning and Aligning
di: Cao, Min, et al.
Pubblicazione: (2025)
di: Cao, Min, et al.
Pubblicazione: (2025)
RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph
di: Malik, Sameer, et al.
Pubblicazione: (2025)
di: Malik, Sameer, et al.
Pubblicazione: (2025)
FastV-RAG: Towards Fast and Fine-Grained Video QA with Retrieval-Augmented Generation
di: Li, Gen, et al.
Pubblicazione: (2026)
di: Li, Gen, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning
di: Lv, Guannan, et al.
Pubblicazione: (2026) -
CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval
di: Jiang, Xintong, et al.
Pubblicazione: (2024) -
CSMCIR: CoT-Enhanced Symmetric Alignment with Memory Bank for Composed Image Retrieval
di: Qian, Zhipeng, et al.
Pubblicazione: (2026) -
TMCIR: Token Merge Benefits Composed Image Retrieval
di: Wang, Chaoyang, et al.
Pubblicazione: (2025) -
MELT: Improve Composed Image Retrieval via the Modification Frequentation-Rarity Balance Network
di: Qiu, Guozhi, et al.
Pubblicazione: (2026)