MCoT-MVS: Multi-level Vision Selection by Multi-modal Chain-of-Thought Reasoning for Composed Image Retrieval

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Ge, Xuri, Wang, Chunhao, Wang, Xindi, Qin, Zheyun, Chen, Zhumin, Xin, Xin
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!