Spherical Linear Interpolation and Text-Anchoring for Zero-shot Composed Image Retrieval
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jang, Young Kyun, Huynh, Dat, Shah, Ashish, Chen, Wen-Kai, Lim, Ser-Nam |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
par: Jang, Young Kyun, et autres
Publié: (2024)
par: Jang, Young Kyun, et autres
Publié: (2024)
Composing Object Relations and Attributes for Image-Text Matching
par: Pham, Khoi, et autres
Publié: (2024)
par: Pham, Khoi, et autres
Publié: (2024)
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval
par: Jang, Young Kyun, et autres
Publié: (2024)
par: Jang, Young Kyun, et autres
Publié: (2024)
Zero-shot Composed Text-Image Retrieval
par: Liu, Yikun, et autres
Publié: (2023)
par: Liu, Yikun, et autres
Publié: (2023)
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
par: Han, Gyuwon, et autres
Publié: (2026)
par: Han, Gyuwon, et autres
Publié: (2026)
Towards Cross-modal Backward-compatible Representation Learning for Vision-Language Models
par: Jang, Young Kyun, et autres
Publié: (2024)
par: Jang, Young Kyun, et autres
Publié: (2024)
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
par: He, Bo, et autres
Publié: (2024)
par: He, Bo, et autres
Publié: (2024)
Data-Efficient Generalization for Zero-shot Composed Image Retrieval
par: Chen, Zining, et autres
Publié: (2025)
par: Chen, Zining, et autres
Publié: (2025)
From Mapping to Composing: A Two-Stage Framework for Zero-shot Composed Image Retrieval
par: Wang, Yabing, et autres
Publié: (2025)
par: Wang, Yabing, et autres
Publié: (2025)
PDV: Prompt Directional Vectors for Zero-shot Composed Image Retrieval
par: Tursun, Osman, et autres
Publié: (2025)
par: Tursun, Osman, et autres
Publié: (2025)
Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval
par: Suo, Yucheng, et autres
Publié: (2024)
par: Suo, Yucheng, et autres
Publié: (2024)
Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
par: Du, Yongchao, et autres
Publié: (2024)
par: Du, Yongchao, et autres
Publié: (2024)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
par: Sun, Shitong, et autres
Publié: (2023)
par: Sun, Shitong, et autres
Publié: (2023)
CoLLM: A Large Language Model for Composed Image Retrieval
par: Huynh, Chuong, et autres
Publié: (2025)
par: Huynh, Chuong, et autres
Publié: (2025)
Language-only Efficient Training of Zero-shot Composed Image Retrieval
par: Gu, Geonmo, et autres
Publié: (2023)
par: Gu, Geonmo, et autres
Publié: (2023)
Modality and Task Adaptation for Enhanced Zero-shot Composed Image Retrieval
par: Li, Haiwen, et autres
Publié: (2024)
par: Li, Haiwen, et autres
Publié: (2024)
Pseudo-triplet Guided Few-shot Composed Image Retrieval
par: Hou, Bohan, et autres
Publié: (2024)
par: Hou, Bohan, et autres
Publié: (2024)
Zero Shot Composed Image Retrieval
par: Kakarla, Santhosh, et autres
Publié: (2025)
par: Kakarla, Santhosh, et autres
Publié: (2025)
Zero-shot Synthetic Video Realism Enhancement via Structure-aware Denoising
par: Wang, Yifan, et autres
Publié: (2025)
par: Wang, Yifan, et autres
Publié: (2025)
Towards Chunk-Wise Generation for Long Videos
par: Zhang, Siyang, et autres
Publié: (2024)
par: Zhang, Siyang, et autres
Publié: (2024)
Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval
par: Yang, Yuxin, et autres
Publié: (2026)
par: Yang, Yuxin, et autres
Publié: (2026)
Scene Co-pilot: Procedural Text to Video Generation with Human in the Loop
par: Qian, Zhaofang, et autres
Publié: (2024)
par: Qian, Zhaofang, et autres
Publié: (2024)
Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs
par: Zhang, Huaying, et autres
Publié: (2024)
par: Zhang, Huaying, et autres
Publié: (2024)
Fine-Grained Zero-Shot Composed Image Retrieval with Complementary Visual-Semantic Integration
par: Ye, Yongcong, et autres
Publié: (2026)
par: Ye, Yongcong, et autres
Publié: (2026)
Fine-grained Textual Inversion Network for Zero-Shot Composed Image Retrieval
par: Lin, Haoqiang, et autres
Publié: (2025)
par: Lin, Haoqiang, et autres
Publié: (2025)
NoiseDiffusion: Correcting Noise for Image Interpolation with Diffusion Models beyond Spherical Linear Interpolation
par: Zheng, PengFei, et autres
Publié: (2024)
par: Zheng, PengFei, et autres
Publié: (2024)
FSViewFusion: Few-Shots View Generation of Novel Objects
par: Hussain, Rukhshanda, et autres
Publié: (2024)
par: Hussain, Rukhshanda, et autres
Publié: (2024)
VideoMerge: Towards Training-free Long Video Generation
par: Zhang, Siyang, et autres
Publié: (2025)
par: Zhang, Siyang, et autres
Publié: (2025)
G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval
par: Lim, Jiyoung, et autres
Publié: (2026)
par: Lim, Jiyoung, et autres
Publié: (2026)
TEMA: Anchor the Image, Follow the Text for Multi-Modification Composed Image Retrieval
par: Li, Zixu, et autres
Publié: (2026)
par: Li, Zixu, et autres
Publié: (2026)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
par: Jang, Young Kyun, et autres
Publié: (2024)
par: Jang, Young Kyun, et autres
Publié: (2024)
Generating a Paracosm for Training-Free Zero-Shot Composed Image Retrieval
par: Wang, Tong, et autres
Publié: (2026)
par: Wang, Tong, et autres
Publié: (2026)
HyCIR: Boosting Zero-Shot Composed Image Retrieval with Synthetic Labels
par: Jiang, Yingying, et autres
Publié: (2024)
par: Jiang, Yingying, et autres
Publié: (2024)
SlerpFace: Face Template Protection via Spherical Linear Interpolation
par: Zhong, Zhizhou, et autres
Publié: (2024)
par: Zhong, Zhizhou, et autres
Publié: (2024)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
par: Tu, Rong-Cheng, et autres
Publié: (2025)
par: Tu, Rong-Cheng, et autres
Publié: (2025)
AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation
par: Liu, Yexin, et autres
Publié: (2025)
par: Liu, Yexin, et autres
Publié: (2025)
Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data
par: Duan, Yiqun, et autres
Publié: (2025)
par: Duan, Yiqun, et autres
Publié: (2025)
Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization
par: Chen, Yiyang, et autres
Publié: (2022)
par: Chen, Yiyang, et autres
Publié: (2022)
Generative Editing in the Joint Vision-Language Space for Zero-Shot Composed Image Retrieval
par: Wang, Xin, et autres
Publié: (2025)
par: Wang, Xin, et autres
Publié: (2025)
VSC: Visual Search Compositional Text-to-Image Diffusion Model
par: Dat, Do Huu, et autres
Publié: (2025)
par: Dat, Do Huu, et autres
Publié: (2025)
Documents similaires
-
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
par: Jang, Young Kyun, et autres
Publié: (2024) -
Composing Object Relations and Attributes for Image-Text Matching
par: Pham, Khoi, et autres
Publié: (2024) -
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval
par: Jang, Young Kyun, et autres
Publié: (2024) -
Zero-shot Composed Text-Image Retrieval
par: Liu, Yikun, et autres
Publié: (2023) -
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
par: Han, Gyuwon, et autres
Publié: (2026)