Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Du, Yongchao, Wang, Min, Zhou, Wengang, Hui, Shuping, Li, Houqiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Zero-shot Composed Text-Image Retrieval
di: Liu, Yikun, et al.
Pubblicazione: (2023)
di: Liu, Yikun, et al.
Pubblicazione: (2023)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
di: Sun, Qi, et al.
Pubblicazione: (2024)
di: Sun, Qi, et al.
Pubblicazione: (2024)
From Mapping to Composing: A Two-Stage Framework for Zero-shot Composed Image Retrieval
di: Wang, Yabing, et al.
Pubblicazione: (2025)
di: Wang, Yabing, et al.
Pubblicazione: (2025)
Data-Efficient Generalization for Zero-shot Composed Image Retrieval
di: Chen, Zining, et al.
Pubblicazione: (2025)
di: Chen, Zining, et al.
Pubblicazione: (2025)
Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval
di: Suo, Yucheng, et al.
Pubblicazione: (2024)
di: Suo, Yucheng, et al.
Pubblicazione: (2024)
PDV: Prompt Directional Vectors for Zero-shot Composed Image Retrieval
di: Tursun, Osman, et al.
Pubblicazione: (2025)
di: Tursun, Osman, et al.
Pubblicazione: (2025)
Instance-aware Exploration-Verification-Exploitation for Instance ImageGoal Navigation
di: Lei, Xiaohan, et al.
Pubblicazione: (2024)
di: Lei, Xiaohan, et al.
Pubblicazione: (2024)
Modality and Task Adaptation for Enhanced Zero-shot Composed Image Retrieval
di: Li, Haiwen, et al.
Pubblicazione: (2024)
di: Li, Haiwen, et al.
Pubblicazione: (2024)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
di: Sun, Shitong, et al.
Pubblicazione: (2023)
di: Sun, Shitong, et al.
Pubblicazione: (2023)
Zero Shot Composed Image Retrieval
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
di: Kakarla, Santhosh, et al.
Pubblicazione: (2025)
Language-only Efficient Training of Zero-shot Composed Image Retrieval
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
SEDS: Semantically Enhanced Dual-Stream Encoder for Sign Language Retrieval
di: Jiang, Longtao, et al.
Pubblicazione: (2024)
di: Jiang, Longtao, et al.
Pubblicazione: (2024)
GaussNav: Gaussian Splatting for Visual Navigation
di: Lei, Xiaohan, et al.
Pubblicazione: (2024)
di: Lei, Xiaohan, et al.
Pubblicazione: (2024)
Spherical Linear Interpolation and Text-Anchoring for Zero-shot Composed Image Retrieval
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
di: Jang, Young Kyun, et al.
Pubblicazione: (2024)
RoFIR: Robust Fisheye Image Rectification Framework Impervious to Optical Center Deviation
di: Liao, Zhaokang, et al.
Pubblicazione: (2024)
di: Liao, Zhaokang, et al.
Pubblicazione: (2024)
TextCoT: Zoom In for Enhanced Multimodal Text-Rich Image Understanding
di: Luan, Bozhi, et al.
Pubblicazione: (2024)
di: Luan, Bozhi, et al.
Pubblicazione: (2024)
DesignDiffusion: High-Quality Text-to-Design Image Generation with Diffusion Models
di: Wang, Zhendong, et al.
Pubblicazione: (2025)
di: Wang, Zhendong, et al.
Pubblicazione: (2025)
Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction
di: Guo, Zhiyang, et al.
Pubblicazione: (2024)
di: Guo, Zhiyang, et al.
Pubblicazione: (2024)
Pseudo-triplet Guided Few-shot Composed Image Retrieval
di: Hou, Bohan, et al.
Pubblicazione: (2024)
di: Hou, Bohan, et al.
Pubblicazione: (2024)
Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs
di: Zhang, Huaying, et al.
Pubblicazione: (2024)
di: Zhang, Huaying, et al.
Pubblicazione: (2024)
Semantic Image Synthesis via Diffusion Models
di: Zhou, Wengang, et al.
Pubblicazione: (2022)
di: Zhou, Wengang, et al.
Pubblicazione: (2022)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
di: Zhao, Weichao, et al.
Pubblicazione: (2024)
di: Zhao, Weichao, et al.
Pubblicazione: (2024)
Fine-Grained Zero-Shot Composed Image Retrieval with Complementary Visual-Semantic Integration
di: Ye, Yongcong, et al.
Pubblicazione: (2026)
di: Ye, Yongcong, et al.
Pubblicazione: (2026)
Multi-Scale Invertible Neural Network for Wide-Range Variable-Rate Learned Image Compression
di: Tu, Hanyue, et al.
Pubblicazione: (2025)
di: Tu, Hanyue, et al.
Pubblicazione: (2025)
AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding
di: Wang, Yonghui, et al.
Pubblicazione: (2024)
di: Wang, Yonghui, et al.
Pubblicazione: (2024)
Revisiting Shadow Detection from a Vision-Language Perspective
di: Wang, Yonghui, et al.
Pubblicazione: (2026)
di: Wang, Yonghui, et al.
Pubblicazione: (2026)
StepVAR: Structure-Texture Guided Pruning for Visual Autoregressive Models
di: Liu, Keli, et al.
Pubblicazione: (2026)
di: Liu, Keli, et al.
Pubblicazione: (2026)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
di: Tu, Rong-Cheng, et al.
Pubblicazione: (2025)
di: Tu, Rong-Cheng, et al.
Pubblicazione: (2025)
Generating a Paracosm for Training-Free Zero-Shot Composed Image Retrieval
di: Wang, Tong, et al.
Pubblicazione: (2026)
di: Wang, Tong, et al.
Pubblicazione: (2026)
HyCIR: Boosting Zero-Shot Composed Image Retrieval with Synthetic Labels
di: Jiang, Yingying, et al.
Pubblicazione: (2024)
di: Jiang, Yingying, et al.
Pubblicazione: (2024)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
di: Sun, Qi, et al.
Pubblicazione: (2024)
di: Sun, Qi, et al.
Pubblicazione: (2024)
Generative Editing in the Joint Vision-Language Space for Zero-Shot Composed Image Retrieval
di: Wang, Xin, et al.
Pubblicazione: (2025)
di: Wang, Xin, et al.
Pubblicazione: (2025)
Video-based Sign Language Recognition without Temporal Segmentation
di: Huang, Jie, et al.
Pubblicazione: (2018)
di: Huang, Jie, et al.
Pubblicazione: (2018)
Structural Action Transformer for 3D Dexterous Manipulation
di: Lei, Xiaohan, et al.
Pubblicazione: (2026)
di: Lei, Xiaohan, et al.
Pubblicazione: (2026)
Modality-Aware Representation Learning for Zero-shot Sketch-based Image Retrieval
di: Lyou, Eunyi, et al.
Pubblicazione: (2024)
di: Lyou, Eunyi, et al.
Pubblicazione: (2024)
MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning
di: Liu, Xiaoyang, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyang, et al.
Pubblicazione: (2024)
Denoise-I2W: Mapping Images to Denoising Words for Accurate Zero-Shot Composed Image Retrieval
di: Tang, Yuanmin, et al.
Pubblicazione: (2024)
di: Tang, Yuanmin, et al.
Pubblicazione: (2024)
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing
di: Xu, Shaodong, et al.
Pubblicazione: (2026)
di: Xu, Shaodong, et al.
Pubblicazione: (2026)
Scaling Prompt Instructed Zero Shot Composed Image Retrieval with Image-Only Data
di: Duan, Yiqun, et al.
Pubblicazione: (2025)
di: Duan, Yiqun, et al.
Pubblicazione: (2025)
Dynamic Multi-level Weighted Alignment Network for Zero-shot Sketch-based Image Retrieval
di: Su, Hanwen, et al.
Pubblicazione: (2025)
di: Su, Hanwen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Zero-shot Composed Text-Image Retrieval
di: Liu, Yikun, et al.
Pubblicazione: (2023) -
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
di: Sun, Qi, et al.
Pubblicazione: (2024) -
From Mapping to Composing: A Two-Stage Framework for Zero-shot Composed Image Retrieval
di: Wang, Yabing, et al.
Pubblicazione: (2025) -
Data-Efficient Generalization for Zero-shot Composed Image Retrieval
di: Chen, Zining, et al.
Pubblicazione: (2025) -
Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval
di: Suo, Yucheng, et al.
Pubblicazione: (2024)