Snap and Diagnose: An Advanced Multimodal Retrieval System for Identifying Plant Diseases in the Wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Tianqi, Chen, Zhi, Yu, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval
von: Kong, Fanheng, et al.
Veröffentlicht: (2025)
von: Kong, Fanheng, et al.
Veröffentlicht: (2025)
Accurate and Scalable Multimodal Pathology Retrieval via Attentive Vision-Language Alignment
von: Wang, Hongyi, et al.
Veröffentlicht: (2025)
von: Wang, Hongyi, et al.
Veröffentlicht: (2025)
From Swath to Full-Disc: Advancing Precipitation Retrieval with Multimodal Knowledge Expansion
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
U-MARVEL: Unveiling Key Factors for Universal Multimodal Retrieval via Embedding Learning with MLLMs
von: Li, Xiaojie, et al.
Veröffentlicht: (2025)
von: Li, Xiaojie, et al.
Veröffentlicht: (2025)
Multimodal Learned Sparse Retrieval with Probabilistic Expansion Control
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)
MR$^2$-Bench: Going Beyond Matching to Reasoning in Multimodal Retrieval
von: Zhou, Junjie, et al.
Veröffentlicht: (2025)
von: Zhou, Junjie, et al.
Veröffentlicht: (2025)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2025)
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2025)
BRIDGE: Multimodal-to-Text Retrieval via Reinforcement-Learned Query Alignment
von: Mounis, Mohamed Darwish, et al.
Veröffentlicht: (2026)
von: Mounis, Mohamed Darwish, et al.
Veröffentlicht: (2026)
LoVR: A Benchmark for Long Video Retrieval in Multimodal Contexts
von: Cai, Qifeng, et al.
Veröffentlicht: (2025)
von: Cai, Qifeng, et al.
Veröffentlicht: (2025)
Beyond Global Similarity: Towards Fine-Grained, Multi-Condition Multimodal Retrieval
von: Lu, Xuan, et al.
Veröffentlicht: (2026)
von: Lu, Xuan, et al.
Veröffentlicht: (2026)
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories
von: Deng, Chenlong, et al.
Veröffentlicht: (2026)
von: Deng, Chenlong, et al.
Veröffentlicht: (2026)
Human-Oriented Image Retrieval System (HORSE): A Neuro-Symbolic Approach to Optimizing Retrieval of Previewed Images
von: Weinberg, Abraham Itzhak
Veröffentlicht: (2025)
von: Weinberg, Abraham Itzhak
Veröffentlicht: (2025)
SHE-Net: Syntax-Hierarchy-Enhanced Text-Video Retrieval
von: Yu, Xuzheng, et al.
Veröffentlicht: (2024)
von: Yu, Xuzheng, et al.
Veröffentlicht: (2024)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
von: Xiao, Zilin, et al.
Veröffentlicht: (2025)
Advancing Re-Ranking with Multimodal Fusion and Target-Oriented Auxiliary Tasks in E-Commerce Search
von: Xu, Enqiang, et al.
Veröffentlicht: (2024)
von: Xu, Enqiang, et al.
Veröffentlicht: (2024)
IISAN: Efficiently Adapting Multimodal Representation for Sequential Recommendation with Decoupled PEFT
von: Fu, Junchen, et al.
Veröffentlicht: (2024)
von: Fu, Junchen, et al.
Veröffentlicht: (2024)
Efficient and Effective Adaptation of Multimodal Foundation Models in Sequential Recommendation
von: Fu, Junchen, et al.
Veröffentlicht: (2024)
von: Fu, Junchen, et al.
Veröffentlicht: (2024)
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
von: Li, Po-han, et al.
Veröffentlicht: (2024)
von: Li, Po-han, et al.
Veröffentlicht: (2024)
Embedding-based Retrieval in Multimodal Content Moderation
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
von: Liang, Hanzhong, et al.
Veröffentlicht: (2025)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
von: Zhao, Shu, et al.
Veröffentlicht: (2025)
OMGM: Orchestrate Multiple Granularities and Modalities for Efficient Multimodal Retrieval
von: Yang, Wei, et al.
Veröffentlicht: (2025)
von: Yang, Wei, et al.
Veröffentlicht: (2025)
Identity-Decoupled Anonymization for Visual Evidence in Multi-modal Retrieval-Augmented Generation
von: Cheng, Zehua, et al.
Veröffentlicht: (2026)
von: Cheng, Zehua, et al.
Veröffentlicht: (2026)
ContextIQ: A Multimodal Expert-Based Video Retrieval System for Contextual Advertising
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2024)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2024)
Indexing Multimodal Language Models for Large-scale Image Retrieval
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
von: Tharwat, Bahey, et al.
Veröffentlicht: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
von: Wan, David, et al.
Veröffentlicht: (2025)
von: Wan, David, et al.
Veröffentlicht: (2025)
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
von: Nozawa, Yuji, et al.
Veröffentlicht: (2025)
von: Nozawa, Yuji, et al.
Veröffentlicht: (2025)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
von: Martin, Alexander, et al.
Veröffentlicht: (2025)
Chain-of-Thought Re-ranking for Image Retrieval Tasks
von: Wu, Shangrong, et al.
Veröffentlicht: (2025)
von: Wu, Shangrong, et al.
Veröffentlicht: (2025)
LongVidSearch: An Agentic Benchmark for Multi-hop Evidence Retrieval Planning in Long Videos
von: Yu, Rongyi, et al.
Veröffentlicht: (2026)
von: Yu, Rongyi, et al.
Veröffentlicht: (2026)
MMDocIR: Benchmarking Multimodal Retrieval for Long Documents
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
Attributes Grouping and Mining Hashing for Fine-Grained Image Retrieval
von: Lu, Xin, et al.
Veröffentlicht: (2023)
von: Lu, Xin, et al.
Veröffentlicht: (2023)
MSAM: Multi-Semantic Adaptive Mining for Cross-Modal Drone Video-Text Retrieval
von: Huang, Jinghao, et al.
Veröffentlicht: (2025)
von: Huang, Jinghao, et al.
Veröffentlicht: (2025)
Towards Text-Image Interleaved Retrieval
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
von: Zhang, Xin, et al.
Veröffentlicht: (2025)
AlbumFill: Album-Guided Reasoning and Retrieval for Personalized Image Completion
von: Tsai, Yu-Ju, et al.
Veröffentlicht: (2026)
von: Tsai, Yu-Ju, et al.
Veröffentlicht: (2026)
Video Editing for Video Retrieval
von: Zhu, Bin, et al.
Veröffentlicht: (2024)
von: Zhu, Bin, et al.
Veröffentlicht: (2024)
Decoding Ancient Oracle Bone Script via Generative Dictionary Retrieval
von: Wu, Yin, et al.
Veröffentlicht: (2026)
von: Wu, Yin, et al.
Veröffentlicht: (2026)
Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
A Resource-Efficient Training Framework for Remote Sensing Text--Image Retrieval
von: Zhang, Weihang, et al.
Veröffentlicht: (2025)
von: Zhang, Weihang, et al.
Veröffentlicht: (2025)
Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval
von: Kong, Fanheng, et al.
Veröffentlicht: (2025) -
Accurate and Scalable Multimodal Pathology Retrieval via Attentive Vision-Language Alignment
von: Wang, Hongyi, et al.
Veröffentlicht: (2025) -
From Swath to Full-Disc: Advancing Precipitation Retrieval with Multimodal Knowledge Expansion
von: Wang, Zheng, et al.
Veröffentlicht: (2025) -
U-MARVEL: Unveiling Key Factors for Universal Multimodal Retrieval via Embedding Learning with MLLMs
von: Li, Xiaojie, et al.
Veröffentlicht: (2025) -
Multimodal Learned Sparse Retrieval with Probabilistic Expansion Control
von: Nguyen, Thong, et al.
Veröffentlicht: (2024)