Multimodal Learned Sparse Retrieval with Probabilistic Expansion Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Nguyen, Thong, Hendriksen, Mariya, Yates, Andrew, de Rijke, Maarten |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multimodal Learned Sparse Retrieval for Image Suggestion
por: Nguyen, Thong, et al.
Publicado: (2024)
por: Nguyen, Thong, et al.
Publicado: (2024)
Benchmark Granularity and Model Robustness for Image-Text Retrieval
por: Hendriksen, Mariya, et al.
Publicado: (2024)
por: Hendriksen, Mariya, et al.
Publicado: (2024)
Demonstrating and Reducing Shortcuts in Vision-Language Representation Learning
por: Bleeker, Maurits, et al.
Publicado: (2024)
por: Bleeker, Maurits, et al.
Publicado: (2024)
Unified Interactive Multimodal Moment Retrieval via Cascaded Embedding-Reranking and Temporal-Aware Score Fusion
por: Thanh, Toan Le Ngo, et al.
Publicado: (2025)
por: Thanh, Toan Le Ngo, et al.
Publicado: (2025)
From Swath to Full-Disc: Advancing Precipitation Retrieval with Multimodal Knowledge Expansion
por: Wang, Zheng, et al.
Publicado: (2025)
por: Wang, Zheng, et al.
Publicado: (2025)
Video-ColBERT: Contextualized Late Interaction for Text-to-Video Retrieval
por: Reddy, Arun, et al.
Publicado: (2025)
por: Reddy, Arun, et al.
Publicado: (2025)
Leveraging Decoder Architectures for Learned Sparse Retrieval
por: Qiao, Jingfen, et al.
Publicado: (2025)
por: Qiao, Jingfen, et al.
Publicado: (2025)
BRIDGE: Multimodal-to-Text Retrieval via Reinforcement-Learned Query Alignment
por: Mounis, Mohamed Darwish, et al.
Publicado: (2026)
por: Mounis, Mohamed Darwish, et al.
Publicado: (2026)
KiseKloset for Fashion Retrieval and Recommendation
por: Phan-Nguyen, Thanh-Tung, et al.
Publicado: (2025)
por: Phan-Nguyen, Thanh-Tung, et al.
Publicado: (2025)
U-MARVEL: Unveiling Key Factors for Universal Multimodal Retrieval via Embedding Learning with MLLMs
por: Li, Xiaojie, et al.
Publicado: (2025)
por: Li, Xiaojie, et al.
Publicado: (2025)
Adapting MLLMs for Nuanced Video Retrieval
por: Bagad, Piyush, et al.
Publicado: (2025)
por: Bagad, Piyush, et al.
Publicado: (2025)
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
por: Tu, Rong-Cheng, et al.
Publicado: (2025)
MMMORRF: Multimodal Multilingual Modularized Reciprocal Rank Fusion
por: Samuel, Saron, et al.
Publicado: (2025)
por: Samuel, Saron, et al.
Publicado: (2025)
Milco: Learned Sparse Retrieval Across Languages via a Multilingual Connector
por: Nguyen, Thong, et al.
Publicado: (2025)
por: Nguyen, Thong, et al.
Publicado: (2025)
LoVR: A Benchmark for Long Video Retrieval in Multimodal Contexts
por: Cai, Qifeng, et al.
Publicado: (2025)
por: Cai, Qifeng, et al.
Publicado: (2025)
MR$^2$-Bench: Going Beyond Matching to Reasoning in Multimodal Retrieval
por: Zhou, Junjie, et al.
Publicado: (2025)
por: Zhou, Junjie, et al.
Publicado: (2025)
Rethinking Sparse Lexical Representations for Image Retrieval in the Age of Rising Multi-Modal Large Language Models
por: Nakata, Kengo, et al.
Publicado: (2024)
por: Nakata, Kengo, et al.
Publicado: (2024)
Snap and Diagnose: An Advanced Multimodal Retrieval System for Identifying Plant Diseases in the Wild
por: Wei, Tianqi, et al.
Publicado: (2024)
por: Wei, Tianqi, et al.
Publicado: (2024)
Beyond Global Similarity: Towards Fine-Grained, Multi-Condition Multimodal Retrieval
por: Lu, Xuan, et al.
Publicado: (2026)
por: Lu, Xuan, et al.
Publicado: (2026)
Accurate and Scalable Multimodal Pathology Retrieval via Attentive Vision-Language Alignment
por: Wang, Hongyi, et al.
Publicado: (2025)
por: Wang, Hongyi, et al.
Publicado: (2025)
DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories
por: Deng, Chenlong, et al.
Publicado: (2026)
por: Deng, Chenlong, et al.
Publicado: (2026)
FIGROTD: A Friendly-to-Handle Dataset for Image Guided Retrieval with Optional Text
por: Le, Hoang-Bao, et al.
Publicado: (2025)
por: Le, Hoang-Bao, et al.
Publicado: (2025)
Neurosymbolic Inference On Foundation Models For Remote Sensing Text-to-image Retrieval With Complex Queries
por: Mezzi, Emanuele, et al.
Publicado: (2025)
por: Mezzi, Emanuele, et al.
Publicado: (2025)
Controlled Retrieval-augmented Context Evaluation for Long-form RAG
por: Ju, Jia-Huei, et al.
Publicado: (2025)
por: Ju, Jia-Huei, et al.
Publicado: (2025)
Effective Inference-Free Retrieval for Learned Sparse Representations
por: Nardini, Franco Maria, et al.
Publicado: (2025)
por: Nardini, Franco Maria, et al.
Publicado: (2025)
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
por: Li, Po-han, et al.
Publicado: (2024)
por: Li, Po-han, et al.
Publicado: (2024)
Windsock is Dancing: Adaptive Multimodal Retrieval-Augmented Generation
por: Zhao, Shu, et al.
Publicado: (2025)
por: Zhao, Shu, et al.
Publicado: (2025)
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
por: Yang, Jiashu, et al.
Publicado: (2026)
por: Yang, Jiashu, et al.
Publicado: (2026)
UNION: A Lightweight Target Representation for Efficient Zero-Shot Image-Guided Retrieval with Optional Textual Queries
por: Le, Hoang-Bao, et al.
Publicado: (2025)
por: Le, Hoang-Bao, et al.
Publicado: (2025)
Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval
por: Kong, Fanheng, et al.
Publicado: (2025)
por: Kong, Fanheng, et al.
Publicado: (2025)
Indexing Multimodal Language Models for Large-scale Image Retrieval
por: Tharwat, Bahey, et al.
Publicado: (2026)
por: Tharwat, Bahey, et al.
Publicado: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
por: Dong, Kuicai, et al.
Publicado: (2025)
por: Dong, Kuicai, et al.
Publicado: (2025)
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
por: Wan, David, et al.
Publicado: (2025)
por: Wan, David, et al.
Publicado: (2025)
Active Learning via Classifier Impact and Greedy Selection for Interactive Image Retrieval
por: Bar, Leah, et al.
Publicado: (2024)
por: Bar, Leah, et al.
Publicado: (2024)
Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
por: Martin, Alexander, et al.
Publicado: (2025)
por: Martin, Alexander, et al.
Publicado: (2025)
Sparton: Fast and Memory-Efficient Triton Kernel for Learned Sparse Retrieval
por: Nguyen, Thong, et al.
Publicado: (2026)
por: Nguyen, Thong, et al.
Publicado: (2026)
Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
por: Wang, Yifan, et al.
Publicado: (2025)
por: Wang, Yifan, et al.
Publicado: (2025)
Embedding-based Retrieval in Multimodal Content Moderation
por: Liang, Hanzhong, et al.
Publicado: (2025)
por: Liang, Hanzhong, et al.
Publicado: (2025)
MetaEmbed: Scaling Multimodal Retrieval at Test-Time with Flexible Late Interaction
por: Xiao, Zilin, et al.
Publicado: (2025)
por: Xiao, Zilin, et al.
Publicado: (2025)
Video Editing for Video Retrieval
por: Zhu, Bin, et al.
Publicado: (2024)
por: Zhu, Bin, et al.
Publicado: (2024)
Ejemplares similares
-
Multimodal Learned Sparse Retrieval for Image Suggestion
por: Nguyen, Thong, et al.
Publicado: (2024) -
Benchmark Granularity and Model Robustness for Image-Text Retrieval
por: Hendriksen, Mariya, et al.
Publicado: (2024) -
Demonstrating and Reducing Shortcuts in Vision-Language Representation Learning
por: Bleeker, Maurits, et al.
Publicado: (2024) -
Unified Interactive Multimodal Moment Retrieval via Cascaded Embedding-Reranking and Temporal-Aware Score Fusion
por: Thanh, Toan Le Ngo, et al.
Publicado: (2025) -
From Swath to Full-Disc: Advancing Precipitation Retrieval with Multimodal Knowledge Expansion
por: Wang, Zheng, et al.
Publicado: (2025)