Accurate and Scalable Multimodal Pathology Retrieval via Attentive Vision-Language Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hongyi, Zhu, Zhengjie, Ma, Jiabo, Wang, Fang, Shi, Yue, Luo, Bo, Wang, Jili, Cai, Qiuyu, Zhang, Xiuming, Chen, Yen-Wei, Lin, Lanfen, Chen, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
by: Hu, Ruofan, et al.
Published: (2025)
by: Hu, Ruofan, et al.
Published: (2025)
When Vision Meets Texts in Listwise Reranking
by: Cai, Hongyi
Published: (2026)
by: Cai, Hongyi
Published: (2026)
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
by: Xu, Jinfeng, et al.
Published: (2026)
by: Xu, Jinfeng, et al.
Published: (2026)
Your Causal Self-Attentive Recommender Hosts a Lonely Neighborhood
by: Wang, Yueqi, et al.
Published: (2024)
by: Wang, Yueqi, et al.
Published: (2024)
CEMG: Collaborative-Enhanced Multimodal Generative Recommendation
by: Lin, Yuzhen, et al.
Published: (2025)
by: Lin, Yuzhen, et al.
Published: (2025)
Towards Semantic Consistency: Dirichlet Energy Driven Robust Multi-Modal Entity Alignment
by: Wang, Yuanyi, et al.
Published: (2024)
by: Wang, Yuanyi, et al.
Published: (2024)
OneRec: Unifying Retrieve and Rank with Generative Recommender and Iterative Preference Alignment
by: Deng, Jiaxin, et al.
Published: (2025)
by: Deng, Jiaxin, et al.
Published: (2025)
Factorized Transport Alignment for Multimodal and Multiview E-commerce Representation Learning
by: Chen, Xiwen, et al.
Published: (2025)
by: Chen, Xiwen, et al.
Published: (2025)
HGAMN: Heterogeneous Graph Attention Matching Network for Multilingual POI Retrieval at Baidu Maps
by: Huang, Jizhou, et al.
Published: (2024)
by: Huang, Jizhou, et al.
Published: (2024)
Multimodal Generative Retrieval Model with Staged Pretraining for Food Delivery on Meituan
by: Chen, Boyu, et al.
Published: (2026)
by: Chen, Boyu, et al.
Published: (2026)
Prospective Preference Enhanced Mixed Attentive Model for Session-based Recommendation
by: Peng, Bo, et al.
Published: (2022)
by: Peng, Bo, et al.
Published: (2022)
ArchSeek: Retrieving Architectural Case Studies Using Vision-Language Models
by: Li, Danrui, et al.
Published: (2025)
by: Li, Danrui, et al.
Published: (2025)
Multimodal Representation Alignment for Cross-modal Information Retrieval
by: Xu, Fan, et al.
Published: (2025)
by: Xu, Fan, et al.
Published: (2025)
Structural and Disentangled Adaptation of Large Vision Language Models for Multimodal Recommendation
by: Rao, Zhongtao, et al.
Published: (2025)
by: Rao, Zhongtao, et al.
Published: (2025)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation
by: Yu, Qinhan, et al.
Published: (2025)
by: Yu, Qinhan, et al.
Published: (2025)
Task Aligned Meta-learning based Augmented Graph for Cold-Start Recommendation
by: Shi, Yuxiang, et al.
Published: (2022)
by: Shi, Yuxiang, et al.
Published: (2022)
ReAlign: Optimizing the Visual Document Retriever with Reasoning-Guided Fine-Grained Alignment
by: Yang, Hao, et al.
Published: (2026)
by: Yang, Hao, et al.
Published: (2026)
Purifying Multimodal Retrieval: Fragment-Level Evidence Selection for RAG
by: Wang, Xihang, et al.
Published: (2026)
by: Wang, Xihang, et al.
Published: (2026)
Data-Efficient Massive Tool Retrieval: A Reinforcement Learning Approach for Query-Tool Alignment with Language Models
by: Zhang, Yuxiang, et al.
Published: (2024)
by: Zhang, Yuxiang, et al.
Published: (2024)
Revisiting Self-Attentive Sequential Recommendation
by: Huang, Zan
Published: (2025)
by: Huang, Zan
Published: (2025)
Beyond Text: Aligning Vision and Language for Multimodal E-Commerce Retrieval
by: Zhang, Qujiaheng, et al.
Published: (2026)
by: Zhang, Qujiaheng, et al.
Published: (2026)
MTFM: A Scalable and Alignment-free Foundation Model for Industrial Recommendation in Meituan
by: Song, Xin, et al.
Published: (2026)
by: Song, Xin, et al.
Published: (2026)
Generating with Fairness: A Modality-Diffused Counterfactual Framework for Incomplete Multimodal Recommendations
by: Li, Jin, et al.
Published: (2025)
by: Li, Jin, et al.
Published: (2025)
UNEX-RL: Reinforcing Long-Term Rewards in Multi-Stage Recommender Systems with UNidirectional EXecution
by: Zhang, Gengrui, et al.
Published: (2024)
by: Zhang, Gengrui, et al.
Published: (2024)
Molar: Multimodal LLMs with Collaborative Filtering Alignment for Enhanced Sequential Recommendation
by: Luo, Yucong, et al.
Published: (2024)
by: Luo, Yucong, et al.
Published: (2024)
JEPOO: Highly Accurate Joint Estimation of Pitch, Onset and Offset for Music Information Retrieval
by: Wei, Haojie, et al.
Published: (2023)
by: Wei, Haojie, et al.
Published: (2023)
Towards Scalable Semantic Representation for Recommendation
by: Zhang, Taolin, et al.
Published: (2024)
by: Zhang, Taolin, et al.
Published: (2024)
Mixed-initiative Query Rewriting in Conversational Passage Retrieval
by: Yang, Dayu, et al.
Published: (2023)
by: Yang, Dayu, et al.
Published: (2023)
The Effectiveness of Graph Contrastive Learning on Mathematical Information Retrieval
by: Wang, Pei-Syuan, et al.
Published: (2024)
by: Wang, Pei-Syuan, et al.
Published: (2024)
FAIR-QR: Enhancing Fairness-aware Information Retrieval through Query Refinement
by: Chen, Fumian, et al.
Published: (2025)
by: Chen, Fumian, et al.
Published: (2025)
Mitigating Recommendation Biases via Group-Alignment and Global-Uniformity in Representation Learning
by: Cai, Miaomiao, et al.
Published: (2025)
by: Cai, Miaomiao, et al.
Published: (2025)
RETLLM: Training and Data-Free MLLMs for Multimodal Information Retrieval
by: Su, Dawei, et al.
Published: (2026)
by: Su, Dawei, et al.
Published: (2026)
Unified Multimodal and Multilingual Retrieval via Multi-Task Learning with NLU Integration
by: Zhang, Xinyuan, et al.
Published: (2026)
by: Zhang, Xinyuan, et al.
Published: (2026)
Boosting Data Utilization for Multilingual Dense Retrieval
by: Huang, Chao, et al.
Published: (2025)
by: Huang, Chao, et al.
Published: (2025)
Unifying Multimodal Retrieval via Document Screenshot Embedding
by: Ma, Xueguang, et al.
Published: (2024)
by: Ma, Xueguang, et al.
Published: (2024)
M2IO-R1: An Efficient RL-Enhanced Reasoning Framework for Multimodal Retrieval Augmented Multimodal Generation
by: Xiao, Zhiyou, et al.
Published: (2025)
by: Xiao, Zhiyou, et al.
Published: (2025)
RREH: Reconstruction Relations Embedded Hashing for Semi-Paired Cross-Modal Retrieval
by: Wang, Jianzong, et al.
Published: (2024)
by: Wang, Jianzong, et al.
Published: (2024)
Popularity-Aware Alignment and Contrast for Mitigating Popularity Bias
by: Cai, Miaomiao, et al.
Published: (2024)
by: Cai, Miaomiao, et al.
Published: (2024)
Beyond Existing Retrievals: Cross-Scenario Incremental Sample Learning Framework
by: Wang, Tao, et al.
Published: (2025)
by: Wang, Tao, et al.
Published: (2025)
Similar Items
-
Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
by: Hu, Ruofan, et al.
Published: (2025) -
When Vision Meets Texts in Listwise Reranking
by: Cai, Hongyi
Published: (2026) -
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
by: Xu, Jinfeng, et al.
Published: (2026) -
Your Causal Self-Attentive Recommender Hosts a Lonely Neighborhood
by: Wang, Yueqi, et al.
Published: (2024) -
CEMG: Collaborative-Enhanced Multimodal Generative Recommendation
by: Lin, Yuzhen, et al.
Published: (2025)