REMOTE: A Unified Multimodal Relation Extraction Framework with Multilevel Optimal Transport and Mixture-of-Experts
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Xinkui, Xu, Yongxiu, Tang, Minghao, Zhang, Shilong, Xu, Hongbo, Xu, Hao, Wang, Yubin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
by: Xu, Jinfeng, et al.
Published: (2026)
by: Xu, Jinfeng, et al.
Published: (2026)
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
by: Luo, Pengfei, et al.
Published: (2025)
by: Luo, Pengfei, et al.
Published: (2025)
Personalized Image Generation with Large Multimodal Models
by: Xu, Yiyan, et al.
Published: (2024)
by: Xu, Yiyan, et al.
Published: (2024)
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
HistLLM: A Unified Framework for LLM-Based Multimodal Recommendation with User History Encoding and Compression
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
Don't Lose Yourself: Boosting Multimodal Recommendation via Reducing Node-neighbor Discrepancy in Graph Convolutional Network
by: Chen, Zheyu, et al.
Published: (2024)
by: Chen, Zheyu, et al.
Published: (2024)
Uni-Retrieval: A Multi-Style Retrieval Framework for STEM's Education
by: Jia, Yanhao, et al.
Published: (2025)
by: Jia, Yanhao, et al.
Published: (2025)
A Unified Optimal Transport Framework for Cross-Modal Retrieval with Noisy Labels
by: Han, Haochen, et al.
Published: (2024)
by: Han, Haochen, et al.
Published: (2024)
Semantic Item Graph Enhancement for Multimodal Recommendation
by: Zhang, Xiaoxiong, et al.
Published: (2025)
by: Zhang, Xiaoxiong, et al.
Published: (2025)
A Survey on Multimodal Recommender Systems: Recent Advances and Future Directions
by: Xu, Jinfeng, et al.
Published: (2025)
by: Xu, Jinfeng, et al.
Published: (2025)
Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning
by: Zhang, Lingzi, et al.
Published: (2023)
by: Zhang, Lingzi, et al.
Published: (2023)
Beyond Static Collision Handling: Adaptive Semantic ID Learning for Multimodal Recommendation at Industrial Scale
by: Pan, Yongsen, et al.
Published: (2026)
by: Pan, Yongsen, et al.
Published: (2026)
Breaking the Curse of Knowledge: Towards Effective Multimodal Recommendation using Knowledge Soft Integration
by: Ouyang, Kai, et al.
Published: (2023)
by: Ouyang, Kai, et al.
Published: (2023)
CLEAR: Null-Space Projection for Cross-Modal De-Redundancy in Multimodal Recommendation
by: Zhan, Hao, et al.
Published: (2026)
by: Zhan, Hao, et al.
Published: (2026)
Multimodal Graph Neural Network for Recommendation with Dynamic De-redundancy and Modality-Guided Feature De-noisy
by: Mo, Feng, et al.
Published: (2024)
by: Mo, Feng, et al.
Published: (2024)
Unified Hallucination Detection for Multimodal Large Language Models
by: Chen, Xiang, et al.
Published: (2024)
by: Chen, Xiang, et al.
Published: (2024)
Towards Unified Multi-Modal Personalization: Large Vision-Language Models for Generative Recommendation and Beyond
by: Wei, Tianxin, et al.
Published: (2024)
by: Wei, Tianxin, et al.
Published: (2024)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
by: Sun, Huatuan, et al.
Published: (2025)
by: Sun, Huatuan, et al.
Published: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
by: Mo, Xian, et al.
Published: (2025)
by: Mo, Xian, et al.
Published: (2025)
CM$^3$: Calibrating Multimodal Recommendation
by: Zhou, Xin, et al.
Published: (2025)
by: Zhou, Xin, et al.
Published: (2025)
Multimodal Learned Sparse Retrieval for Image Suggestion
by: Nguyen, Thong, et al.
Published: (2024)
by: Nguyen, Thong, et al.
Published: (2024)
Automating Steering for Safe Multimodal Large Language Models
by: Wu, Lyucheng, et al.
Published: (2025)
by: Wu, Lyucheng, et al.
Published: (2025)
Attribute-driven Disentangled Representation Learning for Multimodal Recommendation
by: Li, Zhenyang, et al.
Published: (2023)
by: Li, Zhenyang, et al.
Published: (2023)
Multimodal Pretraining, Adaptation, and Generation for Recommendation: A Survey
by: Liu, Qijiong, et al.
Published: (2024)
by: Liu, Qijiong, et al.
Published: (2024)
DREAM: A Dual Representation Learning Model for Multimodal Recommendation
by: Zhang, Kangning, et al.
Published: (2024)
by: Zhang, Kangning, et al.
Published: (2024)
Disentangled Graph Variational Auto-Encoder for Multimodal Recommendation with Interpretability
by: Zhou, Xin, et al.
Published: (2024)
by: Zhou, Xin, et al.
Published: (2024)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
by: Jaspal, Amit, et al.
Published: (2025)
by: Jaspal, Amit, et al.
Published: (2025)
Multimodal Music Recommendation System using LLMs
by: Kandagatla, Srikar Prabhas, et al.
Published: (2026)
by: Kandagatla, Srikar Prabhas, et al.
Published: (2026)
Learning Item Representations Directly from Multimodal Features for Effective Recommendation
by: Zhou, Xin, et al.
Published: (2025)
by: Zhou, Xin, et al.
Published: (2025)
Spectrum-based Modality Representation Fusion Graph Convolutional Network for Multimodal Recommendation
by: Ong, Rongqing Kenneth, et al.
Published: (2024)
by: Ong, Rongqing Kenneth, et al.
Published: (2024)
Does Multimodality Improve Recommender Systems as Expected? A Critical Analysis and Future Directions
by: Zhou, Hongyu, et al.
Published: (2025)
by: Zhou, Hongyu, et al.
Published: (2025)
Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation
by: Ma, Hongjian, et al.
Published: (2026)
by: Ma, Hongjian, et al.
Published: (2026)
MMSRARec: Summarization and Retrieval Augumented Sequential Recommendation Based on Multimodal Large Language Model
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
MCA: Modality Composition Awareness for Robust Composed Multimodal Retrieval
by: Wu, Qiyu, et al.
Published: (2025)
by: Wu, Qiyu, et al.
Published: (2025)
The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval
by: Fu, Junchen, et al.
Published: (2026)
by: Fu, Junchen, et al.
Published: (2026)
From ID-based to ID-free: Rethinking ID Effectiveness in Multimodal Collaborative Filtering Recommendation
by: Li, Guohao, et al.
Published: (2025)
by: Li, Guohao, et al.
Published: (2025)
Unraveling Movie Genres through Cross-Attention Fusion of Bi-Modal Synergy of Poster
by: Nareti, Utsav Kumar, et al.
Published: (2024)
by: Nareti, Utsav Kumar, et al.
Published: (2024)
On the Brittleness of CLIP Text Encoders
by: Tran, Allie, et al.
Published: (2025)
by: Tran, Allie, et al.
Published: (2025)
Improving the Consistency in Cross-Lingual Cross-Modal Retrieval with 1-to-K Contrastive Learning
by: Nie, Zhijie, et al.
Published: (2024)
by: Nie, Zhijie, et al.
Published: (2024)
Adaptive Multi-Agent Reasoning for Text-to-Video Retrieval
by: Wu, Jiaxin, et al.
Published: (2025)
by: Wu, Jiaxin, et al.
Published: (2025)
Similar Items
-
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
by: Xu, Jinfeng, et al.
Published: (2026) -
ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
by: Luo, Pengfei, et al.
Published: (2025) -
Personalized Image Generation with Large Multimodal Models
by: Xu, Yiyan, et al.
Published: (2024) -
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
by: Li, Yang, et al.
Published: (2025) -
HistLLM: A Unified Framework for LLM-Based Multimodal Recommendation with User History Encoding and Compression
by: Zhang, Chen, et al.
Published: (2025)