LoginMEA: Local-to-Global Interaction Network for Multi-modal Entity Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Su, Taoyu, Zhang, Xinghua, Sheng, Jiawei, Zhang, Zhenyu, Liu, Tingwen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity Alignment
di: Su, Taoyu, et al.
Pubblicazione: (2024)
di: Su, Taoyu, et al.
Pubblicazione: (2024)
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
di: Su, Taoyu, et al.
Pubblicazione: (2025)
di: Su, Taoyu, et al.
Pubblicazione: (2025)
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
di: Cui, Shiyao, et al.
Pubblicazione: (2023)
di: Cui, Shiyao, et al.
Pubblicazione: (2023)
Verifying Cross-modal Entity Consistency in News using Vision-language Models
di: Tahmasebi, Sahar, et al.
Pubblicazione: (2025)
di: Tahmasebi, Sahar, et al.
Pubblicazione: (2025)
Unlocking the Power of Large Language Models for Multi-table Entity Matching
di: Tang, Yingkai, et al.
Pubblicazione: (2026)
di: Tang, Yingkai, et al.
Pubblicazione: (2026)
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
di: Wu, Zichen, et al.
Pubblicazione: (2024)
di: Wu, Zichen, et al.
Pubblicazione: (2024)
MMPKUBase: A Comprehensive and High-quality Chinese Multi-modal Knowledge Graph
di: Yi, Xuan, et al.
Pubblicazione: (2024)
di: Yi, Xuan, et al.
Pubblicazione: (2024)
AIM: Let Any Multi-modal Large Language Models Embrace Efficient In-Context Learning
di: Gao, Jun, et al.
Pubblicazione: (2024)
di: Gao, Jun, et al.
Pubblicazione: (2024)
Not All Attention is Needed: Parameter and Computation Efficient Transfer Learning for Multi-modal Large Language Models
di: Wu, Qiong, et al.
Pubblicazione: (2024)
di: Wu, Qiong, et al.
Pubblicazione: (2024)
CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
di: He, Zheqi, et al.
Pubblicazione: (2024)
di: He, Zheqi, et al.
Pubblicazione: (2024)
MAC-SLU: Multi-Intent Automotive Cabin Spoken Language Understanding Benchmark
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
A Benchmark and Robustness Study of In-Context-Learning with Large Language Models in Music Entity Detection
di: Hachmeier, Simon, et al.
Pubblicazione: (2024)
di: Hachmeier, Simon, et al.
Pubblicazione: (2024)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
di: Wang, Xiao, et al.
Pubblicazione: (2025)
di: Wang, Xiao, et al.
Pubblicazione: (2025)
Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
di: Liu, Yangyang, et al.
Pubblicazione: (2025)
di: Liu, Yangyang, et al.
Pubblicazione: (2025)
Conversation Understanding using Relational Temporal Graph Neural Networks with Auxiliary Cross-Modality Interaction
di: Nguyen, Cam-Van Thi, et al.
Pubblicazione: (2023)
di: Nguyen, Cam-Van Thi, et al.
Pubblicazione: (2023)
Collaborative Evolution: Multi-Round Learning Between Large and Small Language Models for Emergent Fake News Detection
di: Zhou, Ziyi, et al.
Pubblicazione: (2025)
di: Zhou, Ziyi, et al.
Pubblicazione: (2025)
RiverEcho: Real-Time Interactive Digital System for Ancient Yellow River Culture
di: Wang, Haofeng, et al.
Pubblicazione: (2025)
di: Wang, Haofeng, et al.
Pubblicazione: (2025)
SkyLink: Unifying Street-Satellite Geo-Localization via UAV-Mediated 3D Scene Alignment
di: Zhang, Hongyang, et al.
Pubblicazione: (2025)
di: Zhang, Hongyang, et al.
Pubblicazione: (2025)
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
di: Wu, Siwei, et al.
Pubblicazione: (2024)
di: Wu, Siwei, et al.
Pubblicazione: (2024)
OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
di: Chen, Lichang, et al.
Pubblicazione: (2024)
di: Chen, Lichang, et al.
Pubblicazione: (2024)
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model
di: Niu, Fuqiang, et al.
Pubblicazione: (2024)
di: Niu, Fuqiang, et al.
Pubblicazione: (2024)
Video Summarization: Towards Entity-Aware Captions
di: Ayyubi, Hammad A., et al.
Pubblicazione: (2023)
di: Ayyubi, Hammad A., et al.
Pubblicazione: (2023)
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation
di: Zhu, Xiaofei, et al.
Pubblicazione: (2024)
di: Zhu, Xiaofei, et al.
Pubblicazione: (2024)
NativE: Multi-modal Knowledge Graph Completion in the Wild
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
di: Zhang, Yichi, et al.
Pubblicazione: (2024)
MIND Your Reasoning: A Meta-Cognitive Intuitive-Reflective Network for Dual-Reasoning in Multimodal Stance Detection
di: Wang, Bingbing, et al.
Pubblicazione: (2025)
di: Wang, Bingbing, et al.
Pubblicazione: (2025)
MMSD-Net: Towards Multi-modal Stuttering Detection
di: Nie, Liangyu, et al.
Pubblicazione: (2024)
di: Nie, Liangyu, et al.
Pubblicazione: (2024)
Distilling Neuro-Symbolic Programs into 3D Multi-modal LLMs
di: Mo, Wentao, et al.
Pubblicazione: (2026)
di: Mo, Wentao, et al.
Pubblicazione: (2026)
MIntRec2.0: A Large-scale Benchmark Dataset for Multimodal Intent Recognition and Out-of-scope Detection in Conversations
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
FineFake: A Knowledge-Enriched Dataset for Fine-Grained Multi-Domain Fake News Detection
di: Zhou, Ziyi, et al.
Pubblicazione: (2024)
di: Zhou, Ziyi, et al.
Pubblicazione: (2024)
RealBench: A Chinese Multi-image Understanding Benchmark Close to Real-world Scenarios
di: Zhao, Fei, et al.
Pubblicazione: (2025)
di: Zhao, Fei, et al.
Pubblicazione: (2025)
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
di: Chen, Tao, et al.
Pubblicazione: (2023)
di: Chen, Tao, et al.
Pubblicazione: (2023)
PediatricsMQA: a Multi-modal Pediatrics Question Answering Benchmark
di: Bahaj, Adil, et al.
Pubblicazione: (2025)
di: Bahaj, Adil, et al.
Pubblicazione: (2025)
Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models
di: Ye, Weihao, et al.
Pubblicazione: (2024)
di: Ye, Weihao, et al.
Pubblicazione: (2024)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
An Evaluation of Interleaved Instruction Tuning on Semantic Reasoning Performance in an Audio MLLM
di: Liu, Jiawei, et al.
Pubblicazione: (2025)
di: Liu, Jiawei, et al.
Pubblicazione: (2025)
Modeling Layout Reading Order as Ordering Relations for Visually-rich Document Understanding
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
di: Fang, Xiang, et al.
Pubblicazione: (2022)
di: Fang, Xiang, et al.
Pubblicazione: (2022)
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection
di: Gu, Yimeng, et al.
Pubblicazione: (2025)
di: Gu, Yimeng, et al.
Pubblicazione: (2025)
MMESGBench: Pioneering Multimodal Understanding and Complex Reasoning Benchmark for ESG Tasks
di: Zhang, Lei, et al.
Pubblicazione: (2025)
di: Zhang, Lei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity Alignment
di: Su, Taoyu, et al.
Pubblicazione: (2024) -
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
di: Su, Taoyu, et al.
Pubblicazione: (2025) -
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
di: Cui, Shiyao, et al.
Pubblicazione: (2023) -
Verifying Cross-modal Entity Consistency in News using Vision-language Models
di: Tahmasebi, Sahar, et al.
Pubblicazione: (2025) -
Unlocking the Power of Large Language Models for Multi-table Entity Matching
di: Tang, Yingkai, et al.
Pubblicazione: (2026)