IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Su, Taoyu, Sheng, Jiawei, Wang, Shicheng, Zhang, Xinghua, Xu, Hongbo, Liu, Tingwen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LoginMEA: Local-to-Global Interaction Network for Multi-modal Entity Alignment
von: Su, Taoyu, et al.
Veröffentlicht: (2024)
von: Su, Taoyu, et al.
Veröffentlicht: (2024)
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
von: Su, Taoyu, et al.
Veröffentlicht: (2025)
von: Su, Taoyu, et al.
Veröffentlicht: (2025)
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
von: Cui, Shiyao, et al.
Veröffentlicht: (2023)
von: Cui, Shiyao, et al.
Veröffentlicht: (2023)
Verifying Cross-modal Entity Consistency in News using Vision-language Models
von: Tahmasebi, Sahar, et al.
Veröffentlicht: (2025)
von: Tahmasebi, Sahar, et al.
Veröffentlicht: (2025)
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
von: Wu, Zichen, et al.
Veröffentlicht: (2024)
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaofei, et al.
Veröffentlicht: (2024)
AIM: Let Any Multi-modal Large Language Models Embrace Efficient In-Context Learning
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
MAC-SLU: Multi-Intent Automotive Cabin Spoken Language Understanding Benchmark
von: Peng, Yuezhang, et al.
Veröffentlicht: (2025)
von: Peng, Yuezhang, et al.
Veröffentlicht: (2025)
Unlocking the Power of Large Language Models for Multi-table Entity Matching
von: Tang, Yingkai, et al.
Veröffentlicht: (2026)
von: Tang, Yingkai, et al.
Veröffentlicht: (2026)
MMPKUBase: A Comprehensive and High-quality Chinese Multi-modal Knowledge Graph
von: Yi, Xuan, et al.
Veröffentlicht: (2024)
von: Yi, Xuan, et al.
Veröffentlicht: (2024)
Not All Attention is Needed: Parameter and Computation Efficient Transfer Learning for Multi-modal Large Language Models
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
von: Wu, Qiong, et al.
Veröffentlicht: (2024)
CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
von: He, Zheqi, et al.
Veröffentlicht: (2024)
von: He, Zheqi, et al.
Veröffentlicht: (2024)
Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
A Benchmark and Robustness Study of In-Context-Learning with Large Language Models in Music Entity Detection
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
von: Hachmeier, Simon, et al.
Veröffentlicht: (2024)
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
von: Wu, Siwei, et al.
Veröffentlicht: (2024)
von: Wu, Siwei, et al.
Veröffentlicht: (2024)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
Listening to the Unspoken: Exploring "365" Aspects of Multimodal Interview Performance Assessment
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
MIntRec2.0: A Large-scale Benchmark Dataset for Multimodal Intent Recognition and Out-of-scope Detection in Conversations
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
Shapley Value-based Contrastive Alignment for Multimodal Information Extraction
von: Luo, Wen, et al.
Veröffentlicht: (2024)
von: Luo, Wen, et al.
Veröffentlicht: (2024)
OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
von: Chen, Lichang, et al.
Veröffentlicht: (2024)
Lost in Overlap: Exploring Logit-based Watermark Collision in LLMs
von: Luo, Yiyang, et al.
Veröffentlicht: (2024)
von: Luo, Yiyang, et al.
Veröffentlicht: (2024)
NativE: Multi-modal Knowledge Graph Completion in the Wild
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
MMESGBench: Pioneering Multimodal Understanding and Complex Reasoning Benchmark for ESG Tasks
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
ChartEditor: A Reinforcement Learning Framework for Robust Chart Editing
von: Chen, Liangyu, et al.
Veröffentlicht: (2025)
von: Chen, Liangyu, et al.
Veröffentlicht: (2025)
MMSD-Net: Towards Multi-modal Stuttering Detection
von: Nie, Liangyu, et al.
Veröffentlicht: (2024)
von: Nie, Liangyu, et al.
Veröffentlicht: (2024)
Distilling Neuro-Symbolic Programs into 3D Multi-modal LLMs
von: Mo, Wentao, et al.
Veröffentlicht: (2026)
von: Mo, Wentao, et al.
Veröffentlicht: (2026)
Collaborative Evolution: Multi-Round Learning Between Large and Small Language Models for Emergent Fake News Detection
von: Zhou, Ziyi, et al.
Veröffentlicht: (2025)
von: Zhou, Ziyi, et al.
Veröffentlicht: (2025)
RealBench: A Chinese Multi-image Understanding Benchmark Close to Real-world Scenarios
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation
von: Zhang, Bo, et al.
Veröffentlicht: (2024)
von: Zhang, Bo, et al.
Veröffentlicht: (2024)
Video Summarization: Towards Entity-Aware Captions
von: Ayyubi, Hammad A., et al.
Veröffentlicht: (2023)
von: Ayyubi, Hammad A., et al.
Veröffentlicht: (2023)
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
von: Chen, Tao, et al.
Veröffentlicht: (2023)
von: Chen, Tao, et al.
Veröffentlicht: (2023)
Multimodal Multi-turn Conversation Stance Detection: A Challenge Dataset and Effective Model
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
von: Niu, Fuqiang, et al.
Veröffentlicht: (2024)
ResearchPulse: Building Method-Experiment Chains through Multi-Document Scientific Inference
von: Chen, Qi, et al.
Veröffentlicht: (2025)
von: Chen, Qi, et al.
Veröffentlicht: (2025)
Towards Pretraining Robust ASR Foundation Model with Acoustic-Aware Data Augmentation
von: Liu, Dancheng, et al.
Veröffentlicht: (2025)
von: Liu, Dancheng, et al.
Veröffentlicht: (2025)
PediatricsMQA: a Multi-modal Pediatrics Question Answering Benchmark
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
von: Bahaj, Adil, et al.
Veröffentlicht: (2025)
MIND Your Reasoning: A Meta-Cognitive Intuitive-Reflective Network for Dual-Reasoning in Multimodal Stance Detection
von: Wang, Bingbing, et al.
Veröffentlicht: (2025)
von: Wang, Bingbing, et al.
Veröffentlicht: (2025)
Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models
von: Ye, Weihao, et al.
Veröffentlicht: (2024)
von: Ye, Weihao, et al.
Veröffentlicht: (2024)
Reverse Region-to-Entity Annotation for Pixel-Level Visual Entity Linking
von: Xu, Zhengfei, et al.
Veröffentlicht: (2024)
von: Xu, Zhengfei, et al.
Veröffentlicht: (2024)
FineFake: A Knowledge-Enriched Dataset for Fine-Grained Multi-Domain Fake News Detection
von: Zhou, Ziyi, et al.
Veröffentlicht: (2024)
von: Zhou, Ziyi, et al.
Veröffentlicht: (2024)
An Evaluation of Interleaved Instruction Tuning on Semantic Reasoning Performance in an Audio MLLM
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LoginMEA: Local-to-Global Interaction Network for Multi-modal Entity Alignment
von: Su, Taoyu, et al.
Veröffentlicht: (2024) -
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
von: Su, Taoyu, et al.
Veröffentlicht: (2025) -
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
von: Cui, Shiyao, et al.
Veröffentlicht: (2023) -
Verifying Cross-modal Entity Consistency in News using Vision-language Models
von: Tahmasebi, Sahar, et al.
Veröffentlicht: (2025) -
Mixture-of-Prompt-Experts for Multi-modal Semantic Understanding
von: Wu, Zichen, et al.
Veröffentlicht: (2024)