Code-MIE: A Code-style Model for Multimodal Information Extraction with Scene Graph and Entity Attribute Knowledge Enhancement
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jiang, Qiu, Ge, Fei, Hao, Xie, Dongdong, Li, Jinbo, Li, Fei, Teng, Chong, Ji, Donghong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement
by: Lin, Shaoqing, et al.
Published: (2025)
by: Lin, Shaoqing, et al.
Published: (2025)
Revisiting Structured Sentiment Analysis as Latent Dependency Graph Parsing
by: Zhou, Chengjie, et al.
Published: (2024)
by: Zhou, Chengjie, et al.
Published: (2024)
M$^{3}$D: A Multimodal, Multilingual and Multitask Dataset for Grounded Document-level Information Extraction
by: Liu, Jiang, et al.
Published: (2024)
by: Liu, Jiang, et al.
Published: (2024)
CMNER: A Chinese Multimodal NER Dataset based on Social Media
by: Ji, Yuanze, et al.
Published: (2024)
by: Ji, Yuanze, et al.
Published: (2024)
Multi-Granular Multimodal Clue Fusion for Meme Understanding
by: Zheng, Li, et al.
Published: (2025)
by: Zheng, Li, et al.
Published: (2025)
Dynamic Emotion and Personality Profiling for Multimodal Deception Detection
by: Zheng, Li, et al.
Published: (2026)
by: Zheng, Li, et al.
Published: (2026)
Enhancing Hyperbole and Metaphor Detection with Their Bidirectional Dynamic Interaction and Emotion Knowledge
by: Zheng, Li, et al.
Published: (2025)
by: Zheng, Li, et al.
Published: (2025)
Enhance-then-Balance Modality Collaboration for Robust Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2026)
by: He, Kang, et al.
Published: (2026)
DALR: Dual-level Alignment Learning for Multimodal Sentence Representation Learning
by: He, Kang, et al.
Published: (2025)
by: He, Kang, et al.
Published: (2025)
PaSE: Prototype-aligned Calibration and Shapley-based Equilibrium for Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2025)
by: He, Kang, et al.
Published: (2025)
Harvesting Events from Multiple Sources: Towards a Cross-Document Event Extraction Paradigm
by: Gao, Qiang, et al.
Published: (2024)
by: Gao, Qiang, et al.
Published: (2024)
Modeling Unified Semantic Discourse Structure for High-quality Headline Generation
by: Xu, Minghui, et al.
Published: (2024)
by: Xu, Minghui, et al.
Published: (2024)
DongbaMIE: A Multimodal Information Extraction Dataset for Evaluating Semantic Understanding of Dongba Pictograms
by: Bi, Xiaojun, et al.
Published: (2025)
by: Bi, Xiaojun, et al.
Published: (2025)
Enhancing Cross-Document Event Coreference Resolution by Discourse Structure and Semantic Information
by: Gao, Qiang, et al.
Published: (2024)
by: Gao, Qiang, et al.
Published: (2024)
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction
by: Yuan, Mengying, et al.
Published: (2025)
by: Yuan, Mengying, et al.
Published: (2025)
Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering
by: Zheng, Li, et al.
Published: (2026)
by: Zheng, Li, et al.
Published: (2026)
Zero-Shot Conversational Stance Detection: Dataset and Approaches
by: Ding, Yuzhe, et al.
Published: (2025)
by: Ding, Yuzhe, et al.
Published: (2025)
TRIDENT: Enhancing Large Language Model Safety with Tri-Dimensional Diversified Red-Teaming Data Synthesis
by: Wu, Xiaorui, et al.
Published: (2025)
by: Wu, Xiaorui, et al.
Published: (2025)
LASQ: A Low-resource Aspect-based Sentiment Quadruple Extraction Dataset
by: Yusufu, Aizihaierjiang, et al.
Published: (2026)
by: Yusufu, Aizihaierjiang, et al.
Published: (2026)
VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models
by: Jiang, Lingjie, et al.
Published: (2025)
by: Jiang, Lingjie, et al.
Published: (2025)
Recognizing Everything from All Modalities at Once: Grounded Multimodal Universal Information Extraction
by: Zhang, Meishan, et al.
Published: (2024)
by: Zhang, Meishan, et al.
Published: (2024)
Scene Understanding Enabled Semantic Communication with Open Channel Coding
by: Xiang, Zhe, et al.
Published: (2025)
by: Xiang, Zhe, et al.
Published: (2025)
EVOREFUSE: Evolutionary Prompt Optimization for Evaluation and Mitigation of LLM Over-Refusal to Pseudo-Malicious Instructions
by: Wu, Xiaorui, et al.
Published: (2025)
by: Wu, Xiaorui, et al.
Published: (2025)
Cardinality Estimation on Hyper-relational Knowledge Graphs
by: Teng, Fei, et al.
Published: (2024)
by: Teng, Fei, et al.
Published: (2024)
BioDeepHash: Mapping Biometrics into a Stable Code
by: Song, Baogang, et al.
Published: (2024)
by: Song, Baogang, et al.
Published: (2024)
Graph Set-colorings And Hypergraphs In Topological Coding
by: Yao, Bing, et al.
Published: (2022)
by: Yao, Bing, et al.
Published: (2022)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
Linking Cryptoasset Attribution Tags to Knowledge Graph Entities: An LLM-based Approach
by: Avice, Régnier, et al.
Published: (2025)
by: Avice, Régnier, et al.
Published: (2025)
Decoding MIE: A Novel Dataset Approach Using Topic Extraction and Affiliation Parsing
by: Bitaraf, Ehsan, et al.
Published: (2024)
by: Bitaraf, Ehsan, et al.
Published: (2024)
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
by: Cui, Shiyao, et al.
Published: (2023)
by: Cui, Shiyao, et al.
Published: (2023)
Chain of Code: Reasoning with a Language Model-Augmented Code Emulator
by: Li, Chengshu, et al.
Published: (2023)
by: Li, Chengshu, et al.
Published: (2023)
Universal Scene Graph Generation
by: Wu, Shengqiong, et al.
Published: (2025)
by: Wu, Shengqiong, et al.
Published: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
Modular Layout Synthesis (MLS): Front-end Code via Structure Normalization and Constrained Generation
by: Liu, Chong, et al.
Published: (2025)
by: Liu, Chong, et al.
Published: (2025)
CodeScore: Evaluating Code Generation by Learning Code Execution
by: Dong, Yihong, et al.
Published: (2023)
by: Dong, Yihong, et al.
Published: (2023)
Coordinated Planning for Stability Enhancement in High IBR-Penetrated Systems
by: Chu, Zhongda, et al.
Published: (2024)
by: Chu, Zhongda, et al.
Published: (2024)
Understanding the Embedding Models on Hyper-relational Knowledge Graph
by: Wang, Yubo, et al.
Published: (2025)
by: Wang, Yubo, et al.
Published: (2025)
Generalized Code Index Modulation Aided AFDM for Spread Spectrum Systems
by: Qian, Mi, et al.
Published: (2025)
by: Qian, Mi, et al.
Published: (2025)
When Graph meets Multimodal: Benchmarking and Meditating on Multimodal Attributed Graphs Learning
by: Yan, Hao, et al.
Published: (2024)
by: Yan, Hao, et al.
Published: (2024)
VSA:Visual-Structural Alignment for UI-to-Code
by: Wu, Xian, et al.
Published: (2025)
by: Wu, Xian, et al.
Published: (2025)
Similar Items
-
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement
by: Lin, Shaoqing, et al.
Published: (2025) -
Revisiting Structured Sentiment Analysis as Latent Dependency Graph Parsing
by: Zhou, Chengjie, et al.
Published: (2024) -
M$^{3}$D: A Multimodal, Multilingual and Multitask Dataset for Grounded Document-level Information Extraction
by: Liu, Jiang, et al.
Published: (2024) -
CMNER: A Chinese Multimodal NER Dataset based on Social Media
by: Ji, Yuanze, et al.
Published: (2024) -
Multi-Granular Multimodal Clue Fusion for Meme Understanding
by: Zheng, Li, et al.
Published: (2025)