DongbaMIE: A Multimodal Information Extraction Dataset for Evaluating Semantic Understanding of Dongba Pictograms
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bi, Xiaojun, Li, Shuo, Xing, Junyao, Wang, Ziyue, Luo, Fuwen, Qiao, Weizheng, Han, Lu, Sun, Ziwei, Li, Peng, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Caption Generation for Dongba Paintings via Prompt Learning and Semantic Fusion
von: Qian, Shuangwu, et al.
Veröffentlicht: (2026)
von: Qian, Shuangwu, et al.
Veröffentlicht: (2026)
Effects of Different Particle Sizes of Ultrafine Indica Rice Flour and Lactobacillus casei on the Fermentation and Flavor Characteristics of Danling Dongba
von: Hong Jiang, et al.
Veröffentlicht: (2025)
von: Hong Jiang, et al.
Veröffentlicht: (2025)
Decoding MIE: A Novel Dataset Approach Using Topic Extraction and Affiliation Parsing
von: Bitaraf, Ehsan, et al.
Veröffentlicht: (2024)
von: Bitaraf, Ehsan, et al.
Veröffentlicht: (2024)
Code-MIE: A Code-style Model for Multimodal Information Extraction with Scene Graph and Entity Attribute Knowledge Enhancement
von: Liu, Jiang, et al.
Veröffentlicht: (2026)
von: Liu, Jiang, et al.
Veröffentlicht: (2026)
Thinking with Visual Abstract: Enhancing Multimodal Reasoning via Visual Abstraction
von: Liu, Dairu, et al.
Veröffentlicht: (2025)
von: Liu, Dairu, et al.
Veröffentlicht: (2025)
ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
Evaluating Time Awareness and Cross-modal Active Perception of Large Models via 4D Escape Room Task
von: Dong, Yurui, et al.
Veröffentlicht: (2026)
von: Dong, Yurui, et al.
Veröffentlicht: (2026)
EscapeCraft: A 3D Room Escape Environment for Benchmarking Complex Multimodal Reasoning Ability
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
Robust Multilingual Text-to-Pictogram Mapping for Scalable Reading Rehabilitation
von: Jhilal, Soufiane, et al.
Veröffentlicht: (2026)
von: Jhilal, Soufiane, et al.
Veröffentlicht: (2026)
Bridging Vision, Language, and Mathematics: Pictographic Character Reconstruction with Bézier Curves
von: Wan, Zihao, et al.
Veröffentlicht: (2025)
von: Wan, Zihao, et al.
Veröffentlicht: (2025)
Browse and Concentrate: Comprehending Multimodal Content via prior-LLM Context Fusion
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
von: Wang, Ziyue, et al.
Veröffentlicht: (2024)
Model Composition for Multimodal Large Language Models
von: Chen, Chi, et al.
Veröffentlicht: (2024)
von: Chen, Chi, et al.
Veröffentlicht: (2024)
InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
von: Wang, Yi, et al.
Veröffentlicht: (2023)
von: Wang, Yi, et al.
Veröffentlicht: (2023)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding
von: Luo, Fuwen, et al.
Veröffentlicht: (2025)
von: Luo, Fuwen, et al.
Veröffentlicht: (2025)
Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors
von: Luo, Fuwen, et al.
Veröffentlicht: (2026)
von: Luo, Fuwen, et al.
Veröffentlicht: (2026)
Perspective Transition of Large Language Models for Solving Subjective Tasks
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
von: Wang, Xiaolong, et al.
Veröffentlicht: (2025)
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2023)
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2023)
EIVEN: Efficient Implicit Attribute Value Extraction using Multimodal LLM
von: Zou, Henry Peng, et al.
Veröffentlicht: (2024)
von: Zou, Henry Peng, et al.
Veröffentlicht: (2024)
$\mathbb{G}_m$-Equivariant Degenerations of del Pezzo Surfaces
von: Peng, Junyao
Veröffentlicht: (2025)
von: Peng, Junyao
Veröffentlicht: (2025)
Asymptotics of stability thresholds
von: Peng, Junyao
Veröffentlicht: (2025)
von: Peng, Junyao
Veröffentlicht: (2025)
Enhancing Steganographic Text Extraction: Evaluating the Impact of NLP Models on Accuracy and Semantic Coherence
von: Li, Mingyang, et al.
Veröffentlicht: (2024)
von: Li, Mingyang, et al.
Veröffentlicht: (2024)
StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding
von: Lin, Junming, et al.
Veröffentlicht: (2024)
von: Lin, Junming, et al.
Veröffentlicht: (2024)
CODIS: Benchmarking Context-Dependent Visual Comprehension for Multimodal Large Language Models
von: Luo, Fuwen, et al.
Veröffentlicht: (2024)
von: Luo, Fuwen, et al.
Veröffentlicht: (2024)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Table-Filling via Mean Teacher for Cross-domain Aspect Sentiment Triplet Extraction
von: Peng, Kun, et al.
Veröffentlicht: (2024)
von: Peng, Kun, et al.
Veröffentlicht: (2024)
Unsolvable Problem Detection: Robust Understanding Evaluation for Large Multimodal Models
von: Miyai, Atsuyuki, et al.
Veröffentlicht: (2024)
von: Miyai, Atsuyuki, et al.
Veröffentlicht: (2024)
MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
DIE ARBEIT DES ÜBERSETZENS: RILKE UND MICHELANGELO („SE ’L MIE ROZZO MARTELLO‘‘)
von: Astrid Dröse, et al.
Veröffentlicht: (2025)
von: Astrid Dröse, et al.
Veröffentlicht: (2025)
MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping
von: Shan, Xiaojun, et al.
Veröffentlicht: (2025)
von: Shan, Xiaojun, et al.
Veröffentlicht: (2025)
NeuBM: Mitigating Model Bias in Graph Neural Networks through Neutral Input Calibration
von: Gu, Jiawei, et al.
Veröffentlicht: (2025)
von: Gu, Jiawei, et al.
Veröffentlicht: (2025)
Spot Risks Before Speaking! Unraveling Safety Attention Heads in Large Vision-Language Models
von: Zheng, Ziwei, et al.
Veröffentlicht: (2025)
von: Zheng, Ziwei, et al.
Veröffentlicht: (2025)
Towards Multimodal Lifelong Understanding: A Dataset and Agentic Baseline
von: Chen, Guo, et al.
Veröffentlicht: (2026)
von: Chen, Guo, et al.
Veröffentlicht: (2026)
Crossroads of Continents: Automated Artifact Extraction for Cultural Adaptation with Large Multimodal Models
von: Mukherjee, Anjishnu, et al.
Veröffentlicht: (2024)
von: Mukherjee, Anjishnu, et al.
Veröffentlicht: (2024)
Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models
von: Wang, Bin, et al.
Veröffentlicht: (2025)
von: Wang, Bin, et al.
Veröffentlicht: (2025)
The Waiting Room Questionnaire, a Patient‐Reported Outcome Instrument Including Pictograms for Facilitating the Diagnosis of Functional Gastrointestinal Disorders
von: Nick Goelen, et al.
Veröffentlicht: (2025)
von: Nick Goelen, et al.
Veröffentlicht: (2025)
A Comparative Study of Mass Extraction Schemes and $π^\pm-ρ^\pm$ Mixing
von: Wang, Ziyue
Veröffentlicht: (2026)
von: Wang, Ziyue
Veröffentlicht: (2026)
Image Matters: A New Dataset and Empirical Study for Multimodal Hyperbole Detection
von: Zhang, Huixuan, et al.
Veröffentlicht: (2023)
von: Zhang, Huixuan, et al.
Veröffentlicht: (2023)
Automated Glaucoma Report Generation via Dual-Attention Semantic Parallel-LSTM and Multimodal Clinical Data Integration
von: Huang, Cheng, et al.
Veröffentlicht: (2025)
von: Huang, Cheng, et al.
Veröffentlicht: (2025)
Understanding the Role of LLMs in Multimodal Evaluation Benchmarks
von: Jiang, Botian, et al.
Veröffentlicht: (2024)
von: Jiang, Botian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Caption Generation for Dongba Paintings via Prompt Learning and Semantic Fusion
von: Qian, Shuangwu, et al.
Veröffentlicht: (2026) -
Effects of Different Particle Sizes of Ultrafine Indica Rice Flour and Lactobacillus casei on the Fermentation and Flavor Characteristics of Danling Dongba
von: Hong Jiang, et al.
Veröffentlicht: (2025) -
Decoding MIE: A Novel Dataset Approach Using Topic Extraction and Affiliation Parsing
von: Bitaraf, Ehsan, et al.
Veröffentlicht: (2024) -
Code-MIE: A Code-style Model for Multimodal Information Extraction with Scene Graph and Entity Attribute Knowledge Enhancement
von: Liu, Jiang, et al.
Veröffentlicht: (2026) -
Thinking with Visual Abstract: Enhancing Multimodal Reasoning via Visual Abstraction
von: Liu, Dairu, et al.
Veröffentlicht: (2025)