Game-MUG: Multimodal Oriented Game Situation Understanding and Commentary Generation Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Zhihao, Cao, Feiqi, Mo, Yingbin, Zhang, Yiran, Poon, Josiah, Han, Caren |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Large Language Models and Tunings: Vision, Language, Sensors, Audio, and Beyond
von: Han, Soyeon Caren, et al.
Veröffentlicht: (2024)
von: Han, Soyeon Caren, et al.
Veröffentlicht: (2024)
3M: Multi-modal Multi-task Multi-teacher Learning for Game Event Detection
von: Ng, Thye Shan, et al.
Veröffentlicht: (2024)
von: Ng, Thye Shan, et al.
Veröffentlicht: (2024)
PEACH: Pretrained-embedding Explanation Across Contextual and Hierarchical Structure
von: Cao, Feiqi, et al.
Veröffentlicht: (2024)
von: Cao, Feiqi, et al.
Veröffentlicht: (2024)
ChuLo: Chunk-Level Key Information Representation for Long Document Understanding
von: Li, Yan, et al.
Veröffentlicht: (2024)
von: Li, Yan, et al.
Veröffentlicht: (2024)
3M-Health: Multimodal Multi-Teacher Knowledge Distillation for Mental Health Detection
von: Cabral, Rina Carines, et al.
Veröffentlicht: (2024)
von: Cabral, Rina Carines, et al.
Veröffentlicht: (2024)
3MVRD: Multimodal Multi-task Multi-teacher Visually-Rich Form Document Understanding
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
In-game Toxic Language Detection: Shared Task and Attention Residuals
von: Jia, Yuanzhe, et al.
Veröffentlicht: (2022)
von: Jia, Yuanzhe, et al.
Veröffentlicht: (2022)
Local Interpretations for Explainable Natural Language Processing: A Survey
von: Luo, Siwen, et al.
Veröffentlicht: (2021)
von: Luo, Siwen, et al.
Veröffentlicht: (2021)
GEM-VPC: A dual Graph-Enhanced Multimodal integration for Video Paragraph Captioning
von: Wang, Eileen, et al.
Veröffentlicht: (2024)
von: Wang, Eileen, et al.
Veröffentlicht: (2024)
From Generation to Detection: A Multimodal Multi-Task Dataset for Benchmarking Health Misinformation
von: Zhang, Zhihao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2025)
LinguaGame: A Linguistically Grounded Game-Theoretic Paradigm for Multi-Agent Dialogue Generation
von: Ye, Yuxiao, et al.
Veröffentlicht: (2026)
von: Ye, Yuxiao, et al.
Veröffentlicht: (2026)
TriG-NER: Triplet-Grid Framework for Discontinuous Named Entity Recognition
von: Cabral, Rina Carines, et al.
Veröffentlicht: (2024)
von: Cabral, Rina Carines, et al.
Veröffentlicht: (2024)
Characterizing Language Use in a Collaborative Situated Game
von: Tomlin, Nicholas, et al.
Veröffentlicht: (2025)
von: Tomlin, Nicholas, et al.
Veröffentlicht: (2025)
Real-Time Generation of Game Video Commentary with Multimodal LLMs: Pause-Aware Decoding Approaches
von: Afzal, Anum, et al.
Veröffentlicht: (2026)
von: Afzal, Anum, et al.
Veröffentlicht: (2026)
VRD-IU: Lessons from Visually Rich Document Intelligence and Understanding
von: Ding, Yihao, et al.
Veröffentlicht: (2025)
von: Ding, Yihao, et al.
Veröffentlicht: (2025)
From Multimodal Perception to Strategic Reasoning: A Survey on AI-Generated Game Commentary
von: Zheng, Qirui, et al.
Veröffentlicht: (2025)
von: Zheng, Qirui, et al.
Veröffentlicht: (2025)
Commentary Generation from Data Records of Multiplayer Strategy Esports Game
von: Wang, Zihan, et al.
Veröffentlicht: (2022)
von: Wang, Zihan, et al.
Veröffentlicht: (2022)
SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling
von: Wang, Eileen, et al.
Veröffentlicht: (2024)
von: Wang, Eileen, et al.
Veröffentlicht: (2024)
Enhancing Dialogue Generation in Werewolf Game Through Situation Analysis and Persuasion Strategies
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
von: Qi, Zhiyang, et al.
Veröffentlicht: (2024)
MUG-Eval: A Proxy Evaluation Framework for Multilingual Generation Capabilities in Any Language
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
von: Song, Seyoung, et al.
Veröffentlicht: (2025)
TransportationGames: Benchmarking Transportation Knowledge of (Multimodal) Large Language Models
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
Multimodal Shannon Game with Images
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
von: Tong, Jingqi, et al.
Veröffentlicht: (2025)
von: Tong, Jingqi, et al.
Veröffentlicht: (2025)
Graph-Based Multimodal Contrastive Learning for Chart Question Answering
von: Dai, Yue, et al.
Veröffentlicht: (2025)
von: Dai, Yue, et al.
Veröffentlicht: (2025)
PDF-MVQA: A Dataset for Multimodal Information Retrieval in PDF-based Visual Question Answering
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
Compositional Understanding in Signaling Games
von: Freeborn, David Peter Wallis
Veröffentlicht: (2025)
von: Freeborn, David Peter Wallis
Veröffentlicht: (2025)
Cultivating Game Sense for Yourself: Making VLMs Gaming Experts
von: Lu, Wenxuan, et al.
Veröffentlicht: (2025)
von: Lu, Wenxuan, et al.
Veröffentlicht: (2025)
Enhancing Commentary Strategies for Imperfect Information Card Games: A Study of Large Language Models in Guandan Commentary
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
MSG-Chart: Multimodal Scene Graph for ChartQA
von: Dai, Yue, et al.
Veröffentlicht: (2024)
von: Dai, Yue, et al.
Veröffentlicht: (2024)
Multimodal Commonsense Knowledge Distillation for Visual Question Answering
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
GameTileNet: A Semantic Dataset for Low-Resolution Game Art in Procedural Content Generation
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
Saying the Unsaid: Revealing the Hidden Language of Multimodal Systems Through Telephone Games
von: Zhao, Juntu, et al.
Veröffentlicht: (2025)
von: Zhao, Juntu, et al.
Veröffentlicht: (2025)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
von: Zhang, Haoran, et al.
Veröffentlicht: (2024)
SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
von: Hong, Jixiang, et al.
Veröffentlicht: (2025)
Multimodal Situational Safety
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zhou, Kaiwen, et al.
Veröffentlicht: (2024)
Deep Learning based Visually Rich Document Content Understanding: A Survey
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
von: Ding, Yihao, et al.
Veröffentlicht: (2024)
When More Is Less: A Systematic Analysis of Spatial and Commonsense Information for Visual Spatial Reasoning
von: Akasaka, Muku, et al.
Veröffentlicht: (2026)
von: Akasaka, Muku, et al.
Veröffentlicht: (2026)
GameArena: Evaluating LLM Reasoning through Live Computer Games
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
von: Hu, Lanxiang, et al.
Veröffentlicht: (2024)
SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models
von: Hyun, Lee, et al.
Veröffentlicht: (2023)
von: Hyun, Lee, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Multimodal Large Language Models and Tunings: Vision, Language, Sensors, Audio, and Beyond
von: Han, Soyeon Caren, et al.
Veröffentlicht: (2024) -
3M: Multi-modal Multi-task Multi-teacher Learning for Game Event Detection
von: Ng, Thye Shan, et al.
Veröffentlicht: (2024) -
PEACH: Pretrained-embedding Explanation Across Contextual and Hierarchical Structure
von: Cao, Feiqi, et al.
Veröffentlicht: (2024) -
ChuLo: Chunk-Level Key Information Representation for Long Document Understanding
von: Li, Yan, et al.
Veröffentlicht: (2024) -
3M-Health: Multimodal Multi-Teacher Knowledge Distillation for Mental Health Detection
von: Cabral, Rina Carines, et al.
Veröffentlicht: (2024)