Cultivating Game Sense for Yourself: Making VLMs Gaming Experts
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Wenxuan, He, Jiangyang, Zhang, Zhanqiu, Guo, Yiwen, Zang, Tianning |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interaction
by: Chen, Xingwu, et al.
Published: (2026)
by: Chen, Xingwu, et al.
Published: (2026)
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning
by: Shi, Chengshuai, et al.
Published: (2026)
by: Shi, Chengshuai, et al.
Published: (2026)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
by: Tong, Jingqi, et al.
Published: (2025)
by: Tong, Jingqi, et al.
Published: (2025)
Level Up Your Tutorials: VLMs for Game Tutorials Quality Assessment
by: Cambrin, Daniele Rege, et al.
Published: (2024)
by: Cambrin, Daniele Rege, et al.
Published: (2024)
Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go
by: Ma, Yichuan, et al.
Published: (2026)
by: Ma, Yichuan, et al.
Published: (2026)
Multi-Task Model Merging via Adaptive Weight Disentanglement
by: Xiong, Feng, et al.
Published: (2024)
by: Xiong, Feng, et al.
Published: (2024)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
by: Huang, Jen-tse, et al.
Published: (2024)
by: Huang, Jen-tse, et al.
Published: (2024)
Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety
by: Bonagiri, Vamshi Krishna, et al.
Published: (2025)
by: Bonagiri, Vamshi Krishna, et al.
Published: (2025)
Do Images Speak Louder than Words? Investigating the Effect of Textual Misinformation in VLMs
by: Zhang, Chi, et al.
Published: (2026)
by: Zhang, Chi, et al.
Published: (2026)
Game-MUG: Multimodal Oriented Game Situation Understanding and Commentary Generation Dataset
by: Zhang, Zhihao, et al.
Published: (2024)
by: Zhang, Zhihao, et al.
Published: (2024)
DR-Arena: an Automated Evaluation Framework for Deep Research Agents
by: Gao, Yiwen, et al.
Published: (2026)
by: Gao, Yiwen, et al.
Published: (2026)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
by: Jing, Linglin, et al.
Published: (2025)
by: Jing, Linglin, et al.
Published: (2025)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
by: Zhang, Haoran, et al.
Published: (2024)
by: Zhang, Haoran, et al.
Published: (2024)
Framing the Game: How Context Shapes LLM Decision-Making
by: Robinson, Isaac, et al.
Published: (2025)
by: Robinson, Isaac, et al.
Published: (2025)
BEYOND DIALOGUE: A Profile-Dialogue Alignment Framework Towards General Role-Playing Language Model
by: Yu, Yeyong, et al.
Published: (2024)
by: Yu, Yeyong, et al.
Published: (2024)
LinguaGame: A Linguistically Grounded Game-Theoretic Paradigm for Multi-Agent Dialogue Generation
by: Ye, Yuxiao, et al.
Published: (2026)
by: Ye, Yuxiao, et al.
Published: (2026)
GAMEBoT: Transparent Assessment of LLM Reasoning in Games
by: Lin, Wenye, et al.
Published: (2024)
by: Lin, Wenye, et al.
Published: (2024)
EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training
by: Du, Yiyang, et al.
Published: (2026)
by: Du, Yiyang, et al.
Published: (2026)
The Hunger Game Debate: On the Emergence of Over-Competition in Multi-Agent Systems
by: Ma, Xinbei, et al.
Published: (2025)
by: Ma, Xinbei, et al.
Published: (2025)
GameArena: Evaluating LLM Reasoning through Live Computer Games
by: Hu, Lanxiang, et al.
Published: (2024)
by: Hu, Lanxiang, et al.
Published: (2024)
Keeping Yourself is Important in Downstream Tuning Multimodal Large Language Model
by: Huang, Wenke, et al.
Published: (2025)
by: Huang, Wenke, et al.
Published: (2025)
TurnaboutLLM: A Deductive Reasoning Benchmark from Detective Games
by: Yuan, Yuan, et al.
Published: (2025)
by: Yuan, Yuan, et al.
Published: (2025)
Talking to Yourself: Defying Forgetting in Large Language Models
by: Sun, Yutao, et al.
Published: (2026)
by: Sun, Yutao, et al.
Published: (2026)
Skin-in-the-Game: Decision Making via Multi-Stakeholder Alignment in LLMs
by: Sel, Bilgehan, et al.
Published: (2024)
by: Sel, Bilgehan, et al.
Published: (2024)
ChronoPlay: A Framework for Modeling Dual Dynamics and Authenticity in Game RAG Benchmarks
by: He, Liyang, et al.
Published: (2025)
by: He, Liyang, et al.
Published: (2025)
How to Protect Yourself from 5G Radiation? Investigating LLM Responses to Implicit Misinformation
by: Guo, Ruohao, et al.
Published: (2025)
by: Guo, Ruohao, et al.
Published: (2025)
On the Perception Bottleneck of VLMs for Chart Understanding
by: Liu, Junteng, et al.
Published: (2025)
by: Liu, Junteng, et al.
Published: (2025)
Can LLMs Play Ô Ăn Quan Game? A Study of Multi-Step Planning and Decision Making
by: Nguyen, Sang Quang, et al.
Published: (2025)
by: Nguyen, Sang Quang, et al.
Published: (2025)
Multimodal Shannon Game with Images
by: Zouhar, Vilém, et al.
Published: (2023)
by: Zouhar, Vilém, et al.
Published: (2023)
A Text-to-Game Engine for UGC-Based Role-Playing Games
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Game Plot Design with an LLM-powered Assistant: An Empirical Study with Game Designers
by: Alavi, Seyed Hossein, et al.
Published: (2024)
by: Alavi, Seyed Hossein, et al.
Published: (2024)
GIFT: Games as Informal Training for Generalizable LLMs
by: Lyu, Nuoyan, et al.
Published: (2026)
by: Lyu, Nuoyan, et al.
Published: (2026)
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
by: Merino, Tim, et al.
Published: (2024)
by: Merino, Tim, et al.
Published: (2024)
Learn from Downstream and Be Yourself in Multimodal Large Language Model Fine-Tuning
by: Huang, Wenke, et al.
Published: (2024)
by: Huang, Wenke, et al.
Published: (2024)
A Theoretical Game of Attacks via Compositional Skills
by: Wu, Xinbo, et al.
Published: (2026)
by: Wu, Xinbo, et al.
Published: (2026)
TextGames: Learning to Self-Play Text-Based Puzzle Games via Language Model Reasoning
by: Hudi, Frederikus, et al.
Published: (2025)
by: Hudi, Frederikus, et al.
Published: (2025)
Manager: Aggregating Insights from Unimodal Experts in Two-Tower VLMs and MLLMs
by: Xu, Xiao, et al.
Published: (2025)
by: Xu, Xiao, et al.
Published: (2025)
Automatically Detecting Amusing Games in Wordle
by: Luo, Ronaldo, et al.
Published: (2025)
by: Luo, Ronaldo, et al.
Published: (2025)
Collaborative Problem-Solving in an Optimization Game
by: Jeknic, Isidora, et al.
Published: (2025)
by: Jeknic, Isidora, et al.
Published: (2025)
Ethical Considerations of Large Language Models in Game Playing
by: Zhang, Qingquan, et al.
Published: (2025)
by: Zhang, Qingquan, et al.
Published: (2025)
Similar Items
-
Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interaction
by: Chen, Xingwu, et al.
Published: (2026) -
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning
by: Shi, Chengshuai, et al.
Published: (2026) -
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
by: Tong, Jingqi, et al.
Published: (2025) -
Level Up Your Tutorials: VLMs for Game Tutorials Quality Assessment
by: Cambrin, Daniele Rege, et al.
Published: (2024) -
Mixing Expert Knowledge: Bring Human Thoughts Back To the Game of Go
by: Ma, Yichuan, et al.
Published: (2026)