Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Tong, Jingqi, Tang, Jixin, Li, Hangcheng, Mou, Yurong, Zhang, Ming, Zhao, Jun, Wen, Yanbo, Song, Fan, Zhan, Jiahao, Lu, Yuyang, Tao, Chaoran, Guo, Zhiyuan, Yu, Jizhou, Cheng, Tianhao, Xi, Zhiheng, Jiang, Changhao, Yin, Zhangyue, Zheng, Yining, Ge, Weifeng, Chen, Guanhua, Gui, Tao, Qiu, Xipeng, Zhang, Qi, Huang, Xuanjing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AI Can Learn Scientific Taste
di: Tong, Jingqi, et al.
Pubblicazione: (2026)
di: Tong, Jingqi, et al.
Pubblicazione: (2026)
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
di: Li, Bowen, et al.
Pubblicazione: (2026)
di: Li, Bowen, et al.
Pubblicazione: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
di: Hong, Yoosung
Pubblicazione: (2026)
di: Hong, Yoosung
Pubblicazione: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
di: Özkan, Miraç Buğra
Pubblicazione: (2025)
di: Özkan, Miraç Buğra
Pubblicazione: (2025)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
di: Belcamino, Valerio, et al.
Pubblicazione: (2026)
di: Belcamino, Valerio, et al.
Pubblicazione: (2026)
On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs
di: Zhao, Rosie, et al.
Pubblicazione: (2026)
di: Zhao, Rosie, et al.
Pubblicazione: (2026)
SIGN: Schema-Induced Games for Naming
di: Zhang, Ryan, et al.
Pubblicazione: (2025)
di: Zhang, Ryan, et al.
Pubblicazione: (2025)
Deciphering Digital Detectives: Understanding LLM Behaviors and Capabilities in Multi-Agent Mystery Games
di: Wu, Dekun, et al.
Pubblicazione: (2023)
di: Wu, Dekun, et al.
Pubblicazione: (2023)
Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models
di: Tan, Chee Wei, et al.
Pubblicazione: (2026)
di: Tan, Chee Wei, et al.
Pubblicazione: (2026)
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
Hybrid Game Control Envelope Synthesis
di: Kabra, Aditi, et al.
Pubblicazione: (2025)
di: Kabra, Aditi, et al.
Pubblicazione: (2025)
Intention Communication and Hypothesis Likelihood in Game-Theoretic Motion Planning
di: Chahine, Makram, et al.
Pubblicazione: (2022)
di: Chahine, Makram, et al.
Pubblicazione: (2022)
elsciRL: Integrating Language Solutions into Reinforcement Learning Problem Settings
di: Osborne, Philip, et al.
Pubblicazione: (2025)
di: Osborne, Philip, et al.
Pubblicazione: (2025)
FDQN: A Flexible Deep Q-Network Framework for Game Automation
di: Gujavarthy, Prabhath Reddy
Pubblicazione: (2024)
di: Gujavarthy, Prabhath Reddy
Pubblicazione: (2024)
PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data
di: Xiong, Kai, et al.
Pubblicazione: (2025)
di: Xiong, Kai, et al.
Pubblicazione: (2025)
When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't
di: Nemitz, Jonathan, et al.
Pubblicazione: (2026)
di: Nemitz, Jonathan, et al.
Pubblicazione: (2026)
The Skin Game: Revolutionizing Standards for AI Dermatology Model Comparison
di: Miętkiewicz, Łukasz, et al.
Pubblicazione: (2025)
di: Miętkiewicz, Łukasz, et al.
Pubblicazione: (2025)
Game of Thought: Robust Information Seeking with Large Language Models Using Game Theory
di: Cui, Langyuan, et al.
Pubblicazione: (2026)
di: Cui, Langyuan, et al.
Pubblicazione: (2026)
Design Process of a Self Adaptive Smart Serious Games Ecosystem
di: Tao, X., et al.
Pubblicazione: (2025)
di: Tao, X., et al.
Pubblicazione: (2025)
Competing LLM Agents in a Non-Cooperative Game of Opinion Polarisation
di: Qasmi, Amin, et al.
Pubblicazione: (2025)
di: Qasmi, Amin, et al.
Pubblicazione: (2025)
LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning
di: Banfi, Tommaso Felice, et al.
Pubblicazione: (2026)
di: Banfi, Tommaso Felice, et al.
Pubblicazione: (2026)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
di: Turan, Berkant, et al.
Pubblicazione: (2025)
di: Turan, Berkant, et al.
Pubblicazione: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
di: Wen, Yuqiao, et al.
Pubblicazione: (2024)
Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game
di: Li, Lixing
Pubblicazione: (2026)
di: Li, Lixing
Pubblicazione: (2026)
Learning and Sustaining Shared Normative Systems via Bayesian Rule Induction in Markov Games
di: Oldenburg, Ninell, et al.
Pubblicazione: (2024)
di: Oldenburg, Ninell, et al.
Pubblicazione: (2024)
Yanyun-3: Enabling Cross-Platform Strategy Game Operation with Vision-Language Models
di: Wang, Guoyan, et al.
Pubblicazione: (2025)
di: Wang, Guoyan, et al.
Pubblicazione: (2025)
Predicting and improving test-time scaling laws via reward tail-guided search
di: Li, Muheng, et al.
Pubblicazione: (2026)
di: Li, Muheng, et al.
Pubblicazione: (2026)
End-to-end example-based sim-to-real RL policy transfer based on neural stylisation with application to robotic cutting
di: Hathaway, Jamie, et al.
Pubblicazione: (2026)
di: Hathaway, Jamie, et al.
Pubblicazione: (2026)
Rethinking VLMs for Image Forgery Detection and Localization
di: Guo, Shaofeng, et al.
Pubblicazione: (2026)
di: Guo, Shaofeng, et al.
Pubblicazione: (2026)
Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
di: Yan, Lingyong, et al.
Pubblicazione: (2026)
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning
di: Zhao, Yang, et al.
Pubblicazione: (2024)
di: Zhao, Yang, et al.
Pubblicazione: (2024)
Are We Winning the Wrong Game? Revisiting Evaluation Practices for Long-Term Time Series Forecasting
di: Phungtua-eng, Thanapol, et al.
Pubblicazione: (2026)
di: Phungtua-eng, Thanapol, et al.
Pubblicazione: (2026)
Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning
di: Li, Lin, et al.
Pubblicazione: (2026)
di: Li, Lin, et al.
Pubblicazione: (2026)
Evolutionary Data Theory: On the Similarities between Data Problems and Evolutionary Games
di: Wissgott, Philipp
Pubblicazione: (2026)
di: Wissgott, Philipp
Pubblicazione: (2026)
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
di: Ansell, Rebecca, et al.
Pubblicazione: (2026)
di: Ansell, Rebecca, et al.
Pubblicazione: (2026)
DualGFL: Federated Learning with a Dual-Level Coalition-Auction Game
di: Chen, Xiaobing, et al.
Pubblicazione: (2024)
di: Chen, Xiaobing, et al.
Pubblicazione: (2024)
Risk-Sensitive Multi-Agent Reinforcement Learning in Network Aggregative Markov Games
di: Ghaemi, Hafez, et al.
Pubblicazione: (2024)
di: Ghaemi, Hafez, et al.
Pubblicazione: (2024)
Macro-Level Correlational Analysis of Mental Disorders: Economy, Education, Society, and Technology Development
di: Tao, Yingzhi, et al.
Pubblicazione: (2025)
di: Tao, Yingzhi, et al.
Pubblicazione: (2025)
CogCanvas: Verbatim-Grounded Artifact Extraction for Long LLM Conversations
di: An, Tao
Pubblicazione: (2025)
di: An, Tao
Pubblicazione: (2025)
Documenti analoghi
-
AI Can Learn Scientific Taste
di: Tong, Jingqi, et al.
Pubblicazione: (2026) -
Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews
di: Li, Bowen, et al.
Pubblicazione: (2026) -
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
di: Hong, Yoosung
Pubblicazione: (2026) -
Procedural Game Level Design with Deep Reinforcement Learning
di: Özkan, Miraç Buğra
Pubblicazione: (2025) -
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
di: Belcamino, Valerio, et al.
Pubblicazione: (2026)