RPGBENCH: Evaluating Large Language Models as Role-Playing Game Engines
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yu, Pengfei, Shen, Dongming, Meng, Silin, Lee, Jaewon, Yin, Weisu, Cui, Andrea Yaoyun, Xu, Zhenlin, Zhu, Yi, Shi, Xingjian, Li, Mu, Smola, Alex |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Back to Basics: Revisiting ASR in the Age of Voice Agents
par: Tay, Geeyang, et autres
Publié: (2026)
par: Tay, Geeyang, et autres
Publié: (2026)
Do Language Models Have Bayesian Brains? Distinguishing Stochastic and Deterministic Decision Patterns within Large Language Models
par: Cui, Andrea Yaoyun, et autres
Publié: (2025)
par: Cui, Andrea Yaoyun, et autres
Publié: (2025)
EmergentTTS-Eval: Evaluating TTS Models on Complex Prosodic, Expressiveness, and Linguistic Challenges Using Model-as-a-Judge
par: Manku, Ruskin Raj, et autres
Publié: (2025)
par: Manku, Ruskin Raj, et autres
Publié: (2025)
L3Ms -- Lagrange Large Language Models
par: Dhillon, Guneet S., et autres
Publié: (2024)
par: Dhillon, Guneet S., et autres
Publié: (2024)
ProactBench: Beyond What The User Asked For
par: Harfi, Sepehr, et autres
Publié: (2026)
par: Harfi, Sepehr, et autres
Publié: (2026)
A Text-to-Game Engine for UGC-Based Role-Playing Games
par: Zhang, Lei, et autres
Publié: (2024)
par: Zhang, Lei, et autres
Publié: (2024)
The Pokémon Theorem and other Fairness Impossibility Results
par: Smola, Daniel Matsui, et autres
Publié: (2026)
par: Smola, Daniel Matsui, et autres
Publié: (2026)
Helmsman of the Masses? Evaluate the Opinion Leadership of Large Language Models in the Werewolf Game
par: Du, Silin, et autres
Publié: (2024)
par: Du, Silin, et autres
Publié: (2024)
Open Role-Playing with Delta-Engines
par: Wu, Hongqiu, et autres
Publié: (2024)
par: Wu, Hongqiu, et autres
Publié: (2024)
Play-Testing REMind: Evaluating an Educational Robot-Mediated Role-Play Game
par: Sanoubari, Elaheh, et autres
Publié: (2026)
par: Sanoubari, Elaheh, et autres
Publié: (2026)
Role-Playing Evaluation for Large Language Models
par: Boudouri, Yassine El, et autres
Publié: (2025)
par: Boudouri, Yassine El, et autres
Publié: (2025)
Submodular Benchmark Selection
par: Smola, Alexander
Publié: (2026)
par: Smola, Alexander
Publié: (2026)
Formen und Funktionen der Intertextualitaet im Prosawerk von Anton Čechov
par: Smola, Klavdia
Publié: (2019)
par: Smola, Klavdia
Publié: (2019)
HERO-Sign: Hierarchical Tuning and Efficient Compiler-Time GPU Optimizations for SPHINCS+ Signature Generation
par: Zhou, Yaoyun, et autres
Publié: (2025)
par: Zhou, Yaoyun, et autres
Publié: (2025)
TARQ: Tail-Aware Reconstruction Quantization for Rare-Word Robust Automatic Speech Recognition
par: Wang, Xinyu, et autres
Publié: (2026)
par: Wang, Xinyu, et autres
Publié: (2026)
Knowledge Graph-enhanced Large Language Model for Incremental Game PlayTesting
par: Mu, Enhong, et autres
Publié: (2025)
par: Mu, Enhong, et autres
Publié: (2025)
Photorealistic Robotic Simulation using Unreal Engine 5 for Agricultural Applications
par: Li, Xingjian, et autres
Publié: (2024)
par: Li, Xingjian, et autres
Publié: (2024)
Ultima and Worldbuilding in the Computer Role-Playing Game
par: Kocurek, Carly, et autres
Publié: (2024)
par: Kocurek, Carly, et autres
Publié: (2024)
Revolutionising Role-Playing Games with ChatGPT
par: Stampfl, Rita, et autres
Publié: (2024)
par: Stampfl, Rita, et autres
Publié: (2024)
Speech-DRAME: A Framework for Human-Aligned Benchmarks in Speech Role-Play
par: Shi, Jiatong, et autres
Publié: (2025)
par: Shi, Jiatong, et autres
Publié: (2025)
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems
par: Tang, Yihong, et autres
Publié: (2024)
par: Tang, Yihong, et autres
Publié: (2024)
Multimodal Chain-of-Thought Reasoning in Language Models
par: Zhang, Zhuosheng, et autres
Publié: (2023)
par: Zhang, Zhuosheng, et autres
Publié: (2023)
Playing with Voices: Tabletop Role-Playing Game Recordings as a Diarization Challenge
par: Remme, Lian, et autres
Publié: (2025)
par: Remme, Lian, et autres
Publié: (2025)
Support-Conditioned Flow Matching Is Kernel Smoothing
par: Smola, Daniel Matsui
Publié: (2026)
par: Smola, Daniel Matsui
Publié: (2026)
Acción e institución en el pensamiento político de Hannah Arendt: lecturas de Sobre la Revolución
par: Julia Gabriela Smola
Publié: (2017)
par: Julia Gabriela Smola
Publié: (2017)
Best Agent Identification for General Game Playing
par: Stephenson, Matthew, et autres
Publié: (2025)
par: Stephenson, Matthew, et autres
Publié: (2025)
Role-Playing Simulation Games using ChatGPT
par: Stampfl, Rita, et autres
Publié: (2024)
par: Stampfl, Rita, et autres
Publié: (2024)
Bimodal Oxidation Electrochemiluminescence Mechanism of Coreactant‐Embedded Covalent Organic Frameworks via Postsynthetic Modification
par: Xiaoxiao Meng, et autres
Publié: (2024)
par: Xiaoxiao Meng, et autres
Publié: (2024)
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models
par: Tao, Meiling, et autres
Publié: (2023)
par: Tao, Meiling, et autres
Publié: (2023)
Playing Large Games with Oracles and AI Debate
par: Chen, Xinyi, et autres
Publié: (2023)
par: Chen, Xinyi, et autres
Publié: (2023)
AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models
par: Li, Wenyu, et autres
Publié: (2025)
par: Li, Wenyu, et autres
Publié: (2025)
Instruction-Driven Game Engines on Large Language Models
par: Wu, Hongqiu, et autres
Publié: (2024)
par: Wu, Hongqiu, et autres
Publié: (2024)
Position: Interactive Generative Video as Next-Generation Game Engine
par: Yu, Jiwen, et autres
Publié: (2025)
par: Yu, Jiwen, et autres
Publié: (2025)
On the Decision-Making Abilities in Role-Playing using Large Language Models
par: Shen, Chenglei, et autres
Publié: (2024)
par: Shen, Chenglei, et autres
Publié: (2024)
LARP: Language-Agent Role Play for Open-World Games
par: Yan, Ming, et autres
Publié: (2023)
par: Yan, Ming, et autres
Publié: (2023)
Hybrid Voting-Based Task Assignment in Role-Playing Games
par: Weiner, Daniel, et autres
Publié: (2025)
par: Weiner, Daniel, et autres
Publié: (2025)
Smooth-Foley: Creating Continuous Sound for Video-to-Audio Generation Under Semantic Guidance
par: Zhang, Yaoyun, et autres
Publié: (2024)
par: Zhang, Yaoyun, et autres
Publié: (2024)
Nanoparticles Enable Twin‐Roll Casting of Ductile Al–Mg–Si Alloys with High Fe Content
par: Bao Wang, et autres
Publié: (2025)
par: Bao Wang, et autres
Publié: (2025)
Towards Position-Robust Talent Recommendation via Large Language Models
par: Du, Silin, et autres
Publié: (2026)
par: Du, Silin, et autres
Publié: (2026)
PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models
par: Shi, Wenlong, et autres
Publié: (2026)
par: Shi, Wenlong, et autres
Publié: (2026)
Documents similaires
-
Back to Basics: Revisiting ASR in the Age of Voice Agents
par: Tay, Geeyang, et autres
Publié: (2026) -
Do Language Models Have Bayesian Brains? Distinguishing Stochastic and Deterministic Decision Patterns within Large Language Models
par: Cui, Andrea Yaoyun, et autres
Publié: (2025) -
EmergentTTS-Eval: Evaluating TTS Models on Complex Prosodic, Expressiveness, and Linguistic Challenges Using Model-as-a-Judge
par: Manku, Ruskin Raj, et autres
Publié: (2025) -
L3Ms -- Lagrange Large Language Models
par: Dhillon, Guneet S., et autres
Publié: (2024) -
ProactBench: Beyond What The User Asked For
par: Harfi, Sepehr, et autres
Publié: (2026)