LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Xiao, Chang, Yang, Brenda Z. |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
par: Zhang, Chen, et autres
Publié: (2024)
par: Zhang, Chen, et autres
Publié: (2024)
AI Agents for Inventory Control: Human-LLM-OR Complementarity
par: Baek, Jackie, et autres
Publié: (2026)
par: Baek, Jackie, et autres
Publié: (2026)
When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems
par: Cheng, Z., et autres
Publié: (2026)
par: Cheng, Z., et autres
Publié: (2026)
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
par: Justus, Aju Ani, et autres
Publié: (2025)
par: Justus, Aju Ani, et autres
Publié: (2025)
Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
par: Xiao, Chang, et autres
Publié: (2025)
par: Xiao, Chang, et autres
Publié: (2025)
Strength Estimation and Human-Like Strength Adjustment in Games
par: Chen, Chun Jung, et autres
Publié: (2025)
par: Chen, Chun Jung, et autres
Publié: (2025)
Can Interpretability Layouts Influence Human Perception of Offensive Sentences?
par: Santos, Thiago Freitas dos, et autres
Publié: (2024)
par: Santos, Thiago Freitas dos, et autres
Publié: (2024)
Label-Free Subjective Player Experience Modelling via Let's Play Videos
par: Goel, Dave, et autres
Publié: (2024)
par: Goel, Dave, et autres
Publié: (2024)
Detection of adversarial intent in Human-AI teams using LLMs
par: Musaffar, Abed K., et autres
Publié: (2026)
par: Musaffar, Abed K., et autres
Publié: (2026)
Predictive AI Can Support Human Learning while Preserving Error Diversity
par: He, Vivianna Fang, et autres
Publié: (2025)
par: He, Vivianna Fang, et autres
Publié: (2025)
Exploring the Panorama of Anxiety Levels: A Multi-Scenario Study Based on Human-Centric Anxiety Level Detection and Personalized Guidance
par: Xian, Longdi, et autres
Publié: (2025)
par: Xian, Longdi, et autres
Publié: (2025)
Robo-Saber: Generating and Simulating Virtual Reality Players
par: Kim, Nam Hee, et autres
Publié: (2026)
par: Kim, Nam Hee, et autres
Publié: (2026)
Can we use LLMs to bootstrap reinforcement learning? -- A case study in digital health behavior change
par: Albers, Nele, et autres
Publié: (2025)
par: Albers, Nele, et autres
Publié: (2025)
ADEPTS: A Capability Framework for Human-Centered Agent Design
par: D'Oro, Pierluca, et autres
Publié: (2025)
par: D'Oro, Pierluca, et autres
Publié: (2025)
Toward Human-AI Alignment in Large-Scale Multi-Player Games
par: Sharma, Sugandha, et autres
Publié: (2024)
par: Sharma, Sugandha, et autres
Publié: (2024)
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
par: Atreya, Alankar, et autres
Publié: (2026)
par: Atreya, Alankar, et autres
Publié: (2026)
Personas Evolved: Designing Ethical LLM-Based Conversational Agent Personalities
par: Desai, Smit, et autres
Publié: (2025)
par: Desai, Smit, et autres
Publié: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals
par: Park, Joon Sung, et autres
Publié: (2024)
par: Park, Joon Sung, et autres
Publié: (2024)
Why and When LLM-Based Assistants Can Go Wrong: Investigating the Effectiveness of Prompt-Based Interactions for Software Help-Seeking
par: Khurana, Anjali, et autres
Publié: (2024)
par: Khurana, Anjali, et autres
Publié: (2024)
The Odyssey of the Fittest: Can Agents Survive and Still Be Good?
par: Waldner, Dylan, et autres
Publié: (2025)
par: Waldner, Dylan, et autres
Publié: (2025)
Agent Laboratory: Using LLM Agents as Research Assistants
par: Schmidgall, Samuel, et autres
Publié: (2025)
par: Schmidgall, Samuel, et autres
Publié: (2025)
Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
par: Zhang, Yizhe, et autres
Publié: (2023)
par: Zhang, Yizhe, et autres
Publié: (2023)
Hierarchical Multi-Armed Bandits for the Concurrent Intelligent Tutoring of Concepts and Problems of Varying Difficulty Levels
par: Castleman, Blake, et autres
Publié: (2024)
par: Castleman, Blake, et autres
Publié: (2024)
To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
Exploring Emotions in Multi-componential Space using Interactive VR Games
par: Somarathna, Rukshani, et autres
Publié: (2024)
par: Somarathna, Rukshani, et autres
Publié: (2024)
Future Research Avenues for Artificial Intelligence in Digital Gaming: An Exploratory Report
par: Dablander, Markus
Publié: (2024)
par: Dablander, Markus
Publié: (2024)
Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?
par: Kim, Jane Paik
Publié: (2026)
par: Kim, Jane Paik
Publié: (2026)
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
par: Zhou, Xinglin, et autres
Publié: (2024)
par: Zhou, Xinglin, et autres
Publié: (2024)
Off-Policy Selection for Initiating Human-Centric Experimental Design
par: Gao, Ge, et autres
Publié: (2024)
par: Gao, Ge, et autres
Publié: (2024)
Efficient Human-in-the-Loop Active Learning: A Novel Framework for Data Labeling in AI Systems
par: Huang, Yiran, et autres
Publié: (2024)
par: Huang, Yiran, et autres
Publié: (2024)
ClickAgent: Enhancing UI Location Capabilities of Autonomous Agents
par: Hoscilowicz, Jakub, et autres
Publié: (2024)
par: Hoscilowicz, Jakub, et autres
Publié: (2024)
What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
par: Vafa, Keyon, et autres
Publié: (2025)
par: Vafa, Keyon, et autres
Publié: (2025)
Enabling On-Device LLMs Personalization with Smartphone Sensing
par: Zhang, Shiquan, et autres
Publié: (2024)
par: Zhang, Shiquan, et autres
Publié: (2024)
Designing for Human-Agent Alignment: Understanding what humans want from their agents
par: Goyal, Nitesh, et autres
Publié: (2024)
par: Goyal, Nitesh, et autres
Publié: (2024)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
par: Yu, Guanghui, et autres
Publié: (2024)
par: Yu, Guanghui, et autres
Publié: (2024)
Human-In-the-Loop Software Development Agents
par: Takerngsaksiri, Wannita, et autres
Publié: (2024)
par: Takerngsaksiri, Wannita, et autres
Publié: (2024)
Harmonic LLMs are Trustworthy
par: Kersting, Nicholas S., et autres
Publié: (2024)
par: Kersting, Nicholas S., et autres
Publié: (2024)
Domain-Grounded Evaluation of LLMs in International Student Knowledge
par: Daitx, Claudinei, et autres
Publié: (2025)
par: Daitx, Claudinei, et autres
Publié: (2025)
Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025
par: Thakkar, Nitya, et autres
Publié: (2025)
par: Thakkar, Nitya, et autres
Publié: (2025)
Documents similaires
-
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
par: Zhang, Chen, et autres
Publié: (2024) -
AI Agents for Inventory Control: Human-LLM-OR Complementarity
par: Baek, Jackie, et autres
Publié: (2026) -
When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems
par: Cheng, Z., et autres
Publié: (2026) -
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
par: Justus, Aju Ani, et autres
Publié: (2025) -
Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
par: Xiao, Chang, et autres
Publié: (2025)