LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Chang, Yang, Brenda Z. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
por: Zhang, Chen, et al.
Publicado: (2024)
por: Zhang, Chen, et al.
Publicado: (2024)
AI Agents for Inventory Control: Human-LLM-OR Complementarity
por: Baek, Jackie, et al.
Publicado: (2026)
por: Baek, Jackie, et al.
Publicado: (2026)
When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems
por: Cheng, Z., et al.
Publicado: (2026)
por: Cheng, Z., et al.
Publicado: (2026)
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
por: Justus, Aju Ani, et al.
Publicado: (2025)
por: Justus, Aju Ani, et al.
Publicado: (2025)
Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
por: Xiao, Chang, et al.
Publicado: (2025)
por: Xiao, Chang, et al.
Publicado: (2025)
Strength Estimation and Human-Like Strength Adjustment in Games
por: Chen, Chun Jung, et al.
Publicado: (2025)
por: Chen, Chun Jung, et al.
Publicado: (2025)
Can Interpretability Layouts Influence Human Perception of Offensive Sentences?
por: Santos, Thiago Freitas dos, et al.
Publicado: (2024)
por: Santos, Thiago Freitas dos, et al.
Publicado: (2024)
Label-Free Subjective Player Experience Modelling via Let's Play Videos
por: Goel, Dave, et al.
Publicado: (2024)
por: Goel, Dave, et al.
Publicado: (2024)
Detection of adversarial intent in Human-AI teams using LLMs
por: Musaffar, Abed K., et al.
Publicado: (2026)
por: Musaffar, Abed K., et al.
Publicado: (2026)
Predictive AI Can Support Human Learning while Preserving Error Diversity
por: He, Vivianna Fang, et al.
Publicado: (2025)
por: He, Vivianna Fang, et al.
Publicado: (2025)
Exploring the Panorama of Anxiety Levels: A Multi-Scenario Study Based on Human-Centric Anxiety Level Detection and Personalized Guidance
por: Xian, Longdi, et al.
Publicado: (2025)
por: Xian, Longdi, et al.
Publicado: (2025)
Robo-Saber: Generating and Simulating Virtual Reality Players
por: Kim, Nam Hee, et al.
Publicado: (2026)
por: Kim, Nam Hee, et al.
Publicado: (2026)
Can we use LLMs to bootstrap reinforcement learning? -- A case study in digital health behavior change
por: Albers, Nele, et al.
Publicado: (2025)
por: Albers, Nele, et al.
Publicado: (2025)
ADEPTS: A Capability Framework for Human-Centered Agent Design
por: D'Oro, Pierluca, et al.
Publicado: (2025)
por: D'Oro, Pierluca, et al.
Publicado: (2025)
Toward Human-AI Alignment in Large-Scale Multi-Player Games
por: Sharma, Sugandha, et al.
Publicado: (2024)
por: Sharma, Sugandha, et al.
Publicado: (2024)
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
por: Atreya, Alankar, et al.
Publicado: (2026)
por: Atreya, Alankar, et al.
Publicado: (2026)
Personas Evolved: Designing Ethical LLM-Based Conversational Agent Personalities
por: Desai, Smit, et al.
Publicado: (2025)
por: Desai, Smit, et al.
Publicado: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
por: Si, Chenglei, et al.
Publicado: (2024)
por: Si, Chenglei, et al.
Publicado: (2024)
LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals
por: Park, Joon Sung, et al.
Publicado: (2024)
por: Park, Joon Sung, et al.
Publicado: (2024)
Why and When LLM-Based Assistants Can Go Wrong: Investigating the Effectiveness of Prompt-Based Interactions for Software Help-Seeking
por: Khurana, Anjali, et al.
Publicado: (2024)
por: Khurana, Anjali, et al.
Publicado: (2024)
The Odyssey of the Fittest: Can Agents Survive and Still Be Good?
por: Waldner, Dylan, et al.
Publicado: (2025)
por: Waldner, Dylan, et al.
Publicado: (2025)
Agent Laboratory: Using LLM Agents as Research Assistants
por: Schmidgall, Samuel, et al.
Publicado: (2025)
por: Schmidgall, Samuel, et al.
Publicado: (2025)
Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
por: Zhang, Yizhe, et al.
Publicado: (2023)
por: Zhang, Yizhe, et al.
Publicado: (2023)
Hierarchical Multi-Armed Bandits for the Concurrent Intelligent Tutoring of Concepts and Problems of Varying Difficulty Levels
por: Castleman, Blake, et al.
Publicado: (2024)
por: Castleman, Blake, et al.
Publicado: (2024)
To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
por: Vishwarupe, Varad, et al.
Publicado: (2026)
por: Vishwarupe, Varad, et al.
Publicado: (2026)
Exploring Emotions in Multi-componential Space using Interactive VR Games
por: Somarathna, Rukshani, et al.
Publicado: (2024)
por: Somarathna, Rukshani, et al.
Publicado: (2024)
Future Research Avenues for Artificial Intelligence in Digital Gaming: An Exploratory Report
por: Dablander, Markus
Publicado: (2024)
por: Dablander, Markus
Publicado: (2024)
Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?
por: Kim, Jane Paik
Publicado: (2026)
por: Kim, Jane Paik
Publicado: (2026)
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
por: Zhou, Xinglin, et al.
Publicado: (2024)
por: Zhou, Xinglin, et al.
Publicado: (2024)
Off-Policy Selection for Initiating Human-Centric Experimental Design
por: Gao, Ge, et al.
Publicado: (2024)
por: Gao, Ge, et al.
Publicado: (2024)
Efficient Human-in-the-Loop Active Learning: A Novel Framework for Data Labeling in AI Systems
por: Huang, Yiran, et al.
Publicado: (2024)
por: Huang, Yiran, et al.
Publicado: (2024)
ClickAgent: Enhancing UI Location Capabilities of Autonomous Agents
por: Hoscilowicz, Jakub, et al.
Publicado: (2024)
por: Hoscilowicz, Jakub, et al.
Publicado: (2024)
What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
por: Vafa, Keyon, et al.
Publicado: (2025)
por: Vafa, Keyon, et al.
Publicado: (2025)
Enabling On-Device LLMs Personalization with Smartphone Sensing
por: Zhang, Shiquan, et al.
Publicado: (2024)
por: Zhang, Shiquan, et al.
Publicado: (2024)
Designing for Human-Agent Alignment: Understanding what humans want from their agents
por: Goyal, Nitesh, et al.
Publicado: (2024)
por: Goyal, Nitesh, et al.
Publicado: (2024)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
por: Yu, Guanghui, et al.
Publicado: (2024)
por: Yu, Guanghui, et al.
Publicado: (2024)
Human-In-the-Loop Software Development Agents
por: Takerngsaksiri, Wannita, et al.
Publicado: (2024)
por: Takerngsaksiri, Wannita, et al.
Publicado: (2024)
Harmonic LLMs are Trustworthy
por: Kersting, Nicholas S., et al.
Publicado: (2024)
por: Kersting, Nicholas S., et al.
Publicado: (2024)
Domain-Grounded Evaluation of LLMs in International Student Knowledge
por: Daitx, Claudinei, et al.
Publicado: (2025)
por: Daitx, Claudinei, et al.
Publicado: (2025)
Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025
por: Thakkar, Nitya, et al.
Publicado: (2025)
por: Thakkar, Nitya, et al.
Publicado: (2025)
Ejemplares similares
-
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
por: Zhang, Chen, et al.
Publicado: (2024) -
AI Agents for Inventory Control: Human-LLM-OR Complementarity
por: Baek, Jackie, et al.
Publicado: (2026) -
When the Loop Closes: Architectural Limits of In-Context Isolation, Metacognitive Co-option, and the Two-Target Design Problem in Human-LLM Systems
por: Cheng, Z., et al.
Publicado: (2026) -
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
por: Justus, Aju Ani, et al.
Publicado: (2025) -
Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
por: Xiao, Chang, et al.
Publicado: (2025)