Quantal Response Equilibrium as a Measure of Strategic Sophistication: Theory and Validation for LLM Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Pechon-Elkins, Mateo, Chun, Jon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Framework for Graph-Conditioned Hierarchical Shapley Attribution in Patent Valuation
by: Bose, Joy
Published: (2026)
by: Bose, Joy
Published: (2026)
Efficient Method for Finding Optimal Strategies in Chopstick Auctions with Uniform Objects Values
by: Kaźmierowski, Stanisław, et al.
Published: (2024)
by: Kaźmierowski, Stanisław, et al.
Published: (2024)
Coalition Formation in LLM Agent Networks: Stability Analysis and Convergence Guarantees
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
Can LLMs Assess Personality? Validating Conversational AI for Trait Profiling
by: Matšenas, Andrius, et al.
Published: (2026)
by: Matšenas, Andrius, et al.
Published: (2026)
When would online platforms pay data dividends
by: Kudva, Sukanya, et al.
Published: (2022)
by: Kudva, Sukanya, et al.
Published: (2022)
Playing the Player: A Heuristic Framework for Adaptive Poker AI
by: Paterson, Andrew, et al.
Published: (2025)
by: Paterson, Andrew, et al.
Published: (2025)
Formal Verification of Diffusion Auctions
by: Galimullin, Rustam, et al.
Published: (2025)
by: Galimullin, Rustam, et al.
Published: (2025)
ADIOS: Antibody Development via Opponent Shaping
by: Towers, Sebastian, et al.
Published: (2024)
by: Towers, Sebastian, et al.
Published: (2024)
Privacy as Commodity: MFG-RegretNet for Large-Scale Privacy Trading in Federated Learning
by: Sun, Kangkang, et al.
Published: (2026)
by: Sun, Kangkang, et al.
Published: (2026)
Evaluating Large Language Models in a Complex Hidden Role Game
by: Bauer, Niklas
Published: (2026)
by: Bauer, Niklas
Published: (2026)
Automated Market Makers for Decentralized Finance (DeFi)
by: Wang, Yongge
Published: (2020)
by: Wang, Yongge
Published: (2020)
Matching Markets Meet LLMs: Algorithmic Reasoning with Ranked Preferences
by: Hosseini, Hadi, et al.
Published: (2025)
by: Hosseini, Hadi, et al.
Published: (2025)
VGC-Bench: Towards Mastering Diverse Team Strategies in Competitive Pokémon
by: Angliss, Cameron, et al.
Published: (2025)
by: Angliss, Cameron, et al.
Published: (2025)
Hierarchical Battery-Aware Game Algorithm for ISL Power Allocation in LEO Mega-Constellations
by: Sun, Kangkang, et al.
Published: (2026)
by: Sun, Kangkang, et al.
Published: (2026)
Heterogeneous Mean Field Game Framework for LEO Satellite-Assisted V2X Networks
by: Sun, Kangkang, et al.
Published: (2026)
by: Sun, Kangkang, et al.
Published: (2026)
The Learning Approach to Games
by: İşeri, Melih, et al.
Published: (2025)
by: İşeri, Melih, et al.
Published: (2025)
What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control
by: Lekeas, Paraskevas V., et al.
Published: (2026)
by: Lekeas, Paraskevas V., et al.
Published: (2026)
Game Intelligence: Theory and Computation
by: Seven, Mehmet Mars
Published: (2023)
by: Seven, Mehmet Mars
Published: (2023)
Generalized Quantal Response Equilibrium: Existence and Efficient Learning
by: Shukla, Apurv, et al.
Published: (2025)
by: Shukla, Apurv, et al.
Published: (2025)
Common $p$-Belief with Plausibility Measures: Extended Abstract
by: Pacuit, Eric, et al.
Published: (2025)
by: Pacuit, Eric, et al.
Published: (2025)
Learning to Maximize Gains From Trade in Small Markets
by: Babaioff, Moshe, et al.
Published: (2024)
by: Babaioff, Moshe, et al.
Published: (2024)
AI Agents for the Dhumbal Card Game: A Comparative Study
by: Malla, Sahaj Raj
Published: (2025)
by: Malla, Sahaj Raj
Published: (2025)
Beyond the Sum: Unlocking AI Agents Potential Through Market Forces
by: Sanabria, Jordi Montes, et al.
Published: (2024)
by: Sanabria, Jordi Montes, et al.
Published: (2024)
LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning
by: Banfi, Tommaso Felice, et al.
Published: (2026)
by: Banfi, Tommaso Felice, et al.
Published: (2026)
MenuNet: A Strategy-Proof Mechanism for Matching Markets
by: Sun, Zhaohong, et al.
Published: (2026)
by: Sun, Zhaohong, et al.
Published: (2026)
How Many Votes is a Lie Worth? Measuring Strategyproofness through Resource Augmentation
by: Berker, Ratip Emin, et al.
Published: (2026)
by: Berker, Ratip Emin, et al.
Published: (2026)
The Illusion of Collusion
by: Douglas, Connor, et al.
Published: (2024)
by: Douglas, Connor, et al.
Published: (2024)
A High-Performance External Validity Index for Clustering with a Large Number of Clusters
by: Karbasian, Mohammad Yasin, et al.
Published: (2024)
by: Karbasian, Mohammad Yasin, et al.
Published: (2024)
Zero Carbon V2X Tariffs for Non-Domestic Customers
by: Shamash, Elisheva S, et al.
Published: (2025)
by: Shamash, Elisheva S, et al.
Published: (2025)
An Equilibrium Analysis of the Arad-Rubinstein Game
by: Ewerhart, Christian, et al.
Published: (2024)
by: Ewerhart, Christian, et al.
Published: (2024)
Selling Privacy in Blockchain Transactions
by: Chionas, Georgios, et al.
Published: (2025)
by: Chionas, Georgios, et al.
Published: (2025)
Decision Aggregation under Quantal Response
by: Huang, Zhihuan, et al.
Published: (2026)
by: Huang, Zhihuan, et al.
Published: (2026)
Discovering Differences in Strategic Behavior Between Humans and LLMs
by: Wang, Caroline, et al.
Published: (2026)
by: Wang, Caroline, et al.
Published: (2026)
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
by: von Cossel, Oskar
Published: (2026)
by: von Cossel, Oskar
Published: (2026)
Multi-agent reinforcement learning in the all-or-nothing public goods game on networks
by: Meylahn, Benedikt Valentin
Published: (2024)
by: Meylahn, Benedikt Valentin
Published: (2024)
Can LLMs Identify Tax Abuse?
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
LLMs Simulate Big Five Personality Traits: Further Evidence
by: Sorokovikova, Aleksandra, et al.
Published: (2024)
by: Sorokovikova, Aleksandra, et al.
Published: (2024)
Scaling In, Not Up? Testing Thick Citation Context Analysis with GPT-5 and Fragile Prompts
by: Simons, Arno
Published: (2026)
by: Simons, Arno
Published: (2026)
Decision Traces: What Multi-System Data Fusion Reveals About Institutional Knowledge in Enterprise Hiring
by: Shafiq, Saad Bin
Published: (2026)
by: Shafiq, Saad Bin
Published: (2026)
Automatic Item Generation for Personality Situational Judgment Tests with Large Language Models
by: Li, Chang-Jin, et al.
Published: (2024)
by: Li, Chang-Jin, et al.
Published: (2024)
Similar Items
-
A Framework for Graph-Conditioned Hierarchical Shapley Attribution in Patent Valuation
by: Bose, Joy
Published: (2026) -
Efficient Method for Finding Optimal Strategies in Chopstick Auctions with Uniform Objects Values
by: Kaźmierowski, Stanisław, et al.
Published: (2024) -
Coalition Formation in LLM Agent Networks: Stability Analysis and Convergence Guarantees
by: Guo, Dongxin, et al.
Published: (2026) -
Can LLMs Assess Personality? Validating Conversational AI for Trait Profiling
by: Matšenas, Andrius, et al.
Published: (2026) -
When would online platforms pay data dividends
by: Kudva, Sukanya, et al.
Published: (2022)