Quantal Response Equilibrium as a Measure of Strategic Sophistication: Theory and Validation for LLM Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Pechon-Elkins, Mateo, Chun, Jon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Framework for Graph-Conditioned Hierarchical Shapley Attribution in Patent Valuation
di: Bose, Joy
Pubblicazione: (2026)
di: Bose, Joy
Pubblicazione: (2026)
Efficient Method for Finding Optimal Strategies in Chopstick Auctions with Uniform Objects Values
di: Kaźmierowski, Stanisław, et al.
Pubblicazione: (2024)
di: Kaźmierowski, Stanisław, et al.
Pubblicazione: (2024)
Coalition Formation in LLM Agent Networks: Stability Analysis and Convergence Guarantees
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Can LLMs Assess Personality? Validating Conversational AI for Trait Profiling
di: Matšenas, Andrius, et al.
Pubblicazione: (2026)
di: Matšenas, Andrius, et al.
Pubblicazione: (2026)
When would online platforms pay data dividends
di: Kudva, Sukanya, et al.
Pubblicazione: (2022)
di: Kudva, Sukanya, et al.
Pubblicazione: (2022)
Playing the Player: A Heuristic Framework for Adaptive Poker AI
di: Paterson, Andrew, et al.
Pubblicazione: (2025)
di: Paterson, Andrew, et al.
Pubblicazione: (2025)
Formal Verification of Diffusion Auctions
di: Galimullin, Rustam, et al.
Pubblicazione: (2025)
di: Galimullin, Rustam, et al.
Pubblicazione: (2025)
ADIOS: Antibody Development via Opponent Shaping
di: Towers, Sebastian, et al.
Pubblicazione: (2024)
di: Towers, Sebastian, et al.
Pubblicazione: (2024)
Privacy as Commodity: MFG-RegretNet for Large-Scale Privacy Trading in Federated Learning
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
Evaluating Large Language Models in a Complex Hidden Role Game
di: Bauer, Niklas
Pubblicazione: (2026)
di: Bauer, Niklas
Pubblicazione: (2026)
Automated Market Makers for Decentralized Finance (DeFi)
di: Wang, Yongge
Pubblicazione: (2020)
di: Wang, Yongge
Pubblicazione: (2020)
Matching Markets Meet LLMs: Algorithmic Reasoning with Ranked Preferences
di: Hosseini, Hadi, et al.
Pubblicazione: (2025)
di: Hosseini, Hadi, et al.
Pubblicazione: (2025)
VGC-Bench: Towards Mastering Diverse Team Strategies in Competitive Pokémon
di: Angliss, Cameron, et al.
Pubblicazione: (2025)
di: Angliss, Cameron, et al.
Pubblicazione: (2025)
Hierarchical Battery-Aware Game Algorithm for ISL Power Allocation in LEO Mega-Constellations
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
Heterogeneous Mean Field Game Framework for LEO Satellite-Assisted V2X Networks
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
di: Sun, Kangkang, et al.
Pubblicazione: (2026)
The Learning Approach to Games
di: İşeri, Melih, et al.
Pubblicazione: (2025)
di: İşeri, Melih, et al.
Pubblicazione: (2025)
What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control
di: Lekeas, Paraskevas V., et al.
Pubblicazione: (2026)
di: Lekeas, Paraskevas V., et al.
Pubblicazione: (2026)
Game Intelligence: Theory and Computation
di: Seven, Mehmet Mars
Pubblicazione: (2023)
di: Seven, Mehmet Mars
Pubblicazione: (2023)
Generalized Quantal Response Equilibrium: Existence and Efficient Learning
di: Shukla, Apurv, et al.
Pubblicazione: (2025)
di: Shukla, Apurv, et al.
Pubblicazione: (2025)
Common $p$-Belief with Plausibility Measures: Extended Abstract
di: Pacuit, Eric, et al.
Pubblicazione: (2025)
di: Pacuit, Eric, et al.
Pubblicazione: (2025)
Learning to Maximize Gains From Trade in Small Markets
di: Babaioff, Moshe, et al.
Pubblicazione: (2024)
di: Babaioff, Moshe, et al.
Pubblicazione: (2024)
AI Agents for the Dhumbal Card Game: A Comparative Study
di: Malla, Sahaj Raj
Pubblicazione: (2025)
di: Malla, Sahaj Raj
Pubblicazione: (2025)
Beyond the Sum: Unlocking AI Agents Potential Through Market Forces
di: Sanabria, Jordi Montes, et al.
Pubblicazione: (2024)
di: Sanabria, Jordi Montes, et al.
Pubblicazione: (2024)
LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning
di: Banfi, Tommaso Felice, et al.
Pubblicazione: (2026)
di: Banfi, Tommaso Felice, et al.
Pubblicazione: (2026)
MenuNet: A Strategy-Proof Mechanism for Matching Markets
di: Sun, Zhaohong, et al.
Pubblicazione: (2026)
di: Sun, Zhaohong, et al.
Pubblicazione: (2026)
How Many Votes is a Lie Worth? Measuring Strategyproofness through Resource Augmentation
di: Berker, Ratip Emin, et al.
Pubblicazione: (2026)
di: Berker, Ratip Emin, et al.
Pubblicazione: (2026)
The Illusion of Collusion
di: Douglas, Connor, et al.
Pubblicazione: (2024)
di: Douglas, Connor, et al.
Pubblicazione: (2024)
A High-Performance External Validity Index for Clustering with a Large Number of Clusters
di: Karbasian, Mohammad Yasin, et al.
Pubblicazione: (2024)
di: Karbasian, Mohammad Yasin, et al.
Pubblicazione: (2024)
Zero Carbon V2X Tariffs for Non-Domestic Customers
di: Shamash, Elisheva S, et al.
Pubblicazione: (2025)
di: Shamash, Elisheva S, et al.
Pubblicazione: (2025)
An Equilibrium Analysis of the Arad-Rubinstein Game
di: Ewerhart, Christian, et al.
Pubblicazione: (2024)
di: Ewerhart, Christian, et al.
Pubblicazione: (2024)
Selling Privacy in Blockchain Transactions
di: Chionas, Georgios, et al.
Pubblicazione: (2025)
di: Chionas, Georgios, et al.
Pubblicazione: (2025)
Decision Aggregation under Quantal Response
di: Huang, Zhihuan, et al.
Pubblicazione: (2026)
di: Huang, Zhihuan, et al.
Pubblicazione: (2026)
Discovering Differences in Strategic Behavior Between Humans and LLMs
di: Wang, Caroline, et al.
Pubblicazione: (2026)
di: Wang, Caroline, et al.
Pubblicazione: (2026)
Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech
di: von Cossel, Oskar
Pubblicazione: (2026)
di: von Cossel, Oskar
Pubblicazione: (2026)
Multi-agent reinforcement learning in the all-or-nothing public goods game on networks
di: Meylahn, Benedikt Valentin
Pubblicazione: (2024)
di: Meylahn, Benedikt Valentin
Pubblicazione: (2024)
Can LLMs Identify Tax Abuse?
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
LLMs Simulate Big Five Personality Traits: Further Evidence
di: Sorokovikova, Aleksandra, et al.
Pubblicazione: (2024)
di: Sorokovikova, Aleksandra, et al.
Pubblicazione: (2024)
Scaling In, Not Up? Testing Thick Citation Context Analysis with GPT-5 and Fragile Prompts
di: Simons, Arno
Pubblicazione: (2026)
di: Simons, Arno
Pubblicazione: (2026)
Decision Traces: What Multi-System Data Fusion Reveals About Institutional Knowledge in Enterprise Hiring
di: Shafiq, Saad Bin
Pubblicazione: (2026)
di: Shafiq, Saad Bin
Pubblicazione: (2026)
Automatic Item Generation for Personality Situational Judgment Tests with Large Language Models
di: Li, Chang-Jin, et al.
Pubblicazione: (2024)
di: Li, Chang-Jin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Framework for Graph-Conditioned Hierarchical Shapley Attribution in Patent Valuation
di: Bose, Joy
Pubblicazione: (2026) -
Efficient Method for Finding Optimal Strategies in Chopstick Auctions with Uniform Objects Values
di: Kaźmierowski, Stanisław, et al.
Pubblicazione: (2024) -
Coalition Formation in LLM Agent Networks: Stability Analysis and Convergence Guarantees
di: Guo, Dongxin, et al.
Pubblicazione: (2026) -
Can LLMs Assess Personality? Validating Conversational AI for Trait Profiling
di: Matšenas, Andrius, et al.
Pubblicazione: (2026) -
When would online platforms pay data dividends
di: Kudva, Sukanya, et al.
Pubblicazione: (2022)