Who's Gaming the System? A Causally-Motivated Approach for Detecting Strategic Adaptation
Fuente:
arXiv
Saved in:
| Main Authors: | Chang, Trenton, Warrenburg, Lindsay, Park, Sae-Hwan, Parikh, Ravi B., Makar, Maggie, Wiens, Jenna |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective
by: Zapzalka, Dylan, et al.
Published: (2025)
by: Zapzalka, Dylan, et al.
Published: (2025)
From Biased Selective Labels to Pseudo-Labels: An Expectation-Maximization Framework for Learning from Biased Decisions
by: Chang, Trenton, et al.
Published: (2024)
by: Chang, Trenton, et al.
Published: (2024)
Conditional Front-door Adjustment for Heterogeneous Treatment Assignment Effect Estimation Under Non-adherence
by: Chen, Winston, et al.
Published: (2025)
by: Chen, Winston, et al.
Published: (2025)
Measuring Model Performance in the Presence of an Intervention
by: Chen, Winston, et al.
Published: (2025)
by: Chen, Winston, et al.
Published: (2025)
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
by: Chang, Trenton, et al.
Published: (2025)
by: Chang, Trenton, et al.
Published: (2025)
Improvisational Games as a Benchmark for Social Intelligence of AI Agents: The Case of Connections
by: Parikh, Gaurav Rajesh, et al.
Published: (2026)
by: Parikh, Gaurav Rajesh, et al.
Published: (2026)
Explaining Decisions of Agents in Mixed-Motive Games
by: Orner, Maayan, et al.
Published: (2024)
by: Orner, Maayan, et al.
Published: (2024)
Causal Strategic Learning with Competitive Selection
by: Vo, Kiet Q. H., et al.
Published: (2023)
by: Vo, Kiet Q. H., et al.
Published: (2023)
Who Does Your Algorithm Fail? Investigating Age and Ethnic Bias in the MAMA-MIA Dataset
by: Parikh, Aditya, et al.
Published: (2025)
by: Parikh, Aditya, et al.
Published: (2025)
DEPICT: Diffusion-Enabled Permutation Importance for Image Classification Tasks
by: Jabbour, Sarah, et al.
Published: (2024)
by: Jabbour, Sarah, et al.
Published: (2024)
Detecting Model Drifts in Non-Stationary Environment Using Edit Operation Measures
by: Lee, Chang-Hwan, et al.
Published: (2025)
by: Lee, Chang-Hwan, et al.
Published: (2025)
Adaptive Punishment for Cooperation in Mixed-Motive Games
by: Tang, Min, et al.
Published: (2026)
by: Tang, Min, et al.
Published: (2026)
Using Large Language Models to Categorize Strategic Situations and Decipher Motivations Behind Human Behaviors
by: Xie, Yutong, et al.
Published: (2025)
by: Xie, Yutong, et al.
Published: (2025)
Cross-Architecture Model Diffing with Crosscoders: Unsupervised Discovery of Differences Between LLMs
by: Jiralerspong, Thomas, et al.
Published: (2026)
by: Jiralerspong, Thomas, et al.
Published: (2026)
Toward Reasoning on the Boundary: A Mixup-based Approach for Graph Anomaly Detection
by: Kim, Hwan, et al.
Published: (2024)
by: Kim, Hwan, et al.
Published: (2024)
Hypothesis Testing the Circuit Hypothesis in LLMs
by: Shi, Claudia, et al.
Published: (2024)
by: Shi, Claudia, et al.
Published: (2024)
Learning to Balance Altruism and Self-interest Based on Empathy in Mixed-Motive Games
by: Kong, Fanqi, et al.
Published: (2024)
by: Kong, Fanqi, et al.
Published: (2024)
Automatic Question Generation for Intuitive Learning Utilizing Causal Graph Guided Chain of Thought Reasoning
by: Wang, Nicholas X., et al.
Published: (2026)
by: Wang, Nicholas X., et al.
Published: (2026)
Generating Novelty in Open-World Multi-Agent Strategic Board Games
by: Kejriwal, Mayank, et al.
Published: (2025)
by: Kejriwal, Mayank, et al.
Published: (2025)
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
by: He, Yidong, et al.
Published: (2026)
by: He, Yidong, et al.
Published: (2026)
On the Limits of Selective AI Prediction: A Case Study in Clinical Decision Making
by: Jabbour, Sarah, et al.
Published: (2025)
by: Jabbour, Sarah, et al.
Published: (2025)
Trajectory Adaptation using Large Language Models
by: Maurya, Anurag, et al.
Published: (2025)
by: Maurya, Anurag, et al.
Published: (2025)
Exploring Large Language Models for Word Games:Who is the Spy?
by: Wei, Chentian, et al.
Published: (2025)
by: Wei, Chentian, et al.
Published: (2025)
Strategic Over-Parameterization for Generalizable Low-Rank Adaptation
by: Gao, Jing, et al.
Published: (2026)
by: Gao, Jing, et al.
Published: (2026)
Multiverse: Language-Conditioned Multi-Game Level Blending via Shared Representation
by: Baek, In-Chang, et al.
Published: (2026)
by: Baek, In-Chang, et al.
Published: (2026)
Do Persona-Infused LLMs Affect Performance in a Strategic Reasoning Game?
by: Licato, John, et al.
Published: (2025)
by: Licato, John, et al.
Published: (2025)
SLOW: Strategic Logical-inference Open Workspace for Cognitive Adaptation in AI Tutoring
by: Wei, Yuang, et al.
Published: (2026)
by: Wei, Yuang, et al.
Published: (2026)
Debiasing Reward Models via Causally Motivated Inference-Time Intervention
by: Shinoda, Kazutoshi, et al.
Published: (2026)
by: Shinoda, Kazutoshi, et al.
Published: (2026)
Individual Causal Inference with Structural Causal Model
by: Chang, Daniel T.
Published: (2025)
by: Chang, Daniel T.
Published: (2025)
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents
by: Costarelli, Anthony, et al.
Published: (2024)
by: Costarelli, Anthony, et al.
Published: (2024)
Strategic Insights from Simulation Gaming of AI Race Dynamics
by: Gruetzemacher, Ross, et al.
Published: (2024)
by: Gruetzemacher, Ross, et al.
Published: (2024)
Who's Your Judge? On the Detectability of LLM-Generated Judgments
by: Li, Dawei, et al.
Published: (2025)
by: Li, Dawei, et al.
Published: (2025)
Strategic Communication under Threat: Learning Information Trade-offs in Pursuit-Evasion Games
by: La Gatta, Valerio, et al.
Published: (2025)
by: La Gatta, Valerio, et al.
Published: (2025)
Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization
by: Xu, Zelai, et al.
Published: (2025)
by: Xu, Zelai, et al.
Published: (2025)
M3-BENCH: Process-Aware Evaluation of LLM Agents' Social Behaviors in Mixed-Motive Games
by: Xie, Sixiong, et al.
Published: (2026)
by: Xie, Sixiong, et al.
Published: (2026)
Correcting Human Labels for Rater Effects in AI Evaluation: An Item Response Theory Approach
by: Casabianca, Jodi M., et al.
Published: (2026)
by: Casabianca, Jodi M., et al.
Published: (2026)
Multi-Agent Strategic Games with LLMs
by: Chupilkin, Maxim
Published: (2026)
by: Chupilkin, Maxim
Published: (2026)
MISID: A Multimodal Multi-turn Dataset for Complex Intent Recognition in Strategic Deception Games
by: Lin, Shufang, et al.
Published: (2026)
by: Lin, Shufang, et al.
Published: (2026)
WGSR-Bench: Wargame-based Game-theoretic Strategic Reasoning Benchmark for Large Language Models
by: Yin, Qiyue, et al.
Published: (2025)
by: Yin, Qiyue, et al.
Published: (2025)
Efficient Adaptation in Mixed-Motive Environments via Hierarchical Opponent Modeling and Planning
by: Huang, Yizhe, et al.
Published: (2024)
by: Huang, Yizhe, et al.
Published: (2024)
Similar Items
-
Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective
by: Zapzalka, Dylan, et al.
Published: (2025) -
From Biased Selective Labels to Pseudo-Labels: An Expectation-Maximization Framework for Learning from Biased Decisions
by: Chang, Trenton, et al.
Published: (2024) -
Conditional Front-door Adjustment for Heterogeneous Treatment Assignment Effect Estimation Under Non-adherence
by: Chen, Winston, et al.
Published: (2025) -
Measuring Model Performance in the Presence of an Intervention
by: Chen, Winston, et al.
Published: (2025) -
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
by: Chang, Trenton, et al.
Published: (2025)