Who's Gaming the System? A Causally-Motivated Approach for Detecting Strategic Adaptation
Fuente:
arXiv
Guardado en:
| Autores principales: | Chang, Trenton, Warrenburg, Lindsay, Park, Sae-Hwan, Parikh, Ravi B., Makar, Maggie, Wiens, Jenna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective
por: Zapzalka, Dylan, et al.
Publicado: (2025)
por: Zapzalka, Dylan, et al.
Publicado: (2025)
From Biased Selective Labels to Pseudo-Labels: An Expectation-Maximization Framework for Learning from Biased Decisions
por: Chang, Trenton, et al.
Publicado: (2024)
por: Chang, Trenton, et al.
Publicado: (2024)
Conditional Front-door Adjustment for Heterogeneous Treatment Assignment Effect Estimation Under Non-adherence
por: Chen, Winston, et al.
Publicado: (2025)
por: Chen, Winston, et al.
Publicado: (2025)
Measuring Model Performance in the Presence of an Intervention
por: Chen, Winston, et al.
Publicado: (2025)
por: Chen, Winston, et al.
Publicado: (2025)
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
por: Chang, Trenton, et al.
Publicado: (2025)
por: Chang, Trenton, et al.
Publicado: (2025)
Improvisational Games as a Benchmark for Social Intelligence of AI Agents: The Case of Connections
por: Parikh, Gaurav Rajesh, et al.
Publicado: (2026)
por: Parikh, Gaurav Rajesh, et al.
Publicado: (2026)
Explaining Decisions of Agents in Mixed-Motive Games
por: Orner, Maayan, et al.
Publicado: (2024)
por: Orner, Maayan, et al.
Publicado: (2024)
Causal Strategic Learning with Competitive Selection
por: Vo, Kiet Q. H., et al.
Publicado: (2023)
por: Vo, Kiet Q. H., et al.
Publicado: (2023)
Who Does Your Algorithm Fail? Investigating Age and Ethnic Bias in the MAMA-MIA Dataset
por: Parikh, Aditya, et al.
Publicado: (2025)
por: Parikh, Aditya, et al.
Publicado: (2025)
DEPICT: Diffusion-Enabled Permutation Importance for Image Classification Tasks
por: Jabbour, Sarah, et al.
Publicado: (2024)
por: Jabbour, Sarah, et al.
Publicado: (2024)
Detecting Model Drifts in Non-Stationary Environment Using Edit Operation Measures
por: Lee, Chang-Hwan, et al.
Publicado: (2025)
por: Lee, Chang-Hwan, et al.
Publicado: (2025)
Adaptive Punishment for Cooperation in Mixed-Motive Games
por: Tang, Min, et al.
Publicado: (2026)
por: Tang, Min, et al.
Publicado: (2026)
Using Large Language Models to Categorize Strategic Situations and Decipher Motivations Behind Human Behaviors
por: Xie, Yutong, et al.
Publicado: (2025)
por: Xie, Yutong, et al.
Publicado: (2025)
Cross-Architecture Model Diffing with Crosscoders: Unsupervised Discovery of Differences Between LLMs
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
por: Jiralerspong, Thomas, et al.
Publicado: (2026)
Toward Reasoning on the Boundary: A Mixup-based Approach for Graph Anomaly Detection
por: Kim, Hwan, et al.
Publicado: (2024)
por: Kim, Hwan, et al.
Publicado: (2024)
Learning to Balance Altruism and Self-interest Based on Empathy in Mixed-Motive Games
por: Kong, Fanqi, et al.
Publicado: (2024)
por: Kong, Fanqi, et al.
Publicado: (2024)
Hypothesis Testing the Circuit Hypothesis in LLMs
por: Shi, Claudia, et al.
Publicado: (2024)
por: Shi, Claudia, et al.
Publicado: (2024)
Automatic Question Generation for Intuitive Learning Utilizing Causal Graph Guided Chain of Thought Reasoning
por: Wang, Nicholas X., et al.
Publicado: (2026)
por: Wang, Nicholas X., et al.
Publicado: (2026)
Generating Novelty in Open-World Multi-Agent Strategic Board Games
por: Kejriwal, Mayank, et al.
Publicado: (2025)
por: Kejriwal, Mayank, et al.
Publicado: (2025)
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
por: He, Yidong, et al.
Publicado: (2026)
por: He, Yidong, et al.
Publicado: (2026)
On the Limits of Selective AI Prediction: A Case Study in Clinical Decision Making
por: Jabbour, Sarah, et al.
Publicado: (2025)
por: Jabbour, Sarah, et al.
Publicado: (2025)
Exploring Large Language Models for Word Games:Who is the Spy?
por: Wei, Chentian, et al.
Publicado: (2025)
por: Wei, Chentian, et al.
Publicado: (2025)
Trajectory Adaptation using Large Language Models
por: Maurya, Anurag, et al.
Publicado: (2025)
por: Maurya, Anurag, et al.
Publicado: (2025)
Strategic Over-Parameterization for Generalizable Low-Rank Adaptation
por: Gao, Jing, et al.
Publicado: (2026)
por: Gao, Jing, et al.
Publicado: (2026)
Do Persona-Infused LLMs Affect Performance in a Strategic Reasoning Game?
por: Licato, John, et al.
Publicado: (2025)
por: Licato, John, et al.
Publicado: (2025)
Multiverse: Language-Conditioned Multi-Game Level Blending via Shared Representation
por: Baek, In-Chang, et al.
Publicado: (2026)
por: Baek, In-Chang, et al.
Publicado: (2026)
SLOW: Strategic Logical-inference Open Workspace for Cognitive Adaptation in AI Tutoring
por: Wei, Yuang, et al.
Publicado: (2026)
por: Wei, Yuang, et al.
Publicado: (2026)
Debiasing Reward Models via Causally Motivated Inference-Time Intervention
por: Shinoda, Kazutoshi, et al.
Publicado: (2026)
por: Shinoda, Kazutoshi, et al.
Publicado: (2026)
Individual Causal Inference with Structural Causal Model
por: Chang, Daniel T.
Publicado: (2025)
por: Chang, Daniel T.
Publicado: (2025)
Who's Your Judge? On the Detectability of LLM-Generated Judgments
por: Li, Dawei, et al.
Publicado: (2025)
por: Li, Dawei, et al.
Publicado: (2025)
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents
por: Costarelli, Anthony, et al.
Publicado: (2024)
por: Costarelli, Anthony, et al.
Publicado: (2024)
Strategic Insights from Simulation Gaming of AI Race Dynamics
por: Gruetzemacher, Ross, et al.
Publicado: (2024)
por: Gruetzemacher, Ross, et al.
Publicado: (2024)
Strategic Communication under Threat: Learning Information Trade-offs in Pursuit-Evasion Games
por: La Gatta, Valerio, et al.
Publicado: (2025)
por: La Gatta, Valerio, et al.
Publicado: (2025)
Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization
por: Xu, Zelai, et al.
Publicado: (2025)
por: Xu, Zelai, et al.
Publicado: (2025)
M3-BENCH: Process-Aware Evaluation of LLM Agents' Social Behaviors in Mixed-Motive Games
por: Xie, Sixiong, et al.
Publicado: (2026)
por: Xie, Sixiong, et al.
Publicado: (2026)
Correcting Human Labels for Rater Effects in AI Evaluation: An Item Response Theory Approach
por: Casabianca, Jodi M., et al.
Publicado: (2026)
por: Casabianca, Jodi M., et al.
Publicado: (2026)
Multi-Agent Strategic Games with LLMs
por: Chupilkin, Maxim
Publicado: (2026)
por: Chupilkin, Maxim
Publicado: (2026)
MISID: A Multimodal Multi-turn Dataset for Complex Intent Recognition in Strategic Deception Games
por: Lin, Shufang, et al.
Publicado: (2026)
por: Lin, Shufang, et al.
Publicado: (2026)
WGSR-Bench: Wargame-based Game-theoretic Strategic Reasoning Benchmark for Large Language Models
por: Yin, Qiyue, et al.
Publicado: (2025)
por: Yin, Qiyue, et al.
Publicado: (2025)
Efficient Adaptation in Mixed-Motive Environments via Hierarchical Opponent Modeling and Planning
por: Huang, Yizhe, et al.
Publicado: (2024)
por: Huang, Yizhe, et al.
Publicado: (2024)
Ejemplares similares
-
Estimating Misreporting in the Presence of Genuine Modification: A Causal Perspective
por: Zapzalka, Dylan, et al.
Publicado: (2025) -
From Biased Selective Labels to Pseudo-Labels: An Expectation-Maximization Framework for Learning from Biased Decisions
por: Chang, Trenton, et al.
Publicado: (2024) -
Conditional Front-door Adjustment for Heterogeneous Treatment Assignment Effect Estimation Under Non-adherence
por: Chen, Winston, et al.
Publicado: (2025) -
Measuring Model Performance in the Presence of an Intervention
por: Chen, Winston, et al.
Publicado: (2025) -
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
por: Chang, Trenton, et al.
Publicado: (2025)