Who Gets the Reward, Who Gets the Blame? Evaluation-Aligned Training Signals for Multi-LLM Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Chih-Hsuan, Mallick, Tanwi, Chen, Le, Raghavan, Krishnan, Wells, Azton, Gueroudji, Amal, Foster, Ian T., Thakur, Rajeev |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
You Can't Always Get What You Want: Games of Ordered Preference
by: Lee, Dong Ho, et al.
Published: (2024)
by: Lee, Dong Ho, et al.
Published: (2024)
Distributionally Robust Markov Games with Average Reward
by: Roch, Zachary, et al.
Published: (2025)
by: Roch, Zachary, et al.
Published: (2025)
MAPPO-LCR: Multi-Agent Proximal Policy Optimization with Local Cooperation Reward in Spatial Public Goods Games
by: Yang, Zhaoqilin, et al.
Published: (2025)
by: Yang, Zhaoqilin, et al.
Published: (2025)
Approximating Nash Equilibria in Normal-Form Games via Stochastic Optimization
by: Gemp, Ian, et al.
Published: (2023)
by: Gemp, Ian, et al.
Published: (2023)
Stochastic Games for Interactive Manipulation Domains
by: Muvvala, Karan, et al.
Published: (2024)
by: Muvvala, Karan, et al.
Published: (2024)
No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows
by: Raghavan, Siddeshwar, et al.
Published: (2026)
by: Raghavan, Siddeshwar, et al.
Published: (2026)
Learning in Conjectural Stackelberg Games
by: Morri, Francesco, et al.
Published: (2025)
by: Morri, Francesco, et al.
Published: (2025)
Steering Noncooperative Games Through Conjecture Design
by: Morri, Francesco, et al.
Published: (2025)
by: Morri, Francesco, et al.
Published: (2025)
Synthesis of Reward Machines for Multi-Agent Equilibrium Design (Full Version)
by: Najib, Muhammad, et al.
Published: (2024)
by: Najib, Muhammad, et al.
Published: (2024)
Approximating the Core via Iterative Coalition Sampling
by: Gemp, Ian, et al.
Published: (2024)
by: Gemp, Ian, et al.
Published: (2024)
Learning in Strategic Queuing Systems with Small Buffers
by: Abel, Ariana, et al.
Published: (2025)
by: Abel, Ariana, et al.
Published: (2025)
$\aleph$-IPOMDP: Mitigating Deception in a Cognitive Hierarchy with Off-Policy Counterfactual Anomaly Detection
by: Alon, Nitay, et al.
Published: (2024)
by: Alon, Nitay, et al.
Published: (2024)
The Multi-Stage Assignment Problem: A Fairness Perspective
by: J, Vibulan, et al.
Published: (2025)
by: J, Vibulan, et al.
Published: (2025)
Constant-Memory Strategies in Stochastic Games: Best Responses and Equilibria
by: Zhu, Fengming, et al.
Published: (2025)
by: Zhu, Fengming, et al.
Published: (2025)
Efficient Investment in Multi-Agent Models of Public Transportation
by: Bullinger, Martin, et al.
Published: (2026)
by: Bullinger, Martin, et al.
Published: (2026)
Binary Decisions in DAOs: Accountability and Belief Aggregation via Linear Opinion Pools
by: Braz, Nuno, et al.
Published: (2026)
by: Braz, Nuno, et al.
Published: (2026)
Auctioning Escape Permits for Multiple Correlated Pollutants Using CMRA
by: Goyal, Keshav, et al.
Published: (2024)
by: Goyal, Keshav, et al.
Published: (2024)
Leadership Inference for Multi-Agent Interactions
by: Khan, Hamzah, et al.
Published: (2023)
by: Khan, Hamzah, et al.
Published: (2023)
DelAC: A Multi-agent Reinforcement Learning of Team-Symmetric Stochastic Games
by: Lee, Duan-Shin, et al.
Published: (2026)
by: Lee, Duan-Shin, et al.
Published: (2026)
On Condorcet's Jury Theorem with Abstention
by: Meir, Reshef, et al.
Published: (2025)
by: Meir, Reshef, et al.
Published: (2025)
Optimal Strategy Revision in Population Games: A Mean Field Game Theory Perspective
by: Barreiro-Gomez, Julian, et al.
Published: (2025)
by: Barreiro-Gomez, Julian, et al.
Published: (2025)
From Necklaces to Coalitions: Fair and Self-Interested Distribution of Coalition Value Calculations
by: Payne, Terry R., et al.
Published: (2026)
by: Payne, Terry R., et al.
Published: (2026)
A potentialization algorithm for games with applications to multi-agent learning in repeated games
by: Lakheshar, Philipp, et al.
Published: (2026)
by: Lakheshar, Philipp, et al.
Published: (2026)
The Danger Of Arrogance: Welfare Equilibra As A Solution To Stackelberg Self-Play In Non-Coincidental Games
by: Levi, Jake, et al.
Published: (2024)
by: Levi, Jake, et al.
Published: (2024)
Existence and Verification of Nash Equilibria in Non-Cooperative Contribution Games with Resource Contention
by: Troquard, Nicolas
Published: (2024)
by: Troquard, Nicolas
Published: (2024)
The (Computational) Social Choice Take on Indivisible Participatory Budgeting
by: Rey, Simon, et al.
Published: (2023)
by: Rey, Simon, et al.
Published: (2023)
Grounded Predictions of Teamwork as a One-Shot Game: A Multiagent Multi-Armed Bandits Approach
by: Gómez, Alejandra López de Aberasturi, et al.
Published: (2024)
by: Gómez, Alejandra López de Aberasturi, et al.
Published: (2024)
Adaptation Procedure in Misinformation Games
by: Varsos, Konstantinos, et al.
Published: (2024)
by: Varsos, Konstantinos, et al.
Published: (2024)
Modeling Prejudice and Its Effect on Societal Prosperity
by: Mohan, Deep Inder, et al.
Published: (2021)
by: Mohan, Deep Inder, et al.
Published: (2021)
Smooth Games of Configuration in the Linear-Quadratic Setting
by: Milzman, Jesse, et al.
Published: (2025)
by: Milzman, Jesse, et al.
Published: (2025)
Nash Q-Network for Multi-Agent Cybersecurity Simulation
by: Xie, Qintong, et al.
Published: (2025)
by: Xie, Qintong, et al.
Published: (2025)
Distribution through Repeated Market with Buying Rights
by: Sychrovský, David, et al.
Published: (2025)
by: Sychrovský, David, et al.
Published: (2025)
Two-Sided Manipulation Games in Stable Matching Markets
by: Hosseini, Hadi, et al.
Published: (2025)
by: Hosseini, Hadi, et al.
Published: (2025)
Identifying Imperfect Clones in Elections
by: Faliszewski, Piotr, et al.
Published: (2025)
by: Faliszewski, Piotr, et al.
Published: (2025)
Persuading Stable Matching
by: Shaki, Jonathan, et al.
Published: (2025)
by: Shaki, Jonathan, et al.
Published: (2025)
The Condorcet Dimension of Metric Spaces
by: Lassota, Alexandra, et al.
Published: (2024)
by: Lassota, Alexandra, et al.
Published: (2024)
Equilibria in routing games with connected autonomous vehicles will not be strong, as exclusive clubs may form
by: Kucharski, Rafał, et al.
Published: (2025)
by: Kucharski, Rafał, et al.
Published: (2025)
Can We Volunteer Out of the Peer Review Crisis?
by: Tang, Theo, et al.
Published: (2026)
by: Tang, Theo, et al.
Published: (2026)
Leveraging Team Correlation for Approximating Equilibrium in Two-Team Zero-Sum Games
by: Liu, Naming, et al.
Published: (2024)
by: Liu, Naming, et al.
Published: (2024)
Beyond Theorems: A Counterexample to Potential Markov Game Criteria
by: Fardno, Fatemeh, et al.
Published: (2024)
by: Fardno, Fatemeh, et al.
Published: (2024)
Similar Items
-
You Can't Always Get What You Want: Games of Ordered Preference
by: Lee, Dong Ho, et al.
Published: (2024) -
Distributionally Robust Markov Games with Average Reward
by: Roch, Zachary, et al.
Published: (2025) -
MAPPO-LCR: Multi-Agent Proximal Policy Optimization with Local Cooperation Reward in Spatial Public Goods Games
by: Yang, Zhaoqilin, et al.
Published: (2025) -
Approximating Nash Equilibria in Normal-Form Games via Stochastic Optimization
by: Gemp, Ian, et al.
Published: (2023) -
Stochastic Games for Interactive Manipulation Domains
by: Muvvala, Karan, et al.
Published: (2024)