In-Context Credit Assignment via the Core
Fuente:
arXiv
Salvato in:
| Autori principali: | Harris, Keegan, Prasad, Siddharth, Trockman, Asher |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
di: Ji, Mengda, et al.
Pubblicazione: (2025)
di: Ji, Mengda, et al.
Pubblicazione: (2025)
LLM-Powered Preference Elicitation in Combinatorial Assignment
di: Soumalias, Ermis, et al.
Pubblicazione: (2025)
di: Soumalias, Ermis, et al.
Pubblicazione: (2025)
Agent-Temporal Credit Assignment for Optimal Policy Preservation in Sparse Multi-Agent Reinforcement Learning
di: Kapoor, Aditya, et al.
Pubblicazione: (2024)
di: Kapoor, Aditya, et al.
Pubblicazione: (2024)
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
di: Chen, Yang, et al.
Pubblicazione: (2025)
di: Chen, Yang, et al.
Pubblicazione: (2025)
Learning in Structured Stackelberg Games
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
The Core in Max-Loss Non-Centroid Clustering Can Be Empty
di: Bredereck, Robert, et al.
Pubblicazione: (2025)
di: Bredereck, Robert, et al.
Pubblicazione: (2025)
Regret Minimization in Stackelberg Games with Side Information
di: Harris, Keegan, et al.
Pubblicazione: (2024)
di: Harris, Keegan, et al.
Pubblicazione: (2024)
Ranking Abuse via Strategic Pairwise Data Perturbations
di: Yao, Junyi, et al.
Pubblicazione: (2026)
di: Yao, Junyi, et al.
Pubblicazione: (2026)
Incentivizing Quality Text Generation via Statistical Contracts
di: Saig, Eden, et al.
Pubblicazione: (2024)
di: Saig, Eden, et al.
Pubblicazione: (2024)
Incentivizing Truthful Language Models via Peer Elicitation Games
di: Chen, Baiting, et al.
Pubblicazione: (2025)
di: Chen, Baiting, et al.
Pubblicazione: (2025)
ElicitationGPT: Text Elicitation Mechanisms via Language Models
di: Wu, Yifan, et al.
Pubblicazione: (2024)
di: Wu, Yifan, et al.
Pubblicazione: (2024)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
di: La Malfa, Gabriele, et al.
Pubblicazione: (2026)
di: La Malfa, Gabriele, et al.
Pubblicazione: (2026)
Meta-Computing Enhanced Federated Learning in IIoT: Satisfaction-Aware Incentive Scheme via DRL-Based Stackelberg Game
di: Li, Xiaohuan, et al.
Pubblicazione: (2025)
di: Li, Xiaohuan, et al.
Pubblicazione: (2025)
Algorithmic Persuasion Through Simulation
di: Harris, Keegan, et al.
Pubblicazione: (2023)
di: Harris, Keegan, et al.
Pubblicazione: (2023)
Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)
Incentive-Aware Synthetic Control: Accurate Counterfactual Estimation via Incentivized Exploration
di: Ngo, Daniel, et al.
Pubblicazione: (2023)
di: Ngo, Daniel, et al.
Pubblicazione: (2023)
Recommending Best Paper Awards for ML/AI Conferences via the Isotonic Mechanism
di: Wen, Garrett G., et al.
Pubblicazione: (2026)
di: Wen, Garrett G., et al.
Pubblicazione: (2026)
Puzzle it Out: Local-to-Global World Model for Offline Multi-Agent Reinforcement Learning
di: Li, Sijia, et al.
Pubblicazione: (2026)
di: Li, Sijia, et al.
Pubblicazione: (2026)
The Intelligent Disobedience Game: Formulating Disobedience in Stackelberg Games and Markov Decision Processes
di: Hornig, Benedikt, et al.
Pubblicazione: (2026)
di: Hornig, Benedikt, et al.
Pubblicazione: (2026)
Efficient Ensemble Selection from Binary and Pairwise Feedback
di: Neoh, Tzeh Yuan, et al.
Pubblicazione: (2026)
di: Neoh, Tzeh Yuan, et al.
Pubblicazione: (2026)
GroupSegment-SHAP: Shapley Value Explanations with Group-Segment Players for Multivariate Time Series
di: Kim, Jinwoong, et al.
Pubblicazione: (2026)
di: Kim, Jinwoong, et al.
Pubblicazione: (2026)
Pencil Puzzle Bench: A Benchmark for Multi-Step Verifiable Reasoning
di: Waugh, Justin
Pubblicazione: (2026)
di: Waugh, Justin
Pubblicazione: (2026)
NePPO: Near-Potential Policy Optimization for General-Sum Multi-Agent Reinforcement Learning
di: Kalanther, Addison, et al.
Pubblicazione: (2026)
di: Kalanther, Addison, et al.
Pubblicazione: (2026)
Finding Common Ground in a Sea of Alternatives
di: Chooi, Jay, et al.
Pubblicazione: (2026)
di: Chooi, Jay, et al.
Pubblicazione: (2026)
Asymmetric regularization mechanism for GAN training with Variational Inequalities
di: Giagtzoglou, Spyridon C., et al.
Pubblicazione: (2026)
di: Giagtzoglou, Spyridon C., et al.
Pubblicazione: (2026)
Strategic Candidacy in Generative AI Arenas
di: Hays, Chris, et al.
Pubblicazione: (2026)
di: Hays, Chris, et al.
Pubblicazione: (2026)
Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models
di: Hennes, Daniel, et al.
Pubblicazione: (2026)
di: Hennes, Daniel, et al.
Pubblicazione: (2026)
Adaptive Contracts for Cost-Effective AI Delegation
di: Saig, Eden, et al.
Pubblicazione: (2026)
di: Saig, Eden, et al.
Pubblicazione: (2026)
Governing AI Forgetting: Auditing for Machine Unlearning Compliance
di: Lin, Qinqi, et al.
Pubblicazione: (2026)
di: Lin, Qinqi, et al.
Pubblicazione: (2026)
Next-Token Prediction and Regret Minimization
di: Mohri, Mehryar, et al.
Pubblicazione: (2026)
di: Mohri, Mehryar, et al.
Pubblicazione: (2026)
The Optimal Sample Complexity of Linear Contracts
di: Høgsgaard, Mikael Møller
Pubblicazione: (2026)
di: Høgsgaard, Mikael Møller
Pubblicazione: (2026)
Routing, Cascades, and User Choice for LLMs
di: Mahmood, Rafid
Pubblicazione: (2026)
di: Mahmood, Rafid
Pubblicazione: (2026)
Knowledge-Free Correlated Agreement for Incentivizing Federated Learning
di: Witt, Leon, et al.
Pubblicazione: (2026)
di: Witt, Leon, et al.
Pubblicazione: (2026)
When Individually Calibrated Models Become Collectively Miscalibrated
di: Wang, Zhaohui
Pubblicazione: (2026)
di: Wang, Zhaohui
Pubblicazione: (2026)
Real-Time Parallel Counterfactual Regret Minimization
di: Li, Boning, et al.
Pubblicazione: (2026)
di: Li, Boning, et al.
Pubblicazione: (2026)
Sharp Spectral Thresholds for Logit Fixed Points
di: Wang, Tongxi
Pubblicazione: (2026)
di: Wang, Tongxi
Pubblicazione: (2026)
Monopoly Deal: A Benchmark Environment for Bounded One-Sided Response Games
di: Wolf, Will
Pubblicazione: (2025)
di: Wolf, Will
Pubblicazione: (2025)
Learning and Collusion in Multi-unit Auctions
di: Brânzei, Simina, et al.
Pubblicazione: (2023)
di: Brânzei, Simina, et al.
Pubblicazione: (2023)
LiteEFG: An Efficient Python Library for Solving Extensive-form Games
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
di: Liu, Mingyang, et al.
Pubblicazione: (2024)
Self-optimization in distributed manufacturing systems using Modular State-based Stackelberg Games
di: Yuwono, Steve, et al.
Pubblicazione: (2024)
di: Yuwono, Steve, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
di: Ji, Mengda, et al.
Pubblicazione: (2025) -
LLM-Powered Preference Elicitation in Combinatorial Assignment
di: Soumalias, Ermis, et al.
Pubblicazione: (2025) -
Agent-Temporal Credit Assignment for Optimal Policy Preservation in Sparse Multi-Agent Reinforcement Learning
di: Kapoor, Aditya, et al.
Pubblicazione: (2024) -
Meta-Inverse Reinforcement Learning for Mean Field Games via Probabilistic Context Variables
di: Chen, Yang, et al.
Pubblicazione: (2025) -
Learning in Structured Stackelberg Games
di: Balcan, Maria-Florina, et al.
Pubblicazione: (2025)