Stackelberg Learning from Human Feedback: Preference Optimization as a Sequential Game
Fuente:
arXiv
Saved in:
| Main Authors: | Pásztor, Barna, Buening, Thomas Kleine, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bandits with Preference Feedback: A Stackelberg Game Perspective
by: Pásztor, Barna, et al.
Published: (2024)
by: Pásztor, Barna, et al.
Published: (2024)
Distributed Stackelberg Strategies in State-based Potential Games for Autonomous Decentralized Learning Manufacturing Systems
by: Yuwono, Steve, et al.
Published: (2024)
by: Yuwono, Steve, et al.
Published: (2024)
Nash Learning from Human Feedback
by: Munos, Rémi, et al.
Published: (2023)
by: Munos, Rémi, et al.
Published: (2023)
Learning Collusion in Episodic, Inventory-Constrained Markets
by: Friedrich, Paul, et al.
Published: (2024)
by: Friedrich, Paul, et al.
Published: (2024)
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
by: Zaman, Muhammad Aneeq uz, et al.
Published: (2024)
by: Zaman, Muhammad Aneeq uz, et al.
Published: (2024)
Differentiable Bilevel Programming for Stackelberg Congestion Games
by: Li, Jiayang, et al.
Published: (2022)
by: Li, Jiayang, et al.
Published: (2022)
Bounded Rationality Equilibrium Learning in Mean Field Games
by: Eich, Yannick, et al.
Published: (2024)
by: Eich, Yannick, et al.
Published: (2024)
Learning Optimal Tax Design in Nonatomic Congestion Games
by: Cui, Qiwen, et al.
Published: (2024)
by: Cui, Qiwen, et al.
Published: (2024)
Online Learning of Counter Categories and Ratings in PvP Games
by: Lin, Chiu-Chou, et al.
Published: (2025)
by: Lin, Chiu-Chou, et al.
Published: (2025)
Inverse Concave-Utility Reinforcement Learning is Inverse Game Theory
by: Çelikok, Mustafa Mert, et al.
Published: (2024)
by: Çelikok, Mustafa Mert, et al.
Published: (2024)
Preference-Based Multi-Agent Reinforcement Learning: Data Coverage and Algorithmic Techniques
by: Zhang, Natalia, et al.
Published: (2024)
by: Zhang, Natalia, et al.
Published: (2024)
ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
by: Lei, Shiqi, et al.
Published: (2024)
by: Lei, Shiqi, et al.
Published: (2024)
Learning Mean Field Games on Sparse Graphs: A Hybrid Graphex Approach
by: Fabian, Christian, et al.
Published: (2024)
by: Fabian, Christian, et al.
Published: (2024)
A Single Online Agent Can Efficiently Learn Mean Field Games
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
by: Lanier, JB, et al.
Published: (2026)
by: Lanier, JB, et al.
Published: (2026)
Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
by: Li, Zun, et al.
Published: (2023)
by: Li, Zun, et al.
Published: (2023)
Evaluating LLMs in Open-Source Games
by: Sistla, Swadesh, et al.
Published: (2025)
by: Sistla, Swadesh, et al.
Published: (2025)
Observation Interference in Partially Observable Assistance Games
by: Emmons, Scott, et al.
Published: (2024)
by: Emmons, Scott, et al.
Published: (2024)
Offline Fictitious Self-Play for Competitive Games
by: Chen, Jingxiao, et al.
Published: (2024)
by: Chen, Jingxiao, et al.
Published: (2024)
Learning in Conjectural Stackelberg Games
by: Morri, Francesco, et al.
Published: (2025)
by: Morri, Francesco, et al.
Published: (2025)
RuleSmith: Multi-Agent LLMs for Automated Game Balancing
by: Zeng, Ziyao, et al.
Published: (2026)
by: Zeng, Ziyao, et al.
Published: (2026)
Socially-Weighted Alignment: A Game-Theoretic Framework for Multi-Agent LLM Systems
by: Mumcu, Furkan, et al.
Published: (2026)
by: Mumcu, Furkan, et al.
Published: (2026)
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
by: Ji, Mengda, et al.
Published: (2025)
by: Ji, Mengda, et al.
Published: (2025)
Game-theoretic LLM: Agent Workflow for Negotiation Games
by: Hua, Wenyue, et al.
Published: (2024)
by: Hua, Wenyue, et al.
Published: (2024)
Learning Bilateral Team Formation in Cooperative Multi-Agent Reinforcement Learning
by: Moslemi, Koorosh, et al.
Published: (2025)
by: Moslemi, Koorosh, et al.
Published: (2025)
Reinforcement Learning for Hanabi
by: Cohen, Nina, et al.
Published: (2025)
by: Cohen, Nina, et al.
Published: (2025)
Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning
by: Gemp, Ian, et al.
Published: (2024)
by: Gemp, Ian, et al.
Published: (2024)
Learning to Negotiate via Voluntary Commitment
by: Zhu, Shuhui, et al.
Published: (2025)
by: Zhu, Shuhui, et al.
Published: (2025)
The Benefits of Power Regularization in Cooperative Reinforcement Learning
by: Li, Michelle, et al.
Published: (2024)
by: Li, Michelle, et al.
Published: (2024)
Learning Mean Field Control on Sparse Graphs
by: Fabian, Christian, et al.
Published: (2025)
by: Fabian, Christian, et al.
Published: (2025)
Emergent Dominance Hierarchies in Reinforcement Learning Agents
by: Rachum, Ram, et al.
Published: (2024)
by: Rachum, Ram, et al.
Published: (2024)
Welfare and Fairness in Multi-objective Reinforcement Learning
by: Fan, Zimeng, et al.
Published: (2022)
by: Fan, Zimeng, et al.
Published: (2022)
Learning Unanimously Acceptable Lotteries via Queries
by: Choo, Davin, et al.
Published: (2026)
by: Choo, Davin, et al.
Published: (2026)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
by: Kudo, Mikoto, et al.
Published: (2024)
by: Kudo, Mikoto, et al.
Published: (2024)
On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents
by: Loftin, Robert, et al.
Published: (2024)
by: Loftin, Robert, et al.
Published: (2024)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
by: Benjamin, Patrick, et al.
Published: (2024)
by: Benjamin, Patrick, et al.
Published: (2024)
Offline Learning of Nash Stable Coalition Structures with Possibly Overlapping Coalitions
by: Cohen, Saar
Published: (2026)
by: Cohen, Saar
Published: (2026)
Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning
by: Kudo, Mikoto, et al.
Published: (2026)
by: Kudo, Mikoto, et al.
Published: (2026)
A Black-box Approach for Non-stationary Multi-agent Reinforcement Learning
by: Jiang, Haozhe, et al.
Published: (2023)
by: Jiang, Haozhe, et al.
Published: (2023)
AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning
by: Ekpo, Promise, et al.
Published: (2025)
by: Ekpo, Promise, et al.
Published: (2025)
Similar Items
-
Bandits with Preference Feedback: A Stackelberg Game Perspective
by: Pásztor, Barna, et al.
Published: (2024) -
Distributed Stackelberg Strategies in State-based Potential Games for Autonomous Decentralized Learning Manufacturing Systems
by: Yuwono, Steve, et al.
Published: (2024) -
Nash Learning from Human Feedback
by: Munos, Rémi, et al.
Published: (2023) -
Learning Collusion in Episodic, Inventory-Constrained Markets
by: Friedrich, Paul, et al.
Published: (2024) -
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
by: Zaman, Muhammad Aneeq uz, et al.
Published: (2024)