Stackelberg Learning from Human Feedback: Preference Optimization as a Sequential Game
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pásztor, Barna, Buening, Thomas Kleine, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
von: Pásztor, Barna, et al.
Veröffentlicht: (2024)
Distributed Stackelberg Strategies in State-based Potential Games for Autonomous Decentralized Learning Manufacturing Systems
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
von: Yuwono, Steve, et al.
Veröffentlicht: (2024)
Nash Learning from Human Feedback
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
von: Munos, Rémi, et al.
Veröffentlicht: (2023)
Learning Collusion in Episodic, Inventory-Constrained Markets
von: Friedrich, Paul, et al.
Veröffentlicht: (2024)
von: Friedrich, Paul, et al.
Veröffentlicht: (2024)
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
von: Zaman, Muhammad Aneeq uz, et al.
Veröffentlicht: (2024)
von: Zaman, Muhammad Aneeq uz, et al.
Veröffentlicht: (2024)
Differentiable Bilevel Programming for Stackelberg Congestion Games
von: Li, Jiayang, et al.
Veröffentlicht: (2022)
von: Li, Jiayang, et al.
Veröffentlicht: (2022)
Bounded Rationality Equilibrium Learning in Mean Field Games
von: Eich, Yannick, et al.
Veröffentlicht: (2024)
von: Eich, Yannick, et al.
Veröffentlicht: (2024)
Learning Optimal Tax Design in Nonatomic Congestion Games
von: Cui, Qiwen, et al.
Veröffentlicht: (2024)
von: Cui, Qiwen, et al.
Veröffentlicht: (2024)
Online Learning of Counter Categories and Ratings in PvP Games
von: Lin, Chiu-Chou, et al.
Veröffentlicht: (2025)
von: Lin, Chiu-Chou, et al.
Veröffentlicht: (2025)
Inverse Concave-Utility Reinforcement Learning is Inverse Game Theory
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2024)
von: Çelikok, Mustafa Mert, et al.
Veröffentlicht: (2024)
Preference-Based Multi-Agent Reinforcement Learning: Data Coverage and Algorithmic Techniques
von: Zhang, Natalia, et al.
Veröffentlicht: (2024)
von: Zhang, Natalia, et al.
Veröffentlicht: (2024)
ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
von: Lei, Shiqi, et al.
Veröffentlicht: (2024)
von: Lei, Shiqi, et al.
Veröffentlicht: (2024)
Learning Mean Field Games on Sparse Graphs: A Hybrid Graphex Approach
von: Fabian, Christian, et al.
Veröffentlicht: (2024)
von: Fabian, Christian, et al.
Veröffentlicht: (2024)
A Single Online Agent Can Efficiently Learn Mean Field Games
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyu, et al.
Veröffentlicht: (2024)
Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games
von: Lanier, JB, et al.
Veröffentlicht: (2026)
von: Lanier, JB, et al.
Veröffentlicht: (2026)
Combining Tree-Search, Generative Models, and Nash Bargaining Concepts in Game-Theoretic Reinforcement Learning
von: Li, Zun, et al.
Veröffentlicht: (2023)
von: Li, Zun, et al.
Veröffentlicht: (2023)
Evaluating LLMs in Open-Source Games
von: Sistla, Swadesh, et al.
Veröffentlicht: (2025)
von: Sistla, Swadesh, et al.
Veröffentlicht: (2025)
Observation Interference in Partially Observable Assistance Games
von: Emmons, Scott, et al.
Veröffentlicht: (2024)
von: Emmons, Scott, et al.
Veröffentlicht: (2024)
Offline Fictitious Self-Play for Competitive Games
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
von: Chen, Jingxiao, et al.
Veröffentlicht: (2024)
Learning in Conjectural Stackelberg Games
von: Morri, Francesco, et al.
Veröffentlicht: (2025)
von: Morri, Francesco, et al.
Veröffentlicht: (2025)
RuleSmith: Multi-Agent LLMs for Automated Game Balancing
von: Zeng, Ziyao, et al.
Veröffentlicht: (2026)
von: Zeng, Ziyao, et al.
Veröffentlicht: (2026)
Socially-Weighted Alignment: A Game-Theoretic Framework for Multi-Agent LLM Systems
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
von: Mumcu, Furkan, et al.
Veröffentlicht: (2026)
Cooperative Game-Theoretic Credit Assignment for Multi-Agent Policy Gradients via the Core
von: Ji, Mengda, et al.
Veröffentlicht: (2025)
von: Ji, Mengda, et al.
Veröffentlicht: (2025)
Game-theoretic LLM: Agent Workflow for Negotiation Games
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
Learning Bilateral Team Formation in Cooperative Multi-Agent Reinforcement Learning
von: Moslemi, Koorosh, et al.
Veröffentlicht: (2025)
von: Moslemi, Koorosh, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Hanabi
von: Cohen, Nina, et al.
Veröffentlicht: (2025)
von: Cohen, Nina, et al.
Veröffentlicht: (2025)
Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning
von: Gemp, Ian, et al.
Veröffentlicht: (2024)
von: Gemp, Ian, et al.
Veröffentlicht: (2024)
Learning to Negotiate via Voluntary Commitment
von: Zhu, Shuhui, et al.
Veröffentlicht: (2025)
von: Zhu, Shuhui, et al.
Veröffentlicht: (2025)
The Benefits of Power Regularization in Cooperative Reinforcement Learning
von: Li, Michelle, et al.
Veröffentlicht: (2024)
von: Li, Michelle, et al.
Veröffentlicht: (2024)
Learning Mean Field Control on Sparse Graphs
von: Fabian, Christian, et al.
Veröffentlicht: (2025)
von: Fabian, Christian, et al.
Veröffentlicht: (2025)
Emergent Dominance Hierarchies in Reinforcement Learning Agents
von: Rachum, Ram, et al.
Veröffentlicht: (2024)
von: Rachum, Ram, et al.
Veröffentlicht: (2024)
Welfare and Fairness in Multi-objective Reinforcement Learning
von: Fan, Zimeng, et al.
Veröffentlicht: (2022)
von: Fan, Zimeng, et al.
Veröffentlicht: (2022)
Learning Unanimously Acceptable Lotteries via Queries
von: Choo, Davin, et al.
Veröffentlicht: (2026)
von: Choo, Davin, et al.
Veröffentlicht: (2026)
Policy Iteration for Two-Player General-Sum Stochastic Stackelberg Games
von: Kudo, Mikoto, et al.
Veröffentlicht: (2024)
von: Kudo, Mikoto, et al.
Veröffentlicht: (2024)
On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents
von: Loftin, Robert, et al.
Veröffentlicht: (2024)
von: Loftin, Robert, et al.
Veröffentlicht: (2024)
Networked Communication for Mean-Field Games with Function Approximation and Empirical Mean-Field Estimation
von: Benjamin, Patrick, et al.
Veröffentlicht: (2024)
von: Benjamin, Patrick, et al.
Veröffentlicht: (2024)
Offline Learning of Nash Stable Coalition Structures with Possibly Overlapping Coalitions
von: Cohen, Saar
Veröffentlicht: (2026)
von: Cohen, Saar
Veröffentlicht: (2026)
Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning
von: Kudo, Mikoto, et al.
Veröffentlicht: (2026)
von: Kudo, Mikoto, et al.
Veröffentlicht: (2026)
A Black-box Approach for Non-stationary Multi-agent Reinforcement Learning
von: Jiang, Haozhe, et al.
Veröffentlicht: (2023)
von: Jiang, Haozhe, et al.
Veröffentlicht: (2023)
AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning
von: Ekpo, Promise, et al.
Veröffentlicht: (2025)
von: Ekpo, Promise, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bandits with Preference Feedback: A Stackelberg Game Perspective
von: Pásztor, Barna, et al.
Veröffentlicht: (2024) -
Distributed Stackelberg Strategies in State-based Potential Games for Autonomous Decentralized Learning Manufacturing Systems
von: Yuwono, Steve, et al.
Veröffentlicht: (2024) -
Nash Learning from Human Feedback
von: Munos, Rémi, et al.
Veröffentlicht: (2023) -
Learning Collusion in Episodic, Inventory-Constrained Markets
von: Friedrich, Paul, et al.
Veröffentlicht: (2024) -
Policy Optimization finds Nash Equilibrium in Regularized General-Sum LQ Games
von: Zaman, Muhammad Aneeq uz, et al.
Veröffentlicht: (2024)