Robust Reward Design for Markov Decision Processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Shuo, Ma, Haoxiang, Fu, Jie, Han, Shuo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games
von: Yang, Tong, et al.
Veröffentlicht: (2025)
von: Yang, Tong, et al.
Veröffentlicht: (2025)
Strategizing against Q-learners: A Control-theoretical Approach
von: Arslantas, Yuksel, et al.
Veröffentlicht: (2024)
von: Arslantas, Yuksel, et al.
Veröffentlicht: (2024)
Interpretable Price Bounds Estimation with Shape Constraints in Price Optimization
von: Ikeda, Shunnosuke, et al.
Veröffentlicht: (2024)
von: Ikeda, Shunnosuke, et al.
Veröffentlicht: (2024)
Indirect Dynamic Negotiation in the Nash Demand Game
von: Guy, Tatiana V., et al.
Veröffentlicht: (2024)
von: Guy, Tatiana V., et al.
Veröffentlicht: (2024)
A General Equilibrium Theory of Orchestrated AI Agent Systems
von: Garnier, Jean-Philippe
Veröffentlicht: (2026)
von: Garnier, Jean-Philippe
Veröffentlicht: (2026)
Beyond discounted returns: Robust Markov decision processes with average and Blackwell optimality
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2023)
von: Grand-Clément, Julien, et al.
Veröffentlicht: (2023)
Near Optimal Convergence to Coarse Correlated Equilibrium in General-Sum Markov Games
von: Yorulmaz, Asrin Efe, et al.
Veröffentlicht: (2025)
von: Yorulmaz, Asrin Efe, et al.
Veröffentlicht: (2025)
On Non-Cooperative Perfect Information Semi-Markov Games
von: Bakshi, K. G., et al.
Veröffentlicht: (2022)
von: Bakshi, K. G., et al.
Veröffentlicht: (2022)
Markov Chains of Evolutionary Games with a Small Number of Players
von: Kehagias, Athanasios
Veröffentlicht: (2025)
von: Kehagias, Athanasios
Veröffentlicht: (2025)
Analytically Tractable Models for Decision Making under Present Bias
von: Akagi, Yasunori, et al.
Veröffentlicht: (2023)
von: Akagi, Yasunori, et al.
Veröffentlicht: (2023)
On the Equilibrium of a Class of Leader-Follower Games with Decision-Dependent Chance Constraints
von: Wang, Jingxiang, et al.
Veröffentlicht: (2024)
von: Wang, Jingxiang, et al.
Veröffentlicht: (2024)
Decoding Game: On Minimax Optimality of Heuristic Text Generation Strategies
von: Chen, Sijin, et al.
Veröffentlicht: (2024)
von: Chen, Sijin, et al.
Veröffentlicht: (2024)
Graphon Mean Field Games with a Representative Player: Analysis and Learning Algorithm
von: Zhou, Fuzhong, et al.
Veröffentlicht: (2024)
von: Zhou, Fuzhong, et al.
Veröffentlicht: (2024)
Learning in Mean Field Games: A Survey
von: Laurière, Mathieu, et al.
Veröffentlicht: (2022)
von: Laurière, Mathieu, et al.
Veröffentlicht: (2022)
Strategically Robust Aggregative Games
von: Feik, Andreas, et al.
Veröffentlicht: (2026)
von: Feik, Andreas, et al.
Veröffentlicht: (2026)
Robust autobidding for noisy conversion prediction models
von: Pudovikov, Andrey, et al.
Veröffentlicht: (2025)
von: Pudovikov, Andrey, et al.
Veröffentlicht: (2025)
Robust Online Selection with Uncertain Offer Acceptance
von: Perez-Salazar, Sebastian, et al.
Veröffentlicht: (2021)
von: Perez-Salazar, Sebastian, et al.
Veröffentlicht: (2021)
Strategically Robust Game Theory via Optimal Transport
von: Lanzetti, Nicolas, et al.
Veröffentlicht: (2025)
von: Lanzetti, Nicolas, et al.
Veröffentlicht: (2025)
Double Distributionally Robust Bid Shading for First Price Auctions
von: Qu, Yanlin, et al.
Veröffentlicht: (2024)
von: Qu, Yanlin, et al.
Veröffentlicht: (2024)
Robust Bilevel Optimization for Near-Optimal Lower-Level Solutions
von: Besançon, Mathieu, et al.
Veröffentlicht: (2019)
von: Besançon, Mathieu, et al.
Veröffentlicht: (2019)
A Scenario Approach to the Robustness of Nonconvex-Nonconcave Minimax Problems
von: Peng, Huan, et al.
Veröffentlicht: (2025)
von: Peng, Huan, et al.
Veröffentlicht: (2025)
PREFER: Personalized Review Summarization with Online Preference Learning
von: Roy, Millend, et al.
Veröffentlicht: (2026)
von: Roy, Millend, et al.
Veröffentlicht: (2026)
Meta-Learning for Repeated Bayesian Persuasion
von: Turna, Ata Poyraz, et al.
Veröffentlicht: (2026)
von: Turna, Ata Poyraz, et al.
Veröffentlicht: (2026)
Bidding Games on Markov Decision Processes with Quantitative Reachability Objectives
von: Avni, Guy, et al.
Veröffentlicht: (2024)
von: Avni, Guy, et al.
Veröffentlicht: (2024)
Dynamic traffic assignment in a corridor network: Optimum versus Equilibrium
von: Fu, Haoran, et al.
Veröffentlicht: (2021)
von: Fu, Haoran, et al.
Veröffentlicht: (2021)
A Unified Variational Design of Predictive Mirror Descent in Convex Games under Stochastic Feedback
von: Pan, Yunian, et al.
Veröffentlicht: (2026)
von: Pan, Yunian, et al.
Veröffentlicht: (2026)
Online Contention Resolution Schemes for Network Revenue Management and Combinatorial Auctions
von: Ma, Will, et al.
Veröffentlicht: (2024)
von: Ma, Will, et al.
Veröffentlicht: (2024)
Matching Queues, Flexibility and Incentives
von: Yan, Chiwei, et al.
Veröffentlicht: (2020)
von: Yan, Chiwei, et al.
Veröffentlicht: (2020)
Distributionally Robust Nash Equilibrium Seeking with Partial Observations and Distributed Communication
von: Mandal, Nirabhra, et al.
Veröffentlicht: (2026)
von: Mandal, Nirabhra, et al.
Veröffentlicht: (2026)
Disentangling Resilience from Robustness: Contextual Dualism, Interactionism, and Game-Theoretic Paradigms
von: Zhu, Quanyan, et al.
Veröffentlicht: (2024)
von: Zhu, Quanyan, et al.
Veröffentlicht: (2024)
Achieving Logarithmic Regret in KL-Regularized Zero-Sum Markov Games
von: Nayak, Anupam, et al.
Veröffentlicht: (2025)
von: Nayak, Anupam, et al.
Veröffentlicht: (2025)
Near-Optimal Policy Optimization for Correlated Equilibrium in General-Sum Markov Games
von: Cai, Yang, et al.
Veröffentlicht: (2024)
von: Cai, Yang, et al.
Veröffentlicht: (2024)
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games
von: Hu, Anran, et al.
Veröffentlicht: (2024)
von: Hu, Anran, et al.
Veröffentlicht: (2024)
Sharing delay costs in stochastic scheduling problems with delays
von: Gonçalves-Dosantos, J. C., et al.
Veröffentlicht: (2024)
von: Gonçalves-Dosantos, J. C., et al.
Veröffentlicht: (2024)
Branch-and-price with novel cuts, and a new Stackelberg Security Game
von: Bustamante-Faúndez, Pamela, et al.
Veröffentlicht: (2024)
von: Bustamante-Faúndez, Pamela, et al.
Veröffentlicht: (2024)
Symmetrically Fair Allocations of Indivisible Goods
von: Johnston, Connor, et al.
Veröffentlicht: (2024)
von: Johnston, Connor, et al.
Veröffentlicht: (2024)
Optimal Pricing for Linear-Quadratic Games with Nonlinear Interaction Between Agents
von: Cai, Jiamin, et al.
Veröffentlicht: (2024)
von: Cai, Jiamin, et al.
Veröffentlicht: (2024)
Grace Period is All You Need: Individual Fairness without Revenue Loss in Revenue Management
von: Jaillet, Patrick, et al.
Veröffentlicht: (2024)
von: Jaillet, Patrick, et al.
Veröffentlicht: (2024)
Multidimensional Blockchain Fees are (Essentially) Optimal
von: Angeris, Guillermo, et al.
Veröffentlicht: (2024)
von: Angeris, Guillermo, et al.
Veröffentlicht: (2024)
A Differentially Private Energy Trading Mechanism Approaching Social Optimum
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games
von: Yang, Tong, et al.
Veröffentlicht: (2025) -
Strategizing against Q-learners: A Control-theoretical Approach
von: Arslantas, Yuksel, et al.
Veröffentlicht: (2024) -
Interpretable Price Bounds Estimation with Shape Constraints in Price Optimization
von: Ikeda, Shunnosuke, et al.
Veröffentlicht: (2024) -
Indirect Dynamic Negotiation in the Nash Demand Game
von: Guy, Tatiana V., et al.
Veröffentlicht: (2024) -
A General Equilibrium Theory of Orchestrated AI Agent Systems
von: Garnier, Jean-Philippe
Veröffentlicht: (2026)