Solving Long-run Average Reward Robust MDPs via Stochastic Games
Fuente:
arXiv
Saved in:
| Main Authors: | Chatterjee, Krishnendu, Goharshady, Ehsan Kafshdar, Karrabi, Mehrdad, Novotný, Petr, Žikelić, Đorđe |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automated Approach for Solving Infinite-state Polynomial Reachability Games
by: Chatterjee, Krishnendu, et al.
Published: (2026)
by: Chatterjee, Krishnendu, et al.
Published: (2026)
Qualitative Analysis of $ω$-Regular Objectives on Robust MDPs
by: Asadi, Ali, et al.
Published: (2025)
by: Asadi, Ali, et al.
Published: (2025)
Quantified Linear and Polynomial Arithmetic Satisfiability via Template-based Skolemization
by: Chatterjee, Krishnendu, et al.
Published: (2024)
by: Chatterjee, Krishnendu, et al.
Published: (2024)
Sound and Complete Witnesses for Template-based Verification of LTL Properties on Polynomial Programs
by: Chatterjee, Krishnendu, et al.
Published: (2024)
by: Chatterjee, Krishnendu, et al.
Published: (2024)
Refuting Equivalence in Probabilistic Programs with Conditioning
by: Chatterjee, Krishnendu, et al.
Published: (2025)
by: Chatterjee, Krishnendu, et al.
Published: (2025)
Equivalence and Similarity Refutation for Probabilistic Programs
by: Chatterjee, Krishnendu, et al.
Published: (2024)
by: Chatterjee, Krishnendu, et al.
Published: (2024)
SuperDP: Differential Privacy Refutation via Supermartingales
by: Chatterjee, Krishnendu, et al.
Published: (2026)
by: Chatterjee, Krishnendu, et al.
Published: (2026)
PolyQEnt: A Polynomial Quantified Entailment Solver
by: Chatterjee, Krishnendu, et al.
Published: (2024)
by: Chatterjee, Krishnendu, et al.
Published: (2024)
Strongly Polynomial Time Complexity of Policy Iteration for $L_\infty$ Robust MDPs
by: Asadi, Ali, et al.
Published: (2026)
by: Asadi, Ali, et al.
Published: (2026)
Certified Policy Verification and Synthesis for MDPs under Distributional Reach-avoidance Properties
by: Akshay, S., et al.
Published: (2024)
by: Akshay, S., et al.
Published: (2024)
Fully Automated Selfish Mining Analysis in Efficient Proof Systems Blockchains
by: Chatterjee, Krishnendu, et al.
Published: (2024)
by: Chatterjee, Krishnendu, et al.
Published: (2024)
Certificate-Guided Evaluation of Reinforcement Learning Generalization
by: Subramanian, Vignesh, et al.
Published: (2026)
by: Subramanian, Vignesh, et al.
Published: (2026)
Neural Control and Certificate Repair via Runtime Monitoring
by: Yu, Emily, et al.
Published: (2024)
by: Yu, Emily, et al.
Published: (2024)
Game Dynamics and Equilibrium Computation in the Population Protocol Model
by: Alistarh, Dan, et al.
Published: (2023)
by: Alistarh, Dan, et al.
Published: (2023)
Automating the Refinement of Reinforcement Learning Specifications
by: Ambadkar, Tanmay, et al.
Published: (2025)
by: Ambadkar, Tanmay, et al.
Published: (2025)
Omega-regular Verification and Control for Distributional Specifications in MDPs
by: Akshay, S., et al.
Published: (2025)
by: Akshay, S., et al.
Published: (2025)
A Computationally Efficient Algorithm for Infinite-Horizon Average-Reward Linear MDPs
by: Hong, Kihyuk, et al.
Published: (2025)
by: Hong, Kihyuk, et al.
Published: (2025)
Bidding Games with Charging
by: Avni, Guy, et al.
Published: (2024)
by: Avni, Guy, et al.
Published: (2024)
DeepAveragers: Offline Reinforcement Learning by Solving Derived Non-Parametric MDPs
by: Shrestha, Aayam, et al.
Published: (2020)
by: Shrestha, Aayam, et al.
Published: (2020)
Global Convergence for Average Reward Constrained MDPs with Primal-Dual Actor Critic Algorithm
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Principal-Agent Reward Shaping in MDPs
by: Ben-Porat, Omer, et al.
Published: (2023)
by: Ben-Porat, Omer, et al.
Published: (2023)
Predictive Monitoring of Black-Box Dynamical Systems
by: Henzinger, Thomas A., et al.
Published: (2024)
by: Henzinger, Thomas A., et al.
Published: (2024)
Learning General Parameterized Policies for Infinite Horizon Average Reward Constrained MDPs via Primal-Dual Policy Gradient Algorithm
by: Bai, Qinbo, et al.
Published: (2024)
by: Bai, Qinbo, et al.
Published: (2024)
Preserving the Privacy of Reward Functions in MDPs through Deception
by: Chirra, Shashank Reddy, et al.
Published: (2024)
by: Chirra, Shashank Reddy, et al.
Published: (2024)
Taming Infinity one Chunk at a Time: Concisely Represented Strategies in One-Counter MDPs
by: Ajdarów, Michal, et al.
Published: (2025)
by: Ajdarów, Michal, et al.
Published: (2025)
ACPO: A Policy Optimization Algorithm for Average MDPs with Constraints
by: Agnihotri, Akhil, et al.
Published: (2023)
by: Agnihotri, Akhil, et al.
Published: (2023)
Beyond Scalar Rewards: An Axiomatic Framework for Lexicographic MDPs
by: Shakerinava, Mehran, et al.
Published: (2025)
by: Shakerinava, Mehran, et al.
Published: (2025)
Bidding Games on Markov Decision Processes with Quantitative Reachability Objectives
by: Avni, Guy, et al.
Published: (2024)
by: Avni, Guy, et al.
Published: (2024)
Ensuring Safety in an Uncertain Environment: Constrained MDPs via Stochastic Thresholds
by: Zuo, Qian, et al.
Published: (2025)
by: Zuo, Qian, et al.
Published: (2025)
Solving Multi-Model MDPs by Coordinate Ascent and Dynamic Programming
by: Su, Xihong, et al.
Published: (2024)
by: Su, Xihong, et al.
Published: (2024)
Threshold UCT: Cost-Constrained Monte Carlo Tree Search with Pareto Curves
by: Kurečka, Martin, et al.
Published: (2024)
by: Kurečka, Martin, et al.
Published: (2024)
Efficient Solution and Learning of Robust Factored MDPs
by: Schnitzer, Yannik, et al.
Published: (2025)
by: Schnitzer, Yannik, et al.
Published: (2025)
Lower Bound on Howard Policy Iteration for Deterministic Markov Decision Processes
by: Asadi, Ali, et al.
Published: (2025)
by: Asadi, Ali, et al.
Published: (2025)
Efficient $Q$-Learning and Actor-Critic Methods for Robust Average Reward Reinforcement Learning
by: Xu, Yang, et al.
Published: (2025)
by: Xu, Yang, et al.
Published: (2025)
Average-Reward Soft Actor-Critic
by: Adamczyk, Jacob, et al.
Published: (2025)
by: Adamczyk, Jacob, et al.
Published: (2025)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026)
by: Muni, Aneri, et al.
Published: (2026)
WARP: On the Benefits of Weight Averaged Rewarded Policies
by: Ramé, Alexandre, et al.
Published: (2024)
by: Ramé, Alexandre, et al.
Published: (2024)
Relevance-Zone Reduction in Game Solving
by: Lin, Chi-Huang, et al.
Published: (2025)
by: Lin, Chi-Huang, et al.
Published: (2025)
Average Reward Reinforcement Learning for Omega-Regular and Mean-Payoff Objectives
by: Kazemi, Milad, et al.
Published: (2025)
by: Kazemi, Milad, et al.
Published: (2025)
No-Regret Strategy Solving in Imperfect-Information Games via Pre-Trained Embedding
by: Fu, Yanchang, et al.
Published: (2025)
by: Fu, Yanchang, et al.
Published: (2025)
Similar Items
-
Automated Approach for Solving Infinite-state Polynomial Reachability Games
by: Chatterjee, Krishnendu, et al.
Published: (2026) -
Qualitative Analysis of $ω$-Regular Objectives on Robust MDPs
by: Asadi, Ali, et al.
Published: (2025) -
Quantified Linear and Polynomial Arithmetic Satisfiability via Template-based Skolemization
by: Chatterjee, Krishnendu, et al.
Published: (2024) -
Sound and Complete Witnesses for Template-based Verification of LTL Properties on Polynomial Programs
by: Chatterjee, Krishnendu, et al.
Published: (2024) -
Refuting Equivalence in Probabilistic Programs with Conditioning
by: Chatterjee, Krishnendu, et al.
Published: (2025)