Solving robust MDPs as a sequence of static RL problems
Fuente:
arXiv
Saved in:
| Main Authors: | Zouitine, Adil, Geist, Matthieu, Rachelson, Emmanuel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time-Constrained Robust MDPs
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
RRLS : Robust Reinforcement Learning Suite
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
Bootstrapping Expectiles in Reinforcement Learning
by: Clavier, Pierre, et al.
Published: (2024)
by: Clavier, Pierre, et al.
Published: (2024)
Robot Learning: A Tutorial
by: Capuano, Francesco, et al.
Published: (2025)
by: Capuano, Francesco, et al.
Published: (2025)
Learning to Handle Parameter Perturbations in Combinatorial Optimization: an Application to Facility Location
by: Lodi, Andrea, et al.
Published: (2019)
by: Lodi, Andrea, et al.
Published: (2019)
Solving Robust MDPs through No-Regret Dynamics
by: Guha, Etash Kumar
Published: (2023)
by: Guha, Etash Kumar
Published: (2023)
Periodic agent-state based Q-learning for POMDPs
by: Sinha, Amit, et al.
Published: (2024)
by: Sinha, Amit, et al.
Published: (2024)
Convergence of regularized agent-state-based Q-learning in POMDPs
by: Sinha, Amit, et al.
Published: (2025)
by: Sinha, Amit, et al.
Published: (2025)
Efficiently Solving Discounted MDPs with Predictions on Transition Matrices
by: Lyu, Lixing, et al.
Published: (2025)
by: Lyu, Lixing, et al.
Published: (2025)
Towards Minimax Optimality of Model-based Robust Reinforcement Learning
by: Clavier, Pierre, et al.
Published: (2023)
by: Clavier, Pierre, et al.
Published: (2023)
Solving Bayesian inverse problems with diffusion priors and off-policy RL
by: Scimeca, Luca, et al.
Published: (2025)
by: Scimeca, Luca, et al.
Published: (2025)
The Overfocusing Bias of Convolutional Neural Networks: A Saliency-Guided Regularization Approach
by: Bertoin, David, et al.
Published: (2024)
by: Bertoin, David, et al.
Published: (2024)
Closing the Gap between TD Learning and Supervised Learning -- A Generalisation Point of View
by: Ghugare, Raj, et al.
Published: (2024)
by: Ghugare, Raj, et al.
Published: (2024)
End-to-End Efficient RL for Linear Bellman Complete MDPs with Deterministic Transitions
by: Mhammedi, Zakaria, et al.
Published: (2026)
by: Mhammedi, Zakaria, et al.
Published: (2026)
Solving Non-Rectangular Reward-Robust MDPs via Frequency Regularization
by: Gadot, Uri, et al.
Published: (2023)
by: Gadot, Uri, et al.
Published: (2023)
Solving Multi-Model MDPs by Coordinate Ascent and Dynamic Programming
by: Su, Xihong, et al.
Published: (2024)
by: Su, Xihong, et al.
Published: (2024)
Convex Is Back: Solving Belief MDPs With Convexity-Informed Deep Reinforcement Learning
by: Koutas, Daniel, et al.
Published: (2025)
by: Koutas, Daniel, et al.
Published: (2025)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
by: Zhou, Runlin, et al.
Published: (2025)
by: Zhou, Runlin, et al.
Published: (2025)
Provably Efficient RL under Episode-Wise Safety in Constrained MDPs with Linear Function Approximation
by: Kitamura, Toshinori, et al.
Published: (2025)
by: Kitamura, Toshinori, et al.
Published: (2025)
Learning Abstract World Models with a Group-Structured Latent Space
by: Delliaux, Thomas, et al.
Published: (2025)
by: Delliaux, Thomas, et al.
Published: (2025)
A Markov Decision Process for Variable Selection in Branch & Bound
by: Strang, Paul, et al.
Published: (2025)
by: Strang, Paul, et al.
Published: (2025)
Influence branching for learning to solve mixed-integer programs online
by: Strang, Paul, et al.
Published: (2025)
by: Strang, Paul, et al.
Published: (2025)
Planning in Branch-and-Bound: Model-Based Reinforcement Learning for Exact Combinatorial Optimization
by: Strang, Paul, et al.
Published: (2025)
by: Strang, Paul, et al.
Published: (2025)
Learning Equilibria from Data: Provably Efficient Multi-Agent Imitation Learning
by: Freihaut, Till, et al.
Published: (2025)
by: Freihaut, Till, et al.
Published: (2025)
Provably Efficient RL for Linear MDPs under Instantaneous Safety Constraints in Non-Convex Feature Spaces
by: Roknilamouki, Amirhossein, et al.
Published: (2025)
by: Roknilamouki, Amirhossein, et al.
Published: (2025)
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
by: Mak, Hei Yi, et al.
Published: (2024)
by: Mak, Hei Yi, et al.
Published: (2024)
Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion
by: Flet-Berliac, Yannis, et al.
Published: (2024)
by: Flet-Berliac, Yannis, et al.
Published: (2024)
Bring Your Own (Non-Robust) Algorithm to Solve Robust MDPs by Estimating The Worst Kernel
by: Wang, Kaixin, et al.
Published: (2023)
by: Wang, Kaixin, et al.
Published: (2023)
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
DeepAveragers: Offline Reinforcement Learning by Solving Derived Non-Parametric MDPs
by: Shrestha, Aayam, et al.
Published: (2020)
by: Shrestha, Aayam, et al.
Published: (2020)
Exploration by Running Away from the Past
by: Tolguenec, Paul-Antoine Le, et al.
Published: (2024)
by: Tolguenec, Paul-Antoine Le, et al.
Published: (2024)
RL in Latent MDPs is Tractable: Online Guarantees via Off-Policy Evaluation
by: Kwon, Jeongyeol, et al.
Published: (2024)
by: Kwon, Jeongyeol, et al.
Published: (2024)
Leveraging Procedural Generation for Learning Autonomous Peg-in-Hole Assembly in Space
by: Orsula, Andrej, et al.
Published: (2024)
by: Orsula, Andrej, et al.
Published: (2024)
Space Robotics Bench: Robot Learning Beyond Earth
by: Orsula, Andrej, et al.
Published: (2025)
by: Orsula, Andrej, et al.
Published: (2025)
Learning Tool-Aware Adaptive Compliant Control for Autonomous Regolith Excavation
by: Orsula, Andrej, et al.
Published: (2025)
by: Orsula, Andrej, et al.
Published: (2025)
Sim2Dust: Mastering Dynamic Waypoint Tracking on Granular Media
by: Orsula, Andrej, et al.
Published: (2025)
by: Orsula, Andrej, et al.
Published: (2025)
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
by: Shukor, Mustafa, et al.
Published: (2025)
by: Shukor, Mustafa, et al.
Published: (2025)
Capacity-Aware Planning and Scheduling in Budget-Constrained Multi-Agent MDPs: A Meta-RL Approach
by: Vora, Manav, et al.
Published: (2024)
by: Vora, Manav, et al.
Published: (2024)
Multi-agent imitation learning with function approximation: Linear Markov games and beyond
by: Viano, Luca, et al.
Published: (2026)
by: Viano, Luca, et al.
Published: (2026)
MDPs with a State Sensing Cost
by: Kapoor, Vansh, et al.
Published: (2025)
by: Kapoor, Vansh, et al.
Published: (2025)
Similar Items
-
Time-Constrained Robust MDPs
by: Zouitine, Adil, et al.
Published: (2024) -
RRLS : Robust Reinforcement Learning Suite
by: Zouitine, Adil, et al.
Published: (2024) -
Bootstrapping Expectiles in Reinforcement Learning
by: Clavier, Pierre, et al.
Published: (2024) -
Robot Learning: A Tutorial
by: Capuano, Francesco, et al.
Published: (2025) -
Learning to Handle Parameter Perturbations in Combinatorial Optimization: an Application to Facility Location
by: Lodi, Andrea, et al.
Published: (2019)