Saved in:
| Main Authors: | Strang, Paul, Alès, Zacharie, Bissuel, Côme, Juan, Olivier, Kedad-Sidhoum, Safia, Rachelson, Emmanuel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.19348 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Planning in Branch-and-Bound: Model-Based Reinforcement Learning for Exact Combinatorial Optimization
by: Strang, Paul, et al.
Published: (2025)
by: Strang, Paul, et al.
Published: (2025)
Influence branching for learning to solve mixed-integer programs online
by: Strang, Paul, et al.
Published: (2025)
by: Strang, Paul, et al.
Published: (2025)
Finite adaptability in two-stage robust optimization: asymptotic optimality and tractability
by: Kedad-Sidhoum, Safia, et al.
Published: (2023)
by: Kedad-Sidhoum, Safia, et al.
Published: (2023)
Solving robust MDPs as a sequence of static RL problems
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
Learning to Handle Parameter Perturbations in Combinatorial Optimization: an Application to Facility Location
by: Lodi, Andrea, et al.
Published: (2019)
by: Lodi, Andrea, et al.
Published: (2019)
Performance Improvement Bounds for Lipschitz Configurable Markov Decision Processes
by: Metelli, Alberto Maria
Published: (2024)
by: Metelli, Alberto Maria
Published: (2024)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
by: Moon, Sang Bin, et al.
Published: (2024)
by: Moon, Sang Bin, et al.
Published: (2024)
RRLS : Robust Reinforcement Learning Suite
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
Time-Constrained Robust MDPs
by: Zouitine, Adil, et al.
Published: (2024)
by: Zouitine, Adil, et al.
Published: (2024)
Bootstrapping Expectiles in Reinforcement Learning
by: Clavier, Pierre, et al.
Published: (2024)
by: Clavier, Pierre, et al.
Published: (2024)
Monitored Markov Decision Processes
by: Parisi, Simone, et al.
Published: (2024)
by: Parisi, Simone, et al.
Published: (2024)
Exploration by Running Away from the Past
by: Tolguenec, Paul-Antoine Le, et al.
Published: (2024)
by: Tolguenec, Paul-Antoine Le, et al.
Published: (2024)
Reinforcement Learning for Node Selection in Branch-and-Bound
by: Mattick, Alexander, et al.
Published: (2023)
by: Mattick, Alexander, et al.
Published: (2023)
A Unified Theory of Compositionality, Modularity, and Interpretability in Markov Decision Processes
by: Ringstrom, Thomas J., et al.
Published: (2025)
by: Ringstrom, Thomas J., et al.
Published: (2025)
Generalized Linear Markov Decision Process
by: Zhang, Sinian, et al.
Published: (2025)
by: Zhang, Sinian, et al.
Published: (2025)
Federated Control in Markov Decision Processes
by: Jin, Hao, et al.
Published: (2024)
by: Jin, Hao, et al.
Published: (2024)
Interaction-Grounded Learning for Contextual Markov Decision Processes with Personalized Feedback
by: Zhang, Mengxiao, et al.
Published: (2026)
by: Zhang, Mengxiao, et al.
Published: (2026)
Learning in Markov Decision Processes with Exogenous Dynamics
by: Maran, Davide, et al.
Published: (2026)
by: Maran, Davide, et al.
Published: (2026)
Multi-RF Fusion with Multi-GNN Blending for Molecular Property Prediction
by: Bugaud, Zacharie
Published: (2026)
by: Bugaud, Zacharie
Published: (2026)
Optimal Decision Tree Policies for Markov Decision Processes
by: Vos, Daniël, et al.
Published: (2023)
by: Vos, Daniël, et al.
Published: (2023)
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025)
by: Ariu, Kaito, et al.
Published: (2025)
Clustering risk in Non-parametric Hidden Markov and I.I.D. Models
by: Gassiat, Elisabeth, et al.
Published: (2023)
by: Gassiat, Elisabeth, et al.
Published: (2023)
Markov Decision Processes under External Temporal Processes
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021)
by: Angelotti, Giorgio, et al.
Published: (2021)
The regret lower bound for communicating Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
An Orthogonal Learner for Individualized Outcomes in Markov Decision Processes
by: Javurek, Emil, et al.
Published: (2025)
by: Javurek, Emil, et al.
Published: (2025)
Initial Distribution Sensitivity of Constrained Markov Decision Processes
by: Tercan, Alperen, et al.
Published: (2025)
by: Tercan, Alperen, et al.
Published: (2025)
Improving Controller Generalization with Dimensionless Markov Decision Processes
by: Charvet, Valentin, et al.
Published: (2025)
by: Charvet, Valentin, et al.
Published: (2025)
Model-Based Exploration in Monitored Markov Decision Processes
by: Kazemipour, Alireza, et al.
Published: (2025)
by: Kazemipour, Alireza, et al.
Published: (2025)
Horizon-Free Regret for Linear Markov Decision Processes
by: Zhang, Zihan, et al.
Published: (2024)
by: Zhang, Zihan, et al.
Published: (2024)
Learning Utilities from Demonstrations in Markov Decision Processes
by: Lazzati, Filippo, et al.
Published: (2024)
by: Lazzati, Filippo, et al.
Published: (2024)
Achieving Constant Regret in Linear Markov Decision Processes
by: Zhang, Weitong, et al.
Published: (2024)
by: Zhang, Weitong, et al.
Published: (2024)
Concentration of Cumulative Reward in Markov Decision Processes
by: Sayedana, Borna, et al.
Published: (2024)
by: Sayedana, Borna, et al.
Published: (2024)
Policy Gradient for Robust Markov Decision Processes
by: Wang, Qiuhao, et al.
Published: (2024)
by: Wang, Qiuhao, et al.
Published: (2024)
An upper bound on the silhouette evaluation metric for clustering
by: Sträng, Hugo, et al.
Published: (2025)
by: Sträng, Hugo, et al.
Published: (2025)
Learning Abstract World Models with a Group-Structured Latent Space
by: Delliaux, Thomas, et al.
Published: (2025)
by: Delliaux, Thomas, et al.
Published: (2025)
A Theoretical Analysis of State Similarity Between Markov Decision Processes
by: Tao, Zhenyu, et al.
Published: (2025)
by: Tao, Zhenyu, et al.
Published: (2025)
No-Regret Thompson Sampling for Finite-Horizon Markov Decision Processes with Gaussian Processes
by: Bayrooti, Jasmine, et al.
Published: (2025)
by: Bayrooti, Jasmine, et al.
Published: (2025)
Transition Transfer $Q$-Learning for Composite Markov Decision Processes
by: Chai, Jinhang, et al.
Published: (2025)
by: Chai, Jinhang, et al.
Published: (2025)
Logarithmic Regret of Exploration in Average Reward Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
Similar Items
-
Planning in Branch-and-Bound: Model-Based Reinforcement Learning for Exact Combinatorial Optimization
by: Strang, Paul, et al.
Published: (2025) -
Influence branching for learning to solve mixed-integer programs online
by: Strang, Paul, et al.
Published: (2025) -
Finite adaptability in two-stage robust optimization: asymptotic optimality and tractability
by: Kedad-Sidhoum, Safia, et al.
Published: (2023) -
Solving robust MDPs as a sequence of static RL problems
by: Zouitine, Adil, et al.
Published: (2024) -
Learning to Handle Parameter Perturbations in Combinatorial Optimization: an Application to Facility Location
by: Lodi, Andrea, et al.
Published: (2019)