Efficient Solving of Large Single Input Superstate Decomposable Markovian Decision Process
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mahjoub, Youssef Ait El, Fourneau, Jean-Michel, Alouah, Salma |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A slot-based energy storage decision-making approach for optimal Off-Grid telecommunication operator
von: Mahjoub, Youssef Ait El, et al.
Veröffentlicht: (2024)
von: Mahjoub, Youssef Ait El, et al.
Veröffentlicht: (2024)
PG-Flow: Deterministic implicit policy gradients for geometric product-form queueing networks
von: Mahjoub, Youssef Ait El
Veröffentlicht: (2025)
von: Mahjoub, Youssef Ait El
Veröffentlicht: (2025)
The Gittins Index: A Design Principle for Decision-Making Under Uncertainty
von: Scully, Ziv, et al.
Veröffentlicht: (2025)
von: Scully, Ziv, et al.
Veröffentlicht: (2025)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
von: Dai, Yan, et al.
Veröffentlicht: (2024)
von: Dai, Yan, et al.
Veröffentlicht: (2024)
Optimizing Asynchronous Federated Learning: A Delicate Trade-Off Between Model-Parameter Staleness and Update Frequency
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2025)
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2025)
Score-Aware Policy-Gradient and Performance Guarantees using Local Lyapunov Stability
von: Comte, Céline, et al.
Veröffentlicht: (2023)
von: Comte, Céline, et al.
Veröffentlicht: (2023)
Optimization Trade-offs in Asynchronous Federated Learning: A Stochastic Networks Approach
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2026)
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2026)
Enhanced Innovized Repair Operator for Evolutionary Multi- and Many-objective Optimization
von: Mittal, Sukrit, et al.
Veröffentlicht: (2020)
von: Mittal, Sukrit, et al.
Veröffentlicht: (2020)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
Block Decomposable Methods for Large-Scale Optimization Problems
von: Maia, Leandro Farias
Veröffentlicht: (2026)
von: Maia, Leandro Farias
Veröffentlicht: (2026)
Efficient Algorithms for Robust Markov Decision Processes with $s$-Rectangular Ambiguity Sets
von: Ho, Chin Pang, et al.
Veröffentlicht: (2026)
von: Ho, Chin Pang, et al.
Veröffentlicht: (2026)
MicroHD: An Accuracy-Driven Optimization of Hyperdimensional Computing Algorithms for TinyML systems
von: Ponzina, Flavio, et al.
Veröffentlicht: (2024)
von: Ponzina, Flavio, et al.
Veröffentlicht: (2024)
Gradient-Free Approaches is a Key to an Efficient Interaction with Markovian Stochasticity
von: Prokhorov, Boris, et al.
Veröffentlicht: (2026)
von: Prokhorov, Boris, et al.
Veröffentlicht: (2026)
Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
von: Zhang, Weitong, et al.
Veröffentlicht: (2021)
von: Zhang, Weitong, et al.
Veröffentlicht: (2021)
Weakly Time-Coupled Approximation of Markov Decision Processes
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
Online Markov Decision Processes with Terminal Law Constraints
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
Convex Relaxation for Solving Large-Margin Classifiers in Hyperbolic Space
von: Yang, Sheng, et al.
Veröffentlicht: (2024)
von: Yang, Sheng, et al.
Veröffentlicht: (2024)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Flipping-based Policy for Chance-Constrained Markov Decision Processes
von: Shen, Xun, et al.
Veröffentlicht: (2024)
von: Shen, Xun, et al.
Veröffentlicht: (2024)
Stochastic-Constrained Stochastic Optimization with Markovian Data
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
Non-stationary and Varying-discounting Markov Decision Processes for Reinforcement Learning
von: Chen, Zhizuo, et al.
Veröffentlicht: (2025)
von: Chen, Zhizuo, et al.
Veröffentlicht: (2025)
Achieving Instance-dependent Sample Complexity for Constrained Markov Decision Process
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2024)
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2024)
Towards An Unsupervised Learning Scheme for Efficiently Solving Parameterized Mixed-Integer Programs
von: Qu, Shiyuan, et al.
Veröffentlicht: (2024)
von: Qu, Shiyuan, et al.
Veröffentlicht: (2024)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
On Convergence of Average-Reward Q-Learning in Weakly Communicating Markov Decision Processes
von: Wan, Yi, et al.
Veröffentlicht: (2024)
von: Wan, Yi, et al.
Veröffentlicht: (2024)
Risk-sensitive Markov Decision Process and Learning under General Utility Functions
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
Computing the Bias of Constant-step Stochastic Approximation with Markovian Noise
von: Allmeier, Sebastian, et al.
Veröffentlicht: (2024)
von: Allmeier, Sebastian, et al.
Veröffentlicht: (2024)
Bias and Extrapolation in Markovian Linear Stochastic Approximation with Constant Stepsizes
von: Huo, Dongyan, et al.
Veröffentlicht: (2022)
von: Huo, Dongyan, et al.
Veröffentlicht: (2022)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
A Reliability Theory of Compromise Decisions for Large-Scale Stochastic Programs
von: Diao, Shuotao, et al.
Veröffentlicht: (2024)
von: Diao, Shuotao, et al.
Veröffentlicht: (2024)
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
First Order Methods with Markovian Noise: from Acceleration to Variational Inequalities
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2023)
Ranking and Selection with Simultaneous Input Data Collection
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
von: Haque, Shaan Ul, et al.
Veröffentlicht: (2024)
Decision Machines: Congruent Decision Trees
von: Zhang, Jinxiong
Veröffentlicht: (2021)
von: Zhang, Jinxiong
Veröffentlicht: (2021)
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
High-Probability Bounds for SGD under the Polyak-Lojasiewicz Condition with Markovian Noise
von: Kar, Avik, et al.
Veröffentlicht: (2026)
von: Kar, Avik, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A slot-based energy storage decision-making approach for optimal Off-Grid telecommunication operator
von: Mahjoub, Youssef Ait El, et al.
Veröffentlicht: (2024) -
PG-Flow: Deterministic implicit policy gradients for geometric product-form queueing networks
von: Mahjoub, Youssef Ait El
Veröffentlicht: (2025) -
The Gittins Index: A Design Principle for Decision-Making Under Uncertainty
von: Scully, Ziv, et al.
Veröffentlicht: (2025) -
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
von: Dai, Yan, et al.
Veröffentlicht: (2024) -
Optimizing Asynchronous Federated Learning: A Delicate Trade-Off Between Model-Parameter Staleness and Update Frequency
von: Alahyane, Abdelkrim, et al.
Veröffentlicht: (2025)