QPLEX Decision Processes: Formulation via Nonlinear Markov Chains and Optimization via Policy Gradients
Fuente:
arXiv
Guardado en:
| Autores principales: | Dieker, Antonius B., Hackman, Steven T., Wang, Zitong, Yan, Yunhao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Admission Control for A Single Server Waiting Time Process in Heavy Traffic
por: Xie, Bowen, et al.
Publicado: (2022)
por: Xie, Bowen, et al.
Publicado: (2022)
On the overlap times in queues with dependence under a Farlie-Gumbel-Morgenstern copula
por: Dimitriou, Ioannis
Publicado: (2025)
por: Dimitriou, Ioannis
Publicado: (2025)
On vector-valued functional equations with multiple recursive terms
por: Dimitriou, Ioannis, et al.
Publicado: (2025)
por: Dimitriou, Ioannis, et al.
Publicado: (2025)
Useful stochastic bounds in time-varying queues with service and patience times having general joint distribution
por: Bodas, Shreehari Anand, et al.
Publicado: (2024)
por: Bodas, Shreehari Anand, et al.
Publicado: (2024)
On Markov-dependent reflected autoregressive processes and related models
por: Dimitriou, Ioannis
Publicado: (2024)
por: Dimitriou, Ioannis
Publicado: (2024)
On a class of multiplicative Lindley-type recursions with Markov-modulated dependencies
por: Dimitriou, Ioannis
Publicado: (2025)
por: Dimitriou, Ioannis
Publicado: (2025)
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
por: Goldsztajn, Diego, et al.
Publicado: (2024)
por: Goldsztajn, Diego, et al.
Publicado: (2024)
Queues with Rechargeable Servers
por: Fuentes-Quezada, Eliezer, et al.
Publicado: (2026)
por: Fuentes-Quezada, Eliezer, et al.
Publicado: (2026)
Learning payoffs while routing in skill-based queues
por: van Kempen, Sanne, et al.
Publicado: (2024)
por: van Kempen, Sanne, et al.
Publicado: (2024)
The diffusion approximation of the Multiclass Processor Sharing queue
por: Ghazali, Mohamed, et al.
Publicado: (2025)
por: Ghazali, Mohamed, et al.
Publicado: (2025)
The Supermarket Model on a Dynamic Regular Hypergraph
por: Fernley, John, et al.
Publicado: (2025)
por: Fernley, John, et al.
Publicado: (2025)
New Perspectives on the Erlang-A Queue
por: Daw, Andrew, et al.
Publicado: (2017)
por: Daw, Andrew, et al.
Publicado: (2017)
Dynamics of the leftmost particle in heterogeneous semi-infinite exclusion systems
por: Menshikov, Mikhail, et al.
Publicado: (2026)
por: Menshikov, Mikhail, et al.
Publicado: (2026)
Semi-infinite particle systems with exclusion interaction and heterogeneous jump rates
por: Menshikov, Mikhail, et al.
Publicado: (2024)
por: Menshikov, Mikhail, et al.
Publicado: (2024)
One-dimensional particle clouds with elastic collisions
por: Menshikov, Mikhail, et al.
Publicado: (2025)
por: Menshikov, Mikhail, et al.
Publicado: (2025)
Dynamics of finite inhomogeneous particle systems with exclusion interaction
por: Malyshev, Vadim, et al.
Publicado: (2022)
por: Malyshev, Vadim, et al.
Publicado: (2022)
Dynamic programming for the stochastic matching model on general graphs: the case of the `N-graph'
por: Jean, Loïc, et al.
Publicado: (2024)
por: Jean, Loïc, et al.
Publicado: (2024)
A Queueing Model for the Ambulance Ramping Problem with an Offload Zone
por: Zuk, Josef, et al.
Publicado: (2024)
por: Zuk, Josef, et al.
Publicado: (2024)
Utility maximization in multivariate Volterra models
por: Aichinger, Florian, et al.
Publicado: (2021)
por: Aichinger, Florian, et al.
Publicado: (2021)
Constructive approaches to concentration inequalities with independent random variables
por: Moucer, Celine, et al.
Publicado: (2024)
por: Moucer, Celine, et al.
Publicado: (2024)
QPLEX: A Computational Modeling and Analysis Methodology for Stochastic Systems
por: Dieker, Antonius B., et al.
Publicado: (2025)
por: Dieker, Antonius B., et al.
Publicado: (2025)
Fluid-Limits of Fragmented Limit-Order Markets
por: Muhle-Karbe, Johannes, et al.
Publicado: (2024)
por: Muhle-Karbe, Johannes, et al.
Publicado: (2024)
Optimal local interventions in the two-dimensional Abelian sandpile model
por: de Jongh, Maike C., et al.
Publicado: (2026)
por: de Jongh, Maike C., et al.
Publicado: (2026)
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
por: Bäuerle, Nicole, et al.
Publicado: (2026)
por: Bäuerle, Nicole, et al.
Publicado: (2026)
On Averaging and Extrapolation for Gradient Descent
por: Luner, Alan, et al.
Publicado: (2024)
por: Luner, Alan, et al.
Publicado: (2024)
Multi-floor generalization of TASEP
por: Baryshnikov, Yuliy, et al.
Publicado: (2026)
por: Baryshnikov, Yuliy, et al.
Publicado: (2026)
Partially observed controlled Markov chains and optimal control of the Wonham filter
por: Confortola, Fulvia, et al.
Publicado: (2026)
por: Confortola, Fulvia, et al.
Publicado: (2026)
Many-Server Queueing Systems with Heterogeneous Strategic Servers in Heavy Traffic
por: Büke, Burak, et al.
Publicado: (2022)
por: Büke, Burak, et al.
Publicado: (2022)
A Linear Parameter-Varying Framework for the Analysis of Time-Varying Optimization Algorithms
por: Jakob, Fabian, et al.
Publicado: (2025)
por: Jakob, Fabian, et al.
Publicado: (2025)
Analytic queueing model for ambulance services
por: Pury, Pedro A.
Publicado: (2016)
por: Pury, Pedro A.
Publicado: (2016)
Piecewise Deterministic Sampling for Constrained Distributions
por: Demano, Joël Tatang, et al.
Publicado: (2025)
por: Demano, Joël Tatang, et al.
Publicado: (2025)
Controlled Interacting Branching Diffusion Processes: Relaxed Formulation in the Mean-Field Regime
por: Ocello, Antonio
Publicado: (2023)
por: Ocello, Antonio
Publicado: (2023)
Exponential Conic Optimization for Multi-Regime Service System Design under Congestion and Tail-Risk Control
por: Blanco, Víctor, et al.
Publicado: (2026)
por: Blanco, Víctor, et al.
Publicado: (2026)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
por: Bäuerle, Nicole, et al.
Publicado: (2026)
por: Bäuerle, Nicole, et al.
Publicado: (2026)
Control of parallel non-observable queues: asymptotic equivalence and optimality of periodic policies
por: Anselmi, Jonatha, et al.
Publicado: (2014)
por: Anselmi, Jonatha, et al.
Publicado: (2014)
Catastrophic failure and cumulative damage models involving two types of extended exponential distributions
por: Mohri, Hiroaki, et al.
Publicado: (2021)
por: Mohri, Hiroaki, et al.
Publicado: (2021)
Structure, Analysis, and Synthesis of First-Order Algorithms
por: Miller, Jared, et al.
Publicado: (2026)
por: Miller, Jared, et al.
Publicado: (2026)
Demonstration of effective UCB-based routing in skill-based queues on real-world data
por: van Kempen, Sanne, et al.
Publicado: (2025)
por: van Kempen, Sanne, et al.
Publicado: (2025)
Time-dependent queue length distribution in queues fed by $K$ customers in a finite interval
por: Hayashi, Kaito, et al.
Publicado: (2024)
por: Hayashi, Kaito, et al.
Publicado: (2024)
A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
por: Ortner, Ronald
Publicado: (2024)
por: Ortner, Ronald
Publicado: (2024)
Ejemplares similares
-
Admission Control for A Single Server Waiting Time Process in Heavy Traffic
por: Xie, Bowen, et al.
Publicado: (2022) -
On the overlap times in queues with dependence under a Farlie-Gumbel-Morgenstern copula
por: Dimitriou, Ioannis
Publicado: (2025) -
On vector-valued functional equations with multiple recursive terms
por: Dimitriou, Ioannis, et al.
Publicado: (2025) -
Useful stochastic bounds in time-varying queues with service and patience times having general joint distribution
por: Bodas, Shreehari Anand, et al.
Publicado: (2024) -
On Markov-dependent reflected autoregressive processes and related models
por: Dimitriou, Ioannis
Publicado: (2024)