Relaxed Equilibria for Time-Inconsistent Markov Decision Processes
Fuente:
arXiv
Guardado en:
| Autores principales: | Bayraktar, Erhan, Huang, Yu-Jui, Wang, Zhenhua, Zhou, Zhou |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Uniform-in-Time Convergence Rates to a Nonlinear Markov Chain for Mean-Field Interacting Jump Processes
por: Cohen, Asaf, et al.
Publicado: (2025)
por: Cohen, Asaf, et al.
Publicado: (2025)
Markov-bridge representation of ergodic large-deviation principles
por: Renger, D. R. Michiel
Publicado: (2024)
por: Renger, D. R. Michiel
Publicado: (2024)
Information-theoretic minimax and submodular optimization algorithms for multivariate Markov chains
por: Lai, Zheyuan, et al.
Publicado: (2025)
por: Lai, Zheyuan, et al.
Publicado: (2025)
Equilibrium for Time-inconsistent Mean Field Games: A Systematic Analysis by Entropy Regularization
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Housing Decisions under Mobility Risk: A Stochastic Threshold Approach
por: Wu, Hui
Publicado: (2026)
por: Wu, Hui
Publicado: (2026)
Dynkin ghost games with asymmetry and consolation
por: Ekström, Erik, et al.
Publicado: (2024)
por: Ekström, Erik, et al.
Publicado: (2024)
Embedding of reversible Markov matrices
por: Baake, Ellen, et al.
Publicado: (2025)
por: Baake, Ellen, et al.
Publicado: (2025)
Controlled Interacting Branching Diffusion Processes: Relaxed Formulation in the Mean-Field Regime
por: Ocello, Antonio
Publicado: (2023)
por: Ocello, Antonio
Publicado: (2023)
Markov chain entropy games and the geometry of their Nash equilibria
por: Choi, Michael C. H., et al.
Publicado: (2023)
por: Choi, Michael C. H., et al.
Publicado: (2023)
Stackelberg stopping games
por: Zhang, Jingjie, et al.
Publicado: (2025)
por: Zhang, Jingjie, et al.
Publicado: (2025)
Measurized Markov Decision Processes
por: Adelman, Daniel, et al.
Publicado: (2024)
por: Adelman, Daniel, et al.
Publicado: (2024)
Multi-agent learning under uncertainty: Recurrence vs. concentration
por: Lotidis, Kyriakos, et al.
Publicado: (2025)
por: Lotidis, Kyriakos, et al.
Publicado: (2025)
Information-theoretic classification of the cutoff phenomenon in Markov processes
por: Wang, Youjia, et al.
Publicado: (2024)
por: Wang, Youjia, et al.
Publicado: (2024)
Geometry and factorization of multivariate Markov chains with applications to MCMC acceleration and approximate inference
por: Choi, Michael C. H., et al.
Publicado: (2024)
por: Choi, Michael C. H., et al.
Publicado: (2024)
Nonlocal Stochastic Optimal Control for Diffusion Processes: Existence, Maximum Principle and Financial Applications
por: Anita, Stefana-Lucia, et al.
Publicado: (2025)
por: Anita, Stefana-Lucia, et al.
Publicado: (2025)
On the saddle point of a zero-sum stopper vs. singular-controller game
por: Bovo, Andrea, et al.
Publicado: (2024)
por: Bovo, Andrea, et al.
Publicado: (2024)
Representation and Characterization of Quasistationary Distributions for Markov Chains
por: Ben-Ari, Iddo, et al.
Publicado: (2024)
por: Ben-Ari, Iddo, et al.
Publicado: (2024)
A new criterion for recurrence of Markov chains with an infinitely countable set of states
por: Abramov, Vyacheslav M.
Publicado: (2024)
por: Abramov, Vyacheslav M.
Publicado: (2024)
Mean-Field Langevin Diffusions with Density-dependent Temperature
por: Huang, Yu-Jui, et al.
Publicado: (2025)
por: Huang, Yu-Jui, et al.
Publicado: (2025)
A Framework for Exploring Social Interactions in Multiagent Decision-Making for Two-Queue Systems
por: Gaspard, Mallory E., et al.
Publicado: (2026)
por: Gaspard, Mallory E., et al.
Publicado: (2026)
A Tokenized Sovereign Debt Conversion Mechanism for Dynamic Public Debt Reduction
por: Firouzi, Kiarash
Publicado: (2025)
por: Firouzi, Kiarash
Publicado: (2025)
A multiple coupon collection process and its Markov embedding structure
por: Baake, Ellen, et al.
Publicado: (2024)
por: Baake, Ellen, et al.
Publicado: (2024)
Inverse Problems for Ergodicity of Markov Chains
por: Wei, Zhi-Feng
Publicado: (2020)
por: Wei, Zhi-Feng
Publicado: (2020)
Reducible Markov modulation, pole order, and tail behavior in random growth models
por: Beare, Brendan K., et al.
Publicado: (2024)
por: Beare, Brendan K., et al.
Publicado: (2024)
On the time consistent solution to optimal stopping problems with expectation constraint
por: Christensen, Sören, et al.
Publicado: (2023)
por: Christensen, Sören, et al.
Publicado: (2023)
$Γ$-expansion of the measure-current large deviations rate functional of non-reversible finite-state Markov chains
por: Kim, Seonwoo, et al.
Publicado: (2024)
por: Kim, Seonwoo, et al.
Publicado: (2024)
Tails of explosive birth processes and applications to non-linear P ólya urns
por: Gottfried, Thomas, et al.
Publicado: (2024)
por: Gottfried, Thomas, et al.
Publicado: (2024)
On additive averaging kernels for finite Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
A New Approach for the Continuous Time Kyle-Back Strategic Insider Equilibrium Problem
por: Qiao, Bixing, et al.
Publicado: (2025)
por: Qiao, Bixing, et al.
Publicado: (2025)
Controlled Interacting Branching Diffusion Processes: A Viscosity Approach
por: Ocello, Antonio
Publicado: (2026)
por: Ocello, Antonio
Publicado: (2026)
Long-Term Average Impulse Control with Mean Field Interactions
por: Helmes, K. L., et al.
Publicado: (2025)
por: Helmes, K. L., et al.
Publicado: (2025)
Continuous-time mean field Markov decision models
por: Bäuerle, Nicole, et al.
Publicado: (2023)
por: Bäuerle, Nicole, et al.
Publicado: (2023)
Improving the convergence of Markov chains via permutations and projections
por: Choi, Michael C. H., et al.
Publicado: (2024)
por: Choi, Michael C. H., et al.
Publicado: (2024)
Optimising two-block averaging kernels to speed up Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
Role of Intra-specific Competition and Additional Food on Prey-Predator Systems exhibiting Holling Type-IV Functional Response
por: Prakash, D Bhanu, et al.
Publicado: (2025)
por: Prakash, D Bhanu, et al.
Publicado: (2025)
Convergence for linear quadratic potential mean field games
por: Cecchin, Alekos, et al.
Publicado: (2026)
por: Cecchin, Alekos, et al.
Publicado: (2026)
Localization for constrained martingale problems and optimal conditions for uniqueness of reflecting diffusions in 2-dimensional domains
por: Costantini, Cristina, et al.
Publicado: (2022)
por: Costantini, Cristina, et al.
Publicado: (2022)
Splitting infinity: a de Finetti game with state-dependent profit rates and singular control for diffusions
por: Chlebicki, Piotr, et al.
Publicado: (2025)
por: Chlebicki, Piotr, et al.
Publicado: (2025)
On the Mean-Field limit of diffusive games through the master equation: $L^{\infty}$ estimates and extreme value behavior
por: Bayraktar, Erhan, et al.
Publicado: (2024)
por: Bayraktar, Erhan, et al.
Publicado: (2024)
Covering a graph with independent walks
por: Hermon, Jonathan, et al.
Publicado: (2021)
por: Hermon, Jonathan, et al.
Publicado: (2021)
Ejemplares similares
-
Uniform-in-Time Convergence Rates to a Nonlinear Markov Chain for Mean-Field Interacting Jump Processes
por: Cohen, Asaf, et al.
Publicado: (2025) -
Markov-bridge representation of ergodic large-deviation principles
por: Renger, D. R. Michiel
Publicado: (2024) -
Information-theoretic minimax and submodular optimization algorithms for multivariate Markov chains
por: Lai, Zheyuan, et al.
Publicado: (2025) -
Equilibrium for Time-inconsistent Mean Field Games: A Systematic Analysis by Entropy Regularization
por: Bayraktar, Erhan, et al.
Publicado: (2026) -
Housing Decisions under Mobility Risk: A Stochastic Threshold Approach
por: Wu, Hui
Publicado: (2026)