Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
Fuente:
arXiv
Guardado en:
| Autores principales: | Luo, Baiting, Zhang, Yunuo, Dubey, Abhishek, Mukhopadhyay, Ayan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
por: Zhang, Yunuo, et al.
Publicado: (2025)
por: Zhang, Yunuo, et al.
Publicado: (2025)
Decision Making in Non-Stationary Environments with Policy-Augmented Search
por: Pettet, Ava, et al.
Publicado: (2024)
por: Pettet, Ava, et al.
Publicado: (2024)
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
por: Luo, Baiting, et al.
Publicado: (2025)
por: Luo, Baiting, et al.
Publicado: (2025)
NS-Gym: Open-Source Simulation Environments and Benchmarks for Non-Stationary Markov Decision Processes
por: Keplinger, Nathaniel S., et al.
Publicado: (2025)
por: Keplinger, Nathaniel S., et al.
Publicado: (2025)
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
por: Zhang, Yunuo, et al.
Publicado: (2025)
por: Zhang, Yunuo, et al.
Publicado: (2025)
In-Context Planning with Latent Temporal Abstractions
por: Luo, Baiting, et al.
Publicado: (2026)
por: Luo, Baiting, et al.
Publicado: (2026)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
por: Amiri, Mohsen, et al.
Publicado: (2025)
por: Amiri, Mohsen, et al.
Publicado: (2025)
Online Decision-Making Under Uncertainty for Vehicle-to-Building Systems
por: Sen, Rishav, et al.
Publicado: (2026)
por: Sen, Rishav, et al.
Publicado: (2026)
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
por: Morimura, Tetsuro, et al.
Publicado: (2022)
por: Morimura, Tetsuro, et al.
Publicado: (2022)
Optimal Decision Tree Policies for Markov Decision Processes
por: Vos, Daniël, et al.
Publicado: (2023)
por: Vos, Daniël, et al.
Publicado: (2023)
Dual Formulation for Non-Rectangular Lp Robust Markov Decision Processes
por: Kumar, Navdeep, et al.
Publicado: (2025)
por: Kumar, Navdeep, et al.
Publicado: (2025)
Markov Decision Processes under External Temporal Processes
por: Ayyagari, Ranga Shaarad, et al.
Publicado: (2023)
por: Ayyagari, Ranga Shaarad, et al.
Publicado: (2023)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
por: Vora, Kevin, et al.
Publicado: (2025)
por: Vora, Kevin, et al.
Publicado: (2025)
Policy Gradient for Robust Markov Decision Processes
por: Wang, Qiuhao, et al.
Publicado: (2024)
por: Wang, Qiuhao, et al.
Publicado: (2024)
Think Before You Act: Decision Transformers with Working Memory
por: Kang, Jikun, et al.
Publicado: (2023)
por: Kang, Jikun, et al.
Publicado: (2023)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
por: Moon, Sang Bin, et al.
Publicado: (2024)
por: Moon, Sang Bin, et al.
Publicado: (2024)
Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
por: Sanokowski, Sebastian, et al.
Publicado: (2025)
Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision Processes
por: Infante, Guillermo, et al.
Publicado: (2021)
por: Infante, Guillermo, et al.
Publicado: (2021)
Almost Sure Convergence of Differential Temporal Difference Learning for Average Reward Markov Decision Processes
por: Blaser, Ethan, et al.
Publicado: (2026)
por: Blaser, Ethan, et al.
Publicado: (2026)
Solving Robust Markov Decision Processes: Generic, Reliable, Efficient
por: Meggendorfer, Tobias, et al.
Publicado: (2024)
por: Meggendorfer, Tobias, et al.
Publicado: (2024)
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes
por: Infante, Guillermo, et al.
Publicado: (2024)
por: Infante, Guillermo, et al.
Publicado: (2024)
SPOT: Scalable Policy Optimization with Trees for Markov Decision Processes
por: Xiong, Xuyuan, et al.
Publicado: (2025)
por: Xiong, Xuyuan, et al.
Publicado: (2025)
Reinforcement Learning-based Approach for Vehicle-to-Building Charging with Heterogeneous Agents and Long Term Rewards
por: Liu, Fangqi, et al.
Publicado: (2025)
por: Liu, Fangqi, et al.
Publicado: (2025)
Linear Mixture Distributionally Robust Markov Decision Processes
por: Liu, Zhishuai, et al.
Publicado: (2025)
por: Liu, Zhishuai, et al.
Publicado: (2025)
Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making
por: Feng, Fan, et al.
Publicado: (2026)
por: Feng, Fan, et al.
Publicado: (2026)
OCMDP: Observation-Constrained Markov Decision Process
por: Wang, Taiyi, et al.
Publicado: (2024)
por: Wang, Taiyi, et al.
Publicado: (2024)
A Cantor-Kantorovich Metric Between Markov Decision Processes with Application to Transfer Learning
por: Banse, Adrien, et al.
Publicado: (2024)
por: Banse, Adrien, et al.
Publicado: (2024)
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
por: Bennett, Andrew, et al.
Publicado: (2024)
por: Bennett, Andrew, et al.
Publicado: (2024)
Homomorphic Mappings for Value-Preserving State Aggregation in Markov Decision Processes
por: Zhao, Shuo, et al.
Publicado: (2025)
por: Zhao, Shuo, et al.
Publicado: (2025)
A Unified Theory of Compositionality, Modularity, and Interpretability in Markov Decision Processes
por: Ringstrom, Thomas J., et al.
Publicado: (2025)
por: Ringstrom, Thomas J., et al.
Publicado: (2025)
Forecasting and Mitigating Disruptions in Public Bus Transit Services
por: Han, Chaeeun, et al.
Publicado: (2024)
por: Han, Chaeeun, et al.
Publicado: (2024)
Surrogate Interpretable Graph for Random Decision Forests
por: Dubey, Akshat, et al.
Publicado: (2025)
por: Dubey, Akshat, et al.
Publicado: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
por: Foffano, Daniele, et al.
Publicado: (2023)
por: Foffano, Daniele, et al.
Publicado: (2023)
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes
por: Ireland, David, et al.
Publicado: (2024)
por: Ireland, David, et al.
Publicado: (2024)
MATE: Solving Contextual Markov Decision Processes with Memory of Accumulated Transition Embeddings
por: Hwang, Himchan, et al.
Publicado: (2026)
por: Hwang, Himchan, et al.
Publicado: (2026)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
por: Gu, Jingwen, et al.
Publicado: (2025)
por: Gu, Jingwen, et al.
Publicado: (2025)
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
Decision Transformer vs. Decision Mamba: Analysing the Complexity of Sequential Decision Making in Atari Games
por: Yan, Ke
Publicado: (2024)
por: Yan, Ke
Publicado: (2024)
Unveiling the Decision-Making Process in Reinforcement Learning with Genetic Programming
por: Eberhardinger, Manuel, et al.
Publicado: (2024)
por: Eberhardinger, Manuel, et al.
Publicado: (2024)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
por: Zimmer, Matthieu, et al.
Publicado: (2025)
por: Zimmer, Matthieu, et al.
Publicado: (2025)
Ejemplares similares
-
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
por: Zhang, Yunuo, et al.
Publicado: (2025) -
Decision Making in Non-Stationary Environments with Policy-Augmented Search
por: Pettet, Ava, et al.
Publicado: (2024) -
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
por: Luo, Baiting, et al.
Publicado: (2025) -
NS-Gym: Open-Source Simulation Environments and Benchmarks for Non-Stationary Markov Decision Processes
por: Keplinger, Nathaniel S., et al.
Publicado: (2025) -
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
por: Zhang, Yunuo, et al.
Publicado: (2025)