Bayesian learning of the optimal action-value function in a Markov decision process
Fuente:
arXiv
Salvato in:
| Autori principali: | Guo, Jiaqi, Ho, Chon Wai, Singh, Sumeetpal S. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bayesian Deep Learning with Multilevel Trace-class Neural Networks
di: Chada, Neil K., et al.
Pubblicazione: (2022)
di: Chada, Neil K., et al.
Pubblicazione: (2022)
Asymptotically optimal regret in communicating Markov decision processes
di: Boone, Victor
Pubblicazione: (2025)
di: Boone, Victor
Pubblicazione: (2025)
Spatio-temporal modeling and forecasting with Fourier neural operators
di: Nag, Pratik, et al.
Pubblicazione: (2026)
di: Nag, Pratik, et al.
Pubblicazione: (2026)
Relaxed Gaussian process interpolation: a goal-oriented approach to Bayesian optimization
di: Petit, Sébastien, et al.
Pubblicazione: (2022)
di: Petit, Sébastien, et al.
Pubblicazione: (2022)
Planning in entropy-regularized Markov decision processes and games
di: Grill, Jean-Bastien, et al.
Pubblicazione: (2026)
di: Grill, Jean-Bastien, et al.
Pubblicazione: (2026)
Average-reward reinforcement learning in semi-Markov decision processes via relative value iteration
di: Yu, Huizhen, et al.
Pubblicazione: (2025)
di: Yu, Huizhen, et al.
Pubblicazione: (2025)
Respecting the limit:Bayesian optimization with a bound on the optimal value
di: Wang, Hanyang, et al.
Pubblicazione: (2024)
di: Wang, Hanyang, et al.
Pubblicazione: (2024)
Nonparametric learning of covariate-based Markov jump processes using RKHS techniques
di: Han, Yuchen, et al.
Pubblicazione: (2025)
di: Han, Yuchen, et al.
Pubblicazione: (2025)
Bilevel optimization for learning hyperparameters: Application to solving PDEs and inverse problems with Gaussian processes
di: Nelsen, Nicholas H., et al.
Pubblicazione: (2025)
di: Nelsen, Nicholas H., et al.
Pubblicazione: (2025)
A safe exploration approach to constrained Markov decision processes
di: Ni, Tingting, et al.
Pubblicazione: (2023)
di: Ni, Tingting, et al.
Pubblicazione: (2023)
Bayesian preference elicitation for decision support in multiobjective optimization
di: Huber, Felix, et al.
Pubblicazione: (2025)
di: Huber, Felix, et al.
Pubblicazione: (2025)
Stochastic first-order methods for average-reward Markov decision processes
di: Li, Tianjiao, et al.
Pubblicazione: (2022)
di: Li, Tianjiao, et al.
Pubblicazione: (2022)
A Bayesian explanation of machine learning models based on modes and functional ANOVA
di: Long, Quan
Pubblicazione: (2024)
di: Long, Quan
Pubblicazione: (2024)
Approximate learning of parsimonious Bayesian context trees
di: Ghani, Daniyar, et al.
Pubblicazione: (2024)
di: Ghani, Daniyar, et al.
Pubblicazione: (2024)
Fast post-process Bayesian inference with Variational Sparse Bayesian Quadrature
di: Li, Chengkun, et al.
Pubblicazione: (2023)
di: Li, Chengkun, et al.
Pubblicazione: (2023)
Towards a turnkey approach to unbiased Monte Carlo estimation of smooth functions of expectations
di: Chopin, Nicolas, et al.
Pubblicazione: (2024)
di: Chopin, Nicolas, et al.
Pubblicazione: (2024)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
di: Gao, Xuefeng, et al.
Pubblicazione: (2022)
di: Gao, Xuefeng, et al.
Pubblicazione: (2022)
Comparison of parallel SMC and MCMC for Bayesian deep learning
di: Liang, Xinzhu, et al.
Pubblicazione: (2024)
di: Liang, Xinzhu, et al.
Pubblicazione: (2024)
Gaussian process learning of nonlinear dynamics
di: Ye, Dongwei, et al.
Pubblicazione: (2023)
di: Ye, Dongwei, et al.
Pubblicazione: (2023)
Piecewise Deterministic Markov Processes for Bayesian Inference of PDE Coefficients
di: Riccius, Leon, et al.
Pubblicazione: (2026)
di: Riccius, Leon, et al.
Pubblicazione: (2026)
CausalLM is not optimal for in-context learning
di: Ding, Nan, et al.
Pubblicazione: (2023)
di: Ding, Nan, et al.
Pubblicazione: (2023)
Surrogate modeling for Bayesian optimization beyond a single Gaussian process
di: Lu, Qin, et al.
Pubblicazione: (2022)
di: Lu, Qin, et al.
Pubblicazione: (2022)
Approximation of differential entropy in Bayesian optimal experimental design
di: Chen, Chuntao, et al.
Pubblicazione: (2025)
di: Chen, Chuntao, et al.
Pubblicazione: (2025)
Quantum framework for Reinforcement Learning: Integrating Markov decision process, quantum arithmetic, and trajectory search
di: Su, Thet Htar, et al.
Pubblicazione: (2024)
di: Su, Thet Htar, et al.
Pubblicazione: (2024)
Transfer learning of state-based potential games for process optimization in decentralized manufacturing systems
di: Yuwono, Steve, et al.
Pubblicazione: (2024)
di: Yuwono, Steve, et al.
Pubblicazione: (2024)
Multi-agent imitation learning with function approximation: Linear Markov games and beyond
di: Viano, Luca, et al.
Pubblicazione: (2026)
di: Viano, Luca, et al.
Pubblicazione: (2026)
Data-driven robust Markov decision processes on Borel spaces: performance guarantees via an axiomatic approach
di: Ramani, Sivaramakrishnan
Pubblicazione: (2026)
di: Ramani, Sivaramakrishnan
Pubblicazione: (2026)
Leveraging heterogeneity for identifiability: Bayesian order-based learning of multiple DAGs
di: Chang, Hyunwoong, et al.
Pubblicazione: (2026)
di: Chang, Hyunwoong, et al.
Pubblicazione: (2026)
Enhancing ASD detection accuracy: a combined approach of machine learning and deep learning models with natural language processing
di: Rubio-Martín, Sergio, et al.
Pubblicazione: (2024)
di: Rubio-Martín, Sergio, et al.
Pubblicazione: (2024)
A novel gradient-based method for decision trees optimizing arbitrary differential loss functions
di: Konstantinov, Andrei V., et al.
Pubblicazione: (2025)
di: Konstantinov, Andrei V., et al.
Pubblicazione: (2025)
Re-examining learning linear functions in context
di: Naim, Omar, et al.
Pubblicazione: (2024)
di: Naim, Omar, et al.
Pubblicazione: (2024)
Stochastic set-valued optimization and its application to robust learning
di: Giovannelli, Tommaso, et al.
Pubblicazione: (2026)
di: Giovannelli, Tommaso, et al.
Pubblicazione: (2026)
Coreset Markov Chain Monte Carlo
di: Chen, Naitong, et al.
Pubblicazione: (2023)
di: Chen, Naitong, et al.
Pubblicazione: (2023)
Adaptive operator learning for infinite-dimensional Bayesian inverse problems
di: Gao, Zhiwei, et al.
Pubblicazione: (2023)
di: Gao, Zhiwei, et al.
Pubblicazione: (2023)
Intrinsic effective sample size for manifold-valued Markov chain Monte Carlo via kernel discrepancy
di: You, Kisung
Pubblicazione: (2026)
di: You, Kisung
Pubblicazione: (2026)
Predicting human decisions with behavioral theories and machine learning
di: Plonsky, Ori, et al.
Pubblicazione: (2019)
di: Plonsky, Ori, et al.
Pubblicazione: (2019)
Joint modeling for learning decision-making dynamics in behavioral experiments
di: Bian, Yuan, et al.
Pubblicazione: (2025)
di: Bian, Yuan, et al.
Pubblicazione: (2025)
Bayesian optimization as a flexible and efficient design framework for sustainable process systems
di: Paulson, Joel A., et al.
Pubblicazione: (2024)
di: Paulson, Joel A., et al.
Pubblicazione: (2024)
Survey on reinforcement learning for language processing
di: Uc-Cetina, Victor, et al.
Pubblicazione: (2021)
di: Uc-Cetina, Victor, et al.
Pubblicazione: (2021)
Towards best practices in low-dimensional semi-supervised latent Bayesian optimization for the design of antimicrobial peptides
di: Menard, Jyler, et al.
Pubblicazione: (2025)
di: Menard, Jyler, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Bayesian Deep Learning with Multilevel Trace-class Neural Networks
di: Chada, Neil K., et al.
Pubblicazione: (2022) -
Asymptotically optimal regret in communicating Markov decision processes
di: Boone, Victor
Pubblicazione: (2025) -
Spatio-temporal modeling and forecasting with Fourier neural operators
di: Nag, Pratik, et al.
Pubblicazione: (2026) -
Relaxed Gaussian process interpolation: a goal-oriented approach to Bayesian optimization
di: Petit, Sébastien, et al.
Pubblicazione: (2022) -
Planning in entropy-regularized Markov decision processes and games
di: Grill, Jean-Bastien, et al.
Pubblicazione: (2026)