Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Benechehab, Abdelhakim, Thomas, Albert, Kégl, Balázs |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
por: Benechehab, Abdelhakim, et al.
Publicado: (2024)
por: Benechehab, Abdelhakim, et al.
Publicado: (2024)
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
por: Paolo, Giuseppe, et al.
Publicado: (2025)
por: Paolo, Giuseppe, et al.
Publicado: (2025)
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
por: Benechehab, Abdelhakim, et al.
Publicado: (2025)
por: Benechehab, Abdelhakim, et al.
Publicado: (2025)
Guided Safe Shooting: model based reinforcement learning with safety constraints
por: Paolo, Giuseppe, et al.
Publicado: (2022)
por: Paolo, Giuseppe, et al.
Publicado: (2022)
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
por: Hili, Youssef Attia El, et al.
Publicado: (2025)
por: Hili, Youssef Attia El, et al.
Publicado: (2025)
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation
por: Benechehab, Abdelhakim, et al.
Publicado: (2025)
por: Benechehab, Abdelhakim, et al.
Publicado: (2025)
Zero-shot Model-based Reinforcement Learning using Large Language Models
por: Benechehab, Abdelhakim, et al.
Publicado: (2024)
por: Benechehab, Abdelhakim, et al.
Publicado: (2024)
Causal prompting model-based offline reinforcement learning
por: Yu, Xuehui, et al.
Publicado: (2024)
por: Yu, Xuehui, et al.
Publicado: (2024)
Can LLMs predict the convergence of Stochastic Gradient Descent?
por: Zekri, Oussama, et al.
Publicado: (2024)
por: Zekri, Oussama, et al.
Publicado: (2024)
Expert or not? assessing data quality in offline reinforcement learning
por: Asadulaev, Arip, et al.
Publicado: (2025)
por: Asadulaev, Arip, et al.
Publicado: (2025)
Experimental evaluation of offline reinforcement learning for HVAC control in buildings
por: Wang, Jun, et al.
Publicado: (2024)
por: Wang, Jun, et al.
Publicado: (2024)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
por: Kobayashi, Seijin, et al.
Publicado: (2025)
por: Kobayashi, Seijin, et al.
Publicado: (2025)
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
por: Zhang, Jing, et al.
Publicado: (2024)
por: Zhang, Jing, et al.
Publicado: (2024)
Physics-informed offline reinforcement learning eliminates catastrophic fuel waste in maritime routing
por: Bora, Aniruddha, et al.
Publicado: (2026)
por: Bora, Aniruddha, et al.
Publicado: (2026)
Large Language Models as Markov Chains
por: Zekri, Oussama, et al.
Publicado: (2024)
por: Zekri, Oussama, et al.
Publicado: (2024)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
por: Grosnit, Antoine, et al.
Publicado: (2024)
por: Grosnit, Antoine, et al.
Publicado: (2024)
PyCFRL: A Python library for counterfactually fair offline reinforcement learning via sequential data preprocessing
por: Zhang, Jianhan, et al.
Publicado: (2025)
por: Zhang, Jianhan, et al.
Publicado: (2025)
Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning
por: Fujii, Keisuke, et al.
Publicado: (2025)
por: Fujii, Keisuke, et al.
Publicado: (2025)
Conservative quantum offline model-based optimization
por: Sotirov, Kristian, et al.
Publicado: (2025)
por: Sotirov, Kristian, et al.
Publicado: (2025)
Continual learning with the neural tangent ensemble
por: Benjamin, Ari S., et al.
Publicado: (2024)
por: Benjamin, Ari S., et al.
Publicado: (2024)
Exploring validation metrics for offline model-based optimisation with diffusion models
por: Beckham, Christopher, et al.
Publicado: (2022)
por: Beckham, Christopher, et al.
Publicado: (2022)
Thermalizer: Stable autoregressive neural emulation of spatiotemporal chaos
por: Pedersen, Chris, et al.
Publicado: (2025)
por: Pedersen, Chris, et al.
Publicado: (2025)
Balancing optimism and pessimism in offline-to-online learning
por: Sentenac, Flore, et al.
Publicado: (2025)
por: Sentenac, Flore, et al.
Publicado: (2025)
High-Dimensional Analysis of Bootstrap Ensemble Classifiers
por: Tiomoko, Malik, et al.
Publicado: (2025)
por: Tiomoko, Malik, et al.
Publicado: (2025)
Holographic functions and neural networks
por: Szegedy, Balazs
Publicado: (2026)
por: Szegedy, Balazs
Publicado: (2026)
Deep autoregressive modeling for land use land cover
por: Krapu, Christopher, et al.
Publicado: (2024)
por: Krapu, Christopher, et al.
Publicado: (2024)
Sufficient conditions for offline reactivation in recurrent neural networks
por: Krishna, Nanda H., et al.
Publicado: (2025)
por: Krishna, Nanda H., et al.
Publicado: (2025)
Adaptive prediction theory combining offline and online learning
por: Li, Haizheng, et al.
Publicado: (2025)
por: Li, Haizheng, et al.
Publicado: (2025)
A domain decomposition-based autoregressive deep learning model for unsteady and nonlinear partial differential equations
por: Nidhan, Sheel, et al.
Publicado: (2024)
por: Nidhan, Sheel, et al.
Publicado: (2024)
Hybrid machine learning models based on physical patterns to accelerate CFD simulations: a short guide on autoregressive models
por: Sengupta, Arindam, et al.
Publicado: (2025)
por: Sengupta, Arindam, et al.
Publicado: (2025)
Development of an offline and online hybrid model for the Integrated Forecasting System
por: Farchi, Alban, et al.
Publicado: (2024)
por: Farchi, Alban, et al.
Publicado: (2024)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
por: Xu, Wei, et al.
Publicado: (2025)
por: Xu, Wei, et al.
Publicado: (2025)
Ergodicity in reinforcement learning
por: Baumann, Dominik, et al.
Publicado: (2026)
por: Baumann, Dominik, et al.
Publicado: (2026)
Flow map learning in nonlinear vector autoregressive models: influence of the feature-library structure on the training error
por: Gross, Markus
Publicado: (2026)
por: Gross, Markus
Publicado: (2026)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
por: Ly, Adrian, et al.
Publicado: (2025)
por: Ly, Adrian, et al.
Publicado: (2025)
Flowsheet synthesis through hierarchical reinforcement learning and graph neural networks
por: Stops, Laura, et al.
Publicado: (2022)
por: Stops, Laura, et al.
Publicado: (2022)
Safe reinforcement learning in uncertain contexts
por: Baumann, Dominik, et al.
Publicado: (2024)
por: Baumann, Dominik, et al.
Publicado: (2024)
Variational Autoencoders for exteroceptive perception in reinforcement learning-based collision avoidance
por: Larsen, Thomas Nakken, et al.
Publicado: (2024)
por: Larsen, Thomas Nakken, et al.
Publicado: (2024)
Silhouettes and quasi residual plots for neural nets and tree-based classifiers
por: Raymaekers, Jakob, et al.
Publicado: (2021)
por: Raymaekers, Jakob, et al.
Publicado: (2021)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
por: Hutson, Miles, et al.
Publicado: (2024)
por: Hutson, Miles, et al.
Publicado: (2024)
Ejemplares similares
-
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
por: Benechehab, Abdelhakim, et al.
Publicado: (2024) -
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
por: Paolo, Giuseppe, et al.
Publicado: (2025) -
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
por: Benechehab, Abdelhakim, et al.
Publicado: (2025) -
Guided Safe Shooting: model based reinforcement learning with safety constraints
por: Paolo, Giuseppe, et al.
Publicado: (2022) -
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
por: Hili, Youssef Attia El, et al.
Publicado: (2025)