Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Benechehab, Abdelhakim, Thomas, Albert, Kégl, Balázs |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
by: Paolo, Giuseppe, et al.
Published: (2025)
by: Paolo, Giuseppe, et al.
Published: (2025)
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
by: Benechehab, Abdelhakim, et al.
Published: (2025)
by: Benechehab, Abdelhakim, et al.
Published: (2025)
Guided Safe Shooting: model based reinforcement learning with safety constraints
by: Paolo, Giuseppe, et al.
Published: (2022)
by: Paolo, Giuseppe, et al.
Published: (2022)
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
by: Hili, Youssef Attia El, et al.
Published: (2025)
by: Hili, Youssef Attia El, et al.
Published: (2025)
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation
by: Benechehab, Abdelhakim, et al.
Published: (2025)
by: Benechehab, Abdelhakim, et al.
Published: (2025)
Zero-shot Model-based Reinforcement Learning using Large Language Models
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Can LLMs predict the convergence of Stochastic Gradient Descent?
by: Zekri, Oussama, et al.
Published: (2024)
by: Zekri, Oussama, et al.
Published: (2024)
Expert or not? assessing data quality in offline reinforcement learning
by: Asadulaev, Arip, et al.
Published: (2025)
by: Asadulaev, Arip, et al.
Published: (2025)
Experimental evaluation of offline reinforcement learning for HVAC control in buildings
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
by: Kobayashi, Seijin, et al.
Published: (2025)
by: Kobayashi, Seijin, et al.
Published: (2025)
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
by: Zhang, Jing, et al.
Published: (2024)
by: Zhang, Jing, et al.
Published: (2024)
Large Language Models as Markov Chains
by: Zekri, Oussama, et al.
Published: (2024)
by: Zekri, Oussama, et al.
Published: (2024)
Physics-informed offline reinforcement learning eliminates catastrophic fuel waste in maritime routing
by: Bora, Aniruddha, et al.
Published: (2026)
by: Bora, Aniruddha, et al.
Published: (2026)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
by: Grosnit, Antoine, et al.
Published: (2024)
by: Grosnit, Antoine, et al.
Published: (2024)
PyCFRL: A Python library for counterfactually fair offline reinforcement learning via sequential data preprocessing
by: Zhang, Jianhan, et al.
Published: (2025)
by: Zhang, Jianhan, et al.
Published: (2025)
Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning
by: Fujii, Keisuke, et al.
Published: (2025)
by: Fujii, Keisuke, et al.
Published: (2025)
Conservative quantum offline model-based optimization
by: Sotirov, Kristian, et al.
Published: (2025)
by: Sotirov, Kristian, et al.
Published: (2025)
Continual learning with the neural tangent ensemble
by: Benjamin, Ari S., et al.
Published: (2024)
by: Benjamin, Ari S., et al.
Published: (2024)
Exploring validation metrics for offline model-based optimisation with diffusion models
by: Beckham, Christopher, et al.
Published: (2022)
by: Beckham, Christopher, et al.
Published: (2022)
Thermalizer: Stable autoregressive neural emulation of spatiotemporal chaos
by: Pedersen, Chris, et al.
Published: (2025)
by: Pedersen, Chris, et al.
Published: (2025)
Balancing optimism and pessimism in offline-to-online learning
by: Sentenac, Flore, et al.
Published: (2025)
by: Sentenac, Flore, et al.
Published: (2025)
High-Dimensional Analysis of Bootstrap Ensemble Classifiers
by: Tiomoko, Malik, et al.
Published: (2025)
by: Tiomoko, Malik, et al.
Published: (2025)
Holographic functions and neural networks
by: Szegedy, Balazs
Published: (2026)
by: Szegedy, Balazs
Published: (2026)
Deep autoregressive modeling for land use land cover
by: Krapu, Christopher, et al.
Published: (2024)
by: Krapu, Christopher, et al.
Published: (2024)
Sufficient conditions for offline reactivation in recurrent neural networks
by: Krishna, Nanda H., et al.
Published: (2025)
by: Krishna, Nanda H., et al.
Published: (2025)
Adaptive prediction theory combining offline and online learning
by: Li, Haizheng, et al.
Published: (2025)
by: Li, Haizheng, et al.
Published: (2025)
A domain decomposition-based autoregressive deep learning model for unsteady and nonlinear partial differential equations
by: Nidhan, Sheel, et al.
Published: (2024)
by: Nidhan, Sheel, et al.
Published: (2024)
Hybrid machine learning models based on physical patterns to accelerate CFD simulations: a short guide on autoregressive models
by: Sengupta, Arindam, et al.
Published: (2025)
by: Sengupta, Arindam, et al.
Published: (2025)
Development of an offline and online hybrid model for the Integrated Forecasting System
by: Farchi, Alban, et al.
Published: (2024)
by: Farchi, Alban, et al.
Published: (2024)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Flow map learning in nonlinear vector autoregressive models: influence of the feature-library structure on the training error
by: Gross, Markus
Published: (2026)
by: Gross, Markus
Published: (2026)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
by: Ly, Adrian, et al.
Published: (2025)
by: Ly, Adrian, et al.
Published: (2025)
Flowsheet synthesis through hierarchical reinforcement learning and graph neural networks
by: Stops, Laura, et al.
Published: (2022)
by: Stops, Laura, et al.
Published: (2022)
Safe reinforcement learning in uncertain contexts
by: Baumann, Dominik, et al.
Published: (2024)
by: Baumann, Dominik, et al.
Published: (2024)
Variational Autoencoders for exteroceptive perception in reinforcement learning-based collision avoidance
by: Larsen, Thomas Nakken, et al.
Published: (2024)
by: Larsen, Thomas Nakken, et al.
Published: (2024)
Silhouettes and quasi residual plots for neural nets and tree-based classifiers
by: Raymaekers, Jakob, et al.
Published: (2021)
by: Raymaekers, Jakob, et al.
Published: (2021)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
by: Hutson, Miles, et al.
Published: (2024)
by: Hutson, Miles, et al.
Published: (2024)
Similar Items
-
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
by: Benechehab, Abdelhakim, et al.
Published: (2024) -
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
by: Paolo, Giuseppe, et al.
Published: (2025) -
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
by: Benechehab, Abdelhakim, et al.
Published: (2025) -
Guided Safe Shooting: model based reinforcement learning with safety constraints
by: Paolo, Giuseppe, et al.
Published: (2022) -
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
by: Hili, Youssef Attia El, et al.
Published: (2025)