Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Benechehab, Abdelhakim, Thomas, Albert, Kégl, Balázs |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2024)
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2024)
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2025)
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2025)
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2025)
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2025)
Guided Safe Shooting: model based reinforcement learning with safety constraints
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2022)
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2022)
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
von: Hili, Youssef Attia El, et al.
Veröffentlicht: (2025)
von: Hili, Youssef Attia El, et al.
Veröffentlicht: (2025)
From Data to Rewards: a Bilevel Optimization Perspective on Maximum Likelihood Estimation
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2025)
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2025)
Zero-shot Model-based Reinforcement Learning using Large Language Models
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2024)
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2024)
Causal prompting model-based offline reinforcement learning
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
Can LLMs predict the convergence of Stochastic Gradient Descent?
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
Expert or not? assessing data quality in offline reinforcement learning
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
von: Asadulaev, Arip, et al.
Veröffentlicht: (2025)
Experimental evaluation of offline reinforcement learning for HVAC control in buildings
von: Wang, Jun, et al.
Veröffentlicht: (2024)
von: Wang, Jun, et al.
Veröffentlicht: (2024)
Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning
von: Kobayashi, Seijin, et al.
Veröffentlicht: (2025)
von: Kobayashi, Seijin, et al.
Veröffentlicht: (2025)
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
von: Zhang, Jing, et al.
Veröffentlicht: (2024)
von: Zhang, Jing, et al.
Veröffentlicht: (2024)
Physics-informed offline reinforcement learning eliminates catastrophic fuel waste in maritime routing
von: Bora, Aniruddha, et al.
Veröffentlicht: (2026)
von: Bora, Aniruddha, et al.
Veröffentlicht: (2026)
Large Language Models as Markov Chains
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
PyCFRL: A Python library for counterfactually fair offline reinforcement learning via sequential data preprocessing
von: Zhang, Jianhan, et al.
Veröffentlicht: (2025)
von: Zhang, Jianhan, et al.
Veröffentlicht: (2025)
Data-driven simulator of multi-animal behavior with unknown dynamics via offline and online reinforcement learning
von: Fujii, Keisuke, et al.
Veröffentlicht: (2025)
von: Fujii, Keisuke, et al.
Veröffentlicht: (2025)
Conservative quantum offline model-based optimization
von: Sotirov, Kristian, et al.
Veröffentlicht: (2025)
von: Sotirov, Kristian, et al.
Veröffentlicht: (2025)
Continual learning with the neural tangent ensemble
von: Benjamin, Ari S., et al.
Veröffentlicht: (2024)
von: Benjamin, Ari S., et al.
Veröffentlicht: (2024)
Exploring validation metrics for offline model-based optimisation with diffusion models
von: Beckham, Christopher, et al.
Veröffentlicht: (2022)
von: Beckham, Christopher, et al.
Veröffentlicht: (2022)
Thermalizer: Stable autoregressive neural emulation of spatiotemporal chaos
von: Pedersen, Chris, et al.
Veröffentlicht: (2025)
von: Pedersen, Chris, et al.
Veröffentlicht: (2025)
Balancing optimism and pessimism in offline-to-online learning
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
High-Dimensional Analysis of Bootstrap Ensemble Classifiers
von: Tiomoko, Malik, et al.
Veröffentlicht: (2025)
von: Tiomoko, Malik, et al.
Veröffentlicht: (2025)
Holographic functions and neural networks
von: Szegedy, Balazs
Veröffentlicht: (2026)
von: Szegedy, Balazs
Veröffentlicht: (2026)
Deep autoregressive modeling for land use land cover
von: Krapu, Christopher, et al.
Veröffentlicht: (2024)
von: Krapu, Christopher, et al.
Veröffentlicht: (2024)
Sufficient conditions for offline reactivation in recurrent neural networks
von: Krishna, Nanda H., et al.
Veröffentlicht: (2025)
von: Krishna, Nanda H., et al.
Veröffentlicht: (2025)
Adaptive prediction theory combining offline and online learning
von: Li, Haizheng, et al.
Veröffentlicht: (2025)
von: Li, Haizheng, et al.
Veröffentlicht: (2025)
A domain decomposition-based autoregressive deep learning model for unsteady and nonlinear partial differential equations
von: Nidhan, Sheel, et al.
Veröffentlicht: (2024)
von: Nidhan, Sheel, et al.
Veröffentlicht: (2024)
Hybrid machine learning models based on physical patterns to accelerate CFD simulations: a short guide on autoregressive models
von: Sengupta, Arindam, et al.
Veröffentlicht: (2025)
von: Sengupta, Arindam, et al.
Veröffentlicht: (2025)
Development of an offline and online hybrid model for the Integrated Forecasting System
von: Farchi, Alban, et al.
Veröffentlicht: (2024)
von: Farchi, Alban, et al.
Veröffentlicht: (2024)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
von: Xu, Wei, et al.
Veröffentlicht: (2025)
von: Xu, Wei, et al.
Veröffentlicht: (2025)
Ergodicity in reinforcement learning
von: Baumann, Dominik, et al.
Veröffentlicht: (2026)
von: Baumann, Dominik, et al.
Veröffentlicht: (2026)
Flow map learning in nonlinear vector autoregressive models: influence of the feature-library structure on the training error
von: Gross, Markus
Veröffentlicht: (2026)
von: Gross, Markus
Veröffentlicht: (2026)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
Flowsheet synthesis through hierarchical reinforcement learning and graph neural networks
von: Stops, Laura, et al.
Veröffentlicht: (2022)
von: Stops, Laura, et al.
Veröffentlicht: (2022)
Safe reinforcement learning in uncertain contexts
von: Baumann, Dominik, et al.
Veröffentlicht: (2024)
von: Baumann, Dominik, et al.
Veröffentlicht: (2024)
Variational Autoencoders for exteroceptive perception in reinforcement learning-based collision avoidance
von: Larsen, Thomas Nakken, et al.
Veröffentlicht: (2024)
von: Larsen, Thomas Nakken, et al.
Veröffentlicht: (2024)
Silhouettes and quasi residual plots for neural nets and tree-based classifiers
von: Raymaekers, Jakob, et al.
Veröffentlicht: (2021)
von: Raymaekers, Jakob, et al.
Veröffentlicht: (2021)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
von: Hutson, Miles, et al.
Veröffentlicht: (2024)
von: Hutson, Miles, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2024) -
TAG: A Decentralized Framework for Multi-Agent Hierarchical Reinforcement Learning
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2025) -
AdaPTS: Adapting Univariate Foundation Models to Probabilistic Multivariate Time Series Forecasting
von: Benechehab, Abdelhakim, et al.
Veröffentlicht: (2025) -
Guided Safe Shooting: model based reinforcement learning with safety constraints
von: Paolo, Giuseppe, et al.
Veröffentlicht: (2022) -
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
von: Hili, Youssef Attia El, et al.
Veröffentlicht: (2025)