On Rollouts in Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Frauenknecht, Bernd, Subhasish, Devdutt, Solowjow, Friedrich, Trimpe, Sebastian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Trust the Model Where It Trusts Itself -- Model-Based Actor-Critic with Uncertainty-Aware Rollout Adaption
by: Frauenknecht, Bernd, et al.
Published: (2024)
by: Frauenknecht, Bernd, et al.
Published: (2024)
Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty
by: Eisele, Artur, et al.
Published: (2026)
by: Eisele, Artur, et al.
Published: (2026)
Learning to Race in Minutes: Infoprop Dyna on the Mini Wheelbot
by: Subhasish, Devdutt, et al.
Published: (2026)
by: Subhasish, Devdutt, et al.
Published: (2026)
Contextualized Hybrid Ensemble Q-learning: Learning Fast with Control Priors
by: Cramer, Emma, et al.
Published: (2024)
by: Cramer, Emma, et al.
Published: (2024)
Biased Dreams: Limitations to Epistemic Uncertainty Quantification in Latent Space Models
by: Berger, Julia, et al.
Published: (2026)
by: Berger, Julia, et al.
Published: (2026)
On the Consistency of Kernel Methods with Dependent Observations
by: Massiani, Pierre-François, et al.
Published: (2024)
by: Massiani, Pierre-François, et al.
Published: (2024)
Event-Triggered Time-Varying Bayesian Optimization
by: Brunzema, Paul, et al.
Published: (2022)
by: Brunzema, Paul, et al.
Published: (2022)
Uncertainty-Aware Predictive Safety Filters for Probabilistic Neural Network Dynamics
by: Frauenknecht, Bernd, et al.
Published: (2026)
by: Frauenknecht, Bernd, et al.
Published: (2026)
On Foundation Models for Dynamical Systems from Purely Synthetic Data
by: Ziegler, Martin, et al.
Published: (2024)
by: Ziegler, Martin, et al.
Published: (2024)
Safe Value Functions
by: Massiani, Pierre-François, et al.
Published: (2021)
by: Massiani, Pierre-François, et al.
Published: (2021)
Sailing Towards Zero-Shot State Estimation using Foundation Models Combined with a UKF
by: Holtmann, Tobin, et al.
Published: (2025)
by: Holtmann, Tobin, et al.
Published: (2025)
Kernel conditional tests from learning-theoretic bounds
by: Massiani, Pierre-François, et al.
Published: (2025)
by: Massiani, Pierre-François, et al.
Published: (2025)
The Mini Wheelbot Dataset: High-Fidelity Data for Robot Learning
by: Hose, Henrik, et al.
Published: (2026)
by: Hose, Henrik, et al.
Published: (2026)
Data-Driven Observability Analysis for Nonlinear Stochastic Systems
by: Massiani, Pierre-François, et al.
Published: (2023)
by: Massiani, Pierre-François, et al.
Published: (2023)
Learning Hybrid Dynamics Models With Simulator-Informed Latent States
by: Ensinger, Katharina, et al.
Published: (2023)
by: Ensinger, Katharina, et al.
Published: (2023)
Tracking Object Positions in Reinforcement Learning: A Metric for Keypoint Detection (extended version)
by: Cramer, Emma, et al.
Published: (2023)
by: Cramer, Emma, et al.
Published: (2023)
Distributed Event-Based Learning via ADMM
by: Er, Guner Dilsad, et al.
Published: (2024)
by: Er, Guner Dilsad, et al.
Published: (2024)
Viability of Future Actions: Robust Safety in Reinforcement Learning via Entropy Regularization
by: Massiani, Pierre-François, et al.
Published: (2025)
by: Massiani, Pierre-François, et al.
Published: (2025)
MPX: Mixed Precision Training for JAX
by: Gräfe, Alexander, et al.
Published: (2025)
by: Gräfe, Alexander, et al.
Published: (2025)
BayeSQP: Bayesian Optimization through Sequential Quadratic Programming
by: Brunzema, Paul, et al.
Published: (2026)
by: Brunzema, Paul, et al.
Published: (2026)
ImagineBench: Evaluating Reinforcement Learning with Large Language Model Rollouts
by: Pang, Jing-Cheng, et al.
Published: (2025)
by: Pang, Jing-Cheng, et al.
Published: (2025)
Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning
by: Xu, Yixuan Even, et al.
Published: (2025)
by: Xu, Yixuan Even, et al.
Published: (2025)
Leveraging Error Diversity in Group Rollouts for Reinforcement Learning
by: Liu, Wenpu, et al.
Published: (2026)
by: Liu, Wenpu, et al.
Published: (2026)
QuRL: Efficient Reinforcement Learning with Quantized Rollout
by: Li, Yuhang, et al.
Published: (2026)
by: Li, Yuhang, et al.
Published: (2026)
CHEQ-ing the Box: Safe Variable Impedance Learning for Robotic Polishing
by: Cramer, Emma, et al.
Published: (2025)
by: Cramer, Emma, et al.
Published: (2025)
Rollout-Training Co-Design for Efficient LLM-Based Multi-Agent Reinforcement Learning
by: Jiang, Zhida, et al.
Published: (2026)
by: Jiang, Zhida, et al.
Published: (2026)
Guided Cooperation in Hierarchical Reinforcement Learning via Model-based Rollout
by: Wang, Haoran, et al.
Published: (2023)
by: Wang, Haoran, et al.
Published: (2023)
EchoRL: Reinforcement Learning via Rollout Echoing
by: Bi, Jinhe, et al.
Published: (2026)
by: Bi, Jinhe, et al.
Published: (2026)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
by: Lu, Xiaodong, et al.
Published: (2026)
by: Lu, Xiaodong, et al.
Published: (2026)
RockNet: Distributed Learning on Ultra-Low-Power Devices
by: Gräfe, Alexander, et al.
Published: (2025)
by: Gräfe, Alexander, et al.
Published: (2025)
Train Less, Learn More: Adaptive Efficient Rollout Optimization for Group-Based Reinforcement Learning
by: Zhang, Zhi, et al.
Published: (2026)
by: Zhang, Zhi, et al.
Published: (2026)
Superior Computer Chess with Model Predictive Control, Reinforcement Learning, and Rollout
by: Gundawar, Atharva, et al.
Published: (2024)
by: Gundawar, Atharva, et al.
Published: (2024)
SRT: Accelerating Reinforcement Learning via Speculative Rollout with Tree-Structured Cache
by: Chang, Chi-Chih, et al.
Published: (2026)
by: Chang, Chi-Chih, et al.
Published: (2026)
Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning
by: Ding, Zihan, et al.
Published: (2024)
by: Ding, Zihan, et al.
Published: (2024)
Exact Inference for Continuous-Time Gaussian Process Dynamics
by: Ensinger, Katharina, et al.
Published: (2023)
by: Ensinger, Katharina, et al.
Published: (2023)
Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts
by: Pang, Jing-Cheng, et al.
Published: (2024)
by: Pang, Jing-Cheng, et al.
Published: (2024)
DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning
by: Zhao, Hanye, et al.
Published: (2024)
by: Zhao, Hanye, et al.
Published: (2024)
Feedforward Controllers from Learned Dynamic Local Model Networks with Application to Excavator Assistance Functions
by: Greiser, Leon, et al.
Published: (2024)
by: Greiser, Leon, et al.
Published: (2024)
SPEC-RL: Accelerating On-Policy Reinforcement Learning with Speculative Rollouts
by: Liu, Bingshuai, et al.
Published: (2025)
by: Liu, Bingshuai, et al.
Published: (2025)
Similar Items
-
Trust the Model Where It Trusts Itself -- Model-Based Actor-Critic with Uncertainty-Aware Rollout Adaption
by: Frauenknecht, Bernd, et al.
Published: (2024) -
Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty
by: Eisele, Artur, et al.
Published: (2026) -
Learning to Race in Minutes: Infoprop Dyna on the Mini Wheelbot
by: Subhasish, Devdutt, et al.
Published: (2026) -
Contextualized Hybrid Ensemble Q-learning: Learning Fast with Control Priors
by: Cramer, Emma, et al.
Published: (2024) -
Biased Dreams: Limitations to Epistemic Uncertainty Quantification in Latent Space Models
by: Berger, Julia, et al.
Published: (2026)