Optimistic World Models: Efficient Exploration in Model-Based Deep Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mete, Akshay, Sheikh, Shahid Aamir, Lin, Tzu-Hsiang, Kalathil, Dileep, Kumar, P. R. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Structured Reinforcement Learning for Media Streaming at the Wireless Edge
von: Bura, Archana, et al.
Veröffentlicht: (2024)
von: Bura, Archana, et al.
Veröffentlicht: (2024)
PowerMamba: A Deep State Space Model and Comprehensive Benchmark for Time Series Prediction in Electric Power Systems
von: Menati, Ali, et al.
Veröffentlicht: (2024)
von: Menati, Ali, et al.
Veröffentlicht: (2024)
Why Reinforcement Learning in Energy Systems Needs Explanations
von: Butt, Hallah Shahid, et al.
Veröffentlicht: (2024)
von: Butt, Hallah Shahid, et al.
Veröffentlicht: (2024)
Task and Domain Adaptive Reinforcement Learning for Robot Control
von: Liu, Yu Tang, et al.
Veröffentlicht: (2024)
von: Liu, Yu Tang, et al.
Veröffentlicht: (2024)
Selling Demand Response Using Options
von: Muthirayan, Deepan, et al.
Veröffentlicht: (2019)
von: Muthirayan, Deepan, et al.
Veröffentlicht: (2019)
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
von: Bertsekas, Dimitri P.
Veröffentlicht: (2024)
von: Bertsekas, Dimitri P.
Veröffentlicht: (2024)
Safe Deep Model-Based Reinforcement Learning with Lyapunov Functions
von: Zhang, Harry
Veröffentlicht: (2024)
von: Zhang, Harry
Veröffentlicht: (2024)
Model-Free Load Frequency Control of Nonlinear Power Systems Based on Deep Reinforcement Learning
von: Chen, Xiaodi, et al.
Veröffentlicht: (2024)
von: Chen, Xiaodi, et al.
Veröffentlicht: (2024)
Novelty Detection in Reinforcement Learning with World Models
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2023)
von: Zollicoffer, Geigh, et al.
Veröffentlicht: (2023)
Simulation-Based Optimistic Policy Iteration For Multi-Agent MDPs with Kullback-Leibler Control Cost
von: Nakhleh, Khaled, et al.
Veröffentlicht: (2024)
von: Nakhleh, Khaled, et al.
Veröffentlicht: (2024)
DeepSafeMPC: Deep Learning-Based Model Predictive Control for Safe Multi-Agent Reinforcement Learning
von: Wang, Xuefeng, et al.
Veröffentlicht: (2024)
von: Wang, Xuefeng, et al.
Veröffentlicht: (2024)
Traffic Smoothing Controllers for Autonomous Vehicles Using Deep Reinforcement Learning and Real-World Trajectory Data
von: Lichtlé, Nathan, et al.
Veröffentlicht: (2024)
von: Lichtlé, Nathan, et al.
Veröffentlicht: (2024)
Energy-Efficient Thermal Comfort Control in Smart Buildings via Deep Reinforcement Learning
von: Gao, Guanyu, et al.
Veröffentlicht: (2019)
von: Gao, Guanyu, et al.
Veröffentlicht: (2019)
RL2: Reinforce Large Language Model to Assist Safe Reinforcement Learning for Energy Management of Active Distribution Networks
von: Yang, Xu, et al.
Veröffentlicht: (2024)
von: Yang, Xu, et al.
Veröffentlicht: (2024)
Large Artificial Intelligence Model Guided Deep Reinforcement Learning for Resource Allocation in Non Terrestrial Networks
von: Ibrahim, Abdikarim Mohamed, et al.
Veröffentlicht: (2026)
von: Ibrahim, Abdikarim Mohamed, et al.
Veröffentlicht: (2026)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
von: Singh, Nikhil Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Nikhil Kumar, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Partial Parametric Model Knowledge
von: Wang, Shuyuan, et al.
Veröffentlicht: (2023)
von: Wang, Shuyuan, et al.
Veröffentlicht: (2023)
Optimizing Traffic Signal Control using High-Dimensional State Representation and Efficient Deep Reinforcement Learning
von: Francis, Lawrence, et al.
Veröffentlicht: (2024)
von: Francis, Lawrence, et al.
Veröffentlicht: (2024)
Kernel-Based Safe Exploration in Deep Reinforcement Learning
von: Majumdar, Rupak, et al.
Veröffentlicht: (2026)
von: Majumdar, Rupak, et al.
Veröffentlicht: (2026)
Optimistic vs Pessimistic Uncertainty Model Unfalsification
von: Hühnerbein, Jannes, et al.
Veröffentlicht: (2025)
von: Hühnerbein, Jannes, et al.
Veröffentlicht: (2025)
Approximate Model-Based Shielding for Safe Reinforcement Learning
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
von: Goodall, Alexander W., et al.
Veröffentlicht: (2023)
Airship Formations for Animal Motion Capture and Behavior Analysis
von: Price, Eric, et al.
Veröffentlicht: (2024)
von: Price, Eric, et al.
Veröffentlicht: (2024)
Reinforcement Learning based Autonomous Multi-Rotor Landing on Moving Platforms
von: Goldschmid, Pascal, et al.
Veröffentlicht: (2023)
von: Goldschmid, Pascal, et al.
Veröffentlicht: (2023)
Reinforcement Learning with Model Predictive Control for Highway Ramp Metering
von: Airaldi, Filippo, et al.
Veröffentlicht: (2023)
von: Airaldi, Filippo, et al.
Veröffentlicht: (2023)
Model Predictive Control-Guided Reinforcement Learning for Implicit Balancing
von: Madahi, Seyed Soroush Karimi, et al.
Veröffentlicht: (2025)
von: Madahi, Seyed Soroush Karimi, et al.
Veröffentlicht: (2025)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
von: Zolman, Nicholas, et al.
Veröffentlicht: (2024)
von: Zolman, Nicholas, et al.
Veröffentlicht: (2024)
Exploration of Multi-Element Collaborative Research and Application for Modern Power System Based on Generative Large Models
von: Cheng, Lu, et al.
Veröffentlicht: (2025)
von: Cheng, Lu, et al.
Veröffentlicht: (2025)
Modelling, Positioning, and Deep Reinforcement Learning Path Tracking Control of Scaled Robotic Vehicles: Design and Experimental Validation
von: Caponio, Carmine, et al.
Veröffentlicht: (2024)
von: Caponio, Carmine, et al.
Veröffentlicht: (2024)
Model-Based Data-Efficient and Robust Reinforcement Learning
von: Svedlund, Ludvig, et al.
Veröffentlicht: (2026)
von: Svedlund, Ludvig, et al.
Veröffentlicht: (2026)
Energy Control Strategy to Enhance AC Fault Ride-Through in Offshore Wind MMC-HVDC Systems
von: Kumar, Dileep, et al.
Veröffentlicht: (2025)
von: Kumar, Dileep, et al.
Veröffentlicht: (2025)
Enhancing Microgrid Performance Prediction with Attention-based Deep Learning Models
von: Maddineni, Vinod Kumar, et al.
Veröffentlicht: (2024)
von: Maddineni, Vinod Kumar, et al.
Veröffentlicht: (2024)
Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm
von: Qiao, Ting, et al.
Veröffentlicht: (2024)
von: Qiao, Ting, et al.
Veröffentlicht: (2024)
Multi-Goal Dexterous Hand Manipulation using Probabilistic Model-based Reinforcement Learning
von: Jiang, Yingzhuo, et al.
Veröffentlicht: (2025)
von: Jiang, Yingzhuo, et al.
Veröffentlicht: (2025)
Towards Ultra-Reliable 6G in-X Subnetworks: Dynamic Link Adaptation by Deep Reinforcement Learning
von: Salehi, Fateme, et al.
Veröffentlicht: (2025)
von: Salehi, Fateme, et al.
Veröffentlicht: (2025)
Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning
von: Rahmani, Saeed, et al.
Veröffentlicht: (2026)
von: Rahmani, Saeed, et al.
Veröffentlicht: (2026)
Deep Learning Based Simulators for the Phosphorus Removal Process Control in Wastewater Treatment via Deep Reinforcement Learning Algorithms
von: Mohammadi, Esmaeel, et al.
Veröffentlicht: (2024)
von: Mohammadi, Esmaeel, et al.
Veröffentlicht: (2024)
Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility
von: Jüni, Colin, et al.
Veröffentlicht: (2026)
von: Jüni, Colin, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning for Power Grid Multi-Stage Cascading Failure Mitigation
von: Meng, Bo, et al.
Veröffentlicht: (2025)
von: Meng, Bo, et al.
Veröffentlicht: (2025)
Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
von: Huang, Yilie, et al.
Veröffentlicht: (2025)
von: Huang, Yilie, et al.
Veröffentlicht: (2025)
Smart Exploration in Reinforcement Learning using Bounded Uncertainty Models
von: van Hulst, J. S., et al.
Veröffentlicht: (2025)
von: van Hulst, J. S., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Structured Reinforcement Learning for Media Streaming at the Wireless Edge
von: Bura, Archana, et al.
Veröffentlicht: (2024) -
PowerMamba: A Deep State Space Model and Comprehensive Benchmark for Time Series Prediction in Electric Power Systems
von: Menati, Ali, et al.
Veröffentlicht: (2024) -
Why Reinforcement Learning in Energy Systems Needs Explanations
von: Butt, Hallah Shahid, et al.
Veröffentlicht: (2024) -
Task and Domain Adaptive Reinforcement Learning for Robot Control
von: Liu, Yu Tang, et al.
Veröffentlicht: (2024) -
Selling Demand Response Using Options
von: Muthirayan, Deepan, et al.
Veröffentlicht: (2019)