Sample-efficient and Scalable Exploration in Continuous-Time RL
Fuente:
arXiv
Guardado en:
| Autores principales: | Iten, Klemens, Treven, Lenart, Sukhija, Bhavya, Dörfler, Florian, Krause, Andreas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Model-Based Reinforcement Learning for Control under Time-Varying Dynamics
por: Iten, Klemens, et al.
Publicado: (2026)
por: Iten, Klemens, et al.
Publicado: (2026)
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
por: Treven, Lenart, et al.
Publicado: (2024)
por: Treven, Lenart, et al.
Publicado: (2024)
NeoRL: Efficient Exploration for Nonepisodic RL
por: Sukhija, Bhavya, et al.
Publicado: (2024)
por: Sukhija, Bhavya, et al.
Publicado: (2024)
Simulation Priors for Data-Efficient Deep Learning
por: Treven, Lenart, et al.
Publicado: (2025)
por: Treven, Lenart, et al.
Publicado: (2025)
Bridging the Sim-to-Real Gap with Bayesian Inference
por: Rothfuss, Jonas, et al.
Publicado: (2024)
por: Rothfuss, Jonas, et al.
Publicado: (2024)
SOMBRL: Scalable and Optimistic Model-Based RL
por: Sukhija, Bhavya, et al.
Publicado: (2025)
por: Sukhija, Bhavya, et al.
Publicado: (2025)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
por: As, Yarden, et al.
Publicado: (2024)
por: As, Yarden, et al.
Publicado: (2024)
Active Few-Shot Fine-Tuning
por: Hübotter, Jonas, et al.
Publicado: (2024)
por: Hübotter, Jonas, et al.
Publicado: (2024)
Transductive Active Learning: Theory and Applications
por: Hübotter, Jonas, et al.
Publicado: (2024)
por: Hübotter, Jonas, et al.
Publicado: (2024)
TARC: Time-Adaptive Robotic Control
por: Sukhija, Arnav, et al.
Publicado: (2025)
por: Sukhija, Arnav, et al.
Publicado: (2025)
Learning Soft Robotic Dynamics with Active Exploration
por: Zheng, Hehui, et al.
Publicado: (2025)
por: Zheng, Hehui, et al.
Publicado: (2025)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
por: Sukhija, Bhavya, et al.
Publicado: (2024)
por: Sukhija, Bhavya, et al.
Publicado: (2024)
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
por: As, Yarden, et al.
Publicado: (2024)
por: As, Yarden, et al.
Publicado: (2024)
Safe Exploration via Policy Priors
por: Wendl, Manuel, et al.
Publicado: (2026)
por: Wendl, Manuel, et al.
Publicado: (2026)
CaRL: Learning Scalable Planning Policies with Simple Rewards
por: Jaeger, Bernhard, et al.
Publicado: (2025)
por: Jaeger, Bernhard, et al.
Publicado: (2025)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
por: Bhardwaj, Arjun, et al.
Publicado: (2023)
por: Bhardwaj, Arjun, et al.
Publicado: (2023)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Sampling-Based Safe Reinforcement Learning
por: Vignola, Luca, et al.
Publicado: (2026)
por: Vignola, Luca, et al.
Publicado: (2026)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
por: Kim, Changyeon, et al.
Publicado: (2025)
por: Kim, Changyeon, et al.
Publicado: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
por: Choi, Wonhyeok, et al.
Publicado: (2026)
por: Choi, Wonhyeok, et al.
Publicado: (2026)
Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
por: Li, Chenhao, et al.
Publicado: (2025)
por: Li, Chenhao, et al.
Publicado: (2025)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
por: Li, Chenhao, et al.
Publicado: (2025)
por: Li, Chenhao, et al.
Publicado: (2025)
SPAARS: Safer RL Policy Alignment through Abstract Exploration and Refined Exploitation of Action Space
por: K, Swaminathan S, et al.
Publicado: (2026)
por: K, Swaminathan S, et al.
Publicado: (2026)
Scilab-RL: A software framework for efficient reinforcement learning and cognitive modeling research
por: Dohmen, Jan, et al.
Publicado: (2024)
por: Dohmen, Jan, et al.
Publicado: (2024)
Enhanced Importance Sampling through Latent Space Exploration in Normalizing Flows
por: Kruse, Liam A., et al.
Publicado: (2025)
por: Kruse, Liam A., et al.
Publicado: (2025)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
por: Hori, Toshiaki, et al.
Publicado: (2025)
por: Hori, Toshiaki, et al.
Publicado: (2025)
Feature-Based vs. GAN-Based Learning from Demonstrations: When and Why
por: Li, Chenhao, et al.
Publicado: (2025)
por: Li, Chenhao, et al.
Publicado: (2025)
SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration
por: Jin, Yang, et al.
Publicado: (2025)
por: Jin, Yang, et al.
Publicado: (2025)
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning
por: Rowe, Luke, et al.
Publicado: (2024)
por: Rowe, Luke, et al.
Publicado: (2024)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
por: Diaz-Bone, Leander, et al.
Publicado: (2025)
por: Diaz-Bone, Leander, et al.
Publicado: (2025)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
por: Bao, Kaixi, et al.
Publicado: (2025)
por: Bao, Kaixi, et al.
Publicado: (2025)
Efficient Preference-Based Reinforcement Learning: Randomized Exploration Meets Experimental Design
por: Schlaginhaufen, Andreas, et al.
Publicado: (2025)
por: Schlaginhaufen, Andreas, et al.
Publicado: (2025)
WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning
por: Kim, Mintae, et al.
Publicado: (2026)
por: Kim, Mintae, et al.
Publicado: (2026)
Continual Model-Based Reinforcement Learning with Hypernetworks
por: Huang, Yizhou, et al.
Publicado: (2020)
por: Huang, Yizhou, et al.
Publicado: (2020)
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
por: Evers, Thomas, et al.
Publicado: (2026)
por: Evers, Thomas, et al.
Publicado: (2026)
Automatic Environment Shaping is the Next Frontier in RL
por: Park, Younghyo, et al.
Publicado: (2024)
por: Park, Younghyo, et al.
Publicado: (2024)
Optimistic Online LQR via Intrinsic Rewards
por: Bartos, Marcell, et al.
Publicado: (2026)
por: Bartos, Marcell, et al.
Publicado: (2026)
Verifier-free Test-Time Sampling for Vision Language Action Models
por: Jang, Suhyeok, et al.
Publicado: (2025)
por: Jang, Suhyeok, et al.
Publicado: (2025)
Investigating Memory in Model-Free RL with POPGym Arcade
por: Wang, Zekang, et al.
Publicado: (2025)
por: Wang, Zekang, et al.
Publicado: (2025)
Language-Conditioned Offline RL for Multi-Robot Navigation
por: Morad, Steven, et al.
Publicado: (2024)
por: Morad, Steven, et al.
Publicado: (2024)
Ejemplares similares
-
Model-Based Reinforcement Learning for Control under Time-Varying Dynamics
por: Iten, Klemens, et al.
Publicado: (2026) -
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
por: Treven, Lenart, et al.
Publicado: (2024) -
NeoRL: Efficient Exploration for Nonepisodic RL
por: Sukhija, Bhavya, et al.
Publicado: (2024) -
Simulation Priors for Data-Efficient Deep Learning
por: Treven, Lenart, et al.
Publicado: (2025) -
Bridging the Sim-to-Real Gap with Bayesian Inference
por: Rothfuss, Jonas, et al.
Publicado: (2024)