ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | As, Yarden, Sukhija, Bhavya, Treven, Lenart, Sferrazza, Carmelo, Coros, Stelian, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging the Sim-to-Real Gap with Bayesian Inference
von: Rothfuss, Jonas, et al.
Veröffentlicht: (2024)
von: Rothfuss, Jonas, et al.
Veröffentlicht: (2024)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
NeoRL: Efficient Exploration for Nonepisodic RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024)
TARC: Time-Adaptive Robotic Control
von: Sukhija, Arnav, et al.
Veröffentlicht: (2025)
von: Sukhija, Arnav, et al.
Veröffentlicht: (2025)
Transductive Active Learning: Theory and Applications
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
Sample-efficient and Scalable Exploration in Continuous-Time RL
von: Iten, Klemens, et al.
Veröffentlicht: (2025)
von: Iten, Klemens, et al.
Veröffentlicht: (2025)
Active Few-Shot Fine-Tuning
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)
Simulation Priors for Data-Efficient Deep Learning
von: Treven, Lenart, et al.
Veröffentlicht: (2025)
von: Treven, Lenart, et al.
Veröffentlicht: (2025)
Model-Based Reinforcement Learning for Control under Time-Varying Dynamics
von: Iten, Klemens, et al.
Veröffentlicht: (2026)
von: Iten, Klemens, et al.
Veröffentlicht: (2026)
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
von: Treven, Lenart, et al.
Veröffentlicht: (2024)
von: Treven, Lenart, et al.
Veröffentlicht: (2024)
SOMBRL: Scalable and Optimistic Model-Based RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
Safe Exploration via Policy Priors
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
von: Wendl, Manuel, et al.
Veröffentlicht: (2026)
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
MetaLoco: Universal Quadrupedal Locomotion with Meta-Reinforcement Learning and Motion Imitation
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
Sampling-Based Safe Reinforcement Learning
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
von: Vignola, Luca, et al.
Veröffentlicht: (2026)
Problem Space Transformations for Out-of-Distribution Generalisation in Behavioural Cloning
von: Doshi, Kiran, et al.
Veröffentlicht: (2024)
von: Doshi, Kiran, et al.
Veröffentlicht: (2024)
What Matters for Simulation to Online Reinforcement Learning on Real Robots
von: As, Yarden, et al.
Veröffentlicht: (2026)
von: As, Yarden, et al.
Veröffentlicht: (2026)
Learning Soft Robotic Dynamics with Active Exploration
von: Zheng, Hehui, et al.
Veröffentlicht: (2025)
von: Zheng, Hehui, et al.
Veröffentlicht: (2025)
CAIMAN: Causal Action Influence Detection for Sample-efficient Loco-manipulation
von: Yuan, Yuanchen, et al.
Veröffentlicht: (2025)
von: Yuan, Yuanchen, et al.
Veröffentlicht: (2025)
Learning Safety Constraints for Large Language Models
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Beyond Binary: Sim-to-Real Dexterous Manipulation with Physics-Grounded Contact Representation
von: Pan, Jiahe, et al.
Veröffentlicht: (2026)
von: Pan, Jiahe, et al.
Veröffentlicht: (2026)
Rethinking Robustness Assessment: Adversarial Attacks on Learning-based Quadrupedal Locomotion Controllers
von: Shi, Fan, et al.
Veröffentlicht: (2024)
von: Shi, Fan, et al.
Veröffentlicht: (2024)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
von: Bao, Kaixi, et al.
Veröffentlicht: (2025)
von: Bao, Kaixi, et al.
Veröffentlicht: (2025)
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
von: Seo, Younggyo, et al.
Veröffentlicht: (2025)
von: Seo, Younggyo, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
von: Liu, Puze, et al.
Veröffentlicht: (2024)
von: Liu, Puze, et al.
Veröffentlicht: (2024)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
Body Transformer: Leveraging Robot Embodiment for Policy Learning
von: Sferrazza, Carmelo, et al.
Veröffentlicht: (2024)
von: Sferrazza, Carmelo, et al.
Veröffentlicht: (2024)
Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning
von: Günster, Jonas, et al.
Veröffentlicht: (2024)
von: Günster, Jonas, et al.
Veröffentlicht: (2024)
Safe Guaranteed Exploration for Non-linear Systems
von: Prajapat, Manish, et al.
Veröffentlicht: (2024)
von: Prajapat, Manish, et al.
Veröffentlicht: (2024)
Revisiting Safe Exploration in Safe Reinforcement learning
von: Eckel, David, et al.
Veröffentlicht: (2024)
von: Eckel, David, et al.
Veröffentlicht: (2024)
Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning
von: Hu, Jiajun, et al.
Veröffentlicht: (2026)
von: Hu, Jiajun, et al.
Veröffentlicht: (2026)
SPiDR: A Simple Approach for Zero-Shot Safety in Sim-to-Real Transfer
von: As, Yarden, et al.
Veröffentlicht: (2025)
von: As, Yarden, et al.
Veröffentlicht: (2025)
Learning Sim-to-Real Humanoid Locomotion in 15 Minutes
von: Seo, Younggyo, et al.
Veröffentlicht: (2025)
von: Seo, Younggyo, et al.
Veröffentlicht: (2025)
Active Exploration in Bayesian Model-based Reinforcement Learning for Robot Manipulation
von: Plou, Carlos, et al.
Veröffentlicht: (2024)
von: Plou, Carlos, et al.
Veröffentlicht: (2024)
Safe Offline Reinforcement Learning with Real-Time Budget Constraints
von: Lin, Qian, et al.
Veröffentlicht: (2023)
von: Lin, Qian, et al.
Veröffentlicht: (2023)
HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
von: Sferrazza, Carmelo, et al.
Veröffentlicht: (2024)
von: Sferrazza, Carmelo, et al.
Veröffentlicht: (2024)
AnySafe: Adapting Latent Safety Filters at Runtime via Safety Constraint Parameterization in the Latent Space
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
von: Agrawal, Sankalp, et al.
Veröffentlicht: (2025)
Active Fine-Tuning of Multi-Task Policies
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Bridging the Sim-to-Real Gap with Bayesian Inference
von: Rothfuss, Jonas, et al.
Veröffentlicht: (2024) -
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024) -
NeoRL: Efficient Exploration for Nonepisodic RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2024) -
TARC: Time-Adaptive Robotic Control
von: Sukhija, Arnav, et al.
Veröffentlicht: (2025) -
Transductive Active Learning: Theory and Applications
von: Hübotter, Jonas, et al.
Veröffentlicht: (2024)