Simulation Priors for Data-Efficient Deep Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Treven, Lenart, Sukhija, Bhavya, Rothfuss, Jonas, Coros, Stelian, Dörfler, Florian, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging the Sim-to-Real Gap with Bayesian Inference
by: Rothfuss, Jonas, et al.
Published: (2024)
by: Rothfuss, Jonas, et al.
Published: (2024)
NeoRL: Efficient Exploration for Nonepisodic RL
by: Sukhija, Bhavya, et al.
Published: (2024)
by: Sukhija, Bhavya, et al.
Published: (2024)
Sample-efficient and Scalable Exploration in Continuous-Time RL
by: Iten, Klemens, et al.
Published: (2025)
by: Iten, Klemens, et al.
Published: (2025)
Transductive Active Learning: Theory and Applications
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
Active Few-Shot Fine-Tuning
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
by: As, Yarden, et al.
Published: (2024)
by: As, Yarden, et al.
Published: (2024)
TARC: Time-Adaptive Robotic Control
by: Sukhija, Arnav, et al.
Published: (2025)
by: Sukhija, Arnav, et al.
Published: (2025)
When to Sense and Control? A Time-adaptive Approach for Continuous-Time RL
by: Treven, Lenart, et al.
Published: (2024)
by: Treven, Lenart, et al.
Published: (2024)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
by: Bhardwaj, Arjun, et al.
Published: (2023)
by: Bhardwaj, Arjun, et al.
Published: (2023)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
by: Sukhija, Bhavya, et al.
Published: (2024)
by: Sukhija, Bhavya, et al.
Published: (2024)
SOMBRL: Scalable and Optimistic Model-Based RL
by: Sukhija, Bhavya, et al.
Published: (2025)
by: Sukhija, Bhavya, et al.
Published: (2025)
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
by: As, Yarden, et al.
Published: (2024)
by: As, Yarden, et al.
Published: (2024)
Model-Based Reinforcement Learning for Control under Time-Varying Dynamics
by: Iten, Klemens, et al.
Published: (2026)
by: Iten, Klemens, et al.
Published: (2026)
Safe Exploration via Policy Priors
by: Wendl, Manuel, et al.
Published: (2026)
by: Wendl, Manuel, et al.
Published: (2026)
Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning
by: Hu, Jiajun, et al.
Published: (2026)
by: Hu, Jiajun, et al.
Published: (2026)
Beyond Binary: Sim-to-Real Dexterous Manipulation with Physics-Grounded Contact Representation
by: Pan, Jiahe, et al.
Published: (2026)
by: Pan, Jiahe, et al.
Published: (2026)
Epistemically-guided forward-backward exploration
by: Urpí, Núria Armengol, et al.
Published: (2025)
by: Urpí, Núria Armengol, et al.
Published: (2025)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
Efficiently Learning at Test-Time: Active Fine-Tuning of LLMs
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
Momentum-Conserving Graph Neural Networks for Deformable Objects
by: Wang, Jiahong, et al.
Published: (2026)
by: Wang, Jiahong, et al.
Published: (2026)
What Matters for Simulation to Online Reinforcement Learning on Real Robots
by: As, Yarden, et al.
Published: (2026)
by: As, Yarden, et al.
Published: (2026)
Probabilistic Artificial Intelligence
by: Krause, Andreas, et al.
Published: (2025)
by: Krause, Andreas, et al.
Published: (2025)
By Fair Means or Foul: Quantifying Collusion in a Market Simulation with Deep Reinforcement Learning
by: Schlechtinger, Michael, et al.
Published: (2024)
by: Schlechtinger, Michael, et al.
Published: (2024)
Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
VQ-Style: Disentangling Style and Content in Motion with Residual Quantized Representations
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
Learning Soft Robotic Dynamics with Active Exploration
by: Zheng, Hehui, et al.
Published: (2025)
by: Zheng, Hehui, et al.
Published: (2025)
Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging
by: Bertolissi, Ryo, et al.
Published: (2025)
by: Bertolissi, Ryo, et al.
Published: (2025)
Learning Safety Constraints for Large Language Models
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles
by: Menet, Nicolas, et al.
Published: (2026)
by: Menet, Nicolas, et al.
Published: (2026)
Symmetry-Guided Memory Augmentation for Efficient Locomotion Learning
by: Bao, Kaixi, et al.
Published: (2025)
by: Bao, Kaixi, et al.
Published: (2025)
Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
Adaptable Cardiovascular Disease Risk Prediction from Heterogeneous Data using Large Language Models
by: Lübeck, Frederike, et al.
Published: (2025)
by: Lübeck, Frederike, et al.
Published: (2025)
What Does Flow Matching Bring To TD Learning?
by: Agrawalla, Bhavya, et al.
Published: (2026)
by: Agrawalla, Bhavya, et al.
Published: (2026)
ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning
by: Melikidze, Davit, et al.
Published: (2026)
by: Melikidze, Davit, et al.
Published: (2026)
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning
by: van der Vaart, Pascal R., et al.
Published: (2025)
by: van der Vaart, Pascal R., et al.
Published: (2025)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration
by: Wilcoxson, Max, et al.
Published: (2024)
by: Wilcoxson, Max, et al.
Published: (2024)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
Similar Items
-
Bridging the Sim-to-Real Gap with Bayesian Inference
by: Rothfuss, Jonas, et al.
Published: (2024) -
NeoRL: Efficient Exploration for Nonepisodic RL
by: Sukhija, Bhavya, et al.
Published: (2024) -
Sample-efficient and Scalable Exploration in Continuous-Time RL
by: Iten, Klemens, et al.
Published: (2025) -
Transductive Active Learning: Theory and Applications
by: Hübotter, Jonas, et al.
Published: (2024) -
Active Few-Shot Fine-Tuning
by: Hübotter, Jonas, et al.
Published: (2024)