TubeDAgger: Reducing the Number of Expert Interventions with Stochastic Reach-Tubes
Fuente:
arXiv
Saved in:
| Main Authors: | Lemmel, Julian, Kranzl, Manuel, Lamine, Adam, Neubauer, Philipp, Grosu, Radu, Neubauer, Sophie A. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Fine-Tuning of Carbon Emission Predictions using Real-Time Recurrent Learning for State Space Models
by: Lemmel, Julian, et al.
Published: (2025)
by: Lemmel, Julian, et al.
Published: (2025)
Real-Time Recurrent Reinforcement Learning
by: Lemmel, Julian, et al.
Published: (2023)
by: Lemmel, Julian, et al.
Published: (2023)
Scalable Offline Reinforcement Learning for Mean Field Games
by: Brunnbauer, Axel, et al.
Published: (2024)
by: Brunnbauer, Axel, et al.
Published: (2024)
A quantum-classical reinforcement learning model to play Atari games
by: Freinberger, Dominik, et al.
Published: (2024)
by: Freinberger, Dominik, et al.
Published: (2024)
Verification of Neural Reachable Tubes via Scenario Optimization and Conformal Prediction
by: Lin, Albert, et al.
Published: (2023)
by: Lin, Albert, et al.
Published: (2023)
Adaptive Control in Autonomous Driving via Real-Time Recurrent RL
by: Lemmel, Julian, et al.
Published: (2026)
by: Lemmel, Julian, et al.
Published: (2026)
Agentic AI Home Energy Management System: A Large Language Model Framework for Residential Load Scheduling
by: Makroum, Reda El, et al.
Published: (2025)
by: Makroum, Reda El, et al.
Published: (2025)
Conversational Demand Response: Bidirectional Aggregator-Prosumer Coordination through Agentic AI
by: Makroum, Reda El, et al.
Published: (2026)
by: Makroum, Reda El, et al.
Published: (2026)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
by: Li, Jingqi, et al.
Published: (2022)
by: Li, Jingqi, et al.
Published: (2022)
Safe Adaptive Cruise Control Under Perception Uncertainty: A Deep Ensemble and Conformal Tube Model Predictive Control Approach
by: Li, Xiao, et al.
Published: (2024)
by: Li, Xiao, et al.
Published: (2024)
Annealing Optimization for Progressive Learning with Stochastic Approximation
by: Mavridis, Christos, et al.
Published: (2022)
by: Mavridis, Christos, et al.
Published: (2022)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
Conformal Safety Shielding for Imperfect-Perception Agents
by: Scarbro, William, et al.
Published: (2025)
by: Scarbro, William, et al.
Published: (2025)
Stochastic Actor-Critic: Mitigating Overestimation via Temporal Aleatoric Uncertainty
by: Özalp, Uğurcan
Published: (2026)
by: Özalp, Uğurcan
Published: (2026)
Stochastic Learning of Computational Resource Usage as Graph Structured Multimarginal Schrödinger Bridge
by: Bondar, Georgiy A., et al.
Published: (2024)
by: Bondar, Georgiy A., et al.
Published: (2024)
Dynamic Tube MPC: Learning Tube Dynamics with Massively Parallel Simulation for Robust Safety in Practice
by: Compton, William D., et al.
Published: (2024)
by: Compton, William D., et al.
Published: (2024)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
by: Zhu, Feng, et al.
Published: (2025)
by: Zhu, Feng, et al.
Published: (2025)
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation
by: Fabbro, Nicolò Dal, et al.
Published: (2024)
by: Fabbro, Nicolò Dal, et al.
Published: (2024)
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control
by: Skifstad, Julian, et al.
Published: (2026)
by: Skifstad, Julian, et al.
Published: (2026)
Spatiotemporal Tubes for Temporal Reach-Avoid-Stay Tasks in Unknown Systems
by: Das, Ratnangshu, et al.
Published: (2024)
by: Das, Ratnangshu, et al.
Published: (2024)
Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems
by: Jayawardana, Vindula, et al.
Published: (2025)
by: Jayawardana, Vindula, et al.
Published: (2025)
Control Policy Correction Framework for Reinforcement Learning-based Energy Arbitrage Strategies
by: Madahi, Seyed Soroush Karimi, et al.
Published: (2024)
by: Madahi, Seyed Soroush Karimi, et al.
Published: (2024)
Towards Developing Socially Compliant Automated Vehicles: Advances, Expert Insights, and A Conceptual Framework
by: Dong, Yongqi, et al.
Published: (2025)
by: Dong, Yongqi, et al.
Published: (2025)
MA-VAE: Multi-head Attention-based Variational Autoencoder Approach for Anomaly Detection in Multivariate Time-series Applied to Automotive Endurance Powertrain Testing
by: Correia, Lucas, et al.
Published: (2023)
by: Correia, Lucas, et al.
Published: (2023)
Online Model-based Anomaly Detection in Multivariate Time Series: Taxonomy, Survey, Research Challenges and Future Directions
by: Correia, Lucas, et al.
Published: (2024)
by: Correia, Lucas, et al.
Published: (2024)
Smooth Spatiotemporal Tube Synthesis for Prescribed-Time Reach-Avoid-Stay Control
by: Upadhyay, Siddhartha, et al.
Published: (2025)
by: Upadhyay, Siddhartha, et al.
Published: (2025)
Reduce, Reuse, Recycle: Categories for Compositional Reinforcement Learning
by: Bakirtzis, Georgios, et al.
Published: (2024)
by: Bakirtzis, Georgios, et al.
Published: (2024)
The Economic Dispatch of Power-to-Gas Systems with Deep Reinforcement Learning:Tackling the Challenge of Delayed Rewards with Long-Term Energy Storage
by: Sage, Manuel, et al.
Published: (2025)
by: Sage, Manuel, et al.
Published: (2025)
A TinyML Reinforcement Learning Approach for Energy-Efficient Light Control in Low-Cost Greenhouse Systems
by: Salem, Mohamed Abdallah, et al.
Published: (2025)
by: Salem, Mohamed Abdallah, et al.
Published: (2025)
Enhancing Battery Storage Energy Arbitrage with Deep Reinforcement Learning and Time-Series Forecasting
by: Sage, Manuel, et al.
Published: (2024)
by: Sage, Manuel, et al.
Published: (2024)
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model
by: Zhou, Zhehua, et al.
Published: (2024)
by: Zhou, Zhehua, et al.
Published: (2024)
Spatiotemporal Tubes for Probabilistic Temporal Reach-Avoid-Stay Task in Uncertain Dynamic Environment
by: Upadhyay, Siddhartha, et al.
Published: (2025)
by: Upadhyay, Siddhartha, et al.
Published: (2025)
Temporal Reach-Avoid-Stay Control for Differential Drive Systems via Spatiotemporal Tubes
by: Das, Ratnangshu, et al.
Published: (2025)
by: Das, Ratnangshu, et al.
Published: (2025)
Generative Stochastic Optimal Transport: Guided Harmonic Path-Integral Diffusion
by: Chertkov, Michael
Published: (2025)
by: Chertkov, Michael
Published: (2025)
Reinforcement learning meets bioprocess control through behaviour cloning: Real-world deployment in an industrial photobioreactor
by: Gil, Juan D., et al.
Published: (2025)
by: Gil, Juan D., et al.
Published: (2025)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
End-to-End Reinforcement Learning of Curative Curtailment with Partial Measurement Availability
by: Wolf, Hinrikus, et al.
Published: (2024)
by: Wolf, Hinrikus, et al.
Published: (2024)
Synaptic Activation and Dual Liquid Dynamics for Interpretable Bio-Inspired Models
by: Farsang, Mónika, et al.
Published: (2026)
by: Farsang, Mónika, et al.
Published: (2026)
Liquid Resistance Liquid Capacitance Networks
by: Farsang, Mónika, et al.
Published: (2024)
by: Farsang, Mónika, et al.
Published: (2024)
Similar Items
-
Online Fine-Tuning of Carbon Emission Predictions using Real-Time Recurrent Learning for State Space Models
by: Lemmel, Julian, et al.
Published: (2025) -
Real-Time Recurrent Reinforcement Learning
by: Lemmel, Julian, et al.
Published: (2023) -
Scalable Offline Reinforcement Learning for Mean Field Games
by: Brunnbauer, Axel, et al.
Published: (2024) -
A quantum-classical reinforcement learning model to play Atari games
by: Freinberger, Dominik, et al.
Published: (2024) -
Verification of Neural Reachable Tubes via Scenario Optimization and Conformal Prediction
by: Lin, Albert, et al.
Published: (2023)