Learning to explore when mistakes are not allowed
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pecqueux-Guézénec, Charly, Doncieux, Stéphane, Perrin-Gilbert, Nicolas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Pitfalls of Imitation Learning when Actions are Continuous
von: Simchowitz, Max, et al.
Veröffentlicht: (2025)
von: Simchowitz, Max, et al.
Veröffentlicht: (2025)
Safe exploration in reproducing kernel Hilbert spaces
von: Tokmak, Abdullah, et al.
Veröffentlicht: (2025)
von: Tokmak, Abdullah, et al.
Veröffentlicht: (2025)
High Effort, Low Gain: Fundamental Limits of Active Learning for Linear Dynamical Systems
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2025)
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2025)
Operator Models for Continuous-Time Offline Reinforcement Learning
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025)
Sample Complexity Bounds for Linear System Identification from a Finite Set
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
End-to-end guarantees for indirect data-driven control of bilinear systems with finite stochastic data
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
Data-efficient, Explainable and Safe Box Manipulation: Illustrating the Advantages of Physical Priors in Model-Predictive Control
von: Salehi, Achkan, et al.
Veröffentlicht: (2023)
von: Salehi, Achkan, et al.
Veröffentlicht: (2023)
Data-Driven Stochastic Optimal Control in Reproducing Kernel Hilbert Spaces
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2024)
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2024)
Fast and Reliable $N-k$ Contingency Screening with Input-Convex Neural Networks
von: Christianson, Nicolas, et al.
Veröffentlicht: (2024)
von: Christianson, Nicolas, et al.
Veröffentlicht: (2024)
Algorithmic Control Improves Residential Building Energy and EV Management when PV Capacity is High but Battery Capacity is Low
von: Ullner, Lennart, et al.
Veröffentlicht: (2025)
von: Ullner, Lennart, et al.
Veröffentlicht: (2025)
Reference-Free Sampling-Based Model Predictive Control
von: Schramm, Fabian, et al.
Veröffentlicht: (2025)
von: Schramm, Fabian, et al.
Veröffentlicht: (2025)
Safely Learning Dynamical Systems
von: Ahmadi, Amir Ali, et al.
Veröffentlicht: (2023)
von: Ahmadi, Amir Ali, et al.
Veröffentlicht: (2023)
Learning Dissipative Neural Dynamical Systems
von: Xu, Yuezhu, et al.
Veröffentlicht: (2023)
von: Xu, Yuezhu, et al.
Veröffentlicht: (2023)
Robust Online Learning over Networks
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
von: Bastianello, Nicola, et al.
Veröffentlicht: (2023)
Optimization and Learning in Open Multi-Agent Systems
von: Deplano, Diego, et al.
Veröffentlicht: (2025)
von: Deplano, Diego, et al.
Veröffentlicht: (2025)
Why Line Search when you can Plane Search? SO-Friendly Neural Networks allow Per-Iteration Optimization of Learning and Momentum Rates for Every Layer
von: Shea, Betty, et al.
Veröffentlicht: (2024)
von: Shea, Betty, et al.
Veröffentlicht: (2024)
Principled Learning-to-Communicate with Quasi-Classical Information Structures
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2026)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
von: Zolman, Nicholas, et al.
Veröffentlicht: (2024)
von: Zolman, Nicholas, et al.
Veröffentlicht: (2024)
Resilient Constrained Reinforcement Learning
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
von: Syed, Shahbaz P Qadri, et al.
Veröffentlicht: (2025)
Communication-Efficient Learning for Satellite Constellations
von: Tudose, Ruxandra-Stefania, et al.
Veröffentlicht: (2025)
von: Tudose, Ruxandra-Stefania, et al.
Veröffentlicht: (2025)
Safe Online Control-Informed Learning
von: Zhou, Tianyu, et al.
Veröffentlicht: (2025)
von: Zhou, Tianyu, et al.
Veröffentlicht: (2025)
Communication-Efficient Stochastic Distributed Learning
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
Learning to Sparsify Stochastic Linear Bandits
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
von: Wang, Zhengmiao, et al.
Veröffentlicht: (2026)
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Online Learning for Supervisory Switching Control
von: Sun, Haoyuan, et al.
Veröffentlicht: (2026)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2026)
Integration Matters for Learning PDEs with Backward SDEs
von: Park, Sungje, et al.
Veröffentlicht: (2025)
von: Park, Sungje, et al.
Veröffentlicht: (2025)
Jointly Computation- and Communication-Efficient Distributed Learning
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
von: Ren, Xiaoxing, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning via Inverse Optimization
von: Dimanidis, Ioannis, et al.
Veröffentlicht: (2025)
von: Dimanidis, Ioannis, et al.
Veröffentlicht: (2025)
Modular Distributed Nonconvex Learning with Error Feedback
von: Carnevale, Guido, et al.
Veröffentlicht: (2025)
von: Carnevale, Guido, et al.
Veröffentlicht: (2025)
(Un)supervised Learning of Maximal Lyapunov Functions
von: Barreau, Matthieu, et al.
Veröffentlicht: (2024)
von: Barreau, Matthieu, et al.
Veröffentlicht: (2024)
Learning Linear Dynamics from Bilinear Observations
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
Robust Q-Learning under Corrupted Rewards
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2024)
Learning Exactly Linearizable Deep Dynamics Models
von: Moriyasu, Ryuta, et al.
Veröffentlicht: (2023)
von: Moriyasu, Ryuta, et al.
Veröffentlicht: (2023)
Kernel-Based Optimal Control: An Infinitesimal Generator Approach
von: Bevanda, Petar, et al.
Veröffentlicht: (2024)
von: Bevanda, Petar, et al.
Veröffentlicht: (2024)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
von: Yifru, Lunet, et al.
Veröffentlicht: (2024)
von: Yifru, Lunet, et al.
Veröffentlicht: (2024)
Cross-fitted Proximal Learning for Model-Based Reinforcement Learning
von: Venkatesh, Nishanth, et al.
Veröffentlicht: (2026)
von: Venkatesh, Nishanth, et al.
Veröffentlicht: (2026)
Optimism as Risk-Seeking in Multi-Agent Reinforcement Learning
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
Meta-Learning for Physically-Constrained Neural System Identification
von: Chakrabarty, Ankush, et al.
Veröffentlicht: (2025)
von: Chakrabarty, Ankush, et al.
Veröffentlicht: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
von: Manenti, Massimiliano, et al.
Veröffentlicht: (2025)
von: Manenti, Massimiliano, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Pitfalls of Imitation Learning when Actions are Continuous
von: Simchowitz, Max, et al.
Veröffentlicht: (2025) -
Safe exploration in reproducing kernel Hilbert spaces
von: Tokmak, Abdullah, et al.
Veröffentlicht: (2025) -
High Effort, Low Gain: Fundamental Limits of Active Learning for Linear Dynamical Systems
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2025) -
Operator Models for Continuous-Time Offline Reinforcement Learning
von: Hoischen, Nicolas, et al.
Veröffentlicht: (2025) -
Sample Complexity Bounds for Linear System Identification from a Finite Set
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)