Hierarchical Upper Confidence Bounds for Constrained Online Learning
Fuente:
arXiv
Saved in:
| Main Author: | Baheri, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
by: Yifru, Lunet, et al.
Published: (2024)
by: Yifru, Lunet, et al.
Published: (2024)
Optimal Transport-Assisted Risk-Sensitive Q-Learning
by: Shahrooei, Zahra, et al.
Published: (2024)
by: Shahrooei, Zahra, et al.
Published: (2024)
Can Optimal Transport Improve Federated Inverse Reinforcement Learning?
by: Millard, David, et al.
Published: (2026)
by: Millard, David, et al.
Published: (2026)
The Synergy Between Optimal Transport Theory and Multi-Agent Reinforcement Learning
by: Baheri, Ali, et al.
Published: (2024)
by: Baheri, Ali, et al.
Published: (2024)
Metriplectic Conditional Flow Matching for Dissipative Dynamics
by: Baheri, Ali, et al.
Published: (2025)
by: Baheri, Ali, et al.
Published: (2025)
Wasserstein Barycenter Soft Actor-Critic
by: Shahrooei, Zahra, et al.
Published: (2025)
by: Shahrooei, Zahra, et al.
Published: (2025)
Hierarchical Neuro-Symbolic Decision Transformer
by: Baheri, Ali, et al.
Published: (2025)
by: Baheri, Ali, et al.
Published: (2025)
Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning
by: Baheri, Ali, et al.
Published: (2025)
by: Baheri, Ali, et al.
Published: (2025)
Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets
by: Pandian, Shriram Karpoora Sundara, et al.
Published: (2025)
by: Pandian, Shriram Karpoora Sundara, et al.
Published: (2025)
Optimizing Falsification for Learning-Based Control Systems: A Multi-Fidelity Bayesian Approach
by: Shahrooei, Zahra, et al.
Published: (2024)
by: Shahrooei, Zahra, et al.
Published: (2024)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Inference with the Upper Confidence Bound Algorithm
by: Khamaru, Koulik, et al.
Published: (2024)
by: Khamaru, Koulik, et al.
Published: (2024)
A Survey on Reinforcement Learning in Aviation Applications
by: Razzaghi, Pouria, et al.
Published: (2022)
by: Razzaghi, Pouria, et al.
Published: (2022)
Wasserstein-Barycenter Consensus for Cooperative Multi-Agent Reinforcement Learning
by: Baheri, Ali
Published: (2025)
by: Baheri, Ali
Published: (2025)
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning
by: Baheri, Ali
Published: (2025)
by: Baheri, Ali
Published: (2025)
Online Learning for Supervisory Switching Control
by: Sun, Haoyuan, et al.
Published: (2026)
by: Sun, Haoyuan, et al.
Published: (2026)
Goal Reaching with Eikonal-Constrained Hierarchical Quasimetric Reinforcement Learning
by: Giammarino, Vittorio, et al.
Published: (2025)
by: Giammarino, Vittorio, et al.
Published: (2025)
A Deep State-Space Model Compression Method using Upper Bound on Output Error
by: Sakamoto, Hiroki, et al.
Published: (2025)
by: Sakamoto, Hiroki, et al.
Published: (2025)
Hierarchical End-to-End Taylor Bounds for Complete Neural Network Verification
by: Entesari, Taha, et al.
Published: (2026)
by: Entesari, Taha, et al.
Published: (2026)
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
by: Chang, Ting-Jui, et al.
Published: (2024)
by: Chang, Ting-Jui, et al.
Published: (2024)
Resilient Constrained Reinforcement Learning
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Smoothed Online Optimization for Target Tracking: Robust and Learning-Augmented Algorithms
by: Zeynali, Ali, et al.
Published: (2025)
by: Zeynali, Ali, et al.
Published: (2025)
Learning Power Flow with Confidence: A Probabilistic Guarantee Framework for Voltage Risk
by: Pareek, Parikshit, et al.
Published: (2023)
by: Pareek, Parikshit, et al.
Published: (2023)
Robust Online Learning over Networks
by: Bastianello, Nicola, et al.
Published: (2023)
by: Bastianello, Nicola, et al.
Published: (2023)
Physics-Constrained Machine Learning for Chemical Engineering
by: Mukherjee, Angan, et al.
Published: (2025)
by: Mukherjee, Angan, et al.
Published: (2025)
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
by: Zhang, Tianqi, et al.
Published: (2025)
by: Zhang, Tianqi, et al.
Published: (2025)
Meta-Learning for Physically-Constrained Neural System Identification
by: Chakrabarty, Ankush, et al.
Published: (2025)
by: Chakrabarty, Ankush, et al.
Published: (2025)
Safe Online Control-Informed Learning
by: Zhou, Tianyu, et al.
Published: (2025)
by: Zhou, Tianyu, et al.
Published: (2025)
Derandomizing Simultaneous Confidence Regions for Band-Limited Functions by Improved Norm Bounds and Majority-Voting Schemes
by: Csáji, Balázs Csanád, et al.
Published: (2025)
by: Csáji, Balázs Csanád, et al.
Published: (2025)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
by: Schmidt, Carolin, et al.
Published: (2024)
by: Schmidt, Carolin, et al.
Published: (2024)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
Distribution-Free Confidence Ellipsoids for Ridge Regression with PAC Bounds
by: Szentpéteri, Szabolcs, et al.
Published: (2026)
by: Szentpéteri, Szabolcs, et al.
Published: (2026)
Logic-Guided Vector Fields for Constrained Generative Modeling
by: Baheri, Ali
Published: (2026)
by: Baheri, Ali
Published: (2026)
On Robust Reinforcement Learning with Lipschitz-Bounded Policy Networks
by: Barbara, Nicholas H., et al.
Published: (2024)
by: Barbara, Nicholas H., et al.
Published: (2024)
HUANet: Hard-Constrained Unrolled ADMM for Constrained Convex Optimization
by: Tran, Trinh, et al.
Published: (2026)
by: Tran, Trinh, et al.
Published: (2026)
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
by: Takubo, Yuji, et al.
Published: (2021)
by: Takubo, Yuji, et al.
Published: (2021)
Nonparametric Sparse Online Learning of the Koopman Operator
by: Hou, Boya, et al.
Published: (2025)
by: Hou, Boya, et al.
Published: (2025)
Adaptive Decentralized Federated Learning in Energy and Latency Constrained Wireless Networks
by: Yan, Zhigang, et al.
Published: (2024)
by: Yan, Zhigang, et al.
Published: (2024)
ECO: Energy-Constrained Operator Learning for Chaotic Dynamics with Boundedness Guarantees
by: Goertzen, Andrea, et al.
Published: (2025)
by: Goertzen, Andrea, et al.
Published: (2025)
Improving Stochastic Action-Constrained Reinforcement Learning via Truncated Distributions
by: Stolz, Roland, et al.
Published: (2025)
by: Stolz, Roland, et al.
Published: (2025)
Similar Items
-
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
by: Yifru, Lunet, et al.
Published: (2024) -
Optimal Transport-Assisted Risk-Sensitive Q-Learning
by: Shahrooei, Zahra, et al.
Published: (2024) -
Can Optimal Transport Improve Federated Inverse Reinforcement Learning?
by: Millard, David, et al.
Published: (2026) -
The Synergy Between Optimal Transport Theory and Multi-Agent Reinforcement Learning
by: Baheri, Ali, et al.
Published: (2024) -
Metriplectic Conditional Flow Matching for Dissipative Dynamics
by: Baheri, Ali, et al.
Published: (2025)