Provably Efficient Sample Complexity for Robust CMDP
Fuente:
arXiv
Salvato in:
| Autori principali: | Ganguly, Sourav, Ghosh, Arnob |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
di: Ganguly, Sourav, et al.
Pubblicazione: (2025)
di: Ganguly, Sourav, et al.
Pubblicazione: (2025)
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
di: Pandit, Kartik, et al.
Pubblicazione: (2025)
di: Pandit, Kartik, et al.
Pubblicazione: (2025)
Provably Safe Generative Sampling with Constricting Barrier Functions
di: Gadginmath, Darshan, et al.
Pubblicazione: (2026)
di: Gadginmath, Darshan, et al.
Pubblicazione: (2026)
Imitation-regularized Optimal Transport on Networks: Provable Robustness and Application to Logistics Planning
di: Oishi, Koshi, et al.
Pubblicazione: (2024)
di: Oishi, Koshi, et al.
Pubblicazione: (2024)
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
di: Ganguly, Sourav, et al.
Pubblicazione: (2026)
di: Ganguly, Sourav, et al.
Pubblicazione: (2026)
Policy-based Primal-Dual Methods for Concave CMDP with Variance Reduction
di: Ying, Donghao, et al.
Pubblicazione: (2022)
di: Ying, Donghao, et al.
Pubblicazione: (2022)
Embed to Control Partially Observed Systems: Representation Learning with Provable Sample Efficiency
di: Wang, Lingxiao, et al.
Pubblicazione: (2022)
di: Wang, Lingxiao, et al.
Pubblicazione: (2022)
Sample-Efficient Diffusion-based Control of Complex Physics Systems
di: Chen, Hongyi, et al.
Pubblicazione: (2025)
di: Chen, Hongyi, et al.
Pubblicazione: (2025)
Neural-Rendezvous: Provably Robust Guidance and Control to Encounter Interstellar Objects
di: Tsukamoto, Hiroyasu, et al.
Pubblicazione: (2022)
di: Tsukamoto, Hiroyasu, et al.
Pubblicazione: (2022)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
di: Moghaddam, Amirreza Neshaei, et al.
Pubblicazione: (2025)
di: Moghaddam, Amirreza Neshaei, et al.
Pubblicazione: (2025)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
di: Boveiri, Mohammad, et al.
Pubblicazione: (2024)
di: Boveiri, Mohammad, et al.
Pubblicazione: (2024)
Provably-Stable Neural Network-Based Control of Nonlinear Systems
di: Li, Anran, et al.
Pubblicazione: (2025)
di: Li, Anran, et al.
Pubblicazione: (2025)
Data-Driven Density Steering via the Gromov-Wasserstein Optimal Transport Distance
di: Nakashima, Haruto, et al.
Pubblicazione: (2025)
di: Nakashima, Haruto, et al.
Pubblicazione: (2025)
On the Sample Complexity of Discounted Reinforcement Learning with Optimized Certainty Equivalents
di: Mortensen, Oliver, et al.
Pubblicazione: (2026)
di: Mortensen, Oliver, et al.
Pubblicazione: (2026)
On the Sample Complexity of Imitation Learning for Smoothed Model Predictive Control
di: Pfrommer, Daniel, et al.
Pubblicazione: (2023)
di: Pfrommer, Daniel, et al.
Pubblicazione: (2023)
Local Updates in Distributed Optimization: Provable Acceleration and Topology Effects
di: Wang, Zuang, et al.
Pubblicazione: (2026)
di: Wang, Zuang, et al.
Pubblicazione: (2026)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
di: Muehlebach, Michael, et al.
Pubblicazione: (2025)
di: Muehlebach, Michael, et al.
Pubblicazione: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
di: Moghaddam, Amirreza Neshaei, et al.
Pubblicazione: (2024)
di: Moghaddam, Amirreza Neshaei, et al.
Pubblicazione: (2024)
Estimation Sample Complexity of a Class of Nonlinear Continuous-time Systems
di: Kuang, Simon, et al.
Pubblicazione: (2023)
di: Kuang, Simon, et al.
Pubblicazione: (2023)
Improved Sample Complexity of Imitation Learning for Barrier Model Predictive Control
di: Pfrommer, Daniel, et al.
Pubblicazione: (2024)
di: Pfrommer, Daniel, et al.
Pubblicazione: (2024)
Formation Shape Control using the Gromov-Wasserstein Metric
di: Nakashima, Haruto, et al.
Pubblicazione: (2025)
di: Nakashima, Haruto, et al.
Pubblicazione: (2025)
Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea
di: Krasowski, Hanna, et al.
Pubblicazione: (2024)
di: Krasowski, Hanna, et al.
Pubblicazione: (2024)
Provable Bounds on the Hessian of Neural Networks: Derivative-Preserving Reachability Analysis
di: Sharifi, Sina, et al.
Pubblicazione: (2024)
di: Sharifi, Sina, et al.
Pubblicazione: (2024)
Identification and Adaptive Control of Markov Jump Systems: Sample Complexity and Regret Bounds
di: Sattar, Yahya, et al.
Pubblicazione: (2021)
di: Sattar, Yahya, et al.
Pubblicazione: (2021)
Information Theoretically Optimal Sample Complexity of Learning Dynamical Directed Acyclic Graphs
di: Veedu, Mishfad Shaikh, et al.
Pubblicazione: (2023)
di: Veedu, Mishfad Shaikh, et al.
Pubblicazione: (2023)
Sample Complexity Bounds for Linear System Identification from a Finite Set
di: Chatzikiriakos, Nicolas, et al.
Pubblicazione: (2024)
di: Chatzikiriakos, Nicolas, et al.
Pubblicazione: (2024)
Sample Efficient Certification of Discrete-Time Control Barrier Functions
di: Mulagaleti, Sampath Kumar, et al.
Pubblicazione: (2025)
di: Mulagaleti, Sampath Kumar, et al.
Pubblicazione: (2025)
Provably Optimal Reinforcement Learning under Safety Filtering
di: Oh, Donggeon David, et al.
Pubblicazione: (2025)
di: Oh, Donggeon David, et al.
Pubblicazione: (2025)
SPP-CNN: An Efficient Framework for Network Robustness Prediction
di: Wu, Chengpei, et al.
Pubblicazione: (2023)
di: Wu, Chengpei, et al.
Pubblicazione: (2023)
Provably Bounding Neural Network Preimages
di: Kotha, Suhas, et al.
Pubblicazione: (2023)
di: Kotha, Suhas, et al.
Pubblicazione: (2023)
Achieving $\widetilde{O}(1/ε)$ Sample Complexity for Bilinear Systems Identification under Bounded Noises
di: Yi, Hongyu, et al.
Pubblicazione: (2026)
di: Yi, Hongyu, et al.
Pubblicazione: (2026)
An Efficient Reachability-Based Framework for Provably Safe Autonomous Navigation in Unknown Environments
di: Bajcsy, Andrea, et al.
Pubblicazione: (2019)
di: Bajcsy, Andrea, et al.
Pubblicazione: (2019)
ShieldNN: A Provably Safe NN Filter for Unsafe NN Controllers
di: Ferlez, James, et al.
Pubblicazione: (2020)
di: Ferlez, James, et al.
Pubblicazione: (2020)
Sample-Efficient Linear Representation Learning from Non-IID Non-Isotropic Data
di: Zhang, Thomas T. C. K., et al.
Pubblicazione: (2023)
di: Zhang, Thomas T. C. K., et al.
Pubblicazione: (2023)
Efficient Sampling for Data-Driven Frequency Stability Constraint via Forward-Mode Automatic Differentiation
di: Xu, Wangkun, et al.
Pubblicazione: (2024)
di: Xu, Wangkun, et al.
Pubblicazione: (2024)
Efficient and Robust Freeway Traffic Speed Estimation under Oblique Grid using Vehicle Trajectory Data
di: He, Yang, et al.
Pubblicazione: (2024)
di: He, Yang, et al.
Pubblicazione: (2024)
Transfer Learning Assisted XgBoost For Adaptable Cyberattack Detection In Battery Packs
di: Ghosh, Sanchita, et al.
Pubblicazione: (2025)
di: Ghosh, Sanchita, et al.
Pubblicazione: (2025)
Beyond Freshness and Semantics: A Coupon-Collector Framework for Effective Status Updates
di: Ahmed, Youssef, et al.
Pubblicazione: (2026)
di: Ahmed, Youssef, et al.
Pubblicazione: (2026)
Sample Complexity of the Sign-Perturbed Sums Identification Method: Scalar Case
di: Szentpéteri, Szabolcs, et al.
Pubblicazione: (2024)
di: Szentpéteri, Szabolcs, et al.
Pubblicazione: (2024)
Sample Complexity of the Sign-Perturbed Sums Method
di: Szentpéteri, Szabolcs, et al.
Pubblicazione: (2024)
di: Szentpéteri, Szabolcs, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
di: Ganguly, Sourav, et al.
Pubblicazione: (2025) -
Certifiable Safe RLHF: Fixed-Penalty Constraint Optimization for Safer Language Models
di: Pandit, Kartik, et al.
Pubblicazione: (2025) -
Provably Safe Generative Sampling with Constricting Barrier Functions
di: Gadginmath, Darshan, et al.
Pubblicazione: (2026) -
Imitation-regularized Optimal Transport on Networks: Provable Robustness and Application to Logistics Planning
di: Oishi, Koshi, et al.
Pubblicazione: (2024) -
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
di: Ganguly, Sourav, et al.
Pubblicazione: (2026)