Sample Efficient Reinforcement Learning with Partial Dynamics Knowledge
Fuente:
arXiv
Guardado en:
| Autores principales: | Alharbi, Meshal, Roozbehani, Mardavij, Dahleh, Munther |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Compositional Online Learning for Multi-Objective System Co-Design
por: Alharbi, Meshal, et al.
Publicado: (2026)
por: Alharbi, Meshal, et al.
Publicado: (2026)
VITA: Variational Pretraining of Transformers for Climate-Robust Crop Yield Forecasting
por: Hasan, Adib, et al.
Publicado: (2025)
por: Hasan, Adib, et al.
Publicado: (2025)
WeatherFormer: A Pretrained Encoder Model for Learning Robust Weather Representations from Small Datasets
por: Hasan, Adib, et al.
Publicado: (2024)
por: Hasan, Adib, et al.
Publicado: (2024)
Finite Sample Analysis of Tensor Decomposition for Learning Mixtures of Linear Systems
por: Rui, Maryann, et al.
Publicado: (2024)
por: Rui, Maryann, et al.
Publicado: (2024)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
por: Cai, Qi, et al.
Publicado: (2022)
por: Cai, Qi, et al.
Publicado: (2022)
Sample-Efficient Reinforcement Learning of Koopman eNMPC
por: Mayfrank, Daniel, et al.
Publicado: (2025)
por: Mayfrank, Daniel, et al.
Publicado: (2025)
Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
por: Wan, Jia, et al.
Publicado: (2024)
por: Wan, Jia, et al.
Publicado: (2024)
Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems
por: Suttle, Wesley A., et al.
Publicado: (2024)
por: Suttle, Wesley A., et al.
Publicado: (2024)
Coreset-Based Task Selection for Sample-Efficient Meta-Reinforcement Learning
por: Zhan, Donglin, et al.
Publicado: (2025)
por: Zhan, Donglin, et al.
Publicado: (2025)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
por: Sattar, Yahya, et al.
Publicado: (2025)
por: Sattar, Yahya, et al.
Publicado: (2025)
TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
por: Sorokin, D., et al.
Publicado: (2023)
por: Sorokin, D., et al.
Publicado: (2023)
Efficient Spectral Control of Partially Observed Linear Dynamical Systems
por: Brahmbhatt, Anand, et al.
Publicado: (2025)
por: Brahmbhatt, Anand, et al.
Publicado: (2025)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
por: Chen, Zijun, et al.
Publicado: (2025)
por: Chen, Zijun, et al.
Publicado: (2025)
Population Dynamics Control with Partial Observations
por: Lu, Zhou, et al.
Publicado: (2025)
por: Lu, Zhou, et al.
Publicado: (2025)
Convergence and Sample Complexity of First-Order Methods for Agnostic Reinforcement Learning
por: Sherman, Uri, et al.
Publicado: (2025)
por: Sherman, Uri, et al.
Publicado: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2024)
por: Moghaddam, Amirreza Neshaei, et al.
Publicado: (2024)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
por: Muehlebach, Michael, et al.
Publicado: (2025)
por: Muehlebach, Michael, et al.
Publicado: (2025)
Scalable Co-Design via Linear Design Problems: Compositional Theory and Algorithms
por: Cai, Yubo, et al.
Publicado: (2026)
por: Cai, Yubo, et al.
Publicado: (2026)
Joint MDPs and Reinforcement Learning in Coupled-Dynamics Environments
por: Kaya, Ege C., et al.
Publicado: (2026)
por: Kaya, Ege C., et al.
Publicado: (2026)
Integrating Human Knowledge Through Action Masking in Reinforcement Learning for Operations Research
por: Stappert, Mirko, et al.
Publicado: (2025)
por: Stappert, Mirko, et al.
Publicado: (2025)
Mitigating Information Asymmetry in Two-Stage Contracts with Non-Myopic Agents
por: Dahleh, Munther A., et al.
Publicado: (2024)
por: Dahleh, Munther A., et al.
Publicado: (2024)
Deep Reinforcement Learning for Dynamic Order Picking in Warehouse Operations
por: Mahmoudinazlou, Sasan, et al.
Publicado: (2024)
por: Mahmoudinazlou, Sasan, et al.
Publicado: (2024)
Sample-Efficient Agnostic Boosting
por: Ghai, Udaya, et al.
Publicado: (2024)
por: Ghai, Udaya, et al.
Publicado: (2024)
Reinforcement Learning Approaches for the Orienteering Problem with Stochastic and Dynamic Release Dates
por: Li, Yuanyuan, et al.
Publicado: (2022)
por: Li, Yuanyuan, et al.
Publicado: (2022)
Single- vs. Dual-Policy Reinforcement Learning for Dynamic Bike Rebalancing
por: Liang, Jiaqi, et al.
Publicado: (2024)
por: Liang, Jiaqi, et al.
Publicado: (2024)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
por: Qiu, Shuang, et al.
Publicado: (2024)
por: Qiu, Shuang, et al.
Publicado: (2024)
Information Theoretically Optimal Sample Complexity of Learning Dynamical Directed Acyclic Graphs
por: Veedu, Mishfad Shaikh, et al.
Publicado: (2023)
por: Veedu, Mishfad Shaikh, et al.
Publicado: (2023)
Safe and Efficient Online Convex Optimization with Linear Budget Constraints and Partial Feedback
por: Liu, Shanqi, et al.
Publicado: (2024)
por: Liu, Shanqi, et al.
Publicado: (2024)
A Computation and Communication Efficient Method for Distributed Nonconvex Problems in the Partial Participation Setting
por: Tyurin, Alexander, et al.
Publicado: (2022)
por: Tyurin, Alexander, et al.
Publicado: (2022)
Learning to Solve Optimization Problems Constrained with Partial Differential Equations
por: Guven, Yusuf, et al.
Publicado: (2025)
por: Guven, Yusuf, et al.
Publicado: (2025)
Control and optimization for Neural Partial Differential Equations in Supervised Learning
por: Bensoussan, Alain, et al.
Publicado: (2025)
por: Bensoussan, Alain, et al.
Publicado: (2025)
ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule
por: Huang, Yilie, et al.
Publicado: (2026)
por: Huang, Yilie, et al.
Publicado: (2026)
Resilient Constrained Reinforcement Learning
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning
por: Jia, Yanwei, et al.
Publicado: (2025)
por: Jia, Yanwei, et al.
Publicado: (2025)
Delayed Momentum Aggregation: Communication-efficient Byzantine-robust Federated Learning with Partial Participation
por: Otsuka, Kaoru, et al.
Publicado: (2025)
por: Otsuka, Kaoru, et al.
Publicado: (2025)
Residuals-based Offline Reinforcement Learning
por: Zhu, Qing, et al.
Publicado: (2026)
por: Zhu, Qing, et al.
Publicado: (2026)
A Pontryagin Perspective on Reinforcement Learning
por: Eberhard, Onno, et al.
Publicado: (2024)
por: Eberhard, Onno, et al.
Publicado: (2024)
Foundations of Multivariate Distributional Reinforcement Learning
por: Wiltzer, Harley, et al.
Publicado: (2024)
por: Wiltzer, Harley, et al.
Publicado: (2024)
Operator World Models for Reinforcement Learning
por: Novelli, Pietro, et al.
Publicado: (2024)
por: Novelli, Pietro, et al.
Publicado: (2024)
Finite-Sample-Based Reachability for Safe Control with Gaussian Process Dynamics
por: Prajapat, Manish, et al.
Publicado: (2025)
por: Prajapat, Manish, et al.
Publicado: (2025)
Ejemplares similares
-
Compositional Online Learning for Multi-Objective System Co-Design
por: Alharbi, Meshal, et al.
Publicado: (2026) -
VITA: Variational Pretraining of Transformers for Climate-Robust Crop Yield Forecasting
por: Hasan, Adib, et al.
Publicado: (2025) -
WeatherFormer: A Pretrained Encoder Model for Learning Robust Weather Representations from Small Datasets
por: Hasan, Adib, et al.
Publicado: (2024) -
Finite Sample Analysis of Tensor Decomposition for Learning Mixtures of Linear Systems
por: Rui, Maryann, et al.
Publicado: (2024) -
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
por: Cai, Qi, et al.
Publicado: (2022)