Policy Transfer for Continuous-Time Reinforcement Learning: A (Rough) Differential Equation Approach
Fuente:
arXiv
Guardado en:
| Autores principales: | Guo, Xin, Lyu, Zijiu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
por: Cheng, Ziheng, et al.
Publicado: (2025)
por: Cheng, Ziheng, et al.
Publicado: (2025)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
por: Qiu, Shuang, et al.
Publicado: (2024)
por: Qiu, Shuang, et al.
Publicado: (2024)
Operator Models for Continuous-Time Offline Reinforcement Learning
por: Hoischen, Nicolas, et al.
Publicado: (2025)
por: Hoischen, Nicolas, et al.
Publicado: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
por: Zhang, Chenyu, et al.
Publicado: (2024)
por: Zhang, Chenyu, et al.
Publicado: (2024)
Action Gaps and Advantages in Continuous-Time Distributional Reinforcement Learning
por: Wiltzer, Harley, et al.
Publicado: (2024)
por: Wiltzer, Harley, et al.
Publicado: (2024)
Continuous Q-Score Matching: Diffusion Guided Reinforcement Learning for Continuous-Time Control
por: Hua, Chengxiu, et al.
Publicado: (2025)
por: Hua, Chengxiu, et al.
Publicado: (2025)
Fast Policy Learning for Linear Quadratic Control with Entropy Regularization
por: Guo, Xin, et al.
Publicado: (2023)
por: Guo, Xin, et al.
Publicado: (2023)
A Minibatch-SGD-Based Learning Meta-Policy for Inventory Systems with Myopic Optimal Policy
por: Lyu, Jiameng, et al.
Publicado: (2024)
por: Lyu, Jiameng, et al.
Publicado: (2024)
Learning to Solve Optimization Problems Constrained with Partial Differential Equations
por: Guven, Yusuf, et al.
Publicado: (2025)
por: Guven, Yusuf, et al.
Publicado: (2025)
Control and optimization for Neural Partial Differential Equations in Supervised Learning
por: Bensoussan, Alain, et al.
Publicado: (2025)
por: Bensoussan, Alain, et al.
Publicado: (2025)
Achieve Performatively Optimal Policy for Performative Reinforcement Learning
por: Chen, Ziyi, et al.
Publicado: (2025)
por: Chen, Ziyi, et al.
Publicado: (2025)
Accuracy of Discretely Sampled Stochastic Policies in Continuous-time Reinforcement Learning
por: Jia, Yanwei, et al.
Publicado: (2025)
por: Jia, Yanwei, et al.
Publicado: (2025)
Deep Reinforcement Learning: A Convex Optimization Approach
por: Gattami, Ather
Publicado: (2024)
por: Gattami, Ather
Publicado: (2024)
A Hessian-Aware Stochastic Differential Equation for Modelling SGD
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
Control Theoretic Approach to Fine-Tuning and Transfer Learning
por: Bayram, Erkan, et al.
Publicado: (2024)
por: Bayram, Erkan, et al.
Publicado: (2024)
Non-Parametric Learning of Stochastic Differential Equations with Non-asymptotic Fast Rates of Convergence
por: Bonalli, Riccardo, et al.
Publicado: (2023)
por: Bonalli, Riccardo, et al.
Publicado: (2023)
Single- vs. Dual-Policy Reinforcement Learning for Dynamic Bike Rebalancing
por: Liang, Jiaqi, et al.
Publicado: (2024)
por: Liang, Jiaqi, et al.
Publicado: (2024)
A Differential and Pointwise Control Approach to Reinforcement Learning
por: Nguyen, Minh, et al.
Publicado: (2024)
por: Nguyen, Minh, et al.
Publicado: (2024)
A Minimal-Assumption Analysis of Q-Learning with Time-Varying Policies
por: Nanda, Phalguni, et al.
Publicado: (2025)
por: Nanda, Phalguni, et al.
Publicado: (2025)
Mean-Field Games with Constraints
por: Hu, Anran, et al.
Publicado: (2025)
por: Hu, Anran, et al.
Publicado: (2025)
A Graph-Partitioning Based Continuous Optimization Approach to Semi-supervised Clustering Problems
por: Liu, Wei, et al.
Publicado: (2025)
por: Liu, Wei, et al.
Publicado: (2025)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
por: Carmona, René, et al.
Publicado: (2019)
por: Carmona, René, et al.
Publicado: (2019)
TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
por: Sorokin, D., et al.
Publicado: (2023)
por: Sorokin, D., et al.
Publicado: (2023)
UVIP: Model-Free Approach to Evaluate Reinforcement Learning Algorithms
por: Belomestny, Denis, et al.
Publicado: (2021)
por: Belomestny, Denis, et al.
Publicado: (2021)
Learning When to Restart: Nonstationary Newsvendor from Uncensored to Censored Demand
por: Chen, Xin, et al.
Publicado: (2025)
por: Chen, Xin, et al.
Publicado: (2025)
Deep Reinforcement Learning for Infinite Horizon Mean Field Problems in Continuous Spaces
por: Angiuli, Andrea, et al.
Publicado: (2023)
por: Angiuli, Andrea, et al.
Publicado: (2023)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Sublinear Regret for a Class of Continuous-Time Linear-Quadratic Reinforcement Learning Problems
por: Huang, Yilie, et al.
Publicado: (2024)
por: Huang, Yilie, et al.
Publicado: (2024)
Signature Approach for Contextual Bandits with Nonlinear and Path-dependent Rewards
por: Guo, Xin, et al.
Publicado: (2026)
por: Guo, Xin, et al.
Publicado: (2026)
Reinforcement Learning Approaches for the Orienteering Problem with Stochastic and Dynamic Release Dates
por: Li, Yuanyuan, et al.
Publicado: (2022)
por: Li, Yuanyuan, et al.
Publicado: (2022)
Reinforcement Learning with Random Time Horizons
por: Borrell, Enric Ribera, et al.
Publicado: (2025)
por: Borrell, Enric Ribera, et al.
Publicado: (2025)
Fill-and-Spill: Deep Reinforcement Learning Policy Gradient Methods for Reservoir Operation Decision and Control
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
por: Tabas, Sadegh Sadeghi, et al.
Publicado: (2024)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
por: Shen, Kaichen, et al.
Publicado: (2026)
por: Shen, Kaichen, et al.
Publicado: (2026)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2021)
por: Zeng, Sihan, et al.
Publicado: (2021)
Control, Optimal Transport and Neural Differential Equations in Supervised Learning
por: Phung, Minh-Nhat, et al.
Publicado: (2025)
por: Phung, Minh-Nhat, et al.
Publicado: (2025)
Meta-Learning Linear Quadratic Regulators: A Policy Gradient MAML Approach for Model-free LQR
por: Toso, Leonardo F., et al.
Publicado: (2024)
por: Toso, Leonardo F., et al.
Publicado: (2024)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
por: Long, Kehan, et al.
Publicado: (2025)
por: Long, Kehan, et al.
Publicado: (2025)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
por: Mazumdar, Abhijit, et al.
Publicado: (2024)
por: Mazumdar, Abhijit, et al.
Publicado: (2024)
Ejemplares similares
-
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
por: Cheng, Ziheng, et al.
Publicado: (2025) -
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
por: Qiu, Shuang, et al.
Publicado: (2024) -
Operator Models for Continuous-Time Offline Reinforcement Learning
por: Hoischen, Nicolas, et al.
Publicado: (2025) -
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
por: Zhang, Chenyu, et al.
Publicado: (2024) -
Action Gaps and Advantages in Continuous-Time Distributional Reinforcement Learning
por: Wiltzer, Harley, et al.
Publicado: (2024)