On-Line Policy Iteration with Trajectory-Driven Policy Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yuchao, Chen, Fei, Li, Yingke, Fan, Chuchu, Bertsekas, Dimitri |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Error Bound for Aggregation in Approximate Dynamic Programming
by: Li, Yuchao, et al.
Published: (2025)
by: Li, Yuchao, et al.
Published: (2025)
Semilinear Dynamic Programming: Analysis, Algorithms, and Certainty Equivalence Properties
by: Li, Yuchao, et al.
Published: (2025)
by: Li, Yuchao, et al.
Published: (2025)
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
by: Bertsekas, Dimitri P.
Published: (2024)
by: Bertsekas, Dimitri P.
Published: (2024)
Compressed Traffic Assignment with the Augmented Lagrangian Method
by: Xuesong, et al.
Published: (2026)
by: Xuesong, et al.
Published: (2026)
Feature-Based Belief Aggregation for Partially Observable Markov Decision Problems
by: Li, Yuchao, et al.
Published: (2025)
by: Li, Yuchao, et al.
Published: (2025)
Data-Driven LQR with Finite-Time Experiments via Extremum-Seeking Policy Iteration
by: Carnevale, Guido, et al.
Published: (2024)
by: Carnevale, Guido, et al.
Published: (2024)
Most Likely Sequence Generation for $n$-Grams, Transformers, HMMs, and Markov Chains, by Using Rollout Algorithms
by: Li, Yuchao, et al.
Published: (2024)
by: Li, Yuchao, et al.
Published: (2024)
A Policy Iteration Algorithm for N-player General-Sum Linear Quadratic Dynamic Games
by: Guan, Yuxiang, et al.
Published: (2024)
by: Guan, Yuxiang, et al.
Published: (2024)
Adaptive Network Security Policies via Belief Aggregation and Rollout
by: Hammar, Kim, et al.
Published: (2025)
by: Hammar, Kim, et al.
Published: (2025)
Incremental Policy Iteration for Unknown Nonlinear Systems with Stability and Performance Guarantees
by: Meng, Qingkai, et al.
Published: (2025)
by: Meng, Qingkai, et al.
Published: (2025)
Physics-Informed Neural Network Policy Iteration: Algorithms, Convergence, and Verification
by: Meng, Yiming, et al.
Published: (2024)
by: Meng, Yiming, et al.
Published: (2024)
Diffusion Policies for Generative Modeling of Spacecraft Trajectories
by: Briden, Julia, et al.
Published: (2025)
by: Briden, Julia, et al.
Published: (2025)
Adaptive Optimal Control of Linear Periodic Systems: An Off-Policy Value Iteration Approach
by: Pang, Bo, et al.
Published: (2019)
by: Pang, Bo, et al.
Published: (2019)
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
Maximizing Reach-Avoid Probabilities for Linear Stochastic Systems via Control Architectures
by: Schmid, Niklas, et al.
Published: (2026)
by: Schmid, Niklas, et al.
Published: (2026)
Data-Enabled Policy and Value Iteration for Continuous-Time Linear Quadratic Output Feedback Control
by: Xie, Jun, et al.
Published: (2026)
by: Xie, Jun, et al.
Published: (2026)
Direct Data-Driven Linear Quadratic Tracking via Policy Optimization
by: Kang, Shubo, et al.
Published: (2026)
by: Kang, Shubo, et al.
Published: (2026)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
by: Ding, Dongsheng, et al.
Published: (2023)
by: Ding, Dongsheng, et al.
Published: (2023)
Ergodic-risk Criterion for Stochastically Stabilizing Policy Optimization
by: Talebi, Shahriar, et al.
Published: (2024)
by: Talebi, Shahriar, et al.
Published: (2024)
Superior Computer Chess with Model Predictive Control, Reinforcement Learning, and Rollout
by: Gundawar, Atharva, et al.
Published: (2024)
by: Gundawar, Atharva, et al.
Published: (2024)
Ergodic-Risk Constrained Policy Optimization: The Linear Quadratic Case
by: Talebi, Shahriar, et al.
Published: (2025)
by: Talebi, Shahriar, et al.
Published: (2025)
Data-Driven Distributionally Robust Mixed-Integer Control through Lifted Control Policy
by: Ma, Xutao, et al.
Published: (2025)
by: Ma, Xutao, et al.
Published: (2025)
Computing Optimal Joint Chance Constrained Control Policies
by: Schmid, Niklas, et al.
Published: (2023)
by: Schmid, Niklas, et al.
Published: (2023)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
by: Lee, Donghwan
Published: (2026)
by: Lee, Donghwan
Published: (2026)
Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
by: Zhang, Songyuan, et al.
Published: (2025)
by: Zhang, Songyuan, et al.
Published: (2025)
Integrated Investment and Policy Planning for Power Systems via Differentiable Scenario Generation
by: Mieth, Robert
Published: (2026)
by: Mieth, Robert
Published: (2026)
Distributionally Robust Stochastic MPC under Disturbance-Affine Feedback Policies
by: Chen, Xu, et al.
Published: (2026)
by: Chen, Xu, et al.
Published: (2026)
Distributionally Robust System Level Synthesis With Output Feedback Affine Control Policy
by: Li, Yun, et al.
Published: (2025)
by: Li, Yun, et al.
Published: (2025)
Dynamic Transfer Policies for Parallel Queues
by: Chan, Timothy C. Y., et al.
Published: (2024)
by: Chan, Timothy C. Y., et al.
Published: (2024)
Policy Gradient Bounds in Multitask LQR
by: Stamouli, Charis, et al.
Published: (2025)
by: Stamouli, Charis, et al.
Published: (2025)
Policy Optimization of Mixed H2/H-infinity Control: Benign Nonconvexity and Global Optimality
by: Pai, Chih-Fan, et al.
Published: (2026)
by: Pai, Chih-Fan, et al.
Published: (2026)
An Iterative Problem-Driven Scenario Reduction Framework for Stochastic Optimization with Conditional Value-at-Risk
by: Zhuang, Yingrui, et al.
Published: (2025)
by: Zhuang, Yingrui, et al.
Published: (2025)
Receding-Horizon Policy Gradient for Polytopic Controller Synthesis
by: Shakeri, Shiva, et al.
Published: (2026)
by: Shakeri, Shiva, et al.
Published: (2026)
Performance Bounds for Rollout Policies in Stochastic Shortest Path Problems
by: Hansson, Anders, et al.
Published: (2026)
by: Hansson, Anders, et al.
Published: (2026)
Policy Optimization with Differentiable MPC: Convergence Analysis under Uncertainty
by: Zuliani, Riccardo, et al.
Published: (2026)
by: Zuliani, Riccardo, et al.
Published: (2026)
Optimality of Linear Policies in Distributionally Robust Linear Quadratic Control
by: Taşkesen, Bahar, et al.
Published: (2025)
by: Taşkesen, Bahar, et al.
Published: (2025)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Stochastic MPC with Online-optimized Policies and Closed-loop Guarantees
by: Bartos, Marcell, et al.
Published: (2025)
by: Bartos, Marcell, et al.
Published: (2025)
Policy Gradient Adaptive Control for the LQR: Indirect and Direct Approaches
by: Zhao, Feiran, et al.
Published: (2025)
by: Zhao, Feiran, et al.
Published: (2025)
Policy Optimization in Robust Control: Weak Convexity and Subgradient Methods
by: Watanabe, Yuto, et al.
Published: (2025)
by: Watanabe, Yuto, et al.
Published: (2025)
Similar Items
-
An Error Bound for Aggregation in Approximate Dynamic Programming
by: Li, Yuchao, et al.
Published: (2025) -
Semilinear Dynamic Programming: Analysis, Algorithms, and Certainty Equivalence Properties
by: Li, Yuchao, et al.
Published: (2025) -
Model Predictive Control and Reinforcement Learning: A Unified Framework Based on Dynamic Programming
by: Bertsekas, Dimitri P.
Published: (2024) -
Compressed Traffic Assignment with the Augmented Lagrangian Method
by: Xuesong, et al.
Published: (2026) -
Feature-Based Belief Aggregation for Partially Observable Markov Decision Problems
by: Li, Yuchao, et al.
Published: (2025)