Second-Order Policy Gradient Methods for the Linear Quadratic Regulator
Fuente:
arXiv
Saved in:
| Main Authors: | Valaei, Amirreza, Kordabad, Arash Bahari, Soudjani, Sadegh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sum-of-Squares Certificates for Almost-Sure Reachability of Stochastic Polynomial Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Almost Sure Reachability in Continuous-time Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2026)
by: Kordabad, Arash Bahari, et al.
Published: (2026)
Intent-Aware MPC for Aircraft Detect-and-Avoid with Response Delay: A Comparative Study with ACAS Xu
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
On Certificates for Almost Sure Reachability in Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Robust Model Predictive Control for Aircraft Intent-Aware Collision Avoidance
by: Kordabad, Arash Bahari, et al.
Published: (2024)
by: Kordabad, Arash Bahari, et al.
Published: (2024)
Distributionally Robust Control for Chance-Constrained Signal Temporal Logic Specifications
by: Kordabad, Arash Bahari, et al.
Published: (2024)
by: Kordabad, Arash Bahari, et al.
Published: (2024)
Data-Driven Distributionally Robust Control for Interacting Agents under Logical Constraints
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Multi-Agent Temporal Logic Planning via Penalty Functions and Block-Coordinate Optimization
by: Vlahakis, Eleftherios E., et al.
Published: (2026)
by: Vlahakis, Eleftherios E., et al.
Published: (2026)
Global Convergence of Policy Gradient Methods for ReLU Controllers in Linear Quadratic Regulation
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
by: Rodriguez-Gil, Jhojan A., et al.
Published: (2026)
Policy Synthesis for Interval MDPs via Polyhedral Lyapunov Functions
by: Monir, Negar, et al.
Published: (2026)
by: Monir, Negar, et al.
Published: (2026)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
Convergence of Flow-Policy Gradient Learning for Linear Quadratic Regulator Problems
by: Yaghmaie, Farnaz Adib, et al.
Published: (2025)
by: Yaghmaie, Farnaz Adib, et al.
Published: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
Data-Driven Yet Formal Policy Synthesis for Stochastic Nonlinear Dynamical Systems
by: Nazeri, Mahdi, et al.
Published: (2025)
by: Nazeri, Mahdi, et al.
Published: (2025)
A Globally Convergent Policy Gradient Method for Linear Quadratic Gaussian (LQG) Control
by: Sadamoto, Tomonori, et al.
Published: (2023)
by: Sadamoto, Tomonori, et al.
Published: (2023)
Robust Control of Uncertain Switched Affine Systems via Scenario Optimization
by: Monir, Negar, et al.
Published: (2025)
by: Monir, Negar, et al.
Published: (2025)
Barrier Certificates for Uncertain Temporal Specifications
by: Mamduhi, Mohammad H., et al.
Published: (2026)
by: Mamduhi, Mohammad H., et al.
Published: (2026)
Logic-based Resilience Computation of Power Systems Against Frequency Requirements
by: Monir, Negar, et al.
Published: (2025)
by: Monir, Negar, et al.
Published: (2025)
Formal Control of New England 39-Bus Test System: An Assume-Guarantee Approach
by: Wooding, Ben, et al.
Published: (2023)
by: Wooding, Ben, et al.
Published: (2023)
Logic-based Knowledge Awareness for Autonomous Agents in Continuous Spaces
by: Ghosh, Arabinda, et al.
Published: (2024)
by: Ghosh, Arabinda, et al.
Published: (2024)
Temporal Logic Resilience for Dynamical Systems
by: Saoud, Adnane, et al.
Published: (2024)
by: Saoud, Adnane, et al.
Published: (2024)
Formal Control for Uncertain Systems via Contract-Based Probabilistic Surrogates (Extended Version)
by: Schön, Oliver, et al.
Published: (2025)
by: Schön, Oliver, et al.
Published: (2025)
Model-Agnostic Meta-Policy Optimization via Zeroth-Order Estimation: A Linear Quadratic Regulator Perspective
by: Pan, Yunian, et al.
Published: (2025)
by: Pan, Yunian, et al.
Published: (2025)
Stochastic Generalized Dynamic Games with Coupled Chance Constraints
by: Yadollahi, Seyed Shahram, et al.
Published: (2025)
by: Yadollahi, Seyed Shahram, et al.
Published: (2025)
Model-Agnostic Zeroth-Order Policy Optimization for Meta-Learning of Ergodic Linear Quadratic Regulators
by: Pan, Yunian, et al.
Published: (2024)
by: Pan, Yunian, et al.
Published: (2024)
Kernel-Based Safe Exploration in Deep Reinforcement Learning
by: Majumdar, Rupak, et al.
Published: (2026)
by: Majumdar, Rupak, et al.
Published: (2026)
Safety Certification is Classification
by: Schön, Oliver, et al.
Published: (2026)
by: Schön, Oliver, et al.
Published: (2026)
Data-Driven Distributionally Robust Safety Verification Using Barrier Certificates and Conditional Mean Embeddings
by: Schön, Oliver, et al.
Published: (2024)
by: Schön, Oliver, et al.
Published: (2024)
Optimality Conditions for Model Predictive Control: Rethinking Predictive Model Design
by: Anand, Akhil S, et al.
Published: (2024)
by: Anand, Akhil S, et al.
Published: (2024)
Formal Verification of Unknown Stochastic Systems via Non-parametric Estimation
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Gradient Dominance in the Linear Quadratic Regulator: A Unified Analysis for Continuous-Time and Discrete-Time Systems
by: Watanabe, Yuto, et al.
Published: (2026)
by: Watanabe, Yuto, et al.
Published: (2026)
Kernel-Based Learning of Safety Barriers
by: Schön, Oliver, et al.
Published: (2026)
by: Schön, Oliver, et al.
Published: (2026)
Asynchronous Parallel Policy Gradient Methods for the Linear Quadratic Regulator
by: Sha, Xingyu, et al.
Published: (2024)
by: Sha, Xingyu, et al.
Published: (2024)
Optimality of Linear Policies in Distributionally Robust Linear Quadratic Control
by: Taşkesen, Bahar, et al.
Published: (2025)
by: Taşkesen, Bahar, et al.
Published: (2025)
Convergence and Robustness of Value and Policy Iteration for the Linear Quadratic Regulator
by: Song, Bowen, et al.
Published: (2024)
by: Song, Bowen, et al.
Published: (2024)
Data-driven Linear Quadratic Integral Control: A Convex Formulation and Policy Gradient Approach
by: Gießler, Armin, et al.
Published: (2026)
by: Gießler, Armin, et al.
Published: (2026)
Conditions for Complete Decentralization of the Linear Quadratic Regulator
by: McCurdy, Addie, et al.
Published: (2026)
by: McCurdy, Addie, et al.
Published: (2026)
Distributed Policy Gradient for Linear Quadratic Networked Control with Limited Communication Range
by: Yan, Yuzi, et al.
Published: (2024)
by: Yan, Yuzi, et al.
Published: (2024)
Verification of Quantum Circuits through Discrete-Time Barrier Certificates
by: Lewis, Marco, et al.
Published: (2024)
by: Lewis, Marco, et al.
Published: (2024)
Similar Items
-
Sum-of-Squares Certificates for Almost-Sure Reachability of Stochastic Polynomial Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025) -
Almost Sure Reachability in Continuous-time Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2026) -
Intent-Aware MPC for Aircraft Detect-and-Avoid with Response Delay: A Comparative Study with ACAS Xu
by: Kordabad, Arash Bahari, et al.
Published: (2025) -
Quasi-Newton Compatible Actor-Critic for Deterministic Policies
by: Kordabad, Arash Bahari, et al.
Published: (2025) -
On Certificates for Almost Sure Reachability in Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025)