Quasi-Newton Compatible Actor-Critic for Deterministic Policies
Fuente:
arXiv
Saved in:
| Main Authors: | Kordabad, Arash Bahari, Brandner, Dean, Gros, Sebastien, Lucia, Sergio, Soudjani, Sadegh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sum-of-Squares Certificates for Almost-Sure Reachability of Stochastic Polynomial Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Data-Driven Distributionally Robust Control for Interacting Agents under Logical Constraints
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
Robust Model Predictive Control for Aircraft Intent-Aware Collision Avoidance
by: Kordabad, Arash Bahari, et al.
Published: (2024)
by: Kordabad, Arash Bahari, et al.
Published: (2024)
Computationally efficient Gauss-Newton reinforcement learning for model predictive control
by: Brandner, Dean, et al.
Published: (2025)
by: Brandner, Dean, et al.
Published: (2025)
Distributionally Robust Control for Chance-Constrained Signal Temporal Logic Specifications
by: Kordabad, Arash Bahari, et al.
Published: (2024)
by: Kordabad, Arash Bahari, et al.
Published: (2024)
Second-Order Policy Gradient Methods for the Linear Quadratic Regulator
by: Valaei, Amirreza, et al.
Published: (2025)
by: Valaei, Amirreza, et al.
Published: (2025)
Reinforced Model Predictive Control via Trust-Region Quasi-Newton Policy Optimization
by: Brandner, Dean, et al.
Published: (2024)
by: Brandner, Dean, et al.
Published: (2024)
Almost Sure Reachability in Continuous-time Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2026)
by: Kordabad, Arash Bahari, et al.
Published: (2026)
Optimality Conditions for Model Predictive Control: Rethinking Predictive Model Design
by: Anand, Akhil S, et al.
Published: (2024)
by: Anand, Akhil S, et al.
Published: (2024)
Intent-Aware MPC for Aircraft Detect-and-Avoid with Response Delay: A Comparative Study with ACAS Xu
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
On Certificates for Almost Sure Reachability in Stochastic Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025)
by: Kordabad, Arash Bahari, et al.
Published: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
Optimizing Operation Recipes with Reinforcement Learning for Safe and Interpretable Control of Chemical Processes
by: Brandner, Dean, et al.
Published: (2025)
by: Brandner, Dean, et al.
Published: (2025)
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
Robust Control of Uncertain Switched Affine Systems via Scenario Optimization
by: Monir, Negar, et al.
Published: (2025)
by: Monir, Negar, et al.
Published: (2025)
Policy Synthesis for Interval MDPs via Polyhedral Lyapunov Functions
by: Monir, Negar, et al.
Published: (2026)
by: Monir, Negar, et al.
Published: (2026)
Barrier Certificates for Uncertain Temporal Specifications
by: Mamduhi, Mohammad H., et al.
Published: (2026)
by: Mamduhi, Mohammad H., et al.
Published: (2026)
Multi-Agent Temporal Logic Planning via Penalty Functions and Block-Coordinate Optimization
by: Vlahakis, Eleftherios E., et al.
Published: (2026)
by: Vlahakis, Eleftherios E., et al.
Published: (2026)
Deterministic Trajectory Optimization through Probabilistic Optimal Control
by: Filabadi, Mohammad Mahmoudi, et al.
Published: (2024)
by: Filabadi, Mohammad Mahmoudi, et al.
Published: (2024)
Gradient Estimation and Variance Reduction in Stochastic and Deterministic Models
by: Keane, Ronan
Published: (2024)
by: Keane, Ronan
Published: (2024)
ACING: Actor-Critic for Instruction Learning in Black-Box LLMs
by: Kharrat, Salma, et al.
Published: (2024)
by: Kharrat, Salma, et al.
Published: (2024)
Random Features Approximation for Control-Affine Systems
by: Kazemian, Kimia, et al.
Published: (2024)
by: Kazemian, Kimia, et al.
Published: (2024)
Learning Linear Dynamics from Bilinear Observations
by: Sattar, Yahya, et al.
Published: (2024)
by: Sattar, Yahya, et al.
Published: (2024)
Flipping-based Policy for Chance-Constrained Markov Decision Processes
by: Shen, Xun, et al.
Published: (2024)
by: Shen, Xun, et al.
Published: (2024)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
Principled Learning-to-Communicate with Quasi-Classical Information Structures
by: Liu, Xiangyu, et al.
Published: (2026)
by: Liu, Xiangyu, et al.
Published: (2026)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
Maximally Resilient Controllers under Temporal Logic Specifications
by: Si, Youssef Ait, et al.
Published: (2025)
by: Si, Youssef Ait, et al.
Published: (2025)
Dual Control of Linear Systems from Bilinear Observations with Belief Space Model Predictive Control
by: Cao, Daniel, et al.
Published: (2026)
by: Cao, Daniel, et al.
Published: (2026)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
On Building Myopic MPC Policies using Supervised Learning
by: Orrico, Christopher A., et al.
Published: (2024)
by: Orrico, Christopher A., et al.
Published: (2024)
Kernel-Based Safe Exploration in Deep Reinforcement Learning
by: Majumdar, Rupak, et al.
Published: (2026)
by: Majumdar, Rupak, et al.
Published: (2026)
Data-Driven Distributionally Robust Safety Verification Using Barrier Certificates and Conditional Mean Embeddings
by: Schön, Oliver, et al.
Published: (2024)
by: Schön, Oliver, et al.
Published: (2024)
Optimal Parameter Adaptation for Safety-Critical Control via Safe Barrier Bayesian Optimization
by: Wang, Shengbo, et al.
Published: (2025)
by: Wang, Shengbo, et al.
Published: (2025)
On Policy Stochasticity in Mutual Information Optimal Control of Linear Systems
by: Enami, Shoju, et al.
Published: (2025)
by: Enami, Shoju, et al.
Published: (2025)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Model-Free Output Feedback Stabilization via Policy Gradient Methods
by: Zhang, Ankang, et al.
Published: (2026)
by: Zhang, Ankang, et al.
Published: (2026)
Sharpened Lazy Incremental Quasi-Newton Method
by: Lahoti, Aakash, et al.
Published: (2023)
by: Lahoti, Aakash, et al.
Published: (2023)
Distributed Control of Network Systems in the Space of Stabilizing Graph Neural Network Policies
by: Cao, John, et al.
Published: (2025)
by: Cao, John, et al.
Published: (2025)
Similar Items
-
Sum-of-Squares Certificates for Almost-Sure Reachability of Stochastic Polynomial Systems
by: Kordabad, Arash Bahari, et al.
Published: (2025) -
Data-Driven Distributionally Robust Control for Interacting Agents under Logical Constraints
by: Kordabad, Arash Bahari, et al.
Published: (2025) -
Robust Model Predictive Control for Aircraft Intent-Aware Collision Avoidance
by: Kordabad, Arash Bahari, et al.
Published: (2024) -
Computationally efficient Gauss-Newton reinforcement learning for model predictive control
by: Brandner, Dean, et al.
Published: (2025) -
Distributionally Robust Control for Chance-Constrained Signal Temporal Logic Specifications
by: Kordabad, Arash Bahari, et al.
Published: (2024)