A Moreau Envelope Approach for LQR Meta-Policy Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Aravind, Ashwin, Toghani, Mohammad Taha, Uribe, César A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
por: Cui, Leilei, et al.
Publicado: (2025)
por: Cui, Leilei, et al.
Publicado: (2025)
Policy Gradient for Continuous-Time Mean-Field Control
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Compositional Construction of Barrier Functions for Switched Impulsive Systems
por: Bieker, Katharina, et al.
Publicado: (2024)
por: Bieker, Katharina, et al.
Publicado: (2024)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
por: Bäuerle, Nicole, et al.
Publicado: (2026)
por: Bäuerle, Nicole, et al.
Publicado: (2026)
Transport maps as flows of control-affine systems
por: Caponigro, Marco, et al.
Publicado: (2024)
por: Caponigro, Marco, et al.
Publicado: (2024)
Continuous time Stochastic optimal control under discrete time partial observations
por: Bayer, Christian, et al.
Publicado: (2024)
por: Bayer, Christian, et al.
Publicado: (2024)
On the continuity and smoothness of the value function in reinforcement learning and optimal control
por: Harder, Hans, et al.
Publicado: (2024)
por: Harder, Hans, et al.
Publicado: (2024)
Reinforcement Learning for Infinite-Dimensional Systems
por: Zhang, Wei, et al.
Publicado: (2024)
por: Zhang, Wei, et al.
Publicado: (2024)
Approximation of shape optimization problems with non-smooth PDE constraints
por: Betz, Livia
Publicado: (2024)
por: Betz, Livia
Publicado: (2024)
Stabilizability of parabolic equations by switching controls based on point actuators
por: Azmi, Behzad, et al.
Publicado: (2024)
por: Azmi, Behzad, et al.
Publicado: (2024)
On Robust Regulation of PDEs: from Abstract Methods to PDE Controllers
por: Paunonen, Lassi, et al.
Publicado: (2022)
por: Paunonen, Lassi, et al.
Publicado: (2022)
Passive and reciprocal networks: From simple models to simple optimal controllers
por: Pates, Richard
Publicado: (2022)
por: Pates, Richard
Publicado: (2022)
Finite Codimensionality Method in Infinite-dimensional Optimization Problems
por: Liu, Xu, et al.
Publicado: (2021)
por: Liu, Xu, et al.
Publicado: (2021)
LQR control for a system describing the interaction between a floating solid and the surrounding fluid
por: Tucsnak, Marius, et al.
Publicado: (2024)
por: Tucsnak, Marius, et al.
Publicado: (2024)
Technological schemes and control methods in the reconstruction of parallel gas pipeline systems under non-stationary conditions
por: Aliyev, Ilgar
Publicado: (2025)
por: Aliyev, Ilgar
Publicado: (2025)
Input design for the optimal control of networked moments
por: Solimine, Philip, et al.
Publicado: (2021)
por: Solimine, Philip, et al.
Publicado: (2021)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
por: Zheng, Yaowei, et al.
Publicado: (2026)
por: Zheng, Yaowei, et al.
Publicado: (2026)
Lecture Notes on Control System Theory and Design
por: Basar, Tamer, et al.
Publicado: (2020)
por: Basar, Tamer, et al.
Publicado: (2020)
Formalising the intentional stance 2: a coinductive approach
por: McGregor, Simon, et al.
Publicado: (2025)
por: McGregor, Simon, et al.
Publicado: (2025)
Formalising the intentional stance 1: attributing goals and beliefs to stochastic processes
por: McGregor, Simon, et al.
Publicado: (2024)
por: McGregor, Simon, et al.
Publicado: (2024)
Measurement-based Initial Point Smoothing and Control Approach to Quantum Memory Systems
por: Vladimirov, Igor G., et al.
Publicado: (2025)
por: Vladimirov, Igor G., et al.
Publicado: (2025)
Distributed Online Optimization for Multi-Agent Optimal Transport
por: Krishnan, Vishaal, et al.
Publicado: (2018)
por: Krishnan, Vishaal, et al.
Publicado: (2018)
Optimal control of infinite-dimensional dissipative systems
por: Hastir, Anthony, et al.
Publicado: (2026)
por: Hastir, Anthony, et al.
Publicado: (2026)
Legendre-Moment Transform for Linear Ensemble Control and Computation
por: Ning, Xin, et al.
Publicado: (2024)
por: Ning, Xin, et al.
Publicado: (2024)
Optimal control of SDEs with merely measurable drift: an HJB approach
por: Du, Kai, et al.
Publicado: (2025)
por: Du, Kai, et al.
Publicado: (2025)
Nonlinear Conjugate Gradient Methods for PDE Constrained Shape Optimization Based on Steklov-Poincaré-Type Metrics
por: Blauth, Sebastian
Publicado: (2020)
por: Blauth, Sebastian
Publicado: (2020)
Quantitative Soft-to-Hard Terminal Constraint Convergence for the Heat Equation
por: Kwon, Sung-Sik
Publicado: (2026)
por: Kwon, Sung-Sik
Publicado: (2026)
Turnpike and dissipativity in generalized discrete-time stochastic linear-quadratic optimal control
por: Schießl, Jonas, et al.
Publicado: (2023)
por: Schießl, Jonas, et al.
Publicado: (2023)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
por: Luo, Sheng, et al.
Publicado: (2024)
por: Luo, Sheng, et al.
Publicado: (2024)
Turnpike Property of Stochastic Linear-Quadratic Optimal Control Problems in Large Horizons with Regime Switching I: Homogeneous Cases
por: Mei, Hongwei, et al.
Publicado: (2025)
por: Mei, Hongwei, et al.
Publicado: (2025)
Trading with propagators and constraints: applications to optimal execution and battery storage
por: Jaber, Eduardo Abi, et al.
Publicado: (2024)
por: Jaber, Eduardo Abi, et al.
Publicado: (2024)
Controllability and Vector Potential
por: Shankar, Shiva
Publicado: (2019)
por: Shankar, Shiva
Publicado: (2019)
Measure propagation along a $\mathscr{C}^0$-vector field and wave controllability on a rough compact manifold
por: Burq, Nicolas, et al.
Publicado: (2023)
por: Burq, Nicolas, et al.
Publicado: (2023)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
por: Mei, Hongwei, et al.
Publicado: (2025)
por: Mei, Hongwei, et al.
Publicado: (2025)
Boundary controllability of incompressible Euler fluids with Boussinesq heat effects
por: Fernández-Cara, Enrique, et al.
Publicado: (2024)
por: Fernández-Cara, Enrique, et al.
Publicado: (2024)
Dissipativity, reciprocity and passive network synthesis: from Jan Willems' seminal Dissipative Dynamical Systems papers to the present day
por: Hughes, Timothy H., et al.
Publicado: (2021)
por: Hughes, Timothy H., et al.
Publicado: (2021)
Second-Order $Λ$-Sets and Extensions to Non-Smooth, Hybrid, and Stochastic Optimal Control
por: Rashid, Mohammad H. M
Publicado: (2025)
por: Rashid, Mohammad H. M
Publicado: (2025)
Generative optimal transport via forward-backward HJB matching
por: Yang, Haiqian, et al.
Publicado: (2026)
por: Yang, Haiqian, et al.
Publicado: (2026)
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
por: Bayraktar, Erhan, et al.
Publicado: (2026)
por: Bayraktar, Erhan, et al.
Publicado: (2026)
Multi-period Asset-liability Management with Reinforcement Learning in a Regime-Switching Market
por: Gao, Zhongqin, et al.
Publicado: (2025)
por: Gao, Zhongqin, et al.
Publicado: (2025)
Ejemplares similares
-
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
por: Cui, Leilei, et al.
Publicado: (2025) -
Policy Gradient for Continuous-Time Mean-Field Control
por: Bayraktar, Erhan, et al.
Publicado: (2026) -
Compositional Construction of Barrier Functions for Switched Impulsive Systems
por: Bieker, Katharina, et al.
Publicado: (2024) -
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
por: Bäuerle, Nicole, et al.
Publicado: (2026) -
Transport maps as flows of control-affine systems
por: Caponigro, Marco, et al.
Publicado: (2024)