A Moreau Envelope Approach for LQR Meta-Policy Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aravind, Ashwin, Toghani, Mohammad Taha, Uribe, César A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
Policy Gradient for Continuous-Time Mean-Field Control
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
Compositional Construction of Barrier Functions for Switched Impulsive Systems
von: Bieker, Katharina, et al.
Veröffentlicht: (2024)
von: Bieker, Katharina, et al.
Veröffentlicht: (2024)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
Transport maps as flows of control-affine systems
von: Caponigro, Marco, et al.
Veröffentlicht: (2024)
von: Caponigro, Marco, et al.
Veröffentlicht: (2024)
Continuous time Stochastic optimal control under discrete time partial observations
von: Bayer, Christian, et al.
Veröffentlicht: (2024)
von: Bayer, Christian, et al.
Veröffentlicht: (2024)
On the continuity and smoothness of the value function in reinforcement learning and optimal control
von: Harder, Hans, et al.
Veröffentlicht: (2024)
von: Harder, Hans, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Infinite-Dimensional Systems
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Approximation of shape optimization problems with non-smooth PDE constraints
von: Betz, Livia
Veröffentlicht: (2024)
von: Betz, Livia
Veröffentlicht: (2024)
Stabilizability of parabolic equations by switching controls based on point actuators
von: Azmi, Behzad, et al.
Veröffentlicht: (2024)
von: Azmi, Behzad, et al.
Veröffentlicht: (2024)
On Robust Regulation of PDEs: from Abstract Methods to PDE Controllers
von: Paunonen, Lassi, et al.
Veröffentlicht: (2022)
von: Paunonen, Lassi, et al.
Veröffentlicht: (2022)
Passive and reciprocal networks: From simple models to simple optimal controllers
von: Pates, Richard
Veröffentlicht: (2022)
von: Pates, Richard
Veröffentlicht: (2022)
Finite Codimensionality Method in Infinite-dimensional Optimization Problems
von: Liu, Xu, et al.
Veröffentlicht: (2021)
von: Liu, Xu, et al.
Veröffentlicht: (2021)
LQR control for a system describing the interaction between a floating solid and the surrounding fluid
von: Tucsnak, Marius, et al.
Veröffentlicht: (2024)
von: Tucsnak, Marius, et al.
Veröffentlicht: (2024)
Technological schemes and control methods in the reconstruction of parallel gas pipeline systems under non-stationary conditions
von: Aliyev, Ilgar
Veröffentlicht: (2025)
von: Aliyev, Ilgar
Veröffentlicht: (2025)
Input design for the optimal control of networked moments
von: Solimine, Philip, et al.
Veröffentlicht: (2021)
von: Solimine, Philip, et al.
Veröffentlicht: (2021)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
Lecture Notes on Control System Theory and Design
von: Basar, Tamer, et al.
Veröffentlicht: (2020)
von: Basar, Tamer, et al.
Veröffentlicht: (2020)
Formalising the intentional stance 2: a coinductive approach
von: McGregor, Simon, et al.
Veröffentlicht: (2025)
von: McGregor, Simon, et al.
Veröffentlicht: (2025)
Formalising the intentional stance 1: attributing goals and beliefs to stochastic processes
von: McGregor, Simon, et al.
Veröffentlicht: (2024)
von: McGregor, Simon, et al.
Veröffentlicht: (2024)
Measurement-based Initial Point Smoothing and Control Approach to Quantum Memory Systems
von: Vladimirov, Igor G., et al.
Veröffentlicht: (2025)
von: Vladimirov, Igor G., et al.
Veröffentlicht: (2025)
Distributed Online Optimization for Multi-Agent Optimal Transport
von: Krishnan, Vishaal, et al.
Veröffentlicht: (2018)
von: Krishnan, Vishaal, et al.
Veröffentlicht: (2018)
Optimal control of infinite-dimensional dissipative systems
von: Hastir, Anthony, et al.
Veröffentlicht: (2026)
von: Hastir, Anthony, et al.
Veröffentlicht: (2026)
Legendre-Moment Transform for Linear Ensemble Control and Computation
von: Ning, Xin, et al.
Veröffentlicht: (2024)
von: Ning, Xin, et al.
Veröffentlicht: (2024)
Optimal control of SDEs with merely measurable drift: an HJB approach
von: Du, Kai, et al.
Veröffentlicht: (2025)
von: Du, Kai, et al.
Veröffentlicht: (2025)
Nonlinear Conjugate Gradient Methods for PDE Constrained Shape Optimization Based on Steklov-Poincaré-Type Metrics
von: Blauth, Sebastian
Veröffentlicht: (2020)
von: Blauth, Sebastian
Veröffentlicht: (2020)
Quantitative Soft-to-Hard Terminal Constraint Convergence for the Heat Equation
von: Kwon, Sung-Sik
Veröffentlicht: (2026)
von: Kwon, Sung-Sik
Veröffentlicht: (2026)
Turnpike and dissipativity in generalized discrete-time stochastic linear-quadratic optimal control
von: Schießl, Jonas, et al.
Veröffentlicht: (2023)
von: Schießl, Jonas, et al.
Veröffentlicht: (2023)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
von: Luo, Sheng, et al.
Veröffentlicht: (2024)
von: Luo, Sheng, et al.
Veröffentlicht: (2024)
Turnpike Property of Stochastic Linear-Quadratic Optimal Control Problems in Large Horizons with Regime Switching I: Homogeneous Cases
von: Mei, Hongwei, et al.
Veröffentlicht: (2025)
von: Mei, Hongwei, et al.
Veröffentlicht: (2025)
Trading with propagators and constraints: applications to optimal execution and battery storage
von: Jaber, Eduardo Abi, et al.
Veröffentlicht: (2024)
von: Jaber, Eduardo Abi, et al.
Veröffentlicht: (2024)
Controllability and Vector Potential
von: Shankar, Shiva
Veröffentlicht: (2019)
von: Shankar, Shiva
Veröffentlicht: (2019)
Measure propagation along a $\mathscr{C}^0$-vector field and wave controllability on a rough compact manifold
von: Burq, Nicolas, et al.
Veröffentlicht: (2023)
von: Burq, Nicolas, et al.
Veröffentlicht: (2023)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
von: Mei, Hongwei, et al.
Veröffentlicht: (2025)
von: Mei, Hongwei, et al.
Veröffentlicht: (2025)
Boundary controllability of incompressible Euler fluids with Boussinesq heat effects
von: Fernández-Cara, Enrique, et al.
Veröffentlicht: (2024)
von: Fernández-Cara, Enrique, et al.
Veröffentlicht: (2024)
Dissipativity, reciprocity and passive network synthesis: from Jan Willems' seminal Dissipative Dynamical Systems papers to the present day
von: Hughes, Timothy H., et al.
Veröffentlicht: (2021)
von: Hughes, Timothy H., et al.
Veröffentlicht: (2021)
Second-Order $Λ$-Sets and Extensions to Non-Smooth, Hybrid, and Stochastic Optimal Control
von: Rashid, Mohammad H. M
Veröffentlicht: (2025)
von: Rashid, Mohammad H. M
Veröffentlicht: (2025)
Generative optimal transport via forward-backward HJB matching
von: Yang, Haiqian, et al.
Veröffentlicht: (2026)
von: Yang, Haiqian, et al.
Veröffentlicht: (2026)
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026)
Multi-period Asset-liability Management with Reinforcement Learning in a Regime-Switching Market
von: Gao, Zhongqin, et al.
Veröffentlicht: (2025)
von: Gao, Zhongqin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
von: Cui, Leilei, et al.
Veröffentlicht: (2025) -
Policy Gradient for Continuous-Time Mean-Field Control
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2026) -
Compositional Construction of Barrier Functions for Switched Impulsive Systems
von: Bieker, Katharina, et al.
Veröffentlicht: (2024) -
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026) -
Transport maps as flows of control-affine systems
von: Caponigro, Marco, et al.
Veröffentlicht: (2024)