A Moreau Envelope Approach for LQR Meta-Policy Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Aravind, Ashwin, Toghani, Mohammad Taha, Uribe, César A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
di: Cui, Leilei, et al.
Pubblicazione: (2025)
di: Cui, Leilei, et al.
Pubblicazione: (2025)
Policy Gradient for Continuous-Time Mean-Field Control
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
Compositional Construction of Barrier Functions for Switched Impulsive Systems
di: Bieker, Katharina, et al.
Pubblicazione: (2024)
di: Bieker, Katharina, et al.
Pubblicazione: (2024)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
di: Bäuerle, Nicole, et al.
Pubblicazione: (2026)
di: Bäuerle, Nicole, et al.
Pubblicazione: (2026)
Transport maps as flows of control-affine systems
di: Caponigro, Marco, et al.
Pubblicazione: (2024)
di: Caponigro, Marco, et al.
Pubblicazione: (2024)
Continuous time Stochastic optimal control under discrete time partial observations
di: Bayer, Christian, et al.
Pubblicazione: (2024)
di: Bayer, Christian, et al.
Pubblicazione: (2024)
On the continuity and smoothness of the value function in reinforcement learning and optimal control
di: Harder, Hans, et al.
Pubblicazione: (2024)
di: Harder, Hans, et al.
Pubblicazione: (2024)
Reinforcement Learning for Infinite-Dimensional Systems
di: Zhang, Wei, et al.
Pubblicazione: (2024)
di: Zhang, Wei, et al.
Pubblicazione: (2024)
Approximation of shape optimization problems with non-smooth PDE constraints
di: Betz, Livia
Pubblicazione: (2024)
di: Betz, Livia
Pubblicazione: (2024)
Stabilizability of parabolic equations by switching controls based on point actuators
di: Azmi, Behzad, et al.
Pubblicazione: (2024)
di: Azmi, Behzad, et al.
Pubblicazione: (2024)
On Robust Regulation of PDEs: from Abstract Methods to PDE Controllers
di: Paunonen, Lassi, et al.
Pubblicazione: (2022)
di: Paunonen, Lassi, et al.
Pubblicazione: (2022)
Passive and reciprocal networks: From simple models to simple optimal controllers
di: Pates, Richard
Pubblicazione: (2022)
di: Pates, Richard
Pubblicazione: (2022)
Finite Codimensionality Method in Infinite-dimensional Optimization Problems
di: Liu, Xu, et al.
Pubblicazione: (2021)
di: Liu, Xu, et al.
Pubblicazione: (2021)
LQR control for a system describing the interaction between a floating solid and the surrounding fluid
di: Tucsnak, Marius, et al.
Pubblicazione: (2024)
di: Tucsnak, Marius, et al.
Pubblicazione: (2024)
Technological schemes and control methods in the reconstruction of parallel gas pipeline systems under non-stationary conditions
di: Aliyev, Ilgar
Pubblicazione: (2025)
di: Aliyev, Ilgar
Pubblicazione: (2025)
Input design for the optimal control of networked moments
di: Solimine, Philip, et al.
Pubblicazione: (2021)
di: Solimine, Philip, et al.
Pubblicazione: (2021)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
di: Zheng, Yaowei, et al.
Pubblicazione: (2026)
di: Zheng, Yaowei, et al.
Pubblicazione: (2026)
Lecture Notes on Control System Theory and Design
di: Basar, Tamer, et al.
Pubblicazione: (2020)
di: Basar, Tamer, et al.
Pubblicazione: (2020)
Formalising the intentional stance 2: a coinductive approach
di: McGregor, Simon, et al.
Pubblicazione: (2025)
di: McGregor, Simon, et al.
Pubblicazione: (2025)
Formalising the intentional stance 1: attributing goals and beliefs to stochastic processes
di: McGregor, Simon, et al.
Pubblicazione: (2024)
di: McGregor, Simon, et al.
Pubblicazione: (2024)
Measurement-based Initial Point Smoothing and Control Approach to Quantum Memory Systems
di: Vladimirov, Igor G., et al.
Pubblicazione: (2025)
di: Vladimirov, Igor G., et al.
Pubblicazione: (2025)
Distributed Online Optimization for Multi-Agent Optimal Transport
di: Krishnan, Vishaal, et al.
Pubblicazione: (2018)
di: Krishnan, Vishaal, et al.
Pubblicazione: (2018)
Optimal control of infinite-dimensional dissipative systems
di: Hastir, Anthony, et al.
Pubblicazione: (2026)
di: Hastir, Anthony, et al.
Pubblicazione: (2026)
Legendre-Moment Transform for Linear Ensemble Control and Computation
di: Ning, Xin, et al.
Pubblicazione: (2024)
di: Ning, Xin, et al.
Pubblicazione: (2024)
Optimal control of SDEs with merely measurable drift: an HJB approach
di: Du, Kai, et al.
Pubblicazione: (2025)
di: Du, Kai, et al.
Pubblicazione: (2025)
Nonlinear Conjugate Gradient Methods for PDE Constrained Shape Optimization Based on Steklov-Poincaré-Type Metrics
di: Blauth, Sebastian
Pubblicazione: (2020)
di: Blauth, Sebastian
Pubblicazione: (2020)
Quantitative Soft-to-Hard Terminal Constraint Convergence for the Heat Equation
di: Kwon, Sung-Sik
Pubblicazione: (2026)
di: Kwon, Sung-Sik
Pubblicazione: (2026)
Turnpike and dissipativity in generalized discrete-time stochastic linear-quadratic optimal control
di: Schießl, Jonas, et al.
Pubblicazione: (2023)
di: Schießl, Jonas, et al.
Pubblicazione: (2023)
Infinite time horizon stochastic recursive control problems with jumps: dynamic programming and stochastic verification theorems
di: Luo, Sheng, et al.
Pubblicazione: (2024)
di: Luo, Sheng, et al.
Pubblicazione: (2024)
Turnpike Property of Stochastic Linear-Quadratic Optimal Control Problems in Large Horizons with Regime Switching I: Homogeneous Cases
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
Trading with propagators and constraints: applications to optimal execution and battery storage
di: Jaber, Eduardo Abi, et al.
Pubblicazione: (2024)
di: Jaber, Eduardo Abi, et al.
Pubblicazione: (2024)
Controllability and Vector Potential
di: Shankar, Shiva
Pubblicazione: (2019)
di: Shankar, Shiva
Pubblicazione: (2019)
Measure propagation along a $\mathscr{C}^0$-vector field and wave controllability on a rough compact manifold
di: Burq, Nicolas, et al.
Pubblicazione: (2023)
di: Burq, Nicolas, et al.
Pubblicazione: (2023)
Turnpike Property of a Linear-Quadratic Optimal Control Problem in Large Horizons with Regime Switching II: Non-Homogeneous Cases
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
di: Mei, Hongwei, et al.
Pubblicazione: (2025)
Boundary controllability of incompressible Euler fluids with Boussinesq heat effects
di: Fernández-Cara, Enrique, et al.
Pubblicazione: (2024)
di: Fernández-Cara, Enrique, et al.
Pubblicazione: (2024)
Dissipativity, reciprocity and passive network synthesis: from Jan Willems' seminal Dissipative Dynamical Systems papers to the present day
di: Hughes, Timothy H., et al.
Pubblicazione: (2021)
di: Hughes, Timothy H., et al.
Pubblicazione: (2021)
Second-Order $Λ$-Sets and Extensions to Non-Smooth, Hybrid, and Stochastic Optimal Control
di: Rashid, Mohammad H. M
Pubblicazione: (2025)
di: Rashid, Mohammad H. M
Pubblicazione: (2025)
Generative optimal transport via forward-backward HJB matching
di: Yang, Haiqian, et al.
Pubblicazione: (2026)
di: Yang, Haiqian, et al.
Pubblicazione: (2026)
Convergence and turnpike properties of linear-quadratic mean field control problems with common noise
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026)
Multi-period Asset-liability Management with Reinforcement Learning in a Regime-Switching Market
di: Gao, Zhongqin, et al.
Pubblicazione: (2025)
di: Gao, Zhongqin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
di: Cui, Leilei, et al.
Pubblicazione: (2025) -
Policy Gradient for Continuous-Time Mean-Field Control
di: Bayraktar, Erhan, et al.
Pubblicazione: (2026) -
Compositional Construction of Barrier Functions for Switched Impulsive Systems
di: Bieker, Katharina, et al.
Pubblicazione: (2024) -
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
di: Bäuerle, Nicole, et al.
Pubblicazione: (2026) -
Transport maps as flows of control-affine systems
di: Caponigro, Marco, et al.
Pubblicazione: (2024)