On the continuity and smoothness of the value function in reinforcement learning and optimal control
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Harder, Hans, Peitz, Sebastian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Observable Neural ODEs for Identifiable Causal Forecasting in Continuous Time
von: Wendland, Jennifer, et al.
Veröffentlicht: (2026)
von: Wendland, Jennifer, et al.
Veröffentlicht: (2026)
From Features to States: Data-Driven Selection of Measured State Variables via RFE-DMDc
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Data-driven optimal prediction with control
von: Katrutsa, Aleksandr, et al.
Veröffentlicht: (2024)
von: Katrutsa, Aleksandr, et al.
Veröffentlicht: (2024)
A Moreau Envelope Approach for LQR Meta-Policy Estimation
von: Aravind, Ashwin, et al.
Veröffentlicht: (2024)
von: Aravind, Ashwin, et al.
Veröffentlicht: (2024)
Determining Disturbance Recovery Conditions by Inverse Sensitivity Minimization
von: Fisher, Michael W., et al.
Veröffentlicht: (2025)
von: Fisher, Michael W., et al.
Veröffentlicht: (2025)
Curve-Induced Dynamical Systems on Riemannian Manifolds and Lie Groups
von: Bakker, Saray, et al.
Veröffentlicht: (2026)
von: Bakker, Saray, et al.
Veröffentlicht: (2026)
Dynamics and Computational Principles of Echo State Networks: A Mathematical Perspective
von: Singh, Pradeep, et al.
Veröffentlicht: (2025)
von: Singh, Pradeep, et al.
Veröffentlicht: (2025)
From geometry to dynamics: Learning overdamped Langevin dynamics from sparse observations with geometric constraints
von: Maoutsa, Dimitra
Veröffentlicht: (2025)
von: Maoutsa, Dimitra
Veröffentlicht: (2025)
Real-Time Learning of Predictive Dynamic Obstacle Models for Robotic Motion Planning
von: Kombo, Stella, et al.
Veröffentlicht: (2025)
von: Kombo, Stella, et al.
Veröffentlicht: (2025)
On Higher Order Drift and Diffusion Estimates for Stochastic SINDy
von: Wanner, Mathias, et al.
Veröffentlicht: (2023)
von: Wanner, Mathias, et al.
Veröffentlicht: (2023)
Stability of randomly switching stochastic reaction networks with asymptotically linear transition rates
von: Cappelletti, Daniele, et al.
Veröffentlicht: (2025)
von: Cappelletti, Daniele, et al.
Veröffentlicht: (2025)
Data-driven Control of T-Product-based Dynamical Systems
von: He, Ziqin, et al.
Veröffentlicht: (2025)
von: He, Ziqin, et al.
Veröffentlicht: (2025)
Optimal response for stochastic differential equations by local kernel perturbations
von: del Sarto, Gianmarco, et al.
Veröffentlicht: (2025)
von: del Sarto, Gianmarco, et al.
Veröffentlicht: (2025)
Interaction-Aware Model Predictive Decision-Making for Socially-Compliant Autonomous Driving in Mixed Urban Traffic Scenarios
von: Varga, Balint, et al.
Veröffentlicht: (2025)
von: Varga, Balint, et al.
Veröffentlicht: (2025)
On Unstable Fixed Points in Modern Continuous Hopfield Networks
von: Beise, Hans-Peter
Veröffentlicht: (2026)
von: Beise, Hans-Peter
Veröffentlicht: (2026)
Universal bounds on the entropy of toroidal attractors
von: Macías, P. Montealegre, et al.
Veröffentlicht: (2024)
von: Macías, P. Montealegre, et al.
Veröffentlicht: (2024)
Stability properties of Minimal Gated Unit neural networks
von: De Carli, Stefano, et al.
Veröffentlicht: (2026)
von: De Carli, Stefano, et al.
Veröffentlicht: (2026)
Chaos-Free Networks are Stable Recurrent Neural Networks
von: De Carli, Stefano, et al.
Veröffentlicht: (2026)
von: De Carli, Stefano, et al.
Veröffentlicht: (2026)
Echoes of the Past: A Unified Perspective on Fading memory and Echo States
von: Ortega, Juan-Pablo, et al.
Veröffentlicht: (2025)
von: Ortega, Juan-Pablo, et al.
Veröffentlicht: (2025)
Dynamic Hybrid Modeling: Incremental Identification and Model Predictive Control
von: Caspari, Adrian, et al.
Veröffentlicht: (2025)
von: Caspari, Adrian, et al.
Veröffentlicht: (2025)
Constrained Ergodic optimization for generic continuous functions
von: Motonaga, Shoya, et al.
Veröffentlicht: (2022)
von: Motonaga, Shoya, et al.
Veröffentlicht: (2022)
A criterion to detect a nontrivial homology of an invariant set of a flow in $\mathbb{R}^3$
von: Sánchez-Gabites, J. J.
Veröffentlicht: (2024)
von: Sánchez-Gabites, J. J.
Veröffentlicht: (2024)
Distributed State Estimation for Linear Time-invariant Systems with Aperiodic Sampled Measurement
von: Wang, Shimin, et al.
Veröffentlicht: (2022)
von: Wang, Shimin, et al.
Veröffentlicht: (2022)
Next Generation Equation-Free Multiscale Modelling of Crowd Dynamics via Machine Learning
von: Alvarez, Hector Vargas, et al.
Veröffentlicht: (2025)
von: Alvarez, Hector Vargas, et al.
Veröffentlicht: (2025)
A nonparametric learning framework for nonlinear robust output regulation
von: Wang, Shimin, et al.
Veröffentlicht: (2023)
von: Wang, Shimin, et al.
Veröffentlicht: (2023)
HRM-Agent: Training a recurrent reasoning model in dynamic environments using reinforcement learning
von: Dang, Long H, et al.
Veröffentlicht: (2025)
von: Dang, Long H, et al.
Veröffentlicht: (2025)
On the cardinality of measures of maximal relative entropy for smooth skew products
von: Castro, Matheus M., et al.
Veröffentlicht: (2025)
von: Castro, Matheus M., et al.
Veröffentlicht: (2025)
Computing Safety Margins of Parameterized Nonlinear Systems for Vulnerability Assessment via Trajectory Sensitivities
von: Fisher, Michael W.
Veröffentlicht: (2025)
von: Fisher, Michael W.
Veröffentlicht: (2025)
Full flexibility of entropies among ergodic measures for partially hyperbolic diffeomorphisms
von: Díaz, Lorenzo J., et al.
Veröffentlicht: (2025)
von: Díaz, Lorenzo J., et al.
Veröffentlicht: (2025)
Anti-classification for flows on two-tori
von: Goncharuk, Nataliya
Veröffentlicht: (2025)
von: Goncharuk, Nataliya
Veröffentlicht: (2025)
Formalising the intentional stance 2: a coinductive approach
von: McGregor, Simon, et al.
Veröffentlicht: (2025)
von: McGregor, Simon, et al.
Veröffentlicht: (2025)
Formalising the intentional stance 1: attributing goals and beliefs to stochastic processes
von: McGregor, Simon, et al.
Veröffentlicht: (2024)
von: McGregor, Simon, et al.
Veröffentlicht: (2024)
Hierarchical Cyclic Pursuit: Algebraic Curves Containing the Laplacian Spectra
von: Parsegov, Sergei E., et al.
Veröffentlicht: (2022)
von: Parsegov, Sergei E., et al.
Veröffentlicht: (2022)
A Fully Data-Driven Value Iteration for Stochastic LQR: Convergence, Robustness and Stability
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
von: Cui, Leilei, et al.
Veröffentlicht: (2025)
Sliding motions on systems with non-Euclidean state spaces: A differential-geometric perspective
von: Castaños, Fernando
Veröffentlicht: (2025)
von: Castaños, Fernando
Veröffentlicht: (2025)
Optimal response for stochastic differential equations in $\mathbb{T}^d$ with perturbations on the drift term
von: Del Sarto, Gianmarco, et al.
Veröffentlicht: (2026)
von: Del Sarto, Gianmarco, et al.
Veröffentlicht: (2026)
SupplyGraph: A Benchmark Dataset for Supply Chain Planning using Graph Neural Networks
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
A Note on Semantic Diffusion
von: Ryjov, Alexander P., et al.
Veröffentlicht: (2025)
von: Ryjov, Alexander P., et al.
Veröffentlicht: (2025)
A Logvinenko-Sereda theorem for vector-valued functions and application to control theory
von: Bombach, Clemens, et al.
Veröffentlicht: (2024)
von: Bombach, Clemens, et al.
Veröffentlicht: (2024)
Simulating Fokker-Planck equations via mean field control of score-based normalizing flows
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Observable Neural ODEs for Identifiable Causal Forecasting in Continuous Time
von: Wendland, Jennifer, et al.
Veröffentlicht: (2026) -
From Features to States: Data-Driven Selection of Measured State Variables via RFE-DMDc
von: Wang, Haoyu, et al.
Veröffentlicht: (2025) -
Data-driven optimal prediction with control
von: Katrutsa, Aleksandr, et al.
Veröffentlicht: (2024) -
A Moreau Envelope Approach for LQR Meta-Policy Estimation
von: Aravind, Ashwin, et al.
Veröffentlicht: (2024) -
Determining Disturbance Recovery Conditions by Inverse Sensitivity Minimization
von: Fisher, Michael W., et al.
Veröffentlicht: (2025)