Some remarks on gradient dominance and LQR policy optimization
Fuente:
arXiv
Saved in:
| Main Author: | Sontag, Eduardo D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Why should autoencoders work?
by: Kvalheim, Matthew D., et al.
Published: (2023)
by: Kvalheim, Matthew D., et al.
Published: (2023)
Autoencoding Dynamics: Topological Limitations and Capabilities
by: Kvalheim, Matthew D., et al.
Published: (2025)
by: Kvalheim, Matthew D., et al.
Published: (2025)
A policy gradient approach for optimization of smooth risk measures
by: Vijayan, Nithia, et al.
Published: (2022)
by: Vijayan, Nithia, et al.
Published: (2022)
Policy Gradient for LQR with Domain Randomization
by: Fujinami, Tesshu, et al.
Published: (2025)
by: Fujinami, Tesshu, et al.
Published: (2025)
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024)
by: Weissmann, Simon, et al.
Published: (2024)
Equivalence of stochastic and deterministic policy gradients
by: Todorov, Emo
Published: (2025)
by: Todorov, Emo
Published: (2025)
Policy gradient methods for ordinal policies
by: Weinberger, Simón, et al.
Published: (2025)
by: Weinberger, Simón, et al.
Published: (2025)
Learning Genetic Circuit Modules with Neural Networks: Full Version
by: Wang, Jichi, et al.
Published: (2025)
by: Wang, Jichi, et al.
Published: (2025)
Bayesian policy gradient and actor-critic algorithms
by: Ghavamzadeh, Mohammad, et al.
Published: (2026)
by: Ghavamzadeh, Mohammad, et al.
Published: (2026)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
ISOPO: Proximal policy gradients without pi-old
by: Abrahamsen, Nilin
Published: (2025)
by: Abrahamsen, Nilin
Published: (2025)
Trainability issues in quantum policy gradients
by: Sequeira, André, et al.
Published: (2024)
by: Sequeira, André, et al.
Published: (2024)
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
by: Dufort-Labbé, Simon, et al.
Published: (2026)
by: Dufort-Labbé, Simon, et al.
Published: (2026)
q-exponential family for policy optimization
by: Zhu, Lingwei, et al.
Published: (2024)
by: Zhu, Lingwei, et al.
Published: (2024)
Max-Entropy Reinforcement Learning with Flow Matching and A Case Study on LQR
by: Zhang, Yuyang, et al.
Published: (2025)
by: Zhang, Yuyang, et al.
Published: (2025)
A remark on omega limit sets for non-expansive dynamics
by: Duvall, Alon, et al.
Published: (2024)
by: Duvall, Alon, et al.
Published: (2024)
Almost Surely $\sqrt{T}$ Regret for Adaptive LQR
by: Lu, Yiwen, et al.
Published: (2023)
by: Lu, Yiwen, et al.
Published: (2023)
Theoretical Analysis of Weak-to-Strong Generalization
by: Lang, Hunter, et al.
Published: (2024)
by: Lang, Hunter, et al.
Published: (2024)
Prediction-powered Generalization of Causal Inferences
by: Demirel, Ilker, et al.
Published: (2024)
by: Demirel, Ilker, et al.
Published: (2024)
Seq-to-Final: A Benchmark for Tuning from Sequential Distributions to a Final Time Point
by: Ji, Christina X, et al.
Published: (2024)
by: Ji, Christina X, et al.
Published: (2024)
Learning in complex action spaces without policy gradients
by: Tavakoli, Arash, et al.
Published: (2024)
by: Tavakoli, Arash, et al.
Published: (2024)
Control randomisation approach for policy gradient and application to reinforcement learning in optimal switching
by: Denkert, Robert, et al.
Published: (2024)
by: Denkert, Robert, et al.
Published: (2024)
Risk and optimal policies in bandit experiments
by: Adusumilli, Karun
Published: (2021)
by: Adusumilli, Karun
Published: (2021)
A stochastic gradient method for trilevel optimization
by: Giovannelli, Tommaso, et al.
Published: (2025)
by: Giovannelli, Tommaso, et al.
Published: (2025)
A policy gradient approach for Finite Horizon Constrained Markov Decision Processes
by: Guin, Soumyajit, et al.
Published: (2022)
by: Guin, Soumyajit, et al.
Published: (2022)
On the System Theoretic Offline Learning of Continuous-Time LQR with Exogenous Disturbances
by: Mukherjee, Sayak, et al.
Published: (2025)
by: Mukherjee, Sayak, et al.
Published: (2025)
Per-example gradients: a new frontier for understanding and improving optimizers
by: Roulet, Vincent, et al.
Published: (2025)
by: Roulet, Vincent, et al.
Published: (2025)
Overshoot: Taking advantage of future gradients in momentum-based stochastic optimization
by: Kopal, Jakub, et al.
Published: (2025)
by: Kopal, Jakub, et al.
Published: (2025)
Convergence Analysis of Gradient Flow for Overparameterized LQR Formulations
by: de Oliveira, Arthur Castello B., et al.
Published: (2024)
by: de Oliveira, Arthur Castello B., et al.
Published: (2024)
Exact Recovery Guarantees for Parameterized Nonlinear System Identification Problem under Sparse Disturbances or Semi-Oblivious Attacks
by: Zhang, Haixiang, et al.
Published: (2024)
by: Zhang, Haixiang, et al.
Published: (2024)
Dealing with unbounded gradients in stochastic saddle-point optimization
by: Neu, Gergely, et al.
Published: (2024)
by: Neu, Gergely, et al.
Published: (2024)
Small-Disturbance Input-to-State Stability of Perturbed Gradient Flows: Applications to LQR Problem
by: Cui, Leilei, et al.
Published: (2023)
by: Cui, Leilei, et al.
Published: (2023)
On the Convergence of Overparameterized Problems: Inherent Properties of the Compositional Structure of Neural Networks
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
by: de Oliveira, Arthur Castello Branco, et al.
Published: (2025)
Uncovering Bias Mechanisms in Observational Studies
by: Demirel, Ilker, et al.
Published: (2025)
by: Demirel, Ilker, et al.
Published: (2025)
Probably approximately correct high-dimensional causal effect estimation given a valid adjustment set
by: Choo, Davin, et al.
Published: (2024)
by: Choo, Davin, et al.
Published: (2024)
Accelerating trajectory optimization with Sobolev-trained diffusion policies
by: Hellard, Théotime Le, et al.
Published: (2026)
by: Hellard, Théotime Le, et al.
Published: (2026)
Safe Reinforcement Learning-Based Vibration Control: Overcoming Training Risks with LQR Guidance
by: Thorat, Rohan Vitthal, et al.
Published: (2025)
by: Thorat, Rohan Vitthal, et al.
Published: (2025)
A note on convergence of Wasserstein policy optimization
by: Šiška, David, et al.
Published: (2026)
by: Šiška, David, et al.
Published: (2026)
Smoothed functional-based gradient algorithms for off-policy reinforcement learning: A non-asymptotic viewpoint
by: Vijayan, Nithia, et al.
Published: (2021)
by: Vijayan, Nithia, et al.
Published: (2021)
A novel gradient-based method for decision trees optimizing arbitrary differential loss functions
by: Konstantinov, Andrei V., et al.
Published: (2025)
by: Konstantinov, Andrei V., et al.
Published: (2025)
Similar Items
-
Why should autoencoders work?
by: Kvalheim, Matthew D., et al.
Published: (2023) -
Autoencoding Dynamics: Topological Limitations and Capabilities
by: Kvalheim, Matthew D., et al.
Published: (2025) -
A policy gradient approach for optimization of smooth risk measures
by: Vijayan, Nithia, et al.
Published: (2022) -
Policy Gradient for LQR with Domain Randomization
by: Fujinami, Tesshu, et al.
Published: (2025) -
Almost sure convergence rates of stochastic gradient methods under gradient domination
by: Weissmann, Simon, et al.
Published: (2024)