Efficiently Escaping Saddle Points for Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Khorasani, Sadegh, Salehkaleybar, Saber, Kiyavash, Negar, He, Niao, Grossglauser, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
Hierarchical Reinforcement Learning with Targeted Causal Interventions
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
Inference Time Causal Probing in LLMs
by: Khorasani, Sadegh, et al.
Published: (2026)
by: Khorasani, Sadegh, et al.
Published: (2026)
Optimal Local Convergence Rates of Stochastic First-Order Methods under Local $α$-PL
by: Masiha, Saeed, et al.
Published: (2024)
by: Masiha, Saeed, et al.
Published: (2024)
FFCG: Effective and Fast Family Column Generation for Solving Large-Scale Linear Program
by: Hu, Yi-Xiang, et al.
Published: (2024)
by: Hu, Yi-Xiang, et al.
Published: (2024)
Ensemble based Closed-Loop Optimal Control using Physics-Informed Neural Networks
by: Barry-Straume, Jostein, et al.
Published: (2025)
by: Barry-Straume, Jostein, et al.
Published: (2025)
Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Bilinear Mamba-Koopman Neural MPC for Varying Dynamics
by: Pagi, Matan, et al.
Published: (2026)
by: Pagi, Matan, et al.
Published: (2026)
Reinforcement Learning in MDPs with Information-Ordered Policies
by: Zhang, Zhongjun, et al.
Published: (2025)
by: Zhang, Zhongjun, et al.
Published: (2025)
Multi-Objective Optimization and Hyperparameter Tuning With Desirability Functions
by: Bartz-Beielstein, Thomas
Published: (2025)
by: Bartz-Beielstein, Thomas
Published: (2025)
Measuring IIA Violations in Similarity Choices with Bayesian Models
by: Corrêa, Hugo Sales, et al.
Published: (2025)
by: Corrêa, Hugo Sales, et al.
Published: (2025)
The Blueprints of Intelligence: A Functional-Topological Foundation for Perception and Representation
by: Di Santi, Eduardo
Published: (2025)
by: Di Santi, Eduardo
Published: (2025)
Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
by: Ueda, Ryota, et al.
Published: (2025)
by: Ueda, Ryota, et al.
Published: (2025)
Multi-Objective Optimization with Desirability and Morris-Mitchell Criterion
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
by: Yuan, Xinzhe, et al.
Published: (2026)
by: Yuan, Xinzhe, et al.
Published: (2026)
Correcting Stochastic Update Bias in Preconditioned Language Model Optimizers
by: Nayak, Nikhil, et al.
Published: (2026)
by: Nayak, Nikhil, et al.
Published: (2026)
Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling
by: Anand, Emile, et al.
Published: (2026)
by: Anand, Emile, et al.
Published: (2026)
Amortized Molecular Optimization via Group Relative Policy Optimization
by: Javaid, Muhammad bin, et al.
Published: (2026)
by: Javaid, Muhammad bin, et al.
Published: (2026)
Deep Policy Iteration with Integer Programming for Inventory Management
by: Harsha, Pavithra, et al.
Published: (2021)
by: Harsha, Pavithra, et al.
Published: (2021)
ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning
by: Shamass, Faiq
Published: (2026)
by: Shamass, Faiq
Published: (2026)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
Predictive Maintenance in Photovoltaic Plants with a Big Data Approach
by: Betti, Alessandro, et al.
Published: (2019)
by: Betti, Alessandro, et al.
Published: (2019)
Reliability of Single-Level Equality-Constrained Inverse Optimal Control
by: Bečanović, Filip, et al.
Published: (2025)
by: Bečanović, Filip, et al.
Published: (2025)
Scalable Nested Optimization for Deep Learning
by: Lorraine, Jonathan
Published: (2024)
by: Lorraine, Jonathan
Published: (2024)
On the Curse of Memory in Recurrent Neural Networks: Approximation and Optimization Analysis
by: Li, Zhong, et al.
Published: (2020)
by: Li, Zhong, et al.
Published: (2020)
Positive-Only Drifting Policy Optimization
by: Zhang, Qi
Published: (2026)
by: Zhang, Qi
Published: (2026)
Deep Reinforcement Learning for Day-to-day Dynamic Tolling in Tradable Credit Schemes
by: Wu, Xiaoyi, et al.
Published: (2025)
by: Wu, Xiaoyi, et al.
Published: (2025)
Efficient Training of Physics-enhanced Neural ODEs via Direct Collocation and Nonlinear Programming
by: Langenkamp, Linus, et al.
Published: (2025)
by: Langenkamp, Linus, et al.
Published: (2025)
Future Aware Safe Active Learning of Time Varying Systems using Gaussian Processes
by: Lange-Hegermann, Markus, et al.
Published: (2024)
by: Lange-Hegermann, Markus, et al.
Published: (2024)
Identifying Policy Gradient Subspaces
by: Schneider, Jan, et al.
Published: (2024)
by: Schneider, Jan, et al.
Published: (2024)
Physics-Informed Policy Optimization via Analytic Dynamics Regularization
by: Chandra, Namai, et al.
Published: (2026)
by: Chandra, Namai, et al.
Published: (2026)
Black-Box Uniform Stability for Non-Euclidean Empirical Risk Minimization
by: Vary, Simon, et al.
Published: (2024)
by: Vary, Simon, et al.
Published: (2024)
Natural Gradient VI: Guarantees for Non-Conjugate Models
by: Sun, Fangyuan, et al.
Published: (2025)
by: Sun, Fangyuan, et al.
Published: (2025)
Adaptive Candidate Point Thompson Sampling for High-Dimensional Bayesian Optimization
by: Fan, Donney, et al.
Published: (2026)
by: Fan, Donney, et al.
Published: (2026)
Rewarded Region Replay (R3) for Policy Learning with Discrete Action Space
by: Li, Bangzheng, et al.
Published: (2024)
by: Li, Bangzheng, et al.
Published: (2024)
`Just One More Sensor is Enough' -- Iterative Water Leak Localization with Physical Simulation and a Small Number of Pressure Sensors
by: Cholewa, Michał, et al.
Published: (2024)
by: Cholewa, Michał, et al.
Published: (2024)
Decentralized Optimization with Topology-Independent Communication
by: Lin, Ying, et al.
Published: (2025)
by: Lin, Ying, et al.
Published: (2025)
Fractional Policy Gradients: Reinforcement Learning with Long-Term Memory
by: Pawar, Urvi, et al.
Published: (2025)
by: Pawar, Urvi, et al.
Published: (2025)
Evaluating and Learning Robust Bandit Policies Under Uncertain Causal Mechanisms
by: Avery, Katherine, et al.
Published: (2025)
by: Avery, Katherine, et al.
Published: (2025)
Dynamic Hybrid Modeling: Incremental Identification and Model Predictive Control
by: Caspari, Adrian, et al.
Published: (2025)
by: Caspari, Adrian, et al.
Published: (2025)
Similar Items
-
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025) -
Hierarchical Reinforcement Learning with Targeted Causal Interventions
by: Khorasani, Sadegh, et al.
Published: (2025) -
Inference Time Causal Probing in LLMs
by: Khorasani, Sadegh, et al.
Published: (2026) -
Optimal Local Convergence Rates of Stochastic First-Order Methods under Local $α$-PL
by: Masiha, Saeed, et al.
Published: (2024) -
FFCG: Effective and Fast Family Column Generation for Solving Large-Scale Linear Program
by: Hu, Yi-Xiang, et al.
Published: (2024)