A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Zeng, Sihan, Doan, Thinh T., Romberg, Justin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2021
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
por: Zeng, Sihan, et al.
Publicado: (2021)
por: Zeng, Sihan, et al.
Publicado: (2021)
Accelerated Multi-Time-Scale Stochastic Approximation: Optimal Complexity and Applications in Reinforcement Learning and Multi-Agent Games
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Fast Nonlinear Two-Time-Scale Stochastic Approximation: Achieving $O(1/k)$ Finite-Sample Complexity
por: Doan, Thinh T.
Publicado: (2024)
por: Doan, Thinh T.
Publicado: (2024)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
por: Bai, Yitao, et al.
Publicado: (2026)
por: Bai, Yitao, et al.
Publicado: (2026)
Accelerating Multi-Task Temporal Difference Learning under Low-Rank Representation
por: Bai, Yitao, et al.
Publicado: (2025)
por: Bai, Yitao, et al.
Publicado: (2025)
Resilient Two-Time-Scale Local Stochastic Gradient Descent for Byzantine Federated Learning
por: Dutta, Amit, et al.
Publicado: (2024)
por: Dutta, Amit, et al.
Publicado: (2024)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
por: Zeng, Sihan, et al.
Publicado: (2026)
por: Zeng, Sihan, et al.
Publicado: (2026)
Finite-Time Decoupled Convergence in Nonlinear Two-Time-Scale Stochastic Approximation
por: Han, Yuze, et al.
Publicado: (2024)
por: Han, Yuze, et al.
Publicado: (2024)
Convergence Rate in Nonlinear Two-Time-Scale Stochastic Approximation with State (Time)-Dependence
por: Chen, Zixi, et al.
Publicado: (2025)
por: Chen, Zixi, et al.
Publicado: (2025)
Transit Network Design with Two-Level Demand Uncertainties: A Machine Learning and Contextual Stochastic Optimization Framework
por: Guan, Hongzhao, et al.
Publicado: (2026)
por: Guan, Hongzhao, et al.
Publicado: (2026)
Reinforcement Learning for a Discrete-Time Linear-Quadratic Control Problem with an Application
por: Li, Lucky
Publicado: (2024)
por: Li, Lucky
Publicado: (2024)
Tight Finite Time Bounds of Two-Time-Scale Linear Stochastic Approximation with Markovian Noise
por: Haque, Shaan Ul, et al.
Publicado: (2023)
por: Haque, Shaan Ul, et al.
Publicado: (2023)
Machine Learning-Augmented Optimization of Large Bilevel and Two-stage Stochastic Programs: Application to Cycling Network Design
por: Chan, Timothy C. Y., et al.
Publicado: (2022)
por: Chan, Timothy C. Y., et al.
Publicado: (2022)
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
por: Takubo, Yuji, et al.
Publicado: (2021)
por: Takubo, Yuji, et al.
Publicado: (2021)
$O(1/k)$ Finite-Time Bound for Non-Linear Two-Time-Scale Stochastic Approximation
por: Chandak, Siddharth
Publicado: (2025)
por: Chandak, Siddharth
Publicado: (2025)
Finite-Time Bounds for Two-Time-Scale Stochastic Approximation with Arbitrary Norm Contractions and Markovian Noise
por: Chandak, Siddharth, et al.
Publicado: (2025)
por: Chandak, Siddharth, et al.
Publicado: (2025)
SANIA: Polyak-type Optimization Framework Leads to Scale Invariant Stochastic Algorithms
por: Abdukhakimov, Farshed, et al.
Publicado: (2023)
por: Abdukhakimov, Farshed, et al.
Publicado: (2023)
Biased Stochastic First-Order Methods for Conditional Stochastic Optimization and Applications in Meta Learning
por: Hu, Yifan, et al.
Publicado: (2020)
por: Hu, Yifan, et al.
Publicado: (2020)
An Efficient On-Policy Deep Learning Framework for Stochastic Optimal Control
por: Hua, Mengjian, et al.
Publicado: (2024)
por: Hua, Mengjian, et al.
Publicado: (2024)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
por: Mazumdar, Abhijit, et al.
Publicado: (2024)
por: Mazumdar, Abhijit, et al.
Publicado: (2024)
Decoupled Functional Central Limit Theorems for Two-Time-Scale Stochastic Approximation
por: Han, Yuze, et al.
Publicado: (2024)
por: Han, Yuze, et al.
Publicado: (2024)
Nonasymptotic CLT and Error Bounds for Two-Time-Scale Stochastic Approximation
por: Kong, Seo Taek, et al.
Publicado: (2025)
por: Kong, Seo Taek, et al.
Publicado: (2025)
QCQP-Net: Reliably Learning Feasible Alternating Current Optimal Power Flow Solutions Under Constraints
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Bayesian Optimization for Non-Convex Two-Stage Stochastic Optimization Problems
por: Buckingham, Jack M., et al.
Publicado: (2024)
por: Buckingham, Jack M., et al.
Publicado: (2024)
Two-Timescale Optimization Framework for Sparse-Feedback Linear-Quadratic Optimal Control
por: Feng, Lechen, et al.
Publicado: (2024)
por: Feng, Lechen, et al.
Publicado: (2024)
Prescriptive PCA: Dimensionality Reduction for Two-stage Stochastic Optimization
por: He, Long, et al.
Publicado: (2023)
por: He, Long, et al.
Publicado: (2023)
Risk-Aware Safe Reinforcement Learning for Control of Stochastic Linear Systems
por: Esmaeili, Babak, et al.
Publicado: (2025)
por: Esmaeili, Babak, et al.
Publicado: (2025)
Improved Learning Rates for Stochastic Optimization
por: Li, Shaojie, et al.
Publicado: (2021)
por: Li, Shaojie, et al.
Publicado: (2021)
A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance
por: Wolter, Axel Friedrich, et al.
Publicado: (2025)
por: Wolter, Axel Friedrich, et al.
Publicado: (2025)
Convex Chance-Constrained Stochastic Control under Uncertain Specifications with Application to Learning-Based Hybrid Powertrain Control
por: Kato, Teruki, et al.
Publicado: (2026)
por: Kato, Teruki, et al.
Publicado: (2026)
A Provably Convergent Plug-and-Play Framework for Stochastic Bilevel Optimization
por: Chu, Tianshu, et al.
Publicado: (2025)
por: Chu, Tianshu, et al.
Publicado: (2025)
Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
por: Meng, Huiling, et al.
Publicado: (2024)
por: Meng, Huiling, et al.
Publicado: (2024)
A Control Theoretic Framework for Adaptive Gradient Optimizers in Machine Learning
por: Chakrabarti, Kushal, et al.
Publicado: (2022)
por: Chakrabarti, Kushal, et al.
Publicado: (2022)
Heuristics for Combinatorial Optimization via Value-based Reinforcement Learning: A Unified Framework and Analysis
por: Davidovich, Orit, et al.
Publicado: (2025)
por: Davidovich, Orit, et al.
Publicado: (2025)
Deep Learning for Two-Stage Robust Integer Optimization
por: Dumouchelle, Justin, et al.
Publicado: (2023)
por: Dumouchelle, Justin, et al.
Publicado: (2023)
Central Limit Theorem for Two-Timescale Stochastic Approximation with Markovian Noise: Theory and Applications
por: Hu, Jie, et al.
Publicado: (2024)
por: Hu, Jie, et al.
Publicado: (2024)
Continuous Q-Score Matching: Diffusion Guided Reinforcement Learning for Continuous-Time Control
por: Hua, Chengxiu, et al.
Publicado: (2025)
por: Hua, Chengxiu, et al.
Publicado: (2025)
Finite-Time Analysis of Stochastic Nonconvex Nonsmooth Optimization on the Riemannian Manifolds
por: Sahinoglu, Emre, et al.
Publicado: (2025)
por: Sahinoglu, Emre, et al.
Publicado: (2025)
Ejemplares similares
-
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024) -
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024) -
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
por: Zeng, Sihan, et al.
Publicado: (2021) -
Accelerated Multi-Time-Scale Stochastic Approximation: Optimal Complexity and Applications in Reinforcement Learning and Multi-Agent Games
por: Zeng, Sihan, et al.
Publicado: (2024) -
Fast Nonlinear Two-Time-Scale Stochastic Approximation: Achieving $O(1/k)$ Finite-Sample Complexity
por: Doan, Thinh T.
Publicado: (2024)