Achieve Performatively Optimal Policy for Performative Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Ziyi, Huang, Heng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Trade-off in Estimating the Number of Byzantine Clients in Federated Learning
by: Chen, Ziyi, et al.
Published: (2025)
by: Chen, Ziyi, et al.
Published: (2025)
Zeroth-Order Methods for Stochastic Nonconvex Nonsmooth Composite Optimization
by: Chen, Ziyi, et al.
Published: (2025)
by: Chen, Ziyi, et al.
Published: (2025)
Optimal and Order-optimal Gated Priority-based Greedy Policies for Two-layer Multi-item Order Fulfillment
by: Chen, Xi, et al.
Published: (2026)
by: Chen, Xi, et al.
Published: (2026)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
by: Shen, Kaichen, et al.
Published: (2026)
by: Shen, Kaichen, et al.
Published: (2026)
Revisiting Convergence: Shuffling Complexity Beyond Lipschitz Smoothness
by: He, Qi, et al.
Published: (2025)
by: He, Qi, et al.
Published: (2025)
Traversing Pareto Optimal Policies: Provably Efficient Multi-Objective Reinforcement Learning
by: Qiu, Shuang, et al.
Published: (2024)
by: Qiu, Shuang, et al.
Published: (2024)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
by: Dus, Mathias
Published: (2026)
by: Dus, Mathias
Published: (2026)
Distributionally Robust Multi-Objective Optimization
by: Yang, Yufeng, et al.
Published: (2026)
by: Yang, Yufeng, et al.
Published: (2026)
Learning-Based Optimal Control with Performance Guarantees for Unknown Systems with Latent States
by: Lefringhausen, Robert, et al.
Published: (2023)
by: Lefringhausen, Robert, et al.
Published: (2023)
Learning to Optimally Dispatch Power: Performance on a Nation-Wide Real-World Dataset
by: Boero, Ignacio, et al.
Published: (2025)
by: Boero, Ignacio, et al.
Published: (2025)
Achieving Linear Speedup for Composite Federated Learning
by: Huang, Kun, et al.
Published: (2026)
by: Huang, Kun, et al.
Published: (2026)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
A Data-Driven Real-Time Optimal Power Flow Algorithm Using Local Feedback
by: Liang, Heng, et al.
Published: (2025)
by: Liang, Heng, et al.
Published: (2025)
Robust Reinforcement Learning in Finance: Modeling Market Impact with Elliptic Uncertainty Sets
by: Ma, Shaocong, et al.
Published: (2025)
by: Ma, Shaocong, et al.
Published: (2025)
An Efficient On-Policy Deep Learning Framework for Stochastic Optimal Control
by: Hua, Mengjian, et al.
Published: (2024)
by: Hua, Mengjian, et al.
Published: (2024)
Achieving Tractable Minimax Optimal Regret in Average Reward MDPs
by: Boone, Victor, et al.
Published: (2024)
by: Boone, Victor, et al.
Published: (2024)
A Minibatch-SGD-Based Learning Meta-Policy for Inventory Systems with Myopic Optimal Policy
by: Lyu, Jiameng, et al.
Published: (2024)
by: Lyu, Jiameng, et al.
Published: (2024)
Finite-Time Analysis of On-Policy Heterogeneous Federated Reinforcement Learning
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Single- vs. Dual-Policy Reinforcement Learning for Dynamic Bike Rebalancing
by: Liang, Jiaqi, et al.
Published: (2024)
by: Liang, Jiaqi, et al.
Published: (2024)
Non-Uniform Noise-to-Signal Ratio in the REINFORCE Policy-Gradient Estimator
by: Han, Haoyu, et al.
Published: (2026)
by: Han, Haoyu, et al.
Published: (2026)
FedCluster: Boosting the Convergence of Federated Learning via Cluster-Cycling
by: Chen, Cheng, et al.
Published: (2020)
by: Chen, Cheng, et al.
Published: (2020)
On the Optimal Construction of Unbiased Gradient Estimators for Zeroth-Order Optimization
by: Ma, Shaocong, et al.
Published: (2025)
by: Ma, Shaocong, et al.
Published: (2025)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
by: Carmona, René, et al.
Published: (2019)
by: Carmona, René, et al.
Published: (2019)
TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
by: Sorokin, D., et al.
Published: (2023)
by: Sorokin, D., et al.
Published: (2023)
On Performance Guarantees for Federated Learning with Personalized Constraints
by: Ebrahimi, Mohammadjavad, et al.
Published: (2026)
by: Ebrahimi, Mohammadjavad, et al.
Published: (2026)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
by: Han, Yinbin, et al.
Published: (2023)
by: Han, Yinbin, et al.
Published: (2023)
SPABA: A Single-Loop and Probabilistic Stochastic Bilevel Algorithm Achieving Optimal Sample Complexity
by: Chu, Tianshu, et al.
Published: (2024)
by: Chu, Tianshu, et al.
Published: (2024)
Policy Transfer for Continuous-Time Reinforcement Learning: A (Rough) Differential Equation Approach
by: Guo, Xin, et al.
Published: (2025)
by: Guo, Xin, et al.
Published: (2025)
A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation
by: Zhao, Heyang, et al.
Published: (2023)
by: Zhao, Heyang, et al.
Published: (2023)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
by: Zeng, Sihan, et al.
Published: (2024)
by: Zeng, Sihan, et al.
Published: (2024)
Data-Driven Performance Guarantees for Classical and Learned Optimizers
by: Sambharya, Rajiv, et al.
Published: (2024)
by: Sambharya, Rajiv, et al.
Published: (2024)
Fill-and-Spill: Deep Reinforcement Learning Policy Gradient Methods for Reservoir Operation Decision and Control
by: Tabas, Sadegh Sadeghi, et al.
Published: (2024)
by: Tabas, Sadegh Sadeghi, et al.
Published: (2024)
FedSEA: Achieving Benefit of Parallelization in Federated Online Learning
by: Sahu, Harekrushna, et al.
Published: (2026)
by: Sahu, Harekrushna, et al.
Published: (2026)
Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions
by: Long, Kehan, et al.
Published: (2025)
by: Long, Kehan, et al.
Published: (2025)
On Policy Stochasticity in Mutual Information Optimal Control of Linear Systems
by: Enami, Shoju, et al.
Published: (2025)
by: Enami, Shoju, et al.
Published: (2025)
Provably Faster Algorithms for Bilevel Optimization via Without-Replacement Sampling
by: Li, Junyi, et al.
Published: (2024)
by: Li, Junyi, et al.
Published: (2024)
Riemannian Zeroth-Order Gradient Estimation with Structure-Preserving Metrics for Geodesically Incomplete Manifolds
by: Ma, Shaocong, et al.
Published: (2026)
by: Ma, Shaocong, et al.
Published: (2026)
Learning an Optimal Assortment Policy under Observational Data
by: Han, Yuxuan, et al.
Published: (2025)
by: Han, Yuxuan, et al.
Published: (2025)
Similar Items
-
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025) -
Trade-off in Estimating the Number of Byzantine Clients in Federated Learning
by: Chen, Ziyi, et al.
Published: (2025) -
Zeroth-Order Methods for Stochastic Nonconvex Nonsmooth Composite Optimization
by: Chen, Ziyi, et al.
Published: (2025) -
Optimal and Order-optimal Gated Priority-based Greedy Policies for Two-layer Multi-item Order Fulfillment
by: Chen, Xi, et al.
Published: (2026) -
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
by: Shen, Kaichen, et al.
Published: (2026)