Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lei, Yuheng, Lyu, Yao, Zhan, Guojian, Zhang, Tao, Li, Jiangtao, Chen, Jianyu, Li, Shengbo Eben, Zheng, Sifa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Canonical Form of Datatic Description in Control Systems
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
An Explicit Discrete-Time Dynamic Vehicle Model with Assured Numerical Stability
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
von: Zhan, Guojian, et al.
Veröffentlicht: (2024)
Distributional Soft Actor-Critic with Three Refinements
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
von: Zhang, Tianqi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2025)
On the Stability of Datatic Control Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
On the Equilibrium between Feasible Zone and Uncertain Model in Safe Exploration
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
von: Yang, Yujie, et al.
Veröffentlicht: (2026)
Controllability Test for Nonlinear Datatic Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
Algorithm Design and Comparative Test of Natural Gradient Gaussian Approximation Filter
von: Cao, Wenhan, et al.
Veröffentlicht: (2025)
von: Cao, Wenhan, et al.
Veröffentlicht: (2025)
Natural Gradient Gaussian Approximation Filter with Positive Definiteness Guarantee
von: Zhang, Tianyi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2026)
The Feasibility Theory of Constrained Reinforcement Learning: A Tutorial Study
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
ZIVR: An Incremental Variance Reduction Technique For Zeroth-Order Composite Problems
von: Zhang, Silan, et al.
Veröffentlicht: (2026)
von: Zhang, Silan, et al.
Veröffentlicht: (2026)
Variance-Reduced Gradient Estimator for Nonconvex Zeroth-Order Distributed Optimization
von: Mu, Huaiyi, et al.
Veröffentlicht: (2024)
von: Mu, Huaiyi, et al.
Veröffentlicht: (2024)
Model-Agnostic Meta-Policy Optimization via Zeroth-Order Estimation: A Linear Quadratic Regulator Perspective
von: Pan, Yunian, et al.
Veröffentlicht: (2025)
von: Pan, Yunian, et al.
Veröffentlicht: (2025)
Scalable Synthesis of Formally Verified Neural Value Function for Hamilton-Jacobi Reachability Analysis
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
von: Yang, Yujie, et al.
Veröffentlicht: (2024)
Feasible Policy Iteration for Safe Reinforcement Learning
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
von: Yang, Yujie, et al.
Veröffentlicht: (2023)
Discretionary Lane-Change Decision and Control via Parameterized Soft Actor-Critic for Hybrid Action Space
von: Lin, Yuan, et al.
Veröffentlicht: (2024)
von: Lin, Yuan, et al.
Veröffentlicht: (2024)
Nonlinear Bayesian Filtering with Natural Gradient Gaussian Approximation
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
Zeroth-Order Constrained Optimization from a Control Perspective via Feedback Linearization
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
von: Zhang, Runyu, et al.
Veröffentlicht: (2025)
Traffic Signal Cycle Control with Centralized Critic and Decentralized Actors under Varying Intervention Frequencies
von: Wang, Maonan, et al.
Veröffentlicht: (2024)
von: Wang, Maonan, et al.
Veröffentlicht: (2024)
Zeroth-Order Optimization at the Edge of Stability
von: Song, Minhak, et al.
Veröffentlicht: (2026)
von: Song, Minhak, et al.
Veröffentlicht: (2026)
Safe Zeroth-Order Optimization Using Quadratic Local Approximations
von: Guo, Baiwei, et al.
Veröffentlicht: (2023)
von: Guo, Baiwei, et al.
Veröffentlicht: (2023)
Continuous-Time Zeroth-Order Dynamics with Projection Maps: Model-Free Feedback Optimization with Safety Guarantees
von: Chen, Xin, et al.
Veröffentlicht: (2023)
von: Chen, Xin, et al.
Veröffentlicht: (2023)
On the Optimization Landscape of Observer-based Dynamic Linear Quadratic Control
von: Duan, Jingliang, et al.
Veröffentlicht: (2026)
von: Duan, Jingliang, et al.
Veröffentlicht: (2026)
Distributional Soft Actor-Critic with Harmonic Gradient for Safe and Efficient Autonomous Driving in Multi-lane Scenarios
von: Zhang, Feihong, et al.
Veröffentlicht: (2025)
von: Zhang, Feihong, et al.
Veröffentlicht: (2025)
Zeroth-Order Feedback Optimization in Multi-Agent Systems: Tackling Coupled Constraints
von: Duan, Yingpeng, et al.
Veröffentlicht: (2024)
von: Duan, Yingpeng, et al.
Veröffentlicht: (2024)
Query-Efficient Zeroth-Order Algorithms for Nonconvex Constrained Optimization
von: Jin, Ruiyang, et al.
Veröffentlicht: (2025)
von: Jin, Ruiyang, et al.
Veröffentlicht: (2025)
Convolutional Bayesian Filtering
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
von: Cao, Wenhan, et al.
Veröffentlicht: (2024)
Curriculum-Based Soft Actor-Critic for Multi-Section R2R Tension Control
von: Li, Shihao, et al.
Veröffentlicht: (2026)
von: Li, Shihao, et al.
Veröffentlicht: (2026)
Jump-Start Reinforcement Learning with Self-Evolving Priors for Extreme Monopedal Locomotion
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
von: Zheng, Ziang, et al.
Veröffentlicht: (2025)
A Zeroth-Order Proximal Algorithm for Consensus Optimization
von: Wang, Chengan, et al.
Veröffentlicht: (2024)
von: Wang, Chengan, et al.
Veröffentlicht: (2024)
Distributed Stochastic Zeroth-Order Optimization with Compressed Communication
von: Hua, Youqing, et al.
Veröffentlicht: (2025)
von: Hua, Youqing, et al.
Veröffentlicht: (2025)
Online Distributed Zeroth-Order Optimization With Non-Zero-Mean Adverse Noises
von: Qin, Yanfu, et al.
Veröffentlicht: (2025)
von: Qin, Yanfu, et al.
Veröffentlicht: (2025)
Nonlinear-Gain Distributed Zeroth-Order Optimization for Networked Black-Box Control
von: Zhang, Shengjun, et al.
Veröffentlicht: (2026)
von: Zhang, Shengjun, et al.
Veröffentlicht: (2026)
Design and Experimental Test of Datatic Approximate Optimal Filter in Nonlinear Dynamic Systems
von: He, Weixian, et al.
Veröffentlicht: (2025)
von: He, Weixian, et al.
Veröffentlicht: (2025)
Optimal Parameter Adaptation for Safety-Critical Control via Safe Barrier Bayesian Optimization
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Natural Gradient Bayesian Filtering: Geometry-Aware Filter for Dynamical Systems
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
Performance Guarantees for Data-Driven Sequential Decision-Making
von: Li, Bowen, et al.
Veröffentlicht: (2026)
von: Li, Bowen, et al.
Veröffentlicht: (2026)
Can an Actor-Critic Optimization Framework Improve Analog Design Optimization?
von: Dutta, Sounak, et al.
Veröffentlicht: (2026)
von: Dutta, Sounak, et al.
Veröffentlicht: (2026)
Towards Enabling Learning for Time-Varying finite horizon Sequential Decision-Making Problems*
von: Tiwari, Dhananjay, et al.
Veröffentlicht: (2025)
von: Tiwari, Dhananjay, et al.
Veröffentlicht: (2025)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Canonical Form of Datatic Description in Control Systems
von: Zhan, Guojian, et al.
Veröffentlicht: (2024) -
An Explicit Discrete-Time Dynamic Vehicle Model with Assured Numerical Stability
von: Zhan, Guojian, et al.
Veröffentlicht: (2024) -
Distributional Soft Actor-Critic with Three Refinements
von: Duan, Jingliang, et al.
Veröffentlicht: (2023) -
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
von: Zhang, Tianqi, et al.
Veröffentlicht: (2025) -
On the Stability of Datatic Control Systems
von: Yang, Yujie, et al.
Veröffentlicht: (2024)