Boosting Reinforcement Learning with Strongly Delayed Feedback Through Auxiliary Short Delays
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Qingyuan, Zhan, Simon Sinong, Wang, Yixuan, Wang, Yuhui, Lin, Chung-Wei, Lv, Chen, Zhu, Qi, Schmidhuber, Jürgen, Huang, Chao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Directly Forecasting Belief for Reinforcement Learning with Delays
por: Wu, Qingyuan, et al.
Publicado: (2025)
por: Wu, Qingyuan, et al.
Publicado: (2025)
Inverse Delayed Reinforcement Learning
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
Variational Delayed Policy Optimization
por: Wu, Qingyuan, et al.
Publicado: (2024)
por: Wu, Qingyuan, et al.
Publicado: (2024)
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
por: Zhan, Simon Sinong, et al.
Publicado: (2025)
por: Zhan, Simon Sinong, et al.
Publicado: (2025)
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
por: Zhan, Simon Sinong, et al.
Publicado: (2024)
Empowering Autonomous Driving with Large Language Models: A Safety Perspective
por: Wang, Yixuan, et al.
Publicado: (2023)
por: Wang, Yixuan, et al.
Publicado: (2023)
Noncooperative Game in Multi-controller System under Delayed and Asymmetric Information
por: Li, Xin, et al.
Publicado: (2026)
por: Li, Xin, et al.
Publicado: (2026)
Output Feedback to Improve the Delay Margin of Linear Delay Systems
por: Renhong Hu, et al.
Publicado: (2025)
por: Renhong Hu, et al.
Publicado: (2025)
Modal Decomposition of Feedback Delay Networks
por: Schlecht, Sebastian J., et al.
Publicado: (2019)
por: Schlecht, Sebastian J., et al.
Publicado: (2019)
Robustness of Reaction-Diffusion PDEs Predictor-Feedback to Stochastic Delay Perturbations
por: Guan, Dandan, et al.
Publicado: (2023)
por: Guan, Dandan, et al.
Publicado: (2023)
Case Study: Runtime Safety Verification of Neural Network Controlled System
por: Yang, Frank, et al.
Publicado: (2024)
por: Yang, Frank, et al.
Publicado: (2024)
Globally Exponential Synchronization of Delayed Complex Dynamic Networks With Average Impulsive Delay‐Gain
por: Kangping Gao, et al.
Publicado: (2024)
por: Kangping Gao, et al.
Publicado: (2024)
Safe Delay-Adaptive Control of Strict-Feedback Nonlinear Systems with Application in Vehicle Platooning
por: Zhao, Zhenxu, et al.
Publicado: (2024)
por: Zhao, Zhenxu, et al.
Publicado: (2024)
Neural Operator Feedback for a First-Order PIDE with Spatially-Varying State Delay
por: Qi, Jie, et al.
Publicado: (2024)
por: Qi, Jie, et al.
Publicado: (2024)
Data-Driven Safe Output Regulation of Strict-Feedback Linear Systems with Input Delay
por: Zhao, Zhenxu, et al.
Publicado: (2026)
por: Zhao, Zhenxu, et al.
Publicado: (2026)
Neural Operators for Predictor Feedback Control of Nonlinear Delay Systems
por: Bhan, Luke, et al.
Publicado: (2024)
por: Bhan, Luke, et al.
Publicado: (2024)
A Canonical Structure for Constructing Projected First-Order Algorithms With Delayed Feedback
por: Li, Mengmou, et al.
Publicado: (2026)
por: Li, Mengmou, et al.
Publicado: (2026)
Static Output Feedback Stabilization of Linear Systems with Multiple Delays
por: Braghini, Danilo, et al.
Publicado: (2026)
por: Braghini, Danilo, et al.
Publicado: (2026)
Sampled‐Data Control of Nonlinear Systems With Input Delay Using Memoryless Feedback
por: Xueliang Liu, et al.
Publicado: (2026)
por: Xueliang Liu, et al.
Publicado: (2026)
IQC-Based Output-Feedback Control of LPV Systems with Time-Varying Input Delays
por: Wu, Fen
Publicado: (2026)
por: Wu, Fen
Publicado: (2026)
State Feedback Control of State-Delayed LPV Systems using Dynamic IQCs
por: Wu, Fen
Publicado: (2026)
por: Wu, Fen
Publicado: (2026)
Reset Controller Synthesis by Reach-avoid Analysis for Delay Hybrid Systems
por: Su, Han, et al.
Publicado: (2023)
por: Su, Han, et al.
Publicado: (2023)
A Unified Framework for Rethinking Policy Divergence Measures in GRPO
por: Wu, Qingyuan, et al.
Publicado: (2026)
por: Wu, Qingyuan, et al.
Publicado: (2026)
Integrating Delay-Absorption Capability into Flight Departure Delay Prediction
por: Zhou, Jianyang
Publicado: (2025)
por: Zhou, Jianyang
Publicado: (2025)
Delay Analysis of 5G HARQ in the Presence of Decoding and Feedback Latencies
por: Moothedath, Vishnu N, et al.
Publicado: (2025)
por: Moothedath, Vishnu N, et al.
Publicado: (2025)
Switching Controller Synthesis for Hybrid Systems Against STL Formulas
por: Su, Han, et al.
Publicado: (2024)
por: Su, Han, et al.
Publicado: (2024)
Co-Design of Cryptographic Parameters and Delay-Aware Feedback Gain for Encrypted Control Systems
por: Jang, Yeongjun
Publicado: (2026)
por: Jang, Yeongjun
Publicado: (2026)
Predictor-Feedback Stabilization of Linear Switched Systems with State-Dependent Switching and Input Delay
por: Katsanikakis, Andreas, et al.
Publicado: (2026)
por: Katsanikakis, Andreas, et al.
Publicado: (2026)
State Estimator Design: Addressing General Delay Structures with Dissipative Constraints
por: Feng, Qian, et al.
Publicado: (2023)
por: Feng, Qian, et al.
Publicado: (2023)
Sliding Mode Control for Uncertain Systems with Time-Varying Delays via Predictor Feedback and Super-Twisting Observer
por: Pinto, Hardy, et al.
Publicado: (2025)
por: Pinto, Hardy, et al.
Publicado: (2025)
Neural Operator based Reinforcement Learning for Control of first-order PDEs with Spatially-Varying State Delay
por: Hu, Jiaqi, et al.
Publicado: (2025)
por: Hu, Jiaqi, et al.
Publicado: (2025)
Delay‐adaptive control of first‐order hyperbolic partial integro‐differential equations
por: Shanshan Wang, et al.
Publicado: (2024)
por: Shanshan Wang, et al.
Publicado: (2024)
Revisiting Multi-Agent Asynchronous Online Optimization with Delays: the Strongly Convex Case
por: Bao, Lingchan, et al.
Publicado: (2025)
por: Bao, Lingchan, et al.
Publicado: (2025)
Contraction Analysis of Time-Delay Systems
por: Watanabe, Rintaro, et al.
Publicado: (2026)
por: Watanabe, Rintaro, et al.
Publicado: (2026)
Optimal Delay Compensation in Networked Predictive Control
por: Beger, Severin, et al.
Publicado: (2025)
por: Beger, Severin, et al.
Publicado: (2025)
Adaptive Neural Delay Feedback Control of a Typical Robot on a Slope
por: Mingyue Ji, et al.
Publicado: (2024)
por: Mingyue Ji, et al.
Publicado: (2024)
Dynamical State Feedback Control for Linear Input Delay Systems, Part I: Dissipative Stabilization via Semidefinite Programming
por: Feng, Qian, et al.
Publicado: (2023)
por: Feng, Qian, et al.
Publicado: (2023)
Existence and Design of Functional Observers for Time-Delay Systems with Delayed Output Measurements
por: Trinh, Hieu, et al.
Publicado: (2026)
por: Trinh, Hieu, et al.
Publicado: (2026)
Adaptive Fuzzy Tracking Control for Nonlinear State Constrained Pure-Feedback Systems With Input Delay via Dynamic Surface Technique
por: Wu, Ju, et al.
Publicado: (2023)
por: Wu, Ju, et al.
Publicado: (2023)
Delayed finite-dimensional observer-based control of 2D linear parabolic PDEs
por: Wang, Pengfei, et al.
Publicado: (2024)
por: Wang, Pengfei, et al.
Publicado: (2024)
Ejemplares similares
-
Directly Forecasting Belief for Reinforcement Learning with Delays
por: Wu, Qingyuan, et al.
Publicado: (2025) -
Inverse Delayed Reinforcement Learning
por: Zhan, Simon Sinong, et al.
Publicado: (2024) -
Variational Delayed Policy Optimization
por: Wu, Qingyuan, et al.
Publicado: (2024) -
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
por: Zhan, Simon Sinong, et al.
Publicado: (2025) -
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
por: Zhan, Simon Sinong, et al.
Publicado: (2024)