Model predictive control-based value estimation for efficient reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Qizhen, Liu, Kexin, Chen, Lei |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Agent Reinforcement Learning-Based UAV Pathfinding for Obstacle Avoidance in Stochastic Environment
by: Wu, Qizhen, et al.
Published: (2023)
by: Wu, Qizhen, et al.
Published: (2023)
Hierarchical Reinforcement Learning for Swarm Confrontation with High Uncertainty
by: Wu, Qizhen, et al.
Published: (2024)
by: Wu, Qizhen, et al.
Published: (2024)
Computationally efficient Gauss-Newton reinforcement learning for model predictive control
by: Brandner, Dean, et al.
Published: (2025)
by: Brandner, Dean, et al.
Published: (2025)
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
by: Zhang, Jing, et al.
Published: (2024)
by: Zhang, Jing, et al.
Published: (2024)
Dynamic feature selection in medical predictive monitoring by reinforcement learning
by: Chen, Yutong, et al.
Published: (2024)
by: Chen, Yutong, et al.
Published: (2024)
Enhancing sample efficiency in reinforcement-learning-based flow control: replacing the critic with an adaptive reduced-order model
by: Yao, Zesheng, et al.
Published: (2026)
by: Yao, Zesheng, et al.
Published: (2026)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
An efficient deep reinforcement learning environment for flexible job-shop scheduling
by: Wu, Xinquan, et al.
Published: (2025)
by: Wu, Xinquan, et al.
Published: (2025)
Transfer learning strategies for accelerating reinforcement-learning-based flow control
by: Salehi, Saeed
Published: (2025)
by: Salehi, Saeed
Published: (2025)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
by: Hutson, Miles, et al.
Published: (2024)
by: Hutson, Miles, et al.
Published: (2024)
Nonlinear sparse variational Bayesian learning based model predictive control with application to PEMFC temperature control
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
Experimental evaluation of offline reinforcement learning for HVAC control in buildings
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
General Formulation and PCL-Analysis for Restless Bandits with Limited Observability
by: Liu, Keqin, et al.
Published: (2023)
by: Liu, Keqin, et al.
Published: (2023)
Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs
by: Roux, Nicolas Le, et al.
Published: (2025)
by: Roux, Nicolas Le, et al.
Published: (2025)
KFS: KAN based adaptive Frequency Selection learning architecture for long term time series forecasting
by: Wu, Changning, et al.
Published: (2025)
by: Wu, Changning, et al.
Published: (2025)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Catastrophic-risk-aware reinforcement learning with extreme-value-theory-based policy gradients
by: Davar, Parisa, et al.
Published: (2024)
by: Davar, Parisa, et al.
Published: (2024)
Not all tokens are needed(NAT): token efficient reinforcement learning
by: Sang, Hejian, et al.
Published: (2026)
by: Sang, Hejian, et al.
Published: (2026)
Attention on flow control: transformer-based reinforcement learning for lift regulation in highly disturbed flows
by: Liu, Zhecheng, et al.
Published: (2025)
by: Liu, Zhecheng, et al.
Published: (2025)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Explainable deep reinforcement learning reveals energy-efficient control strategies for turbulent drag reduction
by: Tonti, Federica, et al.
Published: (2026)
by: Tonti, Federica, et al.
Published: (2026)
Agile and versatile bipedal robot tracking control through reinforcement learning
by: Li, Jiayi, et al.
Published: (2024)
by: Li, Jiayi, et al.
Published: (2024)
Koopman-based surrogate modeling for reinforcement-learning-control of Rayleigh-Benard convection
by: Plotzki, Tim, et al.
Published: (2026)
by: Plotzki, Tim, et al.
Published: (2026)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Designing an efficient and equitable humanitarian supply chain dynamically via reinforcement learning
by: Jin, Weijia
Published: (2025)
by: Jin, Weijia
Published: (2025)
Revisiting the Efficacy of Signal Decomposition in AI-based Time Series Prediction
by: Jiang, Kexin, et al.
Published: (2024)
by: Jiang, Kexin, et al.
Published: (2024)
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control
by: Sheng, Zihao, et al.
Published: (2024)
by: Sheng, Zihao, et al.
Published: (2024)
SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
by: Crowley, Peter, et al.
Published: (2025)
by: Crowley, Peter, et al.
Published: (2025)
Focal plane wavefront control with model-based reinforcement learning
by: Nousiainen, Jalo, et al.
Published: (2026)
by: Nousiainen, Jalo, et al.
Published: (2026)
A modular framework for stabilizing deep reinforcement learning control
by: Lawrence, Nathan P., et al.
Published: (2023)
by: Lawrence, Nathan P., et al.
Published: (2023)
Autonomous vehicle decision and control through reinforcement learning with traffic flow randomization
by: Lin, Yuan, et al.
Published: (2024)
by: Lin, Yuan, et al.
Published: (2024)
A first realization of reinforcement learning-based closed-loop EEG-TMS
by: Humaidan, Dania, et al.
Published: (2026)
by: Humaidan, Dania, et al.
Published: (2026)
Combining T-learning and DR-learning: a framework for oracle-efficient estimation of causal contrasts
by: van der Laan, Lars, et al.
Published: (2024)
by: van der Laan, Lars, et al.
Published: (2024)
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
A machine learning model for skillful climate system prediction
by: Zhou, Chenguang, et al.
Published: (2025)
by: Zhou, Chenguang, et al.
Published: (2025)
The surprising efficiency of temporal difference learning for rare event prediction
by: Cheng, Xiaoou, et al.
Published: (2024)
by: Cheng, Xiaoou, et al.
Published: (2024)
Discovering highly efficient low-weight quantum error-correcting codes with reinforcement learning
by: He, Austin Yubo, et al.
Published: (2025)
by: He, Austin Yubo, et al.
Published: (2025)
How to craft a deep reinforcement learning policy for wind farm flow control
by: Kadoche, Elie, et al.
Published: (2025)
by: Kadoche, Elie, et al.
Published: (2025)
BAPR: Bayesian amnesic piecewise-robust reinforcement learning for non-stationary continuous control
by: Zhang, Yifan, et al.
Published: (2026)
by: Zhang, Yifan, et al.
Published: (2026)
Similar Items
-
Multi-Agent Reinforcement Learning-Based UAV Pathfinding for Obstacle Avoidance in Stochastic Environment
by: Wu, Qizhen, et al.
Published: (2023) -
Hierarchical Reinforcement Learning for Swarm Confrontation with High Uncertainty
by: Wu, Qizhen, et al.
Published: (2024) -
Computationally efficient Gauss-Newton reinforcement learning for model predictive control
by: Brandner, Dean, et al.
Published: (2025) -
Q-Distribution guided Q-learning for offline reinforcement learning: Uncertainty penalized Q-value via consistency model
by: Zhang, Jing, et al.
Published: (2024) -
Dynamic feature selection in medical predictive monitoring by reinforcement learning
by: Chen, Yutong, et al.
Published: (2024)