Saved in:
| Main Authors: | Bian, Yuexin, Feng, Jie, Wang, Tao, Li, Yijiang, Gao, Sicun, Shi, Yuanyuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.23075 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mollification Effects of Policy Gradient Methods
by: Wang, Tao, et al.
Published: (2024)
by: Wang, Tao, et al.
Published: (2024)
Improving Value Estimation Critically Enhances Vanilla Policy Gradient
by: Wang, Tao, et al.
Published: (2025)
by: Wang, Tao, et al.
Published: (2025)
Hamilton-Jacobi Reachability in Reinforcement Learning: A Survey
by: Ganai, Milan, et al.
Published: (2024)
by: Ganai, Milan, et al.
Published: (2024)
Extremum-Seeking Action Selection for Accelerating Policy Optimization
by: Chang, Ya-Chien, et al.
Published: (2024)
by: Chang, Ya-Chien, et al.
Published: (2024)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
by: Liu, Tenglong, et al.
Published: (2024)
by: Liu, Tenglong, et al.
Published: (2024)
Controllable Motion Generation via Diffusion Modal Coupling
by: Wang, Luobin, et al.
Published: (2025)
by: Wang, Luobin, et al.
Published: (2025)
Certificated Actor-Critic: Hierarchical Reinforcement Learning with Control Barrier Functions for Safe Navigation
by: Xie, Junjun, et al.
Published: (2025)
by: Xie, Junjun, et al.
Published: (2025)
Pretraining in Actor-Critic Reinforcement Learning for Robot Locomotion
by: Fan, Jiale, et al.
Published: (2025)
by: Fan, Jiale, et al.
Published: (2025)
ReActor: Reinforcement Learning for Physics-Aware Motion Retargeting
by: Müller, David, et al.
Published: (2026)
by: Müller, David, et al.
Published: (2026)
Offline Reinforcement Learning with Discrete Diffusion Skills
by: Qiao, RuiXi, et al.
Published: (2025)
by: Qiao, RuiXi, et al.
Published: (2025)
Activation-Descent Regularization for Input Optimization of ReLU Networks
by: Yu, Hongzhan, et al.
Published: (2024)
by: Yu, Hongzhan, et al.
Published: (2024)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
by: Zhuang, Yuan, et al.
Published: (2026)
by: Zhuang, Yuan, et al.
Published: (2026)
Offline Actor-Critic Reinforcement Learning Scales to Large Models
by: Springenberg, Jost Tobias, et al.
Published: (2024)
by: Springenberg, Jost Tobias, et al.
Published: (2024)
Efficient Motion Planning for Manipulators with Control Barrier Function-Induced Neural Controller
by: Yu, Mingxin, et al.
Published: (2024)
by: Yu, Mingxin, et al.
Published: (2024)
Fractal Landscapes in Policy Optimization
by: Wang, Tao, et al.
Published: (2023)
by: Wang, Tao, et al.
Published: (2023)
Safe Human Robot Navigation in Warehouse Scenario
by: Farrell, Seth, et al.
Published: (2025)
by: Farrell, Seth, et al.
Published: (2025)
DiffOP: Reinforcement Learning of Optimization-Based Control Policies via Implicit Policy Gradients
by: Bian, Yuexin, et al.
Published: (2024)
by: Bian, Yuexin, et al.
Published: (2024)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
by: Singh, Nikhil Kumar, et al.
Published: (2024)
by: Singh, Nikhil Kumar, et al.
Published: (2024)
Q-value Regularized Decision ConvFormer for Offline Reinforcement Learning
by: Yan, Teng, et al.
Published: (2024)
by: Yan, Teng, et al.
Published: (2024)
Optimal Actuator Attacks on Autonomous Vehicles Using Reinforcement Learning
by: Wang, Pengyu, et al.
Published: (2025)
by: Wang, Pengyu, et al.
Published: (2025)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
by: Li, Guopeng, et al.
Published: (2026)
by: Li, Guopeng, et al.
Published: (2026)
Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)
by: Mahran, Youssef, et al.
Published: (2025)
by: Mahran, Youssef, et al.
Published: (2025)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
Equivariant Ensembles and Regularization for Reinforcement Learning in Map-based Path Planning
by: Theile, Mirco, et al.
Published: (2024)
by: Theile, Mirco, et al.
Published: (2024)
TOP-ERL: Transformer-based Off-Policy Episodic Reinforcement Learning
by: Li, Ge, et al.
Published: (2024)
by: Li, Ge, et al.
Published: (2024)
HOPE: A Reinforcement Learning-based Hybrid Policy Path Planner for Diverse Parking Scenarios
by: Jiang, Mingyang, et al.
Published: (2024)
by: Jiang, Mingyang, et al.
Published: (2024)
Hysteresis-Aware Neural Network Modeling and Whole-Body Reinforcement Learning Control of Soft Robots
by: Chen, Zongyuan, et al.
Published: (2025)
by: Chen, Zongyuan, et al.
Published: (2025)
Enhancing Sample Efficiency and Exploration in Reinforcement Learning through the Integration of Diffusion Models and Proximal Policy Optimization
by: Gao, Tianci, et al.
Published: (2024)
by: Gao, Tianci, et al.
Published: (2024)
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
by: Lei, Kun, et al.
Published: (2023)
by: Lei, Kun, et al.
Published: (2023)
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
by: Zhang, Tonghe, et al.
Published: (2025)
by: Zhang, Tonghe, et al.
Published: (2025)
3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
by: Ke, Tsung-Wei, et al.
Published: (2024)
by: Ke, Tsung-Wei, et al.
Published: (2024)
Operator learning for energy-efficient building ventilation control with computational fluid dynamics simulation of a real-world classroom
by: Bian, Yuexin, et al.
Published: (2025)
by: Bian, Yuexin, et al.
Published: (2025)
ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
ERPPO: Entropy Regularization-based Proximal Policy Optimization
by: Lee, Changha, et al.
Published: (2026)
by: Lee, Changha, et al.
Published: (2026)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
by: Wagenmaker, Andrew, et al.
Published: (2025)
by: Wagenmaker, Andrew, et al.
Published: (2025)
Oracle-Guided Masked Contrastive Reinforcement Learning for Visuomotor Policies
by: Zhang, Yuhang, et al.
Published: (2025)
by: Zhang, Yuhang, et al.
Published: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
by: Nakanishi, Kosuke, et al.
Published: (2025)
by: Nakanishi, Kosuke, et al.
Published: (2025)
Exterior Penalty Policy Optimization with Penalty Metric Network under Constraints
by: Gao, Shiqing, et al.
Published: (2024)
by: Gao, Shiqing, et al.
Published: (2024)
Similar Items
-
Mollification Effects of Policy Gradient Methods
by: Wang, Tao, et al.
Published: (2024) -
Improving Value Estimation Critically Enhances Vanilla Policy Gradient
by: Wang, Tao, et al.
Published: (2025) -
Hamilton-Jacobi Reachability in Reinforcement Learning: A Survey
by: Ganai, Milan, et al.
Published: (2024) -
Extremum-Seeking Action Selection for Accelerating Policy Optimization
by: Chang, Ya-Chien, et al.
Published: (2024) -
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
by: Liu, Tenglong, et al.
Published: (2024)