SPEQ: Offline Stabilization Phases for Efficient Q-Learning in High Update-To-Data Ratio Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Romeo, Carlo, Macaluso, Girolamo, Sestini, Alessandro, Bagdanov, Andrew D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Benchmark Environment for Offline Reinforcement Learning in Racing Games
von: Macaluso, Girolamo, et al.
Veröffentlicht: (2024)
von: Macaluso, Girolamo, et al.
Veröffentlicht: (2024)
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning with Imputed Rewards
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
von: Romeo, Carlo, et al.
Veröffentlicht: (2024)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)
NTRL: Encounter Generation via Reinforcement Learning for Dynamic Difficulty Adjustment in Dungeons and Dragons
von: Romeo, Carlo, et al.
Veröffentlicht: (2025)
von: Romeo, Carlo, et al.
Veröffentlicht: (2025)
Pretraining a Shared Q-Network for Data-Efficient Offline Reinforcement Learning
von: Park, Jongchan, et al.
Veröffentlicht: (2025)
von: Park, Jongchan, et al.
Veröffentlicht: (2025)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
Mildly Conservative Q-Learning for Offline Reinforcement Learning
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
von: Lyu, Jiafei, et al.
Veröffentlicht: (2022)
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
von: Mao, Yixiu, et al.
Veröffentlicht: (2025)
von: Mao, Yixiu, et al.
Veröffentlicht: (2025)
Exclusively Penalized Q-learning for Offline Reinforcement Learning
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
von: Yeom, Junghyuk, et al.
Veröffentlicht: (2024)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
von: Park, Kwanyoung, et al.
Veröffentlicht: (2024)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2024)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
von: Li, Mingxuan, et al.
Veröffentlicht: (2026)
von: Li, Mingxuan, et al.
Veröffentlicht: (2026)
When is Offline Policy Selection Sample Efficient for Reinforcement Learning?
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
PIQL: Projective Implicit Q-Learning with Support Constraint for Offline Reinforcement Learning
von: Han, Xinchen, et al.
Veröffentlicht: (2025)
von: Han, Xinchen, et al.
Veröffentlicht: (2025)
Simple Ingredients for Offline Reinforcement Learning
von: Cetin, Edoardo, et al.
Veröffentlicht: (2024)
von: Cetin, Edoardo, et al.
Veröffentlicht: (2024)
On the Complexity of Offline Reinforcement Learning with $Q^\star$-Approximation and Partial Coverage
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
FORLER: Federated Offline Reinforcement Learning with Q-Ensemble and Actor Rectification
von: Qiao, Nan, et al.
Veröffentlicht: (2026)
von: Qiao, Nan, et al.
Veröffentlicht: (2026)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
von: Zhao, Kai, et al.
Veröffentlicht: (2023)
von: Zhao, Kai, et al.
Veröffentlicht: (2023)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
von: Zhang, Ziqi, et al.
Veröffentlicht: (2023)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2023)
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
von: Tayal, Mumuksh, et al.
Veröffentlicht: (2026)
Sample Efficient Active Algorithms for Offline Reinforcement Learning
von: Roy, Soumyadeep, et al.
Veröffentlicht: (2026)
von: Roy, Soumyadeep, et al.
Veröffentlicht: (2026)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
von: Alles, Marvin, et al.
Veröffentlicht: (2025)
von: Alles, Marvin, et al.
Veröffentlicht: (2025)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
Adaptive $Q$-Network: On-the-fly Target Selection for Deep Reinforcement Learning
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
von: Vincent, Théo, et al.
Veröffentlicht: (2024)
Pre-training with Synthetic Data Helps Offline Reinforcement Learning
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
Pessimistic Causal Reinforcement Learning with Mediators for Confounded Offline Data
von: Wang, Danyang, et al.
Veröffentlicht: (2024)
von: Wang, Danyang, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning: Role of State Aggregation and Trajectory Data
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
von: Jia, Zeyu, et al.
Veröffentlicht: (2024)
Offline Policy Evaluation for Reinforcement Learning with Adaptively Collected Data
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
von: Madhow, Sunil, et al.
Veröffentlicht: (2023)
FOVA: Offline Federated Reinforcement Learning with Mixed-Quality Data
von: Qiao, Nan, et al.
Veröffentlicht: (2025)
von: Qiao, Nan, et al.
Veröffentlicht: (2025)
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
Preference Elicitation for Offline Reinforcement Learning
von: Pace, Alizée, et al.
Veröffentlicht: (2024)
von: Pace, Alizée, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning with Imbalanced Datasets
von: Jiang, Li, et al.
Veröffentlicht: (2023)
von: Jiang, Li, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
A Benchmark Environment for Offline Reinforcement Learning in Racing Games
von: Macaluso, Girolamo, et al.
Veröffentlicht: (2024) -
SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data
von: Romeo, Carlo, et al.
Veröffentlicht: (2026) -
Offline Reinforcement Learning with Imputed Rewards
von: Romeo, Carlo, et al.
Veröffentlicht: (2024) -
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025) -
ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders
von: Romeo, Carlo, et al.
Veröffentlicht: (2026)