Efficient Reinforcement Learning via Decoupling Exploration and Utilization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Jingpu, Wang, Helin, Zhao, Qirui, Shi, Zhecheng, Song, Zirui, Fang, Miao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Asynchronous and Segmented Bidirectional Encoding for NMT
von: Yang, Jingpu, et al.
Veröffentlicht: (2024)
von: Yang, Jingpu, et al.
Veröffentlicht: (2024)
Guardian: Decoupling Exploration from Safety in Reinforcement Learning
von: Cai, Kaitong, et al.
Veröffentlicht: (2025)
von: Cai, Kaitong, et al.
Veröffentlicht: (2025)
Goal-Guided Efficient Exploration via Large Language Model in Reinforcement Learning
von: Qi, Yajie, et al.
Veröffentlicht: (2025)
von: Qi, Yajie, et al.
Veröffentlicht: (2025)
Exploration in Knowledge Transfer Utilizing Reinforcement Learning
von: Jedlička, Adam, et al.
Veröffentlicht: (2024)
von: Jedlička, Adam, et al.
Veröffentlicht: (2024)
TranDRL: A Transformer-Driven Deep Reinforcement Learning Enabled Prescriptive Maintenance Framework
von: Zhao, Yang, et al.
Veröffentlicht: (2023)
von: Zhao, Yang, et al.
Veröffentlicht: (2023)
Hyper: Hyperparameter Robust Efficient Exploration in Reinforcement Learning
von: Wang, Yiran, et al.
Veröffentlicht: (2024)
von: Wang, Yiran, et al.
Veröffentlicht: (2024)
Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2025)
von: Shi, Taiwei, et al.
Veröffentlicht: (2025)
Preference-Guided Reinforcement Learning for Efficient Exploration
von: Wang, Guojian, et al.
Veröffentlicht: (2024)
von: Wang, Guojian, et al.
Veröffentlicht: (2024)
More Efficient Randomized Exploration for Reinforcement Learning via Approximate Sampling
von: Ishfaq, Haque, et al.
Veröffentlicht: (2024)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2024)
Making Bias Non-Predictive: Training Robust LLM Reasoning via Reinforcement Learning
von: Wang, Qian, et al.
Veröffentlicht: (2026)
von: Wang, Qian, et al.
Veröffentlicht: (2026)
Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
von: Ishfaq, Haque, et al.
Veröffentlicht: (2023)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2023)
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
von: Song, Meichen, et al.
Veröffentlicht: (2026)
von: Song, Meichen, et al.
Veröffentlicht: (2026)
Model-Based Reinforcement Learning for Control of Strongly-Disturbed Unsteady Aerodynamic Flows
von: Liu, Zhecheng, et al.
Veröffentlicht: (2024)
von: Liu, Zhecheng, et al.
Veröffentlicht: (2024)
On Efficient Bayesian Exploration in Model-Based Reinforcement Learning
von: Caron, Alberto, et al.
Veröffentlicht: (2025)
von: Caron, Alberto, et al.
Veröffentlicht: (2025)
BroRL: Scaling Reinforcement Learning via Broadened Exploration
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
Efficient On-Policy Reinforcement Learning via Exploration of Sparse Parameter Space
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Closed-Form Concept Erasure via Double Projections
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
von: Zhang, Chi, et al.
Veröffentlicht: (2026)
Hybrid Belief Reinforcement Learning for Efficient Coordinated Spatial Exploration
von: Rizvi, Danish, et al.
Veröffentlicht: (2026)
von: Rizvi, Danish, et al.
Veröffentlicht: (2026)
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
von: Sun, Ke, et al.
Veröffentlicht: (2021)
von: Sun, Ke, et al.
Veröffentlicht: (2021)
Olympus: A Jumping Quadruped for Planetary Exploration Utilizing Reinforcement Learning for In-Flight Attitude Control
von: Olsen, Jørgen Anker, et al.
Veröffentlicht: (2025)
von: Olsen, Jørgen Anker, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning for Large Language Models with Intrinsic Exploration
von: Sun, Yan, et al.
Veröffentlicht: (2025)
von: Sun, Yan, et al.
Veröffentlicht: (2025)
Solving Robotics Tasks with Prior Demonstration via Exploration-Efficient Deep Reinforcement Learning
von: Shen, Chengyandan, et al.
Veröffentlicht: (2025)
von: Shen, Chengyandan, et al.
Veröffentlicht: (2025)
Breaking the Exploration Bottleneck: Rubric-Scaffolded Reinforcement Learning for General LLM Reasoning
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
von: Zhou, Yang, et al.
Veröffentlicht: (2025)
Active Learning Methods for Efficient Data Utilization and Model Performance Enhancement
von: Tseng, Chiung-Yi, et al.
Veröffentlicht: (2025)
von: Tseng, Chiung-Yi, et al.
Veröffentlicht: (2025)
Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning
von: Han, Shuai, et al.
Veröffentlicht: (2025)
von: Han, Shuai, et al.
Veröffentlicht: (2025)
A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control
von: Kang, Zilin, et al.
Veröffentlicht: (2025)
von: Kang, Zilin, et al.
Veröffentlicht: (2025)
Efficient Learned Data Compression via Dual-Stream Feature Decoupling
von: Ma, Huidong, et al.
Veröffentlicht: (2026)
von: Ma, Huidong, et al.
Veröffentlicht: (2026)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
von: Yue, Bo, et al.
Veröffentlicht: (2024)
von: Yue, Bo, et al.
Veröffentlicht: (2024)
Task-Specific Directions: Definition, Exploration, and Utilization in Parameter Efficient Fine-Tuning
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning
von: Deng, Yue, et al.
Veröffentlicht: (2026)
von: Deng, Yue, et al.
Veröffentlicht: (2026)
Learning-Driven Exploration for Reinforcement Learning
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
OPRIDE: Offline Preference-based Reinforcement Learning via In-Dataset Exploration
von: Yang, Yiqin, et al.
Veröffentlicht: (2026)
von: Yang, Yiqin, et al.
Veröffentlicht: (2026)
Gradient-Based Non-Linear Inverse Learning
von: Abhishake, et al.
Veröffentlicht: (2024)
von: Abhishake, et al.
Veröffentlicht: (2024)
Experiential Reinforcement Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
Sample Efficient Myopic Exploration Through Multitask Reinforcement Learning with Diverse Tasks
von: Xu, Ziping, et al.
Veröffentlicht: (2024)
von: Xu, Ziping, et al.
Veröffentlicht: (2024)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Efficient Episodic Memory Utilization of Cooperative Multi-Agent Reinforcement Learning
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
von: Na, Hyungho, et al.
Veröffentlicht: (2024)
LESSON: Learning to Integrate Exploration Strategies for Reinforcement Learning via an Option Framework
von: Kim, Woojun, et al.
Veröffentlicht: (2023)
von: Kim, Woojun, et al.
Veröffentlicht: (2023)
MIR: Efficient Exploration in Episodic Multi-Agent Reinforcement Learning via Mutual Intrinsic Reward
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
Privacy-Preserving Reinforcement Learning from Human Feedback via Decoupled Reward Modeling
von: Cho, Young Hyun, et al.
Veröffentlicht: (2026)
von: Cho, Young Hyun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Asynchronous and Segmented Bidirectional Encoding for NMT
von: Yang, Jingpu, et al.
Veröffentlicht: (2024) -
Guardian: Decoupling Exploration from Safety in Reinforcement Learning
von: Cai, Kaitong, et al.
Veröffentlicht: (2025) -
Goal-Guided Efficient Exploration via Large Language Model in Reinforcement Learning
von: Qi, Yajie, et al.
Veröffentlicht: (2025) -
Exploration in Knowledge Transfer Utilizing Reinforcement Learning
von: Jedlička, Adam, et al.
Veröffentlicht: (2024) -
TranDRL: A Transformer-Driven Deep Reinforcement Learning Enabled Prescriptive Maintenance Framework
von: Zhao, Yang, et al.
Veröffentlicht: (2023)