CAE: Repurposing the Critic as an Explorer in Deep Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Li, Yexin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
von: Thalagala, Shiron, et al.
Veröffentlicht: (2024)
von: Thalagala, Shiron, et al.
Veröffentlicht: (2024)
DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care
von: Tan, Daniel Jason, et al.
Veröffentlicht: (2025)
von: Tan, Daniel Jason, et al.
Veröffentlicht: (2025)
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
von: Humayoo, Mahammad, et al.
Veröffentlicht: (2018)
von: Humayoo, Mahammad, et al.
Veröffentlicht: (2018)
A Contextual Combinatorial Bandit Approach to Negotiation
von: Li, Yexin, et al.
Veröffentlicht: (2024)
von: Li, Yexin, et al.
Veröffentlicht: (2024)
LLM-Explorer: A Plug-in Reinforcement Learning Policy Exploration Enhancement Driven by Large Language Models
von: Hao, Qianyue, et al.
Veröffentlicht: (2025)
von: Hao, Qianyue, et al.
Veröffentlicht: (2025)
Flow Actor-Critic for Offline Reinforcement Learning
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
Adaptive Data Exploitation in Deep Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
Skill-Critic: Refining Learned Skills for Hierarchical Reinforcement Learning
von: Hao, Ce, et al.
Veröffentlicht: (2023)
von: Hao, Ce, et al.
Veröffentlicht: (2023)
Learning to Select In-Context Demonstration Preferred by Large Language Model
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
An Invitation to Deep Reinforcement Learning
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2023)
von: Jaeger, Bernhard, et al.
Veröffentlicht: (2023)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
Deep Reinforcement Learning from Hierarchical Preference Design
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
von: Zhuang, Yuan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yuan, et al.
Veröffentlicht: (2026)
Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning
von: Ma, Guozheng, et al.
Veröffentlicht: (2025)
von: Ma, Guozheng, et al.
Veröffentlicht: (2025)
Tutorial on Using Machine Learning and Deep Learning Models for Mental Illness Detection
von: Zhang, Yeyubei, et al.
Veröffentlicht: (2025)
von: Zhang, Yeyubei, et al.
Veröffentlicht: (2025)
Causal-Paced Deep Reinforcement Learning
von: Cho, Geonwoo, et al.
Veröffentlicht: (2025)
von: Cho, Geonwoo, et al.
Veröffentlicht: (2025)
Recursive Deep Inverse Reinforcement Learning
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
Behavior-Consistent Deep Reinforcement Learning
von: Hussing, Marcel, et al.
Veröffentlicht: (2026)
von: Hussing, Marcel, et al.
Veröffentlicht: (2026)
Satisficing Exploration for Deep Reinforcement Learning
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
von: Arumugam, Dilip, et al.
Veröffentlicht: (2024)
Rethinking Plasticity in Deep Reinforcement Learning
von: He, Zhiqiang
Veröffentlicht: (2026)
von: He, Zhiqiang
Veröffentlicht: (2026)
Understanding and Diagnosing Deep Reinforcement Learning
von: Korkmaz, Ezgi
Veröffentlicht: (2024)
von: Korkmaz, Ezgi
Veröffentlicht: (2024)
Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning
von: Garcin, Samuel, et al.
Veröffentlicht: (2025)
von: Garcin, Samuel, et al.
Veröffentlicht: (2025)
Reinforcement Learning in Dynamic Treatment Regimes Needs Critical Reexamination
von: Luo, Zhiyao, et al.
Veröffentlicht: (2024)
von: Luo, Zhiyao, et al.
Veröffentlicht: (2024)
Towards Interpretable Deep Reinforcement Learning Models via Inverse Reinforcement Learning
von: Xie, Sean, et al.
Veröffentlicht: (2022)
von: Xie, Sean, et al.
Veröffentlicht: (2022)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
Learning Markov State Abstractions for Deep Reinforcement Learning
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing
von: Li, Yuanpeng, et al.
Veröffentlicht: (2026)
von: Li, Yuanpeng, et al.
Veröffentlicht: (2026)
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2024)
von: Zhao, Kaiyan, et al.
Veröffentlicht: (2024)
Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning
von: Poole, Benjamin, et al.
Veröffentlicht: (2026)
von: Poole, Benjamin, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning with Gradient Eligibility Traces
von: Elelimy, Esraa, et al.
Veröffentlicht: (2025)
von: Elelimy, Esraa, et al.
Veröffentlicht: (2025)
A Survey on Explainable Deep Reinforcement Learning
von: Cheng, Zelei, et al.
Veröffentlicht: (2025)
von: Cheng, Zelei, et al.
Veröffentlicht: (2025)
A Practical Introduction to Deep Reinforcement Learning
von: Sun, Yinghan, et al.
Veröffentlicht: (2025)
von: Sun, Yinghan, et al.
Veröffentlicht: (2025)
On The Presence of Double-Descent in Deep Reinforcement Learning
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
von: Veselý, Viktor, et al.
Veröffentlicht: (2025)
Investigating the Treacherous Turn in Deep Reinforcement Learning
von: Ashcraft, Chace, et al.
Veröffentlicht: (2025)
von: Ashcraft, Chace, et al.
Veröffentlicht: (2025)
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
von: Berseth, Glen
Veröffentlicht: (2025)
von: Berseth, Glen
Veröffentlicht: (2025)
Understanding and Improving Hyperbolic Deep Reinforcement Learning
von: Klein, Timo, et al.
Veröffentlicht: (2025)
von: Klein, Timo, et al.
Veröffentlicht: (2025)
FlowCritic: Bridging Value Estimation with Flow Matching in Reinforcement Learning
von: Zhong, Shan, et al.
Veröffentlicht: (2025)
von: Zhong, Shan, et al.
Veröffentlicht: (2025)
Fast Value Tracking for Deep Reinforcement Learning
von: Shih, Frank, et al.
Veröffentlicht: (2024)
von: Shih, Frank, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025) -
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
von: Thalagala, Shiron, et al.
Veröffentlicht: (2024) -
DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care
von: Tan, Daniel Jason, et al.
Veröffentlicht: (2025) -
Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning
von: Humayoo, Mahammad, et al.
Veröffentlicht: (2018) -
A Contextual Combinatorial Bandit Approach to Negotiation
von: Li, Yexin, et al.
Veröffentlicht: (2024)