S$^2$AC: Energy-Based Reinforcement Learning with Stein Soft Actor Critic
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Messaoud, Safa, Mokeddem, Billel, Xue, Zhenghai, Pang, Linsey, An, Bo, Chen, Haipeng, Chawla, Sanjay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Explaining the role of Intrinsic Dimensionality in Adversarial Training
von: Altinisik, Enes, et al.
Veröffentlicht: (2024)
von: Altinisik, Enes, et al.
Veröffentlicht: (2024)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025)
Actor-Critic Reinforcement Learning with Phased Actor
von: Wu, Ruofan, et al.
Veröffentlicht: (2024)
von: Wu, Ruofan, et al.
Veröffentlicht: (2024)
PAC-Bayesian Soft Actor-Critic Learning
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)
Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
von: Vo, Thanh Vinh, et al.
Veröffentlicht: (2025)
Learning Without Time-Based Embodiment Resets in Soft-Actor Critic
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
von: Farrahi, Homayoon, et al.
Veröffentlicht: (2025)
Two-Stage Constrained Actor-Critic for Short Video Recommendation
von: Cai, Qingpeng, et al.
Veröffentlicht: (2023)
von: Cai, Qingpeng, et al.
Veröffentlicht: (2023)
Safe Langevin Soft Actor Critic
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
von: Keswani, Mahesh, et al.
Veröffentlicht: (2026)
Distributional Soft Actor-Critic with Three Refinements
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
von: Duan, Jingliang, et al.
Veröffentlicht: (2023)
Bounded Exploration with World Model Uncertainty in Soft Actor-Critic Reinforcement Learning Algorithm
von: Qiao, Ting, et al.
Veröffentlicht: (2024)
von: Qiao, Ting, et al.
Veröffentlicht: (2024)
Ask-AC: An Initiative Advisor-in-the-Loop Actor-Critic Framework
von: Liu, Shunyu, et al.
Veröffentlicht: (2022)
von: Liu, Shunyu, et al.
Veröffentlicht: (2022)
Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)
von: Mahran, Youssef, et al.
Veröffentlicht: (2025)
von: Mahran, Youssef, et al.
Veröffentlicht: (2025)
SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning
von: Xue, Zhenghai, et al.
Veröffentlicht: (2025)
von: Xue, Zhenghai, et al.
Veröffentlicht: (2025)
Revisiting Discrete Soft Actor-Critic
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
Average-Reward Soft Actor-Critic
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Wasserstein Barycenter Soft Actor-Critic
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
von: Shahrooei, Zahra, et al.
Veröffentlicht: (2025)
Distributional Soft Actor-Critic with Diffusion Policy
von: Liu, Tong, et al.
Veröffentlicht: (2025)
von: Liu, Tong, et al.
Veröffentlicht: (2025)
Risk-Sensitive Soft Actor-Critic for Robust Deep Reinforcement Learning under Distribution Shifts
von: Enders, Tobias, et al.
Veröffentlicht: (2024)
von: Enders, Tobias, et al.
Veröffentlicht: (2024)
Bidirectional Soft Actor-Critic: Leveraging Forward and Reverse KL Divergence for Efficient Reinforcement Learning
von: Zhang, Yixian, et al.
Veröffentlicht: (2025)
von: Zhang, Yixian, et al.
Veröffentlicht: (2025)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
von: Ishfaq, Haque, et al.
Veröffentlicht: (2025)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
AURO: Reinforcement Learning for Adaptive User Retention Optimization in Recommender Systems
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
von: Xue, Zhenghai, et al.
Veröffentlicht: (2023)
Policy Regularization on Globally Accessible States in Cross-Dynamics Reinforcement Learning
von: Xue, Zhenghai, et al.
Veröffentlicht: (2025)
von: Xue, Zhenghai, et al.
Veröffentlicht: (2025)
Competitive Audio-Language Models with Data-Efficient Single-Stage Training on Public Data
von: Kumar, Gokul Karthik, et al.
Veröffentlicht: (2025)
von: Kumar, Gokul Karthik, et al.
Veröffentlicht: (2025)
Flow Actor-Critic for Offline Reinforcement Learning
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
von: Chae, Jongseong, et al.
Veröffentlicht: (2026)
Cryptocurrency Portfolio Management with Reinforcement Learning: Soft Actor--Critic and Deep Deterministic Policy Gradient Algorithms
von: Paykan, Kamal
Veröffentlicht: (2025)
von: Paykan, Kamal
Veröffentlicht: (2025)
Generative Actor-Critic with Soft Bridge Policies
von: He, Ke, et al.
Veröffentlicht: (2026)
von: He, Ke, et al.
Veröffentlicht: (2026)
Pretraining in Actor-Critic Reinforcement Learning for Robot Locomotion
von: Fan, Jiale, et al.
Veröffentlicht: (2025)
von: Fan, Jiale, et al.
Veröffentlicht: (2025)
Policy-Based Radiative Transfer: Solving the $2$-Level Atom Non-LTE Problem using Soft Actor-Critic Reinforcement Learning
von: Panos, Brandon, et al.
Veröffentlicht: (2025)
von: Panos, Brandon, et al.
Veröffentlicht: (2025)
DSAC-C: Constrained Maximum Entropy for Robust Discrete Soft-Actor Critic
von: Neo, Dexter, et al.
Veröffentlicht: (2023)
von: Neo, Dexter, et al.
Veröffentlicht: (2023)
Group-in-Group Policy Optimization for LLM Agent Training
von: Feng, Lang, et al.
Veröffentlicht: (2025)
von: Feng, Lang, et al.
Veröffentlicht: (2025)
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
von: Chen, Haohui, et al.
Veröffentlicht: (2024)
von: Chen, Haohui, et al.
Veröffentlicht: (2024)
SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement Learning
von: Li, Na, et al.
Veröffentlicht: (2025)
von: Li, Na, et al.
Veröffentlicht: (2025)
Efficient Soft Actor-Critic with LLM-Based Action-Level Guidance for Continuous Control
von: Ma, Hao, et al.
Veröffentlicht: (2026)
von: Ma, Hao, et al.
Veröffentlicht: (2026)
Quantum Advantage Actor-Critic for Reinforcement Learning
von: Kölle, Michael, et al.
Veröffentlicht: (2024)
von: Kölle, Michael, et al.
Veröffentlicht: (2024)
SACn: Soft Actor-Critic with n-step Returns
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
von: Łyskawa, Jakub, et al.
Veröffentlicht: (2025)
RIDGECUT: Learning Graph Partitioning with Rings and Wedges
von: Jiang, Qize, et al.
Veröffentlicht: (2025)
von: Jiang, Qize, et al.
Veröffentlicht: (2025)
Chunking the Critic: A Transformer-based Soft Actor-Critic with N-Step Returns
von: Tian, Dong, et al.
Veröffentlicht: (2025)
von: Tian, Dong, et al.
Veröffentlicht: (2025)
Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
von: Nauman, Michal, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Explaining the role of Intrinsic Dimensionality in Adversarial Training
von: Altinisik, Enes, et al.
Veröffentlicht: (2024) -
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020) -
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning
von: Küçükoğlu, Burcu, et al.
Veröffentlicht: (2025) -
Actor-Critic Reinforcement Learning with Phased Actor
von: Wu, Ruofan, et al.
Veröffentlicht: (2024) -
PAC-Bayesian Soft Actor-Critic Learning
von: Tasdighi, Bahareh, et al.
Veröffentlicht: (2023)