Generalizing soft actor-critic algorithms to discrete action spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Le, Gu, Yong, Zhao, Xin, Zhang, Yanshuo, Zhao, Shu, Jin, Yifei, Wu, Xinxin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement learning for automatic quadrilateral mesh generation: a soft actor-critic approach
by: Pan, Jie, et al.
Published: (2022)
by: Pan, Jie, et al.
Published: (2022)
Feature Fusion Based on Mutual-Cross-Attention Mechanism for EEG Emotion Recognition
by: Zhao, Yimin, et al.
Published: (2024)
by: Zhao, Yimin, et al.
Published: (2024)
FAME: Adaptive Functional Attention with Expert Routing for Function-on-Function Regression
by: Gao, Yifei, et al.
Published: (2025)
by: Gao, Yifei, et al.
Published: (2025)
Teaching RL Agents to Act Better: VLM as Action Advisor for Online Reinforcement Learning
by: Wu, Xiefeng, et al.
Published: (2025)
by: Wu, Xiefeng, et al.
Published: (2025)
Why are hyperbolic neural networks effective? A study on hierarchical representation capability
by: Tan, Shicheng, et al.
Published: (2024)
by: Tan, Shicheng, et al.
Published: (2024)
HT-GNN: Hyper-Temporal Graph Neural Network for Customer Lifetime Value Prediction in Baidu Ads
by: Zhao, Xiaohui, et al.
Published: (2026)
by: Zhao, Xiaohui, et al.
Published: (2026)
Addressing imperfect symmetry: A novel symmetry-learning actor-critic extension
by: Abreu, Miguel, et al.
Published: (2023)
by: Abreu, Miguel, et al.
Published: (2023)
Gradient Compression May Hurt Generalization: A Remedy by Synthetic Data Guided Sharpness Aware Minimization
by: Gu, Yujie, et al.
Published: (2026)
by: Gu, Yujie, et al.
Published: (2026)
Learning in complex action spaces without policy gradients
by: Tavakoli, Arash, et al.
Published: (2024)
by: Tavakoli, Arash, et al.
Published: (2024)
Federated Generative Learning with Foundation Models
by: Zhang, Jie, et al.
Published: (2023)
by: Zhang, Jie, et al.
Published: (2023)
Cross-Modality Controlled Molecule Generation with Diffusion Language Model
by: Zhang, Yunzhe, et al.
Published: (2025)
by: Zhang, Yunzhe, et al.
Published: (2025)
Deconstructing Generative Diversity: An Information Bottleneck Analysis of Discrete Latent Generative Models
by: Wu, Yudi, et al.
Published: (2025)
by: Wu, Yudi, et al.
Published: (2025)
Principled Understanding of Generalization for Generative Transformer Models in Arithmetic Reasoning Tasks
by: Xu, Xingcheng, et al.
Published: (2024)
by: Xu, Xingcheng, et al.
Published: (2024)
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
by: Liu, Siyuan, et al.
Published: (2026)
by: Liu, Siyuan, et al.
Published: (2026)
DIVE: Subgraph Disagreement for Graph Out-of-Distribution Generalization
by: Sun, Xin, et al.
Published: (2024)
by: Sun, Xin, et al.
Published: (2024)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
by: Zhang, Junkai, et al.
Published: (2024)
by: Zhang, Junkai, et al.
Published: (2024)
Group-Adaptive Adversarial Learning for Robust Fake News Detection Against Malicious Comments
by: Tong, Zhao, et al.
Published: (2025)
by: Tong, Zhao, et al.
Published: (2025)
Value-guided action planning with JEPA world models
by: Destrade, Matthieu, et al.
Published: (2025)
by: Destrade, Matthieu, et al.
Published: (2025)
Distribution Preference Optimization: A Fine-grained Perspective for LLM Unlearning
by: Qin, Kai, et al.
Published: (2025)
by: Qin, Kai, et al.
Published: (2025)
EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation
by: Zhang, Shu-Hao, et al.
Published: (2026)
by: Zhang, Shu-Hao, et al.
Published: (2026)
AL-GNN: Privacy-Preserving and Replay-Free Continual Graph Learning via Analytic Learning
by: Zhang, Xuling, et al.
Published: (2025)
by: Zhang, Xuling, et al.
Published: (2025)
OLion: Approaching the Hadamard Ideal by Intersecting Spectral and $\ell_{\infty}$ Implicit Biases
by: Wang, Zixiao, et al.
Published: (2026)
by: Wang, Zixiao, et al.
Published: (2026)
An Augmentation Overlap Theory of Contrastive Learning
by: Zhang, Qi, et al.
Published: (2025)
by: Zhang, Qi, et al.
Published: (2025)
Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI
by: Shu, Bo, et al.
Published: (2024)
by: Shu, Bo, et al.
Published: (2024)
Ordering-based Causal Discovery via Generalized Score Matching
by: Vo, Vy, et al.
Published: (2026)
by: Vo, Vy, et al.
Published: (2026)
Generalizing Hyperedge Expansion for Hyper-relational Knowledge Graph Modeling
by: Liu, Yu, et al.
Published: (2024)
by: Liu, Yu, et al.
Published: (2024)
ViSymRe: Vision Multimodal Symbolic Regression
by: Li, Da, et al.
Published: (2024)
by: Li, Da, et al.
Published: (2024)
PateGail: A Privacy-Preserving Mobility Trajectory Generator with Imitation Learning
by: Wang, Huandong, et al.
Published: (2024)
by: Wang, Huandong, et al.
Published: (2024)
Generative Fuzzy System for Sequence Generation
by: Yang, Hailong, et al.
Published: (2024)
by: Yang, Hailong, et al.
Published: (2024)
Towards General Continuous Memory for Vision-Language Models
by: Wu, Wenyi, et al.
Published: (2025)
by: Wu, Wenyi, et al.
Published: (2025)
Towards a Sharp Analysis of Offline Policy Learning for $f$-Divergence-Regularized Contextual Bandits
by: Zhao, Qingyue, et al.
Published: (2025)
by: Zhao, Qingyue, et al.
Published: (2025)
Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion
by: Sanders, Justin, et al.
Published: (2026)
by: Sanders, Justin, et al.
Published: (2026)
ScenGAN: Attention-Intensive Generative Model for Uncertainty-Aware Renewable Scenario Forecasting
by: Wu, Yifei, et al.
Published: (2025)
by: Wu, Yifei, et al.
Published: (2025)
PID-controlled Langevin Dynamics for Faster Sampling of Generative Models
by: Chen, Hongyi, et al.
Published: (2025)
by: Chen, Hongyi, et al.
Published: (2025)
Hypergraph and Latent ODE Learning for Multimodal Root Cause Localization in Microservices
by: Liu, Xin, et al.
Published: (2026)
by: Liu, Xin, et al.
Published: (2026)
Out-of-Distribution Generalization in Time Series: A Survey
by: Wu, Xin, et al.
Published: (2025)
by: Wu, Xin, et al.
Published: (2025)
GEM: Generative Entropy-Guided Preference Modeling for Few-shot Alignment of LLMs
by: Zhao, Yiyang, et al.
Published: (2025)
by: Zhao, Yiyang, et al.
Published: (2025)
GFRIEND: Generative Few-shot Reward Inference through EfficieNt DPO
by: Zhao, Yiyang, et al.
Published: (2025)
by: Zhao, Yiyang, et al.
Published: (2025)
EvoFA: Evolvable Fast Adaptation for EEG Emotion Recognition
by: Jin, Ming, et al.
Published: (2024)
by: Jin, Ming, et al.
Published: (2024)
Learning Unbiased Cluster Descriptors for Interpretable Imbalanced Concept Drift Detection
by: Zhang, Yiqun, et al.
Published: (2026)
by: Zhang, Yiqun, et al.
Published: (2026)
Similar Items
-
Reinforcement learning for automatic quadrilateral mesh generation: a soft actor-critic approach
by: Pan, Jie, et al.
Published: (2022) -
Feature Fusion Based on Mutual-Cross-Attention Mechanism for EEG Emotion Recognition
by: Zhao, Yimin, et al.
Published: (2024) -
FAME: Adaptive Functional Attention with Expert Routing for Function-on-Function Regression
by: Gao, Yifei, et al.
Published: (2025) -
Teaching RL Agents to Act Better: VLM as Action Advisor for Online Reinforcement Learning
by: Wu, Xiefeng, et al.
Published: (2025) -
Why are hyperbolic neural networks effective? A study on hierarchical representation capability
by: Tan, Shicheng, et al.
Published: (2024)