Distributional Reinforcement Learning with Diffusion Bridge Critics
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Shutong, Zhou, Yimiao, Hu, Ke, Pan, Mokai, Zhong, Shan, Fu, Yanwei, Wang, Jingya, Shi, Ye |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance
by: Ding, Shutong, et al.
Published: (2026)
by: Ding, Shutong, et al.
Published: (2026)
GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
by: Ding, Shutong, et al.
Published: (2025)
by: Ding, Shutong, et al.
Published: (2025)
A Unified and Fast-Sampling Diffusion Bridge Framework via Stochastic Optimal Control
by: Pan, Mokai, et al.
Published: (2025)
by: Pan, Mokai, et al.
Published: (2025)
Diffusion-based Reinforcement Learning via Q-weighted Variational Policy Optimization
by: Ding, Shutong, et al.
Published: (2024)
by: Ding, Shutong, et al.
Published: (2024)
Diffusion-based learning framework for Constrained Nonconvex Optimization with Weighted Bootstrapped Refinement
by: Ding, Shutong, et al.
Published: (2025)
by: Ding, Shutong, et al.
Published: (2025)
Sample from What You See: Visuomotor Policy Learning via Diffusion Bridge with Observation-Embedded Stochastic Differential Equation
by: Liu, Zhaoyang, et al.
Published: (2025)
by: Liu, Zhaoyang, et al.
Published: (2025)
Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrödinger Bridge
by: Gong, Yuehu, et al.
Published: (2026)
by: Gong, Yuehu, et al.
Published: (2026)
FlowCritic: Bridging Value Estimation with Flow Matching in Reinforcement Learning
by: Zhong, Shan, et al.
Published: (2025)
by: Zhong, Shan, et al.
Published: (2025)
Guidance with Spherical Gaussian Constraint for Conditional Diffusion
by: Yang, Lingxiao, et al.
Published: (2024)
by: Yang, Lingxiao, et al.
Published: (2024)
UniDB: A Unified Diffusion Bridge Framework via Stochastic Optimal Control
by: Zhu, Kaizhen, et al.
Published: (2025)
by: Zhu, Kaizhen, et al.
Published: (2025)
Diffusion Bridge or Flow Matching? A Unifying Framework and Comparative Analysis
by: Zhu, Kaizhen, et al.
Published: (2025)
by: Zhu, Kaizhen, et al.
Published: (2025)
Harmonizing Generalization and Personalization in Federated Prompt Learning
by: Cui, Tianyu, et al.
Published: (2024)
by: Cui, Tianyu, et al.
Published: (2024)
DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion
by: Fan, Yahao, et al.
Published: (2025)
by: Fan, Yahao, et al.
Published: (2025)
EvoNash-MARL: A Closed-Loop Multi-Agent Reinforcement Learning Framework for Medium-Horizon Equity Allocation
by: Jia, Chongliu, et al.
Published: (2026)
by: Jia, Chongliu, et al.
Published: (2026)
Adaptive Pruning of Pretrained Transformer via Differential Inclusions
by: Ding, Yizhuo, et al.
Published: (2025)
by: Ding, Yizhuo, et al.
Published: (2025)
Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning
by: Shan, Zikang, et al.
Published: (2026)
by: Shan, Zikang, et al.
Published: (2026)
Federated Distributional Reinforcement Learning with Distributional Critic Regularization
by: Millard, David, et al.
Published: (2026)
by: Millard, David, et al.
Published: (2026)
Stabilizing Reinforcement Learning for Diffusion Language Models
by: Zhong, Jianyuan, et al.
Published: (2026)
by: Zhong, Jianyuan, et al.
Published: (2026)
Actor-Critic Reinforcement Learning with Phased Actor
by: Wu, Ruofan, et al.
Published: (2024)
by: Wu, Ruofan, et al.
Published: (2024)
One-Shot Federated Learning with Classifier-Free Diffusion Models
by: Zaland, Obaidullah, et al.
Published: (2025)
by: Zaland, Obaidullah, et al.
Published: (2025)
RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion
by: Hu, Tianmeng, et al.
Published: (2026)
by: Hu, Tianmeng, et al.
Published: (2026)
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning
by: Wang, Qingjun, et al.
Published: (2026)
by: Wang, Qingjun, et al.
Published: (2026)
Distributional Inverse Reinforcement Learning
by: Wu, Feiyang, et al.
Published: (2025)
by: Wu, Feiyang, et al.
Published: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
Global and Local Prompts Cooperation via Optimal Transport for Federated Learning
by: Li, Hongxia, et al.
Published: (2024)
by: Li, Hongxia, et al.
Published: (2024)
How Does Return Distribution in Distributional Reinforcement Learning Help Optimization?
by: Sun, Ke, et al.
Published: (2022)
by: Sun, Ke, et al.
Published: (2022)
Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models
by: Wu, Shutong, et al.
Published: (2025)
by: Wu, Shutong, et al.
Published: (2025)
SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement Learning
by: Li, Na, et al.
Published: (2025)
by: Li, Na, et al.
Published: (2025)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
by: Ma, Xiaoteng, et al.
Published: (2020)
by: Ma, Xiaoteng, et al.
Published: (2020)
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
by: Sun, Ke, et al.
Published: (2021)
by: Sun, Ke, et al.
Published: (2021)
Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning
by: Fang, Linjiajie, et al.
Published: (2024)
by: Fang, Linjiajie, et al.
Published: (2024)
Generative Actor-Critic with Soft Bridge Policies
by: He, Ke, et al.
Published: (2026)
by: He, Ke, et al.
Published: (2026)
CausalGDP: Causality-Guided Diffusion Policies for Reinforcement Learning
by: Xiao, Xiaofeng, et al.
Published: (2026)
by: Xiao, Xiaofeng, et al.
Published: (2026)
Distributional Reinforcement Learning with Regularized Wasserstein Loss
by: Sun, Ke, et al.
Published: (2022)
by: Sun, Ke, et al.
Published: (2022)
The Distributional Reward Critic Framework for Reinforcement Learning Under Perturbed Rewards
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Distributed Gradient Descent for Functional Learning
by: Yu, Zhan, et al.
Published: (2023)
by: Yu, Zhan, et al.
Published: (2023)
PhaseNAS: Language-Model Driven Architecture Search with Dynamic Phase Adaptation
by: Kong, Fei, et al.
Published: (2025)
by: Kong, Fei, et al.
Published: (2025)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
by: Cui, Mingxuan, et al.
Published: (2025)
by: Cui, Mingxuan, et al.
Published: (2025)
Bridging Dynamics Gaps via Diffusion Schrödinger Bridge for Cross-Domain Reinforcement Learning
by: Zhang, Hanping, et al.
Published: (2026)
by: Zhang, Hanping, et al.
Published: (2026)
One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow
by: Wang, Zeyuan, et al.
Published: (2025)
by: Wang, Zeyuan, et al.
Published: (2025)
Similar Items
-
Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance
by: Ding, Shutong, et al.
Published: (2026) -
GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
by: Ding, Shutong, et al.
Published: (2025) -
A Unified and Fast-Sampling Diffusion Bridge Framework via Stochastic Optimal Control
by: Pan, Mokai, et al.
Published: (2025) -
Diffusion-based Reinforcement Learning via Q-weighted Variational Policy Optimization
by: Ding, Shutong, et al.
Published: (2024) -
Diffusion-based learning framework for Constrained Nonconvex Optimization with Weighted Bootstrapped Refinement
by: Ding, Shutong, et al.
Published: (2025)