Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Huayu, Zheng, Kaiwen, Su, Hang, Zhu, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Score Regularized Policy Optimization through Diffusion Behavior
by: Chen, Huayu, et al.
Published: (2023)
by: Chen, Huayu, et al.
Published: (2023)
Visual Generation Without Guidance
by: Chen, Huayu, et al.
Published: (2025)
by: Chen, Huayu, et al.
Published: (2025)
Exploratory Diffusion Model for Unsupervised Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2025)
by: Ying, Chengyang, et al.
Published: (2025)
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
Improved Techniques for Maximum Likelihood Estimation for Diffusion ODEs
by: Zheng, Kaiwen, et al.
Published: (2023)
by: Zheng, Kaiwen, et al.
Published: (2023)
Toward Guidance-Free AR Visual Generation via Condition Contrastive Alignment
by: Chen, Huayu, et al.
Published: (2024)
by: Chen, Huayu, et al.
Published: (2024)
Vidarc: Embodied Video Diffusion Model for Closed-loop Control
by: Feng, Yao, et al.
Published: (2025)
by: Feng, Yao, et al.
Published: (2025)
Diffusion Bridge Implicit Models
by: Zheng, Kaiwen, et al.
Published: (2024)
by: Zheng, Kaiwen, et al.
Published: (2024)
Noise Contrastive Alignment of Language Models with Explicit Rewards
by: Chen, Huayu, et al.
Published: (2024)
by: Chen, Huayu, et al.
Published: (2024)
HiDe-PET: Continual Learning via Hierarchical Decomposition of Parameter-Efficient Tuning
by: Wang, Liyuan, et al.
Published: (2024)
by: Wang, Liyuan, et al.
Published: (2024)
Consistency Diffusion Bridge Models
by: He, Guande, et al.
Published: (2024)
by: He, Guande, et al.
Published: (2024)
Direct Discriminative Optimization: Your Likelihood-Based Visual Generative Model is Secretly a GAN Discriminator
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
C-GAIL: Stabilizing Generative Adversarial Imitation Learning with Control Theory
by: Luo, Tianjiao, et al.
Published: (2024)
by: Luo, Tianjiao, et al.
Published: (2024)
Schrodinger Bridges Beat Diffusion Models on Text-to-Speech Synthesis
by: Chen, Zehua, et al.
Published: (2023)
by: Chen, Zehua, et al.
Published: (2023)
Towards a General Framework for Continual Learning with Pre-training
by: Wang, Liyuan, et al.
Published: (2023)
by: Wang, Liyuan, et al.
Published: (2023)
RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
by: Liu, Songming, et al.
Published: (2024)
by: Liu, Songming, et al.
Published: (2024)
Continuous Q-Score Matching: Diffusion Guided Reinforcement Learning for Continuous-Time Control
by: Hua, Chengxiu, et al.
Published: (2025)
by: Hua, Chengxiu, et al.
Published: (2025)
Efficient Backpropagation with Variance-Controlled Adaptive Sampling
by: Wang, Ziteng, et al.
Published: (2024)
by: Wang, Ziteng, et al.
Published: (2024)
S-STE: Continuous Pruning Function for Efficient 2:4 Sparse Pre-training
by: Hu, Yuezhou, et al.
Published: (2024)
by: Hu, Yuezhou, et al.
Published: (2024)
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
by: Zheng, Kaiwen, et al.
Published: (2024)
by: Zheng, Kaiwen, et al.
Published: (2024)
A Comprehensive Survey of Continual Learning: Theory, Method and Application
by: Wang, Liyuan, et al.
Published: (2023)
by: Wang, Liyuan, et al.
Published: (2023)
VLM Q-Learning: Aligning Vision-Language Models for Interactive Decision-Making
by: Grigsby, Jake, et al.
Published: (2025)
by: Grigsby, Jake, et al.
Published: (2025)
Why the Maximum Second Derivative of Activations Matters for Adversarial Robustness
by: Yu, Yunrui, et al.
Published: (2026)
by: Yu, Yunrui, et al.
Published: (2026)
Safety Alignment as Continual Learning: Mitigating the Alignment Tax via Orthogonal Gradient Projection
by: Sun, Guanglong, et al.
Published: (2026)
by: Sun, Guanglong, et al.
Published: (2026)
Towards the Worst-case Robustness of Large Language Models
by: Chen, Huanran, et al.
Published: (2025)
by: Chen, Huanran, et al.
Published: (2025)
Efficient, Property-Aligned Fan-Out Retrieval via RL-Compiled Diffusion
by: Jiang, Pengcheng, et al.
Published: (2026)
by: Jiang, Pengcheng, et al.
Published: (2026)
SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents
by: Chen, Xiwen, et al.
Published: (2026)
by: Chen, Xiwen, et al.
Published: (2026)
TurboDiffusion: Accelerating Video Diffusion Models by 100-200 Times
by: Zhang, Jintao, et al.
Published: (2025)
by: Zhang, Jintao, et al.
Published: (2025)
Aligning Flow Map Policies with Optimal Q-Guidance
by: Ziakas, Christos, et al.
Published: (2026)
by: Ziakas, Christos, et al.
Published: (2026)
NFT: Bridging Supervised Learning and Reinforcement Learning in Math Reasoning
by: Chen, Huayu, et al.
Published: (2025)
by: Chen, Huayu, et al.
Published: (2025)
Elucidating the Preconditioning in Consistency Distillation
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
Your Diffusion Model is Secretly a Certifiably Robust Classifier
by: Chen, Huanran, et al.
Published: (2024)
by: Chen, Huanran, et al.
Published: (2024)
Fourier Controller Networks for Real-Time Decision-Making in Embodied Learning
by: Tan, Hengkai, et al.
Published: (2024)
by: Tan, Hengkai, et al.
Published: (2024)
The Behavior and Convergence of Local Bayesian Optimization
by: Wu, Kaiwen, et al.
Published: (2023)
by: Wu, Kaiwen, et al.
Published: (2023)
SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-Shot
by: Tuo, Kaiwen, et al.
Published: (2025)
by: Tuo, Kaiwen, et al.
Published: (2025)
Data-regularized Reinforcement Learning for Diffusion Models at Scale
by: Ye, Haotian, et al.
Published: (2025)
by: Ye, Haotian, et al.
Published: (2025)
SparseDM: Toward Sparse Efficient Diffusion Models
by: Wang, Kafeng, et al.
Published: (2024)
by: Wang, Kafeng, et al.
Published: (2024)
Robust Classification via a Single Diffusion Model
by: Chen, Huanran, et al.
Published: (2023)
by: Chen, Huanran, et al.
Published: (2023)
Diffusion Models as Dataset Distillation Priors
by: Su, Duo, et al.
Published: (2025)
by: Su, Duo, et al.
Published: (2025)
Similar Items
-
Score Regularized Policy Optimization through Diffusion Behavior
by: Chen, Huayu, et al.
Published: (2023) -
Visual Generation Without Guidance
by: Chen, Huayu, et al.
Published: (2025) -
Exploratory Diffusion Model for Unsupervised Reinforcement Learning
by: Ying, Chengyang, et al.
Published: (2025) -
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
by: Zheng, Kaiwen, et al.
Published: (2025) -
Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency
by: Zheng, Kaiwen, et al.
Published: (2025)