Data-regularized Reinforcement Learning for Diffusion Models at Scale
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Haotian, Zheng, Kaiwen, Xu, Jiashu, Li, Puheng, Chen, Huayu, Han, Jiaqi, Liu, Sheng, Zhang, Qinsheng, Mao, Hanzi, Hao, Zekun, Chattopadhyay, Prithvijit, Yang, Dinghao, Feng, Liang, Liao, Maosheng, Bai, Junjie, Liu, Ming-Yu, Zou, James, Ermon, Stefano |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers
von: Han, Jiaqi, et al.
Veröffentlicht: (2025)
von: Han, Jiaqi, et al.
Veröffentlicht: (2025)
Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration
von: Han, Jiaqi, et al.
Veröffentlicht: (2026)
von: Han, Jiaqi, et al.
Veröffentlicht: (2026)
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024)
InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression
von: Ye, Haotian, et al.
Veröffentlicht: (2025)
von: Ye, Haotian, et al.
Veröffentlicht: (2025)
NFT: Bridging Supervised Learning and Reinforcement Learning in Math Reasoning
von: Chen, Huayu, et al.
Veröffentlicht: (2025)
von: Chen, Huayu, et al.
Veröffentlicht: (2025)
Direct Discriminative Optimization: Your Likelihood-Based Visual Generative Model is Secretly a GAN Discriminator
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
Scalable Policy Evaluation with Video World Models
von: Tseng, Wei-Cheng, et al.
Veröffentlicht: (2025)
von: Tseng, Wei-Cheng, et al.
Veröffentlicht: (2025)
TFG: Unified Training-Free Guidance for Diffusion Models
von: Ye, Haotian, et al.
Veröffentlicht: (2024)
von: Ye, Haotian, et al.
Veröffentlicht: (2024)
Editorial for “Deep Learning‐Based Brainstem Segmentation and Multi‐Class Classification for Parkinsonian Syndrome”
von: Prithvijit Chakraborty
Veröffentlicht: (2025)
von: Prithvijit Chakraborty
Veröffentlicht: (2025)
Geometric Trajectory Diffusion Models
von: Han, Jiaqi, et al.
Veröffentlicht: (2024)
von: Han, Jiaqi, et al.
Veröffentlicht: (2024)
Discrete Diffusion Trajectory Alignment via Stepwise Decomposition
von: Han, Jiaqi, et al.
Veröffentlicht: (2025)
von: Han, Jiaqi, et al.
Veröffentlicht: (2025)
FaiREE: Fair Classification with Finite-Sample and Distribution-Free Guarantee
von: Li, Puheng, et al.
Veröffentlicht: (2022)
von: Li, Puheng, et al.
Veröffentlicht: (2022)
AUGCAL: Improving Sim2Real Adaptation by Uncertainty Calibration on Augmented Synthetic Images
von: Chattopadhyay, Prithvijit, et al.
Veröffentlicht: (2023)
von: Chattopadhyay, Prithvijit, et al.
Veröffentlicht: (2023)
Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025)
EdgeRunner: Auto-regressive Auto-encoder for Artistic Mesh Generation
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Tang, Jiaxiang, et al.
Veröffentlicht: (2024)
SkyScenes: A Synthetic Dataset for Aerial Scene Understanding
von: Khose, Sahil, et al.
Veröffentlicht: (2023)
von: Khose, Sahil, et al.
Veröffentlicht: (2023)
Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models
von: Wang, Austin, et al.
Veröffentlicht: (2026)
von: Wang, Austin, et al.
Veröffentlicht: (2026)
GeoAda: Efficiently Finetune Geometric Diffusion Models with Equivariant Adapters
von: Zhao, Wanjia, et al.
Veröffentlicht: (2025)
von: Zhao, Wanjia, et al.
Veröffentlicht: (2025)
We're Not Using Videos Effectively: An Updated Domain Adaptive Video Segmentation Baseline
von: Kareer, Simar, et al.
Veröffentlicht: (2024)
von: Kareer, Simar, et al.
Veröffentlicht: (2024)
RefDrop: Controllable Consistency in Image or Video Generation via Reference Feature Guidance
von: Fan, Jiaojiao, et al.
Veröffentlicht: (2024)
von: Fan, Jiaojiao, et al.
Veröffentlicht: (2024)
Die Ordnung des Theaters. Eine Soziologie der Regie
von: Hänzi, Denis
Veröffentlicht: (2015)
von: Hänzi, Denis
Veröffentlicht: (2015)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking
von: Cundy, Chris, et al.
Veröffentlicht: (2023)
von: Cundy, Chris, et al.
Veröffentlicht: (2023)
REAR: Rethinking Visual Autoregressive Models via Generator-Tokenizer Consistency Regularization
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
von: He, Qiyuan, et al.
Veröffentlicht: (2025)
Divergence Minimization Preference Optimization for Diffusion Model Alignment
von: Li, Binxu, et al.
Veröffentlicht: (2025)
von: Li, Binxu, et al.
Veröffentlicht: (2025)
$f$-PO: Generalizing Preference Optimization with $f$-divergence Minimization
von: Han, Jiaqi, et al.
Veröffentlicht: (2024)
von: Han, Jiaqi, et al.
Veröffentlicht: (2024)
CPSample: Classifier Protected Sampling for Guarding Training Data During Diffusion
von: Kazdan, Joshua, et al.
Veröffentlicht: (2024)
von: Kazdan, Joshua, et al.
Veröffentlicht: (2024)
Robust Sampling for Active Statistical Inference
von: Li, Puheng, et al.
Veröffentlicht: (2025)
von: Li, Puheng, et al.
Veröffentlicht: (2025)
Analyzing the Role of Permutation Invariance in Linear Mode Connectivity
von: Zhan, Keyao, et al.
Veröffentlicht: (2025)
von: Zhan, Keyao, et al.
Veröffentlicht: (2025)
Exploring Neural Network Landscapes: Star-Shaped and Geodesic Connectivity
von: Lin, Zhanran, et al.
Veröffentlicht: (2024)
von: Lin, Zhanran, et al.
Veröffentlicht: (2024)
78‐3: High Picture Quality of LCD via WHVA Technology
von: Jing Liu, et al.
Veröffentlicht: (2025)
von: Jing Liu, et al.
Veröffentlicht: (2025)
Conformal Scalar-Flat Metrics with Prescribed Boundary Mean Curvature
von: Shen, Jiashu, et al.
Veröffentlicht: (2024)
von: Shen, Jiashu, et al.
Veröffentlicht: (2024)
In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
von: Liu, Sheng, et al.
Veröffentlicht: (2023)
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
von: Liu, Sheng, et al.
Veröffentlicht: (2024)
Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
DistriFusion: Distributed Parallel Inference for High-Resolution Diffusion Models
von: Li, Muyang, et al.
Veröffentlicht: (2024)
von: Li, Muyang, et al.
Veröffentlicht: (2024)
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition
von: Guo, Hanyu, et al.
Veröffentlicht: (2024)
von: Guo, Hanyu, et al.
Veröffentlicht: (2024)
Privacy-Constrained Policies via Mutual Information Regularized Policy Gradients
von: Cundy, Chris, et al.
Veröffentlicht: (2020)
von: Cundy, Chris, et al.
Veröffentlicht: (2020)
Inductive Moment Matching
von: Zhou, Linqi, et al.
Veröffentlicht: (2025)
von: Zhou, Linqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2025) -
CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers
von: Han, Jiaqi, et al.
Veröffentlicht: (2025) -
Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration
von: Han, Jiaqi, et al.
Veröffentlicht: (2026) -
Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling
von: Zheng, Kaiwen, et al.
Veröffentlicht: (2024) -
InfoTok: Adaptive Discrete Video Tokenizer via Information-Theoretic Compression
von: Ye, Haotian, et al.
Veröffentlicht: (2025)