Fine-Tuning Diffusion Models via Intermediate Distribution Shaping
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Anil, Gautham Govind, Haque, Shaan Ul, Kannen, Nithish, Nagaraj, Dheeraj, Shakkottai, Sanjay, Shanmugam, Karthikeyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Interleaved Gibbs Diffusion: Generating Discrete-Continuous Data with Implicit Constraints
von: Anil, Gautham Govind, et al.
Veröffentlicht: (2025)
von: Anil, Gautham Govind, et al.
Veröffentlicht: (2025)
Efficient Approximate Posterior Sampling with Annealed Langevin Monte Carlo
von: Parulekar, Advait, et al.
Veröffentlicht: (2025)
von: Parulekar, Advait, et al.
Veröffentlicht: (2025)
Glauber Generative Model: Discrete Diffusion Models via Binary Classification
von: Varma, Harshit, et al.
Veröffentlicht: (2024)
von: Varma, Harshit, et al.
Veröffentlicht: (2024)
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
von: Kumar, Ramnath, et al.
Veröffentlicht: (2023)
von: Kumar, Ramnath, et al.
Veröffentlicht: (2023)
Model Editing by Standard Fine-Tuning
von: Gangadhar, Govind, et al.
Veröffentlicht: (2024)
von: Gangadhar, Govind, et al.
Veröffentlicht: (2024)
Entropy Aware Reward Guidance for Diffusion Language Model Alignment
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
Bandits with Mean Bounds
von: Sharma, Nihal, et al.
Veröffentlicht: (2020)
von: Sharma, Nihal, et al.
Veröffentlicht: (2020)
Machine Unlearning under Overparameterization
von: Block, Jacob L., et al.
Veröffentlicht: (2025)
von: Block, Jacob L., et al.
Veröffentlicht: (2025)
Safety Subspaces are Not Linearly Distinct: A Fine-Tuning Case Study
von: Ponkshe, Kaustubh, et al.
Veröffentlicht: (2025)
von: Ponkshe, Kaustubh, et al.
Veröffentlicht: (2025)
Beyond Propagation of Chaos: A Stochastic Algorithm for Mean Field Optimization
von: Tankala, Chandan, et al.
Veröffentlicht: (2025)
von: Tankala, Chandan, et al.
Veröffentlicht: (2025)
PITA: Preference-Guided Inference-Time Alignment for LLM Post-Training
von: Bobbili, Sarat Chandra, et al.
Veröffentlicht: (2025)
von: Bobbili, Sarat Chandra, et al.
Veröffentlicht: (2025)
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
Learning from Label Proportions: Bootstrapping Supervised Learners via Belief Propagation
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
Generating Universal Adversarial Perturbations for Quantum Classifiers
von: Anil, Gautham, et al.
Veröffentlicht: (2024)
von: Anil, Gautham, et al.
Veröffentlicht: (2024)
Stronger Enforcement of Instruction Hierarchy via Augmented Intermediate Representations
von: Kariyappa, Sanjay, et al.
Veröffentlicht: (2025)
von: Kariyappa, Sanjay, et al.
Veröffentlicht: (2025)
Towards a Pretrained Model for Restless Bandits via Multi-arm Generalization
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
von: Zhao, Yunfan, et al.
Veröffentlicht: (2023)
Bandits with Stochastic Experts: Constant Regret, Empirical Experts and Episodes
von: Sharma, Nihal, et al.
Veröffentlicht: (2021)
von: Sharma, Nihal, et al.
Veröffentlicht: (2021)
Combinatorial Multi-armed Bandits: Arm Selection via Group Testing
von: Mukherjee, Arpan, et al.
Veröffentlicht: (2024)
von: Mukherjee, Arpan, et al.
Veröffentlicht: (2024)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
von: Patil, Gandharv, et al.
Veröffentlicht: (2022)
Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport
von: Shah, Shaan, et al.
Veröffentlicht: (2025)
von: Shah, Shaan, et al.
Veröffentlicht: (2025)
Bridging Model-Based Optimization and Generative Modeling via Conservative Fine-Tuning of Diffusion Models
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
Reward Sharpness-Aware Fine-Tuning for Diffusion Models
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2026)
von: Kim, Kwanyoung, et al.
Veröffentlicht: (2026)
Fine-Tuning Diffusion Models for Molecular Generation via Reinforcement Learning and Fast Sampling
von: Lin, Guang, et al.
Veröffentlicht: (2026)
von: Lin, Guang, et al.
Veröffentlicht: (2026)
Fairness under Covariate Shift: Improving Fairness-Accuracy tradeoff with few Unlabeled Test Samples
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
von: Behari, Nikhil, et al.
Veröffentlicht: (2024)
Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein Design
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
von: Wang, Chenyu, et al.
Veröffentlicht: (2024)
PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment
von: Verma, Richa, et al.
Veröffentlicht: (2026)
von: Verma, Richa, et al.
Veröffentlicht: (2026)
In-Context Learning with Transformers: Softmax Attention Adapts to Function Lipschitzness
von: Collins, Liam, et al.
Veröffentlicht: (2024)
von: Collins, Liam, et al.
Veröffentlicht: (2024)
Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
Diffusion Fine-Tuning via Reparameterized Policy Gradient of the Soft Q-Function
von: Kang, Hyeongyu, et al.
Veröffentlicht: (2025)
von: Kang, Hyeongyu, et al.
Veröffentlicht: (2025)
DLPO: Diffusion Model Loss-Guided Reinforcement Learning for Fine-Tuning Text-to-Speech Diffusion Models
von: Chen, Jingyi, et al.
Veröffentlicht: (2024)
von: Chen, Jingyi, et al.
Veröffentlicht: (2024)
Feedback Efficient Online Fine-Tuning of Diffusion Models
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
SVFT: Parameter-Efficient Fine-Tuning with Singular Vectors
von: Lingam, Vijay, et al.
Veröffentlicht: (2024)
von: Lingam, Vijay, et al.
Veröffentlicht: (2024)
Fine-tuning Flow Matching Generative Models with Intermediate Feedback
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
Fine-Tuning Discrete Diffusion Models with Policy Gradient Methods
von: Zekri, Oussama, et al.
Veröffentlicht: (2025)
von: Zekri, Oussama, et al.
Veröffentlicht: (2025)
Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Control
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
von: Wang, Xinyu, et al.
Veröffentlicht: (2026)
Constrained Posterior Sampling: Time Series Generation with Hard Constraints
von: Narasimhan, Sai Shankar, et al.
Veröffentlicht: (2024)
von: Narasimhan, Sai Shankar, et al.
Veröffentlicht: (2024)
How Robust is Model Editing after Fine-Tuning? An Empirical Study on Text-to-Image Diffusion Models
von: He, Feng, et al.
Veröffentlicht: (2025)
von: He, Feng, et al.
Veröffentlicht: (2025)
Thompson Sampling via Fine-Tuning of LLMs
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
Evaluating Generalization and Representation Stability in Small LMs via Prompting, Fine-Tuning and Out-of-Distribution Prompts
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
von: Raja, Rahul, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Interleaved Gibbs Diffusion: Generating Discrete-Continuous Data with Implicit Constraints
von: Anil, Gautham Govind, et al.
Veröffentlicht: (2025) -
Efficient Approximate Posterior Sampling with Annealed Langevin Monte Carlo
von: Parulekar, Advait, et al.
Veröffentlicht: (2025) -
Glauber Generative Model: Discrete Diffusion Models via Binary Classification
von: Varma, Harshit, et al.
Veröffentlicht: (2024) -
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
von: Kumar, Ramnath, et al.
Veröffentlicht: (2023) -
Model Editing by Standard Fine-Tuning
von: Gangadhar, Govind, et al.
Veröffentlicht: (2024)