Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control
Fuente:
arXiv
Saved in:
| Main Authors: | Uehara, Masatoshi, Zhao, Yulai, Black, Kevin, Hajiramezanali, Ehsan, Scalia, Gabriele, Diamant, Nathaniel Lee, Tseng, Alex M, Biancalani, Tommaso, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Feedback Efficient Online Fine-Tuning of Diffusion Models
by: Uehara, Masatoshi, et al.
Published: (2024)
by: Uehara, Masatoshi, et al.
Published: (2024)
Adding Conditional Control to Diffusion Models with Reinforcement Learning
by: Zhao, Yulai, et al.
Published: (2024)
by: Zhao, Yulai, et al.
Published: (2024)
Bridging Model-Based Optimization and Generative Modeling via Conservative Fine-Tuning of Diffusion Models
by: Uehara, Masatoshi, et al.
Published: (2024)
by: Uehara, Masatoshi, et al.
Published: (2024)
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
by: Uehara, Masatoshi, et al.
Published: (2024)
by: Uehara, Masatoshi, et al.
Published: (2024)
Iterative Distillation for Reward-Guided Fine-Tuning of Diffusion Models in Biomolecular Design
by: Su, Xingyu, et al.
Published: (2025)
by: Su, Xingyu, et al.
Published: (2025)
Dynamic Search for Inference-Time Alignment in Diffusion Models
by: Li, Xiner, et al.
Published: (2025)
by: Li, Xiner, et al.
Published: (2025)
Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review
by: Uehara, Masatoshi, et al.
Published: (2025)
by: Uehara, Masatoshi, et al.
Published: (2025)
Derivative-Free Guidance in Continuous and Discrete Diffusion Models with Soft Value-Based Decoding
by: Li, Xiner, et al.
Published: (2024)
by: Li, Xiner, et al.
Published: (2024)
Reward-Guided Iterative Refinement in Diffusion Models at Test-Time with Applications to Protein and DNA Design
by: Uehara, Masatoshi, et al.
Published: (2025)
by: Uehara, Masatoshi, et al.
Published: (2025)
MolCap-Arena: A Comprehensive Captioning Benchmark on Language-Enhanced Molecular Property Prediction
by: Edwards, Carl, et al.
Published: (2024)
by: Edwards, Carl, et al.
Published: (2024)
RAG-Enhanced Collaborative LLM Agents for Drug Discovery
by: Lee, Namkyeong, et al.
Published: (2025)
by: Lee, Namkyeong, et al.
Published: (2025)
Hierarchically branched diffusion models leverage dataset structure for class-conditional generation
by: Tseng, Alex M., et al.
Published: (2022)
by: Tseng, Alex M., et al.
Published: (2022)
A mechanistically interpretable neural network for regulatory genomics
by: Tseng, Alex M., et al.
Published: (2024)
by: Tseng, Alex M., et al.
Published: (2024)
Accurate and Efficient Structural Ensemble Generation of Macrocyclic Peptides using Internal Coordinate Diffusion
by: Grambow, Colin A., et al.
Published: (2023)
by: Grambow, Colin A., et al.
Published: (2023)
Cell Morphology-Guided Small Molecule Generation with GFlowNets
by: Lu, Stephen Zhewen, et al.
Published: (2024)
by: Lu, Stephen Zhewen, et al.
Published: (2024)
Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein Design
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
Efficient Fine-Tuning of Single-Cell Foundation Models Enables Zero-Shot Molecular Perturbation Prediction
by: Maleki, Sepideh, et al.
Published: (2024)
by: Maleki, Sepideh, et al.
Published: (2024)
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
by: Kuba, Jakub Grudzien, et al.
Published: (2024)
by: Kuba, Jakub Grudzien, et al.
Published: (2024)
Real-Time Execution of Action Chunking Flow Policies
by: Black, Kevin, et al.
Published: (2025)
by: Black, Kevin, et al.
Published: (2025)
Training-Time Action Conditioning for Efficient Real-Time Chunking
by: Black, Kevin, et al.
Published: (2025)
by: Black, Kevin, et al.
Published: (2025)
Training Diffusion Models with Reinforcement Learning
by: Black, Kevin, et al.
Published: (2023)
by: Black, Kevin, et al.
Published: (2023)
DC-W2S: Dual-Consensus Weak-to-Strong Training for Reliable Process Reward Modeling in Biological Reasoning
by: Chan, Chi-Min, et al.
Published: (2026)
by: Chan, Chi-Min, et al.
Published: (2026)
Flow Density Control: Generative Optimization Beyond Entropy-Regularized Fine-Tuning
by: De Santi, Riccardo, et al.
Published: (2025)
by: De Santi, Riccardo, et al.
Published: (2025)
Diffusion Guidance Is a Controllable Policy Improvement Operator
by: Frans, Kevin, et al.
Published: (2025)
by: Frans, Kevin, et al.
Published: (2025)
Q-SFT: Q-Learning for Language Models via Supervised Fine-Tuning
by: Hong, Joey, et al.
Published: (2024)
by: Hong, Joey, et al.
Published: (2024)
Regularized DeepIV with Model Selection
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Policy Gradient with Adaptive Entropy Annealing for Continual Fine-Tuning
by: Zhang, Yaqian, et al.
Published: (2026)
by: Zhang, Yaqian, et al.
Published: (2026)
Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data
by: Zhou, Zhiyuan, et al.
Published: (2024)
by: Zhou, Zhiyuan, et al.
Published: (2024)
AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents
by: De Brouwer, Edward, et al.
Published: (2026)
by: De Brouwer, Edward, et al.
Published: (2026)
Directly Fine-Tuning Diffusion Models on Differentiable Rewards
by: Clark, Kevin, et al.
Published: (2023)
by: Clark, Kevin, et al.
Published: (2023)
One Step Diffusion via Shortcut Models
by: Frans, Kevin, et al.
Published: (2024)
by: Frans, Kevin, et al.
Published: (2024)
Distributional Offline Policy Evaluation with Predictive Error Guarantees
by: Wu, Runzhe, et al.
Published: (2023)
by: Wu, Runzhe, et al.
Published: (2023)
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control
by: Zhang, Jiazheng, et al.
Published: (2026)
by: Zhang, Jiazheng, et al.
Published: (2026)
Contextualizing biological perturbation experiments through language
by: Wu, Menghua, et al.
Published: (2025)
by: Wu, Menghua, et al.
Published: (2025)
Universal Relation between Entropy and Kinetics
by: Sorkin, Benjamin, et al.
Published: (2022)
by: Sorkin, Benjamin, et al.
Published: (2022)
Toward the Identifiability of Comparative Deep Generative Models
by: Lopez, Romain, et al.
Published: (2024)
by: Lopez, Romain, et al.
Published: (2024)
Hölder type estimates for Gaussian multiplicative chaos
by: Huang, Yulai
Published: (2025)
by: Huang, Yulai
Published: (2025)
Big Cooperative Learning
by: Cong, Yulai
Published: (2024)
by: Cong, Yulai
Published: (2024)
Prohibicionismo, grupos sociales "a riesgo" y autoritarismo institucional: la censura social hacia los "microtraficantes"
by: Paolo Scalia
Published: (2005)
by: Paolo Scalia
Published: (2005)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
by: Tang, Grace, et al.
Published: (2024)
by: Tang, Grace, et al.
Published: (2024)
Similar Items
-
Feedback Efficient Online Fine-Tuning of Diffusion Models
by: Uehara, Masatoshi, et al.
Published: (2024) -
Adding Conditional Control to Diffusion Models with Reinforcement Learning
by: Zhao, Yulai, et al.
Published: (2024) -
Bridging Model-Based Optimization and Generative Modeling via Conservative Fine-Tuning of Diffusion Models
by: Uehara, Masatoshi, et al.
Published: (2024) -
Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
by: Uehara, Masatoshi, et al.
Published: (2024) -
Iterative Distillation for Reward-Guided Fine-Tuning of Diffusion Models in Biomolecular Design
by: Su, Xingyu, et al.
Published: (2025)