Steer Away From Mode Collisions: Improving Composition In Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Dutta, Debottam, Chen, Jianchong, Rajagopalan, Rajalaxmi, Wei, Yu-Lin, Choudhury, Romit Roy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Image Generation via Human-in-the-loop Bayesian Optimization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
Dependency-Aware Discrete Diffusion for Scene Graph Generation
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
Learning Energy-based Variational Latent Prior for VAEs
by: Dutta, Debottam, et al.
Published: (2025)
by: Dutta, Debottam, et al.
Published: (2025)
Kernel Learning for Sample Constrained Black-Box Optimization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025)
Sample-Constrained Black Box Optimization for Audio Personalization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025)
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
Multi-Source Music Generation with Latent Diffusion
by: Xu, Zhongweiyang, et al.
Published: (2024)
by: Xu, Zhongweiyang, et al.
Published: (2024)
Discrete Langevin-Inspired Posterior Sampling
by: Amballa, Chaitanya, et al.
Published: (2026)
by: Amballa, Chaitanya, et al.
Published: (2026)
Contrastive Diffusion Guidance for Spatial Inverse Problems
by: Basu, Sattwik, et al.
Published: (2025)
by: Basu, Sattwik, et al.
Published: (2025)
Steering Away from Memorization: Reachability-Constrained Reinforcement Learning for Text-to-Image Diffusion
by: Karnik, Sathwik, et al.
Published: (2026)
by: Karnik, Sathwik, et al.
Published: (2026)
Estimating Multi-chirp Parameters using Curvature-guided Langevin Monte Carlo
by: Basu, Sattwik, et al.
Published: (2025)
by: Basu, Sattwik, et al.
Published: (2025)
Attention Shift: Steering AI Away from Unsafe Content
by: Garg, Shivank, et al.
Published: (2024)
by: Garg, Shivank, et al.
Published: (2024)
Can NeRFs See without Cameras?
by: Amballa, Chaitanya, et al.
Published: (2025)
by: Amballa, Chaitanya, et al.
Published: (2025)
Improving Compositional Generation with Diffusion Models Using Lift Scores
by: Yu, Chenning, et al.
Published: (2025)
by: Yu, Chenning, et al.
Published: (2025)
Image Inpainting via Tractable Steering of Diffusion Models
by: Liu, Anji, et al.
Published: (2023)
by: Liu, Anji, et al.
Published: (2023)
Steering Guidance for Personalized Text-to-Image Diffusion Models
by: Park, Sunghyun, et al.
Published: (2025)
by: Park, Sunghyun, et al.
Published: (2025)
ATHENA: Adaptive Test-Time Steering for Improving Count Fidelity in Diffusion Models
by: Sepehri, Mohammad Shahab, et al.
Published: (2026)
by: Sepehri, Mohammad Shahab, et al.
Published: (2026)
Model Steering: Learning with a Reference Model Improves Generalization Bounds and Scaling Laws
by: Wei, Xiyuan, et al.
Published: (2025)
by: Wei, Xiyuan, et al.
Published: (2025)
Improving Predictive Confidence in Medical Imaging via Online Label Smoothing
by: Choudhury, Kushan, et al.
Published: (2025)
by: Choudhury, Kushan, et al.
Published: (2025)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
by: Singhal, Raghav, et al.
Published: (2025)
by: Singhal, Raghav, et al.
Published: (2025)
RealCompo: Balancing Realism and Compositionality Improves Text-to-Image Diffusion Models
by: Zhang, Xinchen, et al.
Published: (2024)
by: Zhang, Xinchen, et al.
Published: (2024)
Localized PCA-Net Neural Operators for Scalable Solution Reconstruction of Elliptic PDEs
by: Dhingra, Mrigank, et al.
Published: (2025)
by: Dhingra, Mrigank, et al.
Published: (2025)
Continual Personalization for Diffusion Models
by: Liao, Yu-Chien, et al.
Published: (2025)
by: Liao, Yu-Chien, et al.
Published: (2025)
Compositional Image Decomposition with Diffusion Models
by: Su, Jocelin, et al.
Published: (2024)
by: Su, Jocelin, et al.
Published: (2024)
SteerVLM: Robust Model Control through Lightweight Activation Steering for Vision Language Models
by: Sivakumar, Anushka, et al.
Published: (2025)
by: Sivakumar, Anushka, et al.
Published: (2025)
Steering Away from Harm: An Adaptive Approach to Defending Vision Language Model Against Jailbreaks
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
SlimDiff: Training-Free, Activation-Guided Hands-free Slimming of Diffusion Models
by: Roy, Arani, et al.
Published: (2025)
by: Roy, Arani, et al.
Published: (2025)
Emergent Natural Language with Communication Games for Improving Image Captioning Capabilities without Additional Data
by: Dutta, Parag, et al.
Published: (2025)
by: Dutta, Parag, et al.
Published: (2025)
Neural Flow Diffusion Models: Learnable Forward Process for Improved Diffusion Modelling
by: Bartosh, Grigory, et al.
Published: (2024)
by: Bartosh, Grigory, et al.
Published: (2024)
Improving Adversarial Energy-Based Model via Diffusion Process
by: Geng, Cong, et al.
Published: (2024)
by: Geng, Cong, et al.
Published: (2024)
CountSteer: Steering Attention for Object Counting in Diffusion Models
by: Boo, Hyemin, et al.
Published: (2025)
by: Boo, Hyemin, et al.
Published: (2025)
Your Pre-trained Diffusion Model Secretly Knows Restoration
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Recognition
by: Rahimi, Parsa, et al.
Published: (2025)
by: Rahimi, Parsa, et al.
Published: (2025)
From Text to Pose to Image: Improving Diffusion Model Control and Quality
by: Bonnet, Clément, et al.
Published: (2024)
by: Bonnet, Clément, et al.
Published: (2024)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
by: Wei, Yuanxin, et al.
Published: (2025)
by: Wei, Yuanxin, et al.
Published: (2025)
DC-Merge: Improving Model Merging with Directional Consistency
by: Zhang, Han-Chen, et al.
Published: (2026)
by: Zhang, Han-Chen, et al.
Published: (2026)
Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models
by: Gan, Woody Haosheng, et al.
Published: (2025)
by: Gan, Woody Haosheng, et al.
Published: (2025)
Adversarially Robust Industrial Anomaly Detection Through Diffusion Model
by: Cao, Yuanpu, et al.
Published: (2024)
by: Cao, Yuanpu, et al.
Published: (2024)
Dynamical Diffusion: Learning Temporal Dynamics with Diffusion Models
by: Guo, Xingzhuo, et al.
Published: (2025)
by: Guo, Xingzhuo, et al.
Published: (2025)
Similar Items
-
Personalized Image Generation via Human-in-the-loop Bayesian Optimization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026) -
Dependency-Aware Discrete Diffusion for Scene Graph Generation
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026) -
Learning Energy-based Variational Latent Prior for VAEs
by: Dutta, Debottam, et al.
Published: (2025) -
Kernel Learning for Sample Constrained Black-Box Optimization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025) -
Sample-Constrained Black Box Optimization for Audio Personalization
by: Rajagopalan, Rajalaxmi, et al.
Published: (2025)