MaxFusion: Plug&Play Multi-Modal Generation in Text-to-Image Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nair, Nithin Gopalakrishnan, Valanarasu, Jeya Maria Jose, Patel, Vishal M |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
Dreamguider: Improved Training free Diffusion-based Conditional Generation
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
Scale-Wise VAR is Secretly Discrete Diffusion
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025)
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2023)
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2023)
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2022)
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2022)
PETALface: Parameter Efficient Transfer Learning for Low-resolution Face Recognition
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
Scaling Transformer-Based Novel View Synthesis Models with Token Disentanglement and Synthetic Data
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2025)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2025)
Auto-Generating Weak Labels for Real & Synthetic Data to Improve Label-Scarce Medical Image Segmentation
von: Deshpande, Tanvi, et al.
Veröffentlicht: (2024)
von: Deshpande, Tanvi, et al.
Veröffentlicht: (2024)
Text-to-Image Rectified Flow as Plug-and-Play Priors
von: Yang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Yang, Xiaofeng, et al.
Veröffentlicht: (2024)
Plug-and-Play Diffusion Distillation
von: Hsiao, Yi-Ting, et al.
Veröffentlicht: (2024)
von: Hsiao, Yi-Ting, et al.
Veröffentlicht: (2024)
Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models
von: Huang, Gexin, et al.
Veröffentlicht: (2026)
von: Huang, Gexin, et al.
Veröffentlicht: (2026)
Text-DiFuse: An Interactive Multi-Modal Image Fusion Framework based on Text-modulated Diffusion Model
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
von: Lee, Saehyung, et al.
Veröffentlicht: (2024)
SpeedUpNet: A Plug-and-Play Adapter Network for Accelerating Text-to-Image Diffusion Models
von: Chai, Weilong, et al.
Veröffentlicht: (2023)
von: Chai, Weilong, et al.
Veröffentlicht: (2023)
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
von: Zeng, Yu, et al.
Veröffentlicht: (2024)
CleanStyle: Plug-and-Play Style Conditioning Purification for Text-to-Image Stylization
von: Feng, Xiaoman, et al.
Veröffentlicht: (2026)
von: Feng, Xiaoman, et al.
Veröffentlicht: (2026)
Plug-and-Play Multi-Concept Adaptive Blending for High-Fidelity Text-to-Image Synthesis
von: Woo, Young-Beom
Veröffentlicht: (2025)
von: Woo, Young-Beom
Veröffentlicht: (2025)
STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
Plug-and-Play Interpretable Responsible Text-to-Image Generation via Dual-Space Multi-facet Concept Control
von: Azam, Basim, et al.
Veröffentlicht: (2025)
von: Azam, Basim, et al.
Veröffentlicht: (2025)
BrushNet: A Plug-and-Play Image Inpainting Model with Decomposed Dual-Branch Diffusion
von: Ju, Xuan, et al.
Veröffentlicht: (2024)
von: Ju, Xuan, et al.
Veröffentlicht: (2024)
ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop
von: Li, Shuangzhi, et al.
Veröffentlicht: (2026)
von: Li, Shuangzhi, et al.
Veröffentlicht: (2026)
PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2023)
von: Zhang, Yiming, et al.
Veröffentlicht: (2023)
Integrating Reweighted Least Squares with Plug-and-Play Diffusion Priors for Noisy Image Restoration
von: Li, Ji, et al.
Veröffentlicht: (2025)
von: Li, Ji, et al.
Veröffentlicht: (2025)
Unlocking Robust Segmentation Across All Age Groups via Continual Learning
von: Liu, Chih-Ying, et al.
Veröffentlicht: (2024)
von: Liu, Chih-Ying, et al.
Veröffentlicht: (2024)
Not All Tokens Need 40 Steps: Heterogeneous Step Allocation in Diffusion Transformers for Efficient Video Generation
von: Chu, Ernie, et al.
Veröffentlicht: (2026)
von: Chu, Ernie, et al.
Veröffentlicht: (2026)
FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
von: Gao, Xiang, et al.
Veröffentlicht: (2024)
TDiff: Thermal Plug-And-Play Prior with Patch-Based Diffusion
von: Dashpute, Piyush, et al.
Veröffentlicht: (2025)
von: Dashpute, Piyush, et al.
Veröffentlicht: (2025)
NFCDS: A Plug-and-Play Noise Frequency-Controlled Diffusion Sampling Strategy for Image Restoration
von: Wang, Zhen, et al.
Veröffentlicht: (2026)
von: Wang, Zhen, et al.
Veröffentlicht: (2026)
Principled Probabilistic Imaging using Diffusion Models as Plug-and-Play Priors
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
P$^2$HCT: Plug-and-Play Hierarchical C2F Transformer for Multi-Scale Feature Fusion
von: Hu, Junyi, et al.
Veröffentlicht: (2025)
von: Hu, Junyi, et al.
Veröffentlicht: (2025)
Time-to-Event Pretraining for 3D Medical Imaging
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
MGML: A Plug-and-Play Meta-Guided Multi-Modal Learning Framework for Incomplete Multimodal Brain Tumor Segmentation
von: Zou, Yulong, et al.
Veröffentlicht: (2025)
von: Zou, Yulong, et al.
Veröffentlicht: (2025)
CBNet: A Plug-and-Play Network for Segmentation-Based Scene Text Detection
von: Zhao, Xi, et al.
Veröffentlicht: (2022)
von: Zhao, Xi, et al.
Veröffentlicht: (2022)
Equivariant Multi-Modality Image Fusion
von: Zhao, Zixiang, et al.
Veröffentlicht: (2023)
von: Zhao, Zixiang, et al.
Veröffentlicht: (2023)
PADS: Plug-and-Play 3D Human Pose Analysis via Diffusion Generative Modeling
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
SR$^{2}$-Net: A General Plug-and-Play Model for Spectral Refinement in Hyperspectral Image Super-Resolution
von: He, Ji-Xuan, et al.
Veröffentlicht: (2026)
von: He, Ji-Xuan, et al.
Veröffentlicht: (2026)
Your Pre-trained Diffusion Model Secretly Knows Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
Plug-and-Play Context Feature Reuse for Efficient Masked Generation
von: Liu, Xuejie, et al.
Veröffentlicht: (2025)
von: Liu, Xuejie, et al.
Veröffentlicht: (2025)
EMMA: Your Text-to-Image Diffusion Model Can Secretly Accept Multi-Modal Prompts
von: Han, Yucheng, et al.
Veröffentlicht: (2024)
von: Han, Yucheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024) -
Dreamguider: Improved Training free Diffusion-based Conditional Generation
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024) -
Scale-Wise VAR is Secretly Discrete Diffusion
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025) -
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2023) -
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)