Resolving Multi-Condition Confusion for Finetuning-Free Personalized Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Qihan, Fu, Siming, Liu, Jinlong, Jiang, Hao, Yu, Yipeng, Song, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PatchDPO: Patch-level DPO for Finetuning-free Personalized Image Generation
by: Huang, Qihan, et al.
Published: (2024)
by: Huang, Qihan, et al.
Published: (2024)
MS-Diffusion: Multi-subject Zero-shot Image Personalization with Layout Guidance
by: Wang, Xierui, et al.
Published: (2024)
by: Wang, Xierui, et al.
Published: (2024)
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
by: Zeng, Yu, et al.
Published: (2024)
by: Zeng, Yu, et al.
Published: (2024)
Finetuning-Free Personalization of Text to Image Generation via Hypernetworks
by: Shrestha, Sagar, et al.
Published: (2025)
by: Shrestha, Sagar, et al.
Published: (2025)
MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement
by: She, Dong, et al.
Published: (2025)
by: She, Dong, et al.
Published: (2025)
Boosting MLLM Reasoning with Text-Debiased Hint-GRPO
by: Huang, Qihan, et al.
Published: (2025)
by: Huang, Qihan, et al.
Published: (2025)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
by: Jin, Qiaoqiao, et al.
Published: (2025)
by: Jin, Qiaoqiao, et al.
Published: (2025)
Evaluating Attribute Confusion in Fashion Text-to-Image Generation
by: Liu, Ziyue, et al.
Published: (2025)
by: Liu, Ziyue, et al.
Published: (2025)
MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation
by: Ma, Xiaoxiao, et al.
Published: (2026)
by: Ma, Xiaoxiao, et al.
Published: (2026)
PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning
by: Liu, Jinlong, et al.
Published: (2026)
by: Liu, Jinlong, et al.
Published: (2026)
ConceptGuard: Continual Personalized Text-to-Image Generation with Forgetting and Confusion Mitigation
by: Guo, Zirun, et al.
Published: (2025)
by: Guo, Zirun, et al.
Published: (2025)
DreamCache: Finetuning-Free Lightweight Personalized Image Generation via Feature Caching
by: Aiello, Emanuele, et al.
Published: (2024)
by: Aiello, Emanuele, et al.
Published: (2024)
Towards Enhanced Image Generation Via Multi-modal Chain of Thought in Unified Generative Models
by: Wang, Yi, et al.
Published: (2025)
by: Wang, Yi, et al.
Published: (2025)
VideoDreamer: Customized Multi-Subject Text-to-Video Generation with Disen-Mix Finetuning on Language-Video Foundation Models
by: Chen, Hong, et al.
Published: (2023)
by: Chen, Hong, et al.
Published: (2023)
LoFA: Learning to Predict Personalized Priors for Fast Adaptation of Visual Generative Models
by: Hao, Yiming, et al.
Published: (2025)
by: Hao, Yiming, et al.
Published: (2025)
Multi-focal Conditioned Latent Diffusion for Person Image Synthesis
by: Liu, Jiaqi, et al.
Published: (2025)
by: Liu, Jiaqi, et al.
Published: (2025)
PixelPonder: Dynamic Patch Adaptation for Enhanced Multi-Conditional Text-to-Image Generation
by: Pan, Yanjie, et al.
Published: (2025)
by: Pan, Yanjie, et al.
Published: (2025)
Impact of Clinical Image Quality on Efficient Foundation Model Finetuning
by: Tang, Yucheng, et al.
Published: (2025)
by: Tang, Yucheng, et al.
Published: (2025)
Multi-View Representation is What You Need for Point-Cloud Pre-Training
by: Yan, Siming, et al.
Published: (2023)
by: Yan, Siming, et al.
Published: (2023)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
by: Ding, Ganggui, et al.
Published: (2024)
by: Ding, Ganggui, et al.
Published: (2024)
LG-CAV: Train Any Concept Activation Vector with Language Guidance
by: Huang, Qihan, et al.
Published: (2024)
by: Huang, Qihan, et al.
Published: (2024)
DynamicControl: Adaptive Condition Selection for Improved Text-to-Image Generation
by: He, Qingdong, et al.
Published: (2024)
by: He, Qingdong, et al.
Published: (2024)
DeconfuseTrack:Dealing with Confusion for Multi-Object Tracking
by: Huang, Cheng, et al.
Published: (2024)
by: Huang, Cheng, et al.
Published: (2024)
SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation
by: Ouyang, Liangyang, et al.
Published: (2026)
by: Ouyang, Liangyang, et al.
Published: (2026)
On the Concept Trustworthiness in Concept Bottleneck Models
by: Huang, Qihan, et al.
Published: (2024)
by: Huang, Qihan, et al.
Published: (2024)
Enhanced Multi-Scale Cross-Attention for Person Image Generation
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
Disentangling to Re-couple: Resolving the Similarity-Controllability Paradox in Subject-Driven Text-to-Image Generation
by: Li, Shuang, et al.
Published: (2026)
by: Li, Shuang, et al.
Published: (2026)
V2X-DGW: Domain Generalization for Multi-agent Perception under Adverse Weather Conditions
by: Li, Baolu, et al.
Published: (2024)
by: Li, Baolu, et al.
Published: (2024)
Controllable Human Image Generation with Personalized Multi-Garments
by: Choi, Yisol, et al.
Published: (2024)
by: Choi, Yisol, et al.
Published: (2024)
When Eyes and Ears Disagree: Can MLLMs Discern Audio-Visual Confusion?
by: Ye, Qilang, et al.
Published: (2025)
by: Ye, Qilang, et al.
Published: (2025)
Tuning-Free Image Customization with Image and Text Guidance
by: Li, Pengzhi, et al.
Published: (2024)
by: Li, Pengzhi, et al.
Published: (2024)
Super-Resolving Blurry Images with Events
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
ProtoPFormer: Concentrating on Prototypical Parts in Vision Transformers for Interpretable Image Recognition
by: Xue, Mengqi, et al.
Published: (2022)
by: Xue, Mengqi, et al.
Published: (2022)
Jointly Conditioned Diffusion Model for Multi-View Pose-Guided Person Image Synthesis
by: Xie, Chengyu, et al.
Published: (2025)
by: Xie, Chengyu, et al.
Published: (2025)
De-Confusing Pseudo-Labels in Source-Free Domain Adaptation
by: Diamant, Idit, et al.
Published: (2024)
by: Diamant, Idit, et al.
Published: (2024)
Context-Aware Autoregressive Models for Multi-Conditional Image Generation
by: Chen, Yixiao, et al.
Published: (2025)
by: Chen, Yixiao, et al.
Published: (2025)
ML-CLIPSim: Multi-Layer CLIP Similarity for Machine-Oriented Image Quality
by: Ding, Feng, et al.
Published: (2026)
by: Ding, Feng, et al.
Published: (2026)
CustomVideoX: 3D Reference Attention Driven Dynamic Adaptation for Zero-Shot Customized Video Diffusion Transformers
by: She, D., et al.
Published: (2025)
by: She, D., et al.
Published: (2025)
Prompt-Free Conditional Diffusion for Multi-object Image Augmentation
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Syn-GRPO: Self-Evolving Data Synthesis for MLLM Perception Reasoning
by: Huang, Qihan, et al.
Published: (2025)
by: Huang, Qihan, et al.
Published: (2025)
Similar Items
-
PatchDPO: Patch-level DPO for Finetuning-free Personalized Image Generation
by: Huang, Qihan, et al.
Published: (2024) -
MS-Diffusion: Multi-subject Zero-shot Image Personalization with Layout Guidance
by: Wang, Xierui, et al.
Published: (2024) -
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
by: Zeng, Yu, et al.
Published: (2024) -
Finetuning-Free Personalization of Text to Image Generation via Hypernetworks
by: Shrestha, Sagar, et al.
Published: (2025) -
MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement
by: She, Dong, et al.
Published: (2025)