MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yuxiang, Ji, Zhilong, Bai, Jinfeng, Zhang, Hongzhi, Zhang, Lei, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
by: Wang, Zihao, et al.
Published: (2026)
by: Wang, Zihao, et al.
Published: (2026)
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
by: Feng, Kailai, et al.
Published: (2024)
by: Feng, Kailai, et al.
Published: (2024)
MDIQA: Unified Image Quality Assessment for Multi-dimensional Evaluation and Restoration
by: Yao, Shunyu, et al.
Published: (2025)
by: Yao, Shunyu, et al.
Published: (2025)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024)
by: Zhang, Yabo, et al.
Published: (2024)
LLM as a Complementary Optimizer to Gradient Descent: A Case Study in Prompt Tuning
by: Guo, Zixian, et al.
Published: (2024)
by: Guo, Zixian, et al.
Published: (2024)
ACE: Anti-Editing Concept Erasure in Text-to-Image Models
by: Wang, Zihao, et al.
Published: (2025)
by: Wang, Zihao, et al.
Published: (2025)
Face2Diffusion for Fast and Editable Face Personalization
by: Shiohara, Kaede, et al.
Published: (2024)
by: Shiohara, Kaede, et al.
Published: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
by: Zhang, Zhilu, et al.
Published: (2024)
by: Zhang, Zhilu, et al.
Published: (2024)
TimeWeaver: Age-Consistent Reference-Based Face Restoration with Identity Preservation
by: Song, Teer, et al.
Published: (2026)
by: Song, Teer, et al.
Published: (2026)
SplatWeaver: Learning to Allocate Gaussian Primitives for Generalizable Novel View Synthesis
by: Wan, Yecong, et al.
Published: (2026)
by: Wan, Yecong, et al.
Published: (2026)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
Improving Image Restoration through Removing Degradations in Textual Representations
by: Lin, Jingbo, et al.
Published: (2023)
by: Lin, Jingbo, et al.
Published: (2023)
PLACE: Adaptive Layout-Semantic Fusion for Semantic Image Synthesis
by: Lv, Zhengyao, et al.
Published: (2024)
by: Lv, Zhengyao, et al.
Published: (2024)
UniRestorer: Universal Image Restoration via Adaptively Estimating Image Degradation at Proper Granularity
by: Lin, Jingbo, et al.
Published: (2024)
by: Lin, Jingbo, et al.
Published: (2024)
Pseudo-Label Guided Real-World Image De-weathering: A Learning Framework with Imperfect Supervision
by: Xu, Heming, et al.
Published: (2025)
by: Xu, Heming, et al.
Published: (2025)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
by: Wang, Zikai, et al.
Published: (2026)
by: Wang, Zikai, et al.
Published: (2026)
DreamSalon: A Staged Diffusion Framework for Preserving Identity-Context in Editable Face Generation
by: Lin, Haonan, et al.
Published: (2024)
by: Lin, Haonan, et al.
Published: (2024)
FlashFace: Human Image Personalization with High-fidelity Identity Preservation
by: Zhang, Shilong, et al.
Published: (2024)
by: Zhang, Shilong, et al.
Published: (2024)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
by: Zhang, Haoze, et al.
Published: (2025)
by: Zhang, Haoze, et al.
Published: (2025)
AnyText: Multilingual Visual Text Generation And Editing
by: Tuo, Yuxiang, et al.
Published: (2023)
by: Tuo, Yuxiang, et al.
Published: (2023)
Explicit Relational Reasoning Network for Scene Text Detection
by: Su, Yuchen, et al.
Published: (2024)
by: Su, Yuchen, et al.
Published: (2024)
Taming Transformer for Emotion-Controllable Talking Face Generation
by: Zhang, Ziqi, et al.
Published: (2025)
by: Zhang, Ziqi, et al.
Published: (2025)
ScrollScape: Unlocking 32K Image Generation With Video Diffusion Priors
by: Yu, Haodong, et al.
Published: (2026)
by: Yu, Haodong, et al.
Published: (2026)
Self-Supervised High Dynamic Range Imaging with Multi-Exposure Images in Dynamic Scenes
by: Zhang, Zhilu, et al.
Published: (2023)
by: Zhang, Zhilu, et al.
Published: (2023)
Towards a Simultaneous and Granular Identity-Expression Control in Personalized Face Generation
by: Liu, Renshuai, et al.
Published: (2024)
by: Liu, Renshuai, et al.
Published: (2024)
CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers
by: Chen, Weidong, et al.
Published: (2026)
by: Chen, Weidong, et al.
Published: (2026)
3D Space as a Scratchpad for Editable Text-to-Image Generation
by: Saha, Oindrila, et al.
Published: (2026)
by: Saha, Oindrila, et al.
Published: (2026)
LLM-Enabled Style and Content Regularization for Personalized Text-to-Image Generation
by: Yu, Anran, et al.
Published: (2025)
by: Yu, Anran, et al.
Published: (2025)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
by: Yang, Feng, et al.
Published: (2025)
by: Yang, Feng, et al.
Published: (2025)
Are Watermarked Images Editable? SafeMark for Watermark-Preserving Text-Guided Image Editing
by: Wu, Xiaodong, et al.
Published: (2026)
by: Wu, Xiaodong, et al.
Published: (2026)
Image Demoiréing Using Dual Camera Fusion on Mobile Phones
by: Mei, Yanting, et al.
Published: (2025)
by: Mei, Yanting, et al.
Published: (2025)
EditID: Training-Free Editable ID Customization for Text-to-Image Generation
by: Li, Guandong, et al.
Published: (2025)
by: Li, Guandong, et al.
Published: (2025)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
by: Dong, Bowen, et al.
Published: (2024)
by: Dong, Bowen, et al.
Published: (2024)
Responsible Visual Editing
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
Visual-O1: Understanding Ambiguous Instructions via Multi-modal Multi-turn Chain-of-thoughts Reasoning
by: Ni, Minheng, et al.
Published: (2024)
by: Ni, Minheng, et al.
Published: (2024)
OmniSplat: Taming Feed-Forward 3D Gaussian Splatting for Omnidirectional Images with Editable Capabilities
by: Lee, Suyoung, et al.
Published: (2024)
by: Lee, Suyoung, et al.
Published: (2024)
AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution
by: Wang, Zehao, et al.
Published: (2026)
by: Wang, Zehao, et al.
Published: (2026)
Source Prompt Disentangled Inversion for Boosting Image Editability with Diffusion Models
by: Li, Ruibin, et al.
Published: (2024)
by: Li, Ruibin, et al.
Published: (2024)
Similar Items
-
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025) -
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
by: Wang, Zihao, et al.
Published: (2026) -
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
by: Feng, Kailai, et al.
Published: (2024) -
MDIQA: Unified Image Quality Assessment for Multi-dimensional Evaluation and Restoration
by: Yao, Shunyu, et al.
Published: (2025) -
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024)