AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Junjie, Tuo, Yuxiang, Chen, Binghui, Zhong, Chongyang, Geng, Yifeng, Bo, Liefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AnyText2: Visual Text Generation and Editing With Customizable Attributes
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024)
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024)
von: He, Junjie, et al.
Veröffentlicht: (2024)
AnyText: Multilingual Visual Text Generation And Editing
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024)
VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
Strictly-ID-Preserved and Controllable Accessory Advertising Image Generation
von: Xue, Youze, et al.
Veröffentlicht: (2024)
von: Xue, Youze, et al.
Veröffentlicht: (2024)
ComFusion: Personalized Subject Generation in Multiple Specific Scenes From Single Image
von: Hong, Yan, et al.
Veröffentlicht: (2024)
von: Hong, Yan, et al.
Veröffentlicht: (2024)
OutfitAnyone: Ultra-high Quality Virtual Try-On for Any Clothing and Any Person
von: Sun, Ke, et al.
Veröffentlicht: (2024)
von: Sun, Ke, et al.
Veröffentlicht: (2024)
I4VGen: Image as Free Stepping Stone for Text-to-Video Generation
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
von: Guo, Xiefan, et al.
Veröffentlicht: (2024)
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
Textual Localization: Decomposing Multi-concept Images for Subject-Driven Text-to-Image Generation
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
von: Qiu, Lingteng, et al.
Veröffentlicht: (2024)
von: Qiu, Lingteng, et al.
Veröffentlicht: (2024)
Towards Fine-grained Interactive Segmentation in Images and Videos
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
ShoeModel: Learning to Wear on the User-specified Shoes via Diffusion Model
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
von: Chen, Binghui, et al.
Veröffentlicht: (2024)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
von: He, Chao, et al.
Veröffentlicht: (2025)
von: He, Chao, et al.
Veröffentlicht: (2025)
StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization
von: Gaur, Gopalji, et al.
Veröffentlicht: (2025)
von: Gaur, Gopalji, et al.
Veröffentlicht: (2025)
DCoAR: Deep Concept Injection into Unified Autoregressive Models for Personalized Text-to-Image Generation
von: Wu, Fangtai, et al.
Veröffentlicht: (2025)
von: Wu, Fangtai, et al.
Veröffentlicht: (2025)
PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards
von: Wang, Shulei, et al.
Veröffentlicht: (2025)
von: Wang, Shulei, et al.
Veröffentlicht: (2025)
Exploring Timeline Control for Facial Motion Generation
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
von: Ma, Yifeng, et al.
Veröffentlicht: (2025)
OmniWeaving: Towards Unified Video Generation with Free-form Composition and Reasoning
von: Pan, Kaihang, et al.
Veröffentlicht: (2026)
von: Pan, Kaihang, et al.
Veröffentlicht: (2026)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
von: He, Huiguo, et al.
Veröffentlicht: (2024)
von: He, Huiguo, et al.
Veröffentlicht: (2024)
MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
Geometric Disentanglement of Text Embeddings for Subject-Consistent Text-to-Image Generation using A Single Prompt
von: Li, Shangxun, et al.
Veröffentlicht: (2025)
von: Li, Shangxun, et al.
Veröffentlicht: (2025)
AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
von: Liu, Lin, et al.
Veröffentlicht: (2025)
von: Liu, Lin, et al.
Veröffentlicht: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
CustomVideo: Customizing Text-to-Video Generation with Multiple Subjects
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
Towards Unified 3D Hair Reconstruction from Single-View Portraits
von: Zheng, Yujian, et al.
Veröffentlicht: (2024)
von: Zheng, Yujian, et al.
Veröffentlicht: (2024)
Subject-Diffusion:Open Domain Personalized Text-to-Image Generation without Test-time Fine-tuning
von: Ma, Jian, et al.
Veröffentlicht: (2023)
von: Ma, Jian, et al.
Veröffentlicht: (2023)
AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea
von: Yu, Qifan, et al.
Veröffentlicht: (2024)
von: Yu, Qifan, et al.
Veröffentlicht: (2024)
Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
von: Jang, Sangwon, et al.
Veröffentlicht: (2024)
ATI: Any Trajectory Instruction for Controllable Video Generation
von: Wang, Angtian, et al.
Veröffentlicht: (2025)
von: Wang, Angtian, et al.
Veröffentlicht: (2025)
Design Your Ad: Personalized Advertising Image and Text Generation with Unified Autoregressive Models
von: Xu, Yexing, et al.
Veröffentlicht: (2026)
von: Xu, Yexing, et al.
Veröffentlicht: (2026)
GAP: Gaussianize Any Point Clouds with Text Guidance
von: Zhang, Weiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Weiqi, et al.
Veröffentlicht: (2025)
UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations
von: Zhao, Yaqi, et al.
Veröffentlicht: (2026)
von: Zhao, Yaqi, et al.
Veröffentlicht: (2026)
Any2Any: Unified Arbitrary Modality Translation for Remote Sensing
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
Single Image, Any Face: Generalisable 3D Face Generation
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
von: Wang, Wenqing, et al.
Veröffentlicht: (2024)
Unified Personalized Understanding, Generating and Editing
von: Zhong, Yu, et al.
Veröffentlicht: (2026)
von: Zhong, Yu, et al.
Veröffentlicht: (2026)
StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
von: Zhou, Zhengguang, et al.
Veröffentlicht: (2024)
von: Zhou, Zhengguang, et al.
Veröffentlicht: (2024)
Unified Prompt Attack Against Text-to-Image Generation Models
von: Peng, Duo, et al.
Veröffentlicht: (2025)
von: Peng, Duo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AnyText2: Visual Text Generation and Editing With Customizable Attributes
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2024) -
UniPortrait: A Unified Framework for Identity-Preserving Single- and Multi-Human Image Personalization
von: He, Junjie, et al.
Veröffentlicht: (2024) -
AnyText: Multilingual Visual Text Generation And Editing
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023) -
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
von: Chen, DuoSheng, et al.
Veröffentlicht: (2024) -
VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
von: Chen, Binghui, et al.
Veröffentlicht: (2024)