StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Zhengguang, Li, Jing, Li, Huaxia, Chen, Nemo, Tang, Xu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Target-Driven Distillation: Consistency Distillation with Target Timestep Selection and Decoupled Guidance
by: Wang, Cunzheng, et al.
Published: (2024)
by: Wang, Cunzheng, et al.
Published: (2024)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
by: Xiang, Qiang, et al.
Published: (2025)
by: Xiang, Qiang, et al.
Published: (2025)
InfinityStory: Unlimited Video Generation with World Consistency and Character-Aware Shot Transitions
by: Elmoghany, Mohamed, et al.
Published: (2026)
by: Elmoghany, Mohamed, et al.
Published: (2026)
ReMix: Towards a Unified View of Consistent Character Generation and Editing
by: Zhou, Benjia, et al.
Published: (2025)
by: Zhou, Benjia, et al.
Published: (2025)
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
by: Wang, Qinghe, et al.
Published: (2024)
by: Wang, Qinghe, et al.
Published: (2024)
PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering
by: Gao, Yifan, et al.
Published: (2025)
by: Gao, Yifan, et al.
Published: (2025)
ReDiStory: Region-Disentangled Diffusion for Consistent Visual Story Generation
by: Sarkar, Ayushman, et al.
Published: (2026)
by: Sarkar, Ayushman, et al.
Published: (2026)
CharaConsist: Fine-Grained Consistent Character Generation
by: Wang, Mengyu, et al.
Published: (2025)
by: Wang, Mengyu, et al.
Published: (2025)
InstantCharacter: Personalize Any Characters with a Scalable Diffusion Transformer Framework
by: Tao, Jiale, et al.
Published: (2025)
by: Tao, Jiale, et al.
Published: (2025)
SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
Flux-Sculptor: Text-Driven Rich-Attribute Portrait Editing through Decomposed Spatial Flow Control
by: He, Tianyao, et al.
Published: (2025)
by: He, Tianyao, et al.
Published: (2025)
Wan-Animate: Unified Character Animation and Replacement with Holistic Replication
by: Cheng, Gang, et al.
Published: (2025)
by: Cheng, Gang, et al.
Published: (2025)
Infinite-Story: A Training-Free Consistent Text-to-Image Generation
by: Park, Jihun, et al.
Published: (2025)
by: Park, Jihun, et al.
Published: (2025)
CharacterShot: Controllable and Consistent 4D Character Animation
by: Gao, Junyao, et al.
Published: (2025)
by: Gao, Junyao, et al.
Published: (2025)
StoryWeaver: A Unified World Model for Knowledge-Enhanced Story Character Customization
by: Zhang, Jinlu, et al.
Published: (2024)
by: Zhang, Jinlu, et al.
Published: (2024)
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
by: Li, Mingxiao, et al.
Published: (2025)
by: Li, Mingxiao, et al.
Published: (2025)
Bringing The Consistency Gap: Explicit Structured Memory for Interleaved Image-Text Generation
by: Lin, Zeteng, et al.
Published: (2025)
by: Lin, Zeteng, et al.
Published: (2025)
DynamicFace: High-Quality and Consistent Face Swapping for Image and Video using Composable 3D Facial Priors
by: Wang, Runqi, et al.
Published: (2025)
by: Wang, Runqi, et al.
Published: (2025)
Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition
by: Zhou, Bangbang, et al.
Published: (2024)
by: Zhou, Bangbang, et al.
Published: (2024)
Bring Your Own Character: A Holistic Solution for Automatic Facial Animation Generation of Customized Characters
by: Bai, Zechen, et al.
Published: (2024)
by: Bai, Zechen, et al.
Published: (2024)
Persistent Story World Simulation with Continuous Character Customization
by: Zhang, Jinlu, et al.
Published: (2026)
by: Zhang, Jinlu, et al.
Published: (2026)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
by: Banerjee, Ayan, et al.
Published: (2025)
by: Banerjee, Ayan, et al.
Published: (2025)
PanoDreamer: Consistent Text to 360-Degree Scene Generation
by: Xiong, Zhexiao, et al.
Published: (2025)
by: Xiong, Zhexiao, et al.
Published: (2025)
TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation
by: Cheng, Junhao, et al.
Published: (2024)
by: Cheng, Junhao, et al.
Published: (2024)
BachVid: Training-Free Video Generation with Consistent Background and Character
by: Yan, Han, et al.
Published: (2025)
by: Yan, Han, et al.
Published: (2025)
Gloria: Consistent Character Video Generation via Content Anchors
by: Yang, Yuhang, et al.
Published: (2026)
by: Yang, Yuhang, et al.
Published: (2026)
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
by: He, Junjie, et al.
Published: (2025)
by: He, Junjie, et al.
Published: (2025)
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
by: Zhou, Yupeng, et al.
Published: (2024)
by: Zhou, Yupeng, et al.
Published: (2024)
Structural Energy Guidance for View-Consistent Text-to-3D Generation
by: Zhang, Qing, et al.
Published: (2026)
by: Zhang, Qing, et al.
Published: (2026)
EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creation
by: Yang, Shiyuan, et al.
Published: (2026)
by: Yang, Shiyuan, et al.
Published: (2026)
StableGarment: Garment-Centric Generation via Stable Diffusion
by: Wang, Rui, et al.
Published: (2024)
by: Wang, Rui, et al.
Published: (2024)
HoloDreamer: Holistic 3D Panoramic World Generation from Text Descriptions
by: Zhou, Haiyang, et al.
Published: (2024)
by: Zhou, Haiyang, et al.
Published: (2024)
ContextAnyone: Context-Aware Diffusion for Character-Consistent Text-to-Video Generation
by: Mai, Ziyang, et al.
Published: (2025)
by: Mai, Ziyang, et al.
Published: (2025)
Privacy Protection Against Personalized Text-to-Image Synthesis via Cross-image Consistency Constraints
by: Wang, Guanyu, et al.
Published: (2025)
by: Wang, Guanyu, et al.
Published: (2025)
Lights, Camera, Consistency: A Multistage Pipeline for Character-Stable AI Video Stories
by: Jain, Chayan, et al.
Published: (2025)
by: Jain, Chayan, et al.
Published: (2025)
Character Mixing for Video Generation
by: Liao, Tingting, et al.
Published: (2025)
by: Liao, Tingting, et al.
Published: (2025)
Towards Reliable and Holistic Visual In-Context Learning Prompt Selection
by: Wu, Wenxiao, et al.
Published: (2025)
by: Wu, Wenxiao, et al.
Published: (2025)
OmniSonic: Towards Universal and Holistic Audio Generation from Video and Text
by: Pian, Weiguo, et al.
Published: (2026)
by: Pian, Weiguo, et al.
Published: (2026)
2K-Characters-10K-Stories: A Quality-Gated Stylized Narrative Dataset with Disentangled Control and Sequence Consistency
by: Yin, Xingxi, et al.
Published: (2025)
by: Yin, Xingxi, et al.
Published: (2025)
Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset
by: Chen, Zhuowei, et al.
Published: (2025)
by: Chen, Zhuowei, et al.
Published: (2025)
Similar Items
-
Target-Driven Distillation: Consistency Distillation with Target Timestep Selection and Decoupled Guidance
by: Wang, Cunzheng, et al.
Published: (2024) -
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
by: Xiang, Qiang, et al.
Published: (2025) -
InfinityStory: Unlimited Video Generation with World Consistency and Character-Aware Shot Transitions
by: Elmoghany, Mohamed, et al.
Published: (2026) -
ReMix: Towards a Unified View of Consistent Character Generation and Editing
by: Zhou, Benjia, et al.
Published: (2025) -
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
by: Wang, Qinghe, et al.
Published: (2024)