Calligrapher: Freestyle Text Image Customization
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Yue, Bai, Qingyan, Ouyang, Hao, Cheng, Ka Leong, Wang, Qiuyu, Liu, Hongyu, Liu, Zichen, Wang, Haofan, Chen, Jingye, Shen, Yujun, Chen, Qifeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Edicho: Consistent Image Editing in the Wild
by: Bai, Qingyan, et al.
Published: (2024)
by: Bai, Qingyan, et al.
Published: (2024)
MagicQuill: An Intelligent Interactive Image Editing System
by: Liu, Zichen, et al.
Published: (2024)
by: Liu, Zichen, et al.
Published: (2024)
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
by: Bai, Qingyan, et al.
Published: (2025)
by: Bai, Qingyan, et al.
Published: (2025)
MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
by: Wang, Hanlin, et al.
Published: (2025)
by: Wang, Hanlin, et al.
Published: (2025)
LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis
by: Wang, Hanlin, et al.
Published: (2024)
by: Wang, Hanlin, et al.
Published: (2024)
Learning Naturally Aggregated Appearance for Efficient 3D Editing
by: Cheng, Ka Leong, et al.
Published: (2023)
by: Cheng, Ka Leong, et al.
Published: (2023)
Real-time 3D-aware Portrait Editing from a Single Image
by: Bai, Qingyan, et al.
Published: (2024)
by: Bai, Qingyan, et al.
Published: (2024)
DepthLab: From Partial to Complete
by: Liu, Zhiheng, et al.
Published: (2024)
by: Liu, Zhiheng, et al.
Published: (2024)
CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
by: Ouyang, Hao, et al.
Published: (2023)
by: Ouyang, Hao, et al.
Published: (2023)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
by: Meng, Yihao, et al.
Published: (2026)
by: Meng, Yihao, et al.
Published: (2026)
AvatarArtist: Open-Domain 4D Avatarization
by: Liu, Hongyu, et al.
Published: (2025)
by: Liu, Hongyu, et al.
Published: (2025)
AniDoc: Animation Creation Made Easier
by: Meng, Yihao, et al.
Published: (2024)
by: Meng, Yihao, et al.
Published: (2024)
Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation
by: Ma, Yue, et al.
Published: (2024)
by: Ma, Yue, et al.
Published: (2024)
MangaNinja: Line Art Colorization with Precise Reference Following
by: Liu, Zhiheng, et al.
Published: (2025)
by: Liu, Zhiheng, et al.
Published: (2025)
InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
by: Liu, Zhiheng, et al.
Published: (2024)
by: Liu, Zhiheng, et al.
Published: (2024)
HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives
by: Meng, Yihao, et al.
Published: (2025)
by: Meng, Yihao, et al.
Published: (2025)
Advancing Open-source World Models
by: Robbyant Team, et al.
Published: (2026)
by: Robbyant Team, et al.
Published: (2026)
TALE: Training-free Cross-domain Image Composition via Adaptive Latent Manipulation and Energy-guided Optimization
by: Pham, Kien T., et al.
Published: (2024)
by: Pham, Kien T., et al.
Published: (2024)
HeadArtist: Text-conditioned 3D Head Generation with Self Score Distillation
by: Liu, Hongyu, et al.
Published: (2023)
by: Liu, Hongyu, et al.
Published: (2023)
Follow-Your-Emoji-Faster: Towards Efficient, Fine-Controllable, and Expressive Freestyle Portrait Animation
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting
by: Yang, Guanrou, et al.
Published: (2025)
by: Yang, Guanrou, et al.
Published: (2025)
Towards Degradation-Robust Reconstruction in Generalizable NeRF
by: Park, Chan Ho, et al.
Published: (2024)
by: Park, Chan Ho, et al.
Published: (2024)
Framer: Interactive Frame Interpolation
by: Wang, Wen, et al.
Published: (2024)
by: Wang, Wen, et al.
Published: (2024)
AvatarPointillist: AutoRegressive 4D Gaussian Avatarization
by: Liu, Hongyu, et al.
Published: (2026)
by: Liu, Hongyu, et al.
Published: (2026)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
by: Lu, Yunhong, et al.
Published: (2025)
by: Lu, Yunhong, et al.
Published: (2025)
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
by: Wang, Haofan, et al.
Published: (2024)
by: Wang, Haofan, et al.
Published: (2024)
InstantStyle-Plus: Style Transfer with Content-Preserving in Text-to-Image Generation
by: Wang, Haofan, et al.
Published: (2024)
by: Wang, Haofan, et al.
Published: (2024)
AnyDoor: Zero-shot Object-level Image Customization
by: Chen, Xi, et al.
Published: (2023)
by: Chen, Xi, et al.
Published: (2023)
CSGO: Content-Style Composition in Text-to-Image Generation
by: Xing, Peng, et al.
Published: (2024)
by: Xing, Peng, et al.
Published: (2024)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
by: Ding, Ganggui, et al.
Published: (2024)
by: Ding, Ganggui, et al.
Published: (2024)
Towards General Text-guided Image Synthesis for Customized Multimodal Brain MRI Generation
by: Wang, Yulin, et al.
Published: (2024)
by: Wang, Yulin, et al.
Published: (2024)
Rethinking Layered Graphic Design Generation with a Top-Down Approach
by: Chen, Jingye, et al.
Published: (2025)
by: Chen, Jingye, et al.
Published: (2025)
LACON: Training Text-to-Image Model from Uncurated Data
by: Liang, Zhiyang, et al.
Published: (2026)
by: Liang, Zhiyang, et al.
Published: (2026)
TIGaussian: Disentangle Gaussians for Spatial-Awared Text-Image-3D Alignment
by: Liu, Jiarun, et al.
Published: (2026)
by: Liu, Jiarun, et al.
Published: (2026)
Tuning-Free Image Customization with Image and Text Guidance
by: Li, Pengzhi, et al.
Published: (2024)
by: Li, Pengzhi, et al.
Published: (2024)
Group Editing: Edit Multiple Images in One Go
by: Ma, Yue, et al.
Published: (2026)
by: Ma, Yue, et al.
Published: (2026)
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
by: Ouyang, Pengxiang, et al.
Published: (2025)
by: Ouyang, Pengxiang, et al.
Published: (2025)
Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation
by: Huang, Siteng, et al.
Published: (2023)
by: Huang, Siteng, et al.
Published: (2023)
Ranni: Taming Text-to-Image Diffusion for Accurate Instruction Following
by: Feng, Yutong, et al.
Published: (2023)
by: Feng, Yutong, et al.
Published: (2023)
Similar Items
-
Edicho: Consistent Image Editing in the Wild
by: Bai, Qingyan, et al.
Published: (2024) -
MagicQuill: An Intelligent Interactive Image Editing System
by: Liu, Zichen, et al.
Published: (2024) -
Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset
by: Bai, Qingyan, et al.
Published: (2025) -
MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues
by: Liu, Zichen, et al.
Published: (2025) -
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
by: Wang, Hanlin, et al.
Published: (2025)