FonTS: Text Rendering with Typography and Style Controls
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Wenda, Song, Yiren, Zhang, Dengming, Liu, Jiaming, Zou, Xingxing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WordCon: Word-level Typography Control in Scene Text Rendering
by: Shi, Wenda, et al.
Published: (2025)
by: Shi, Wenda, et al.
Published: (2025)
AMO Sampler: Enhancing Text Rendering with Overshooting
by: Hu, Xixi, et al.
Published: (2024)
by: Hu, Xixi, et al.
Published: (2024)
AnySurf: Any Surface Generation with Directed Edge
by: Shi, Wenda, et al.
Published: (2026)
by: Shi, Wenda, et al.
Published: (2026)
FastCLIPstyler: Optimisation-free Text-based Image Style Transfer Using Style Representations
by: Suresh, Ananda Padhmanabhan, et al.
Published: (2022)
by: Suresh, Ananda Padhmanabhan, et al.
Published: (2022)
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
by: Lu, Runnan, et al.
Published: (2025)
by: Lu, Runnan, et al.
Published: (2025)
Improving Diffusion Inverse Problem Solving with Decoupled Noise Annealing
by: Zhang, Bingliang, et al.
Published: (2024)
by: Zhang, Bingliang, et al.
Published: (2024)
VisionTS: Visual Masked Autoencoders Are Free-Lunch Zero-Shot Time Series Forecasters
by: Chen, Mouxiang, et al.
Published: (2024)
by: Chen, Mouxiang, et al.
Published: (2024)
CaTS-Bench: Can Language Models Describe Time Series?
by: Zhou, Luca, et al.
Published: (2025)
by: Zhou, Luca, et al.
Published: (2025)
TextCraftor: Your Text Encoder Can be Image Quality Controller
by: Li, Yanyu, et al.
Published: (2024)
by: Li, Yanyu, et al.
Published: (2024)
DiffiT: Diffusion Vision Transformers for Image Generation
by: Hatamizadeh, Ali, et al.
Published: (2023)
by: Hatamizadeh, Ali, et al.
Published: (2023)
Text-to-Image Diffusion Models Cannot Count, and Prompt Refinement Cannot Help
by: Guo, Xuyang, et al.
Published: (2025)
by: Guo, Xuyang, et al.
Published: (2025)
A Comprehensive Survey of Continual Learning: Theory, Method and Application
by: Wang, Liyuan, et al.
Published: (2023)
by: Wang, Liyuan, et al.
Published: (2023)
VQ-Style: Disentangling Style and Content in Motion with Residual Quantized Representations
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
by: Zargarbashi, Fatemeh, et al.
Published: (2026)
Terminal Velocity Matching
by: Zhou, Linqi, et al.
Published: (2025)
by: Zhou, Linqi, et al.
Published: (2025)
Kinetic Typography Diffusion Model
by: Park, Seonmi, et al.
Published: (2024)
by: Park, Seonmi, et al.
Published: (2024)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
by: Qiu, Zeju, et al.
Published: (2023)
by: Qiu, Zeju, et al.
Published: (2023)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023)
by: Ki, Taekyung, et al.
Published: (2023)
MLLM4TS: Leveraging Vision and Multimodal Language Models for General Time-Series Analysis
by: Liu, Qinghua, et al.
Published: (2025)
by: Liu, Qinghua, et al.
Published: (2025)
Can You Count to Nine? A Human Evaluation Benchmark for Counting Limits in Modern Text-to-Video Models
by: Guo, Xuyang, et al.
Published: (2025)
by: Guo, Xuyang, et al.
Published: (2025)
Analyzing domain shift when using additional data for the MICCAI KiTS23 Challenge
by: Stoica, George, et al.
Published: (2023)
by: Stoica, George, et al.
Published: (2023)
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and Generation
by: Bai, Yuhang, et al.
Published: (2024)
by: Bai, Yuhang, et al.
Published: (2024)
Regional Style and Color Transfer
by: Ding, Zhicheng, et al.
Published: (2024)
by: Ding, Zhicheng, et al.
Published: (2024)
Forecasting as Rendering: A 2D Gaussian Splatting Framework for Time Series Forecasting
by: Wang, Yixin, et al.
Published: (2026)
by: Wang, Yixin, et al.
Published: (2026)
Control Your View: High-Resolution Global Semantic Manipulation in Learned Image Compression
by: Liang, Jiaming, et al.
Published: (2026)
by: Liang, Jiaming, et al.
Published: (2026)
FashionR2R: Texture-preserving Rendered-to-Real Image Translation with Diffusion Models
by: Hu, Rui, et al.
Published: (2024)
by: Hu, Rui, et al.
Published: (2024)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
by: Ogezi, Michael, et al.
Published: (2024)
by: Ogezi, Michael, et al.
Published: (2024)
Quantum Implicit Neural Representations
by: Zhao, Jiaming, et al.
Published: (2024)
by: Zhao, Jiaming, et al.
Published: (2024)
TempoControl: Temporal Attention Guidance for Text-to-Video Models
by: Schiber, Shira, et al.
Published: (2025)
by: Schiber, Shira, et al.
Published: (2025)
PreciseCam: Precise Camera Control for Text-to-Image Generation
by: Bernal-Berdun, Edurne, et al.
Published: (2025)
by: Bernal-Berdun, Edurne, et al.
Published: (2025)
RefineNet: Enhancing Text-to-Image Conversion with High-Resolution and Detail Accuracy through Hierarchical Transformers and Progressive Refinement
by: Shi, Fan
Published: (2023)
by: Shi, Fan
Published: (2023)
On the Scalability of Diffusion-based Text-to-Image Generation
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
From Text to Pose to Image: Improving Diffusion Model Control and Quality
by: Bonnet, Clément, et al.
Published: (2024)
by: Bonnet, Clément, et al.
Published: (2024)
Wasserstein Distance Rivals Kullback-Leibler Divergence for Knowledge Distillation
by: Lv, Jiaming, et al.
Published: (2024)
by: Lv, Jiaming, et al.
Published: (2024)
Style-Extracting Diffusion Models for Semi-Supervised Histopathology Segmentation
by: Öttl, Mathias, et al.
Published: (2024)
by: Öttl, Mathias, et al.
Published: (2024)
SCFlow: Implicitly Learning Style and Content Disentanglement with Flow Models
by: Ma, Pingchuan, et al.
Published: (2025)
by: Ma, Pingchuan, et al.
Published: (2025)
Evaluation of Randomization through Style Transfer for Enhanced Domain Generalization
by: Eisenhardt, Dustin, et al.
Published: (2026)
by: Eisenhardt, Dustin, et al.
Published: (2026)
OpenNDD: Open Set Recognition for Neurodevelopmental Disorders Detection
by: Yu, Jiaming, et al.
Published: (2023)
by: Yu, Jiaming, et al.
Published: (2023)
SafeDreamer: Safe Reinforcement Learning with World Models
by: Huang, Weidong, et al.
Published: (2023)
by: Huang, Weidong, et al.
Published: (2023)
Multi-modal Machine Learning for Vehicle Rating Predictions Using Image, Text, and Parametric Data
by: Su, Hanqi, et al.
Published: (2023)
by: Su, Hanqi, et al.
Published: (2023)
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Similar Items
-
WordCon: Word-level Typography Control in Scene Text Rendering
by: Shi, Wenda, et al.
Published: (2025) -
AMO Sampler: Enhancing Text Rendering with Overshooting
by: Hu, Xixi, et al.
Published: (2024) -
AnySurf: Any Surface Generation with Directed Edge
by: Shi, Wenda, et al.
Published: (2026) -
FastCLIPstyler: Optimisation-free Text-based Image Style Transfer Using Style Representations
by: Suresh, Ananda Padhmanabhan, et al.
Published: (2022) -
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
by: Lu, Runnan, et al.
Published: (2025)