Glyph-ByT5: A Customized Text Encoder for Accurate Visual Text Rendering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zeyu, Liang, Weicong, Liang, Zhanhao, Luo, Chong, Li, Ji, Huang, Gao, Yuan, Yuhui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering
von: Shuai, Xincheng, et al.
Veröffentlicht: (2026)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2026)
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
GlyphBanana: Advancing Precise Text Rendering Through Agentic Workflows
von: Yan, Zexuan, et al.
Veröffentlicht: (2026)
von: Yan, Zexuan, et al.
Veröffentlicht: (2026)
BizGen: Advancing Article-level Visual Text Rendering for Infographics Generation
von: Peng, Yuyang, et al.
Veröffentlicht: (2025)
von: Peng, Yuyang, et al.
Veröffentlicht: (2025)
FreeText: Training-Free Text Rendering in Diffusion Transformers via Attention Localization and Spectral Glyph Injection
von: Zhang, Ruiqiang, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiqiang, et al.
Veröffentlicht: (2026)
UniGlyph: Unified Segmentation-Conditioned Diffusion for Precise Visual Text Synthesis
von: Wang, Yuanrui, et al.
Veröffentlicht: (2025)
von: Wang, Yuanrui, et al.
Veröffentlicht: (2025)
HDGlyph: A Hierarchical Disentangled Glyph-Based Framework for Long-Tail Text Rendering in Diffusion Models
von: Zhuang, Shuhan, et al.
Veröffentlicht: (2025)
von: Zhuang, Shuhan, et al.
Veröffentlicht: (2025)
Training-Free Occluded Text Rendering via Glyph Priors and Attention-Guided Semantic Blending
von: Hou, Jingqi, et al.
Veröffentlicht: (2026)
von: Hou, Jingqi, et al.
Veröffentlicht: (2026)
Refining Text-to-Image Generation: Towards Accurate Training-Free Glyph-Enhanced Image Generation
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
RepText: Rendering Visual Text via Replicating
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
Glyph: Scaling Context Windows via Visual-Text Compression
von: Cheng, Jiale, et al.
Veröffentlicht: (2025)
von: Cheng, Jiale, et al.
Veröffentlicht: (2025)
TextPecker: Rewarding Structural Anomaly Quantification for Enhancing Visual Text Rendering
von: Zhu, Hanshen, et al.
Veröffentlicht: (2026)
von: Zhu, Hanshen, et al.
Veröffentlicht: (2026)
LOGO: Video Text Spotting with Language Collaboration and Glyph Perception Model
von: Liu, Hongen, et al.
Veröffentlicht: (2024)
von: Liu, Hongen, et al.
Veröffentlicht: (2024)
PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
Aesthetic Post-Training Diffusion Models from Generic Preferences with Step-by-step Preference Optimization
von: Liang, Zhanhao, et al.
Veröffentlicht: (2024)
von: Liang, Zhanhao, et al.
Veröffentlicht: (2024)
Empowering Backbone Models for Visual Text Generation with Input Granularity Control and Glyph-Aware Training
von: Li, Wenbo, et al.
Veröffentlicht: (2024)
von: Li, Wenbo, et al.
Veröffentlicht: (2024)
TextMaster: A Unified Framework for Realistic Text Editing via Glyph-Style Dual-Control
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
GLDesigner: Leveraging Multi-Modal LLMs as Designer for Enhanced Aesthetic Text Glyph Layouts
von: He, Junwen, et al.
Veröffentlicht: (2024)
von: He, Junwen, et al.
Veröffentlicht: (2024)
First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending
von: Li, Zhenhang, et al.
Veröffentlicht: (2024)
von: Li, Zhenhang, et al.
Veröffentlicht: (2024)
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
von: Ma, Lichen, et al.
Veröffentlicht: (2024)
von: Ma, Lichen, et al.
Veröffentlicht: (2024)
Text-guided Visual Prompt DINO for Generic Segmentation
von: Guan, Yuchen, et al.
Veröffentlicht: (2025)
von: Guan, Yuchen, et al.
Veröffentlicht: (2025)
RealCustom: Narrowing Real Text Word for Real-Time Open-Domain Text-to-Image Customization
von: Huang, Mengqi, et al.
Veröffentlicht: (2024)
von: Huang, Mengqi, et al.
Veröffentlicht: (2024)
WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing
von: Zhang, Hui, et al.
Veröffentlicht: (2026)
von: Zhang, Hui, et al.
Veröffentlicht: (2026)
Enhancing Diffusion Models with Text-Encoder Reinforcement Learning
von: Chen, Chaofeng, et al.
Veröffentlicht: (2023)
von: Chen, Chaofeng, et al.
Veröffentlicht: (2023)
Scaling Down Text Encoders of Text-to-Image Diffusion Models
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
von: Wang, Lifu, et al.
Veröffentlicht: (2025)
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
Text-Animator: Controllable Visual Text Video Generation
von: Liu, Lin, et al.
Veröffentlicht: (2024)
von: Liu, Lin, et al.
Veröffentlicht: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
DiffInk: Glyph- and Style-Aware Latent Diffusion Transformer for Text to Online Handwriting Generation
von: Pan, Wei, et al.
Veröffentlicht: (2025)
von: Pan, Wei, et al.
Veröffentlicht: (2025)
How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
von: Dong, Jiahua, et al.
Veröffentlicht: (2024)
von: Dong, Jiahua, et al.
Veröffentlicht: (2024)
CustomVideo: Customizing Text-to-Video Generation with Multiple Subjects
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
Visual Text Generation in the Wild
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Yuanzhi, et al.
Veröffentlicht: (2024)
TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
Calligrapher: Freestyle Text Image Customization
von: Ma, Yue, et al.
Veröffentlicht: (2025)
von: Ma, Yue, et al.
Veröffentlicht: (2025)
Bring Your Dreams to Life: Continual Text-to-Video Customization
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
von: Dong, Jiahua, et al.
Veröffentlicht: (2025)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
AnyArtisticGlyph: Multilingual Controllable Artistic Glyph Generation
von: Lu, Xiongbo, et al.
Veröffentlicht: (2025)
von: Lu, Xiongbo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering
von: Liu, Zeyu, et al.
Veröffentlicht: (2024) -
GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering
von: Shuai, Xincheng, et al.
Veröffentlicht: (2026) -
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025) -
GlyphBanana: Advancing Precise Text Rendering Through Agentic Workflows
von: Yan, Zexuan, et al.
Veröffentlicht: (2026) -
BizGen: Advancing Article-level Visual Text Rendering for Infographics Generation
von: Peng, Yuyang, et al.
Veröffentlicht: (2025)