First Creating Backgrounds Then Rendering Texts: A New Paradigm for Visual Text Blending
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zhenhang, Shu, Yan, Zeng, Weichao, Yang, Dongbao, Zhou, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing
von: Shu, Yan, et al.
Veröffentlicht: (2024)
von: Shu, Yan, et al.
Veröffentlicht: (2024)
TADoc: Robust Time-Aware Document Image Dewarping
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025)
Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
von: Luo, Minxing, et al.
Veröffentlicht: (2025)
Visual Text Processing: A Comprehensive Review and Unified Evaluation
von: Shu, Yan, et al.
Veröffentlicht: (2025)
von: Shu, Yan, et al.
Veröffentlicht: (2025)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
The Role of Video Generation in Enhancing Data-Limited Action Understanding
von: Li, Wei, et al.
Veröffentlicht: (2025)
von: Li, Wei, et al.
Veröffentlicht: (2025)
RepText: Rendering Visual Text via Replicating
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
Masked and Permuted Implicit Context Learning for Scene Text Recognition
von: Yang, Xiaomeng, et al.
Veröffentlicht: (2023)
von: Yang, Xiaomeng, et al.
Veröffentlicht: (2023)
The Devil is in Fine-tuning and Long-tailed Problems:A New Benchmark for Scene Text Detection
von: Cao, Tianjiao, et al.
Veröffentlicht: (2025)
von: Cao, Tianjiao, et al.
Veröffentlicht: (2025)
Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance
von: Lyu, Jiahao, et al.
Veröffentlicht: (2024)
von: Lyu, Jiahao, et al.
Veröffentlicht: (2024)
TextPecker: Rewarding Structural Anomaly Quantification for Enhancing Visual Text Rendering
von: Zhu, Hanshen, et al.
Veröffentlicht: (2026)
von: Zhu, Hanshen, et al.
Veröffentlicht: (2026)
Glyph-ByT5: A Customized Text Encoder for Accurate Visual Text Rendering
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
VidText: Towards Comprehensive Evaluation for Video Text Understanding
von: Yang, Zhoufaran, et al.
Veröffentlicht: (2025)
von: Yang, Zhoufaran, et al.
Veröffentlicht: (2025)
EmoCaliber: Advancing Reliable Visual Emotion Comprehension via Confidence Verbalization and Calibration
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
Training-Free Occluded Text Rendering via Glyph Priors and Attention-Guided Semantic Blending
von: Hou, Jingqi, et al.
Veröffentlicht: (2026)
von: Hou, Jingqi, et al.
Veröffentlicht: (2026)
Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval
von: Zeng, Gangyan, et al.
Veröffentlicht: (2024)
von: Zeng, Gangyan, et al.
Veröffentlicht: (2024)
Zero-Shot Visual Concept Blending Without Text Guidance
von: Makino, Hiroya, et al.
Veröffentlicht: (2025)
von: Makino, Hiroya, et al.
Veröffentlicht: (2025)
TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering
von: Mao, Dongxing, et al.
Veröffentlicht: (2026)
von: Mao, Dongxing, et al.
Veröffentlicht: (2026)
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
von: Chen, Zeyu, et al.
Veröffentlicht: (2026)
von: Chen, Zeyu, et al.
Veröffentlicht: (2026)
IMTBench: A Multi-Scenario Cross-Modal Collaborative Evaluation Benchmark for In-Image Machine Translation
von: Lyu, Jiahao, et al.
Veröffentlicht: (2026)
von: Lyu, Jiahao, et al.
Veröffentlicht: (2026)
Customizing Visual Emotion Evaluation for MLLMs: An Open-vocabulary, Multifaceted, and Scalable Approach
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
von: Wu, Daiqing, et al.
Veröffentlicht: (2025)
DreamRenderer: Taming Multi-Instance Attribute Control in Large-Scale Text-to-Image Models
von: Zhou, Dewei, et al.
Veröffentlicht: (2025)
von: Zhou, Dewei, et al.
Veröffentlicht: (2025)
BizGen: Advancing Article-level Visual Text Rendering for Infographics Generation
von: Peng, Yuyang, et al.
Veröffentlicht: (2025)
von: Peng, Yuyang, et al.
Veröffentlicht: (2025)
EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
von: Lu, Runnan, et al.
Veröffentlicht: (2025)
TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2026)
TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
Toward Real Text Manipulation Detection: New Dataset and New Solution
von: Luo, Dongliang, et al.
Veröffentlicht: (2023)
von: Luo, Dongliang, et al.
Veröffentlicht: (2023)
StyleBlend: Enhancing Style-Specific Content Creation in Text-to-Image Diffusion Models
von: Chen, Zichong, et al.
Veröffentlicht: (2025)
von: Chen, Zichong, et al.
Veröffentlicht: (2025)
GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering
von: Shuai, Xincheng, et al.
Veröffentlicht: (2026)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2026)
Text-Animator: Controllable Visual Text Video Generation
von: Liu, Lin, et al.
Veröffentlicht: (2024)
von: Liu, Lin, et al.
Veröffentlicht: (2024)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
von: Baek, Kanghyun, et al.
Veröffentlicht: (2025)
Blending Concepts with Text-to-Image Diffusion Models
von: Olearo, Lorenzo, et al.
Veröffentlicht: (2025)
von: Olearo, Lorenzo, et al.
Veröffentlicht: (2025)
When Semantics Mislead Vision: Mitigating Large Multimodal Models Hallucinations in Scene Text Spotting and Understanding
von: Shu, Yan, et al.
Veröffentlicht: (2025)
von: Shu, Yan, et al.
Veröffentlicht: (2025)
TextEditBench: Evaluating Reasoning-aware Text Editing Beyond Rendering
von: Gui, Rui, et al.
Veröffentlicht: (2025)
von: Gui, Rui, et al.
Veröffentlicht: (2025)
UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing
von: Ma, Lichen, et al.
Veröffentlicht: (2026)
von: Ma, Lichen, et al.
Veröffentlicht: (2026)
Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
von: Liu, Zeyu, et al.
Veröffentlicht: (2024)
AnyText: Multilingual Visual Text Generation And Editing
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
von: Tuo, Yuxiang, et al.
Veröffentlicht: (2023)
Imagine for Me: Creative Conceptual Blending of Real Images and Text via Blended Attention
von: Cho, Wonwoong, et al.
Veröffentlicht: (2025)
von: Cho, Wonwoong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024) -
Visual Text Meets Low-level Vision: A Comprehensive Survey on Visual Text Processing
von: Shu, Yan, et al.
Veröffentlicht: (2024) -
TADoc: Robust Time-Aware Document Image Dewarping
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025) -
Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion
von: Zhao, Fangmin, et al.
Veröffentlicht: (2025) -
Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation
von: Luo, Minxing, et al.
Veröffentlicht: (2025)