TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Tianyi, Liu, Jiangqi, Huang, Yifei, Jiang, Shiqi, Shi, Jianshen, Wang, Changbo, Li, Chenhui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MPJudge: Towards Perceptual Assessment of Music-Induced Paintings
von: Jiang, Shiqi, et al.
Veröffentlicht: (2025)
von: Jiang, Shiqi, et al.
Veröffentlicht: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
ExpertGen: Training-Free Expert Guidance for Controllable Text-to-Face Generation
von: Shi, Liang, et al.
Veröffentlicht: (2025)
von: Shi, Liang, et al.
Veröffentlicht: (2025)
EmoGen: Emotional Image Content Generation with Text-to-Image Diffusion Models
von: Yang, Jingyuan, et al.
Veröffentlicht: (2024)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2024)
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
von: Chen, Zeyu, et al.
Veröffentlicht: (2026)
von: Chen, Zeyu, et al.
Veröffentlicht: (2026)
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
von: Li, Yayuan, et al.
Veröffentlicht: (2024)
PPJudge: Towards Human-Aligned Assessment of Artistic Painting Process
von: Jiang, Shiqi, et al.
Veröffentlicht: (2025)
von: Jiang, Shiqi, et al.
Veröffentlicht: (2025)
AACP: Aesthetics assessment of children's paintings based on self-supervised learning
von: Jiang, Shiqi, et al.
Veröffentlicht: (2024)
von: Jiang, Shiqi, et al.
Veröffentlicht: (2024)
Investigating Text Insulation and Attention Mechanisms for Complex Visual Text Generation
von: Tai, Ying, et al.
Veröffentlicht: (2025)
von: Tai, Ying, et al.
Veröffentlicht: (2025)
Synthetic Perception: Can Generated Images Unlock Latent Visual Prior for Text-Centric Reasoning?
von: Huang, Yuesheng, et al.
Veröffentlicht: (2025)
von: Huang, Yuesheng, et al.
Veröffentlicht: (2025)
GenAgent: Scaling Text-to-Image Generation via Agentic Multimodal Reasoning
von: Jiang, Kaixun, et al.
Veröffentlicht: (2026)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2026)
DialogGen: Multi-modal Interactive Dialogue System for Multi-turn Text-to-Image Generation
von: Huang, Minbin, et al.
Veröffentlicht: (2024)
von: Huang, Minbin, et al.
Veröffentlicht: (2024)
TexGen: Text-Guided 3D Texture Generation with Multi-view Sampling and Resampling
von: Huo, Dong, et al.
Veröffentlicht: (2024)
von: Huo, Dong, et al.
Veröffentlicht: (2024)
Robust Message Embedding via Attention Flow-Based Steganography
von: Ye, Huayuan, et al.
Veröffentlicht: (2024)
von: Ye, Huayuan, et al.
Veröffentlicht: (2024)
WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation
von: Zhang, Daoan, et al.
Veröffentlicht: (2025)
von: Zhang, Daoan, et al.
Veröffentlicht: (2025)
Text Data-Centric Image Captioning with Interactive Prompts
von: Wang, Yiyu, et al.
Veröffentlicht: (2024)
von: Wang, Yiyu, et al.
Veröffentlicht: (2024)
Unsupervised Modality Adaptation with Text-to-Image Diffusion Models for Semantic Segmentation
von: Xia, Ruihao, et al.
Veröffentlicht: (2024)
von: Xia, Ruihao, et al.
Veröffentlicht: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
von: Park, NaHyeon, et al.
Veröffentlicht: (2024)
Towards Understanding Cross and Self-Attention in Stable Diffusion for Text-Guided Image Editing
von: Liu, Bingyan, et al.
Veröffentlicht: (2024)
von: Liu, Bingyan, et al.
Veröffentlicht: (2024)
GenExam: A Multidisciplinary Text-to-Image Exam
von: Wang, Zhaokai, et al.
Veröffentlicht: (2025)
von: Wang, Zhaokai, et al.
Veröffentlicht: (2025)
GenMAC: Compositional Text-to-Video Generation with Multi-Agent Collaboration
von: Huang, Kaiyi, et al.
Veröffentlicht: (2024)
von: Huang, Kaiyi, et al.
Veröffentlicht: (2024)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
Anonymization Prompt Learning for Facial Privacy-Preserving Text-to-Image Generation
von: Shi, Liang, et al.
Veröffentlicht: (2024)
von: Shi, Liang, et al.
Veröffentlicht: (2024)
OSPO: Object-Centric Self-Improving Preference Optimization for Text-to-Image Generation
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
von: Oh, Yoonjin, et al.
Veröffentlicht: (2025)
TextSquare: Scaling up Text-Centric Visual Instruction Tuning
von: Tang, Jingqun, et al.
Veröffentlicht: (2024)
von: Tang, Jingqun, et al.
Veröffentlicht: (2024)
SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation
von: Ma, Yingzi, et al.
Veröffentlicht: (2026)
von: Ma, Yingzi, et al.
Veröffentlicht: (2026)
MathGen: Revealing the Illusion of Mathematical Competence through Text-to-Image Generation
von: Liu, Ruiyao, et al.
Veröffentlicht: (2026)
von: Liu, Ruiyao, et al.
Veröffentlicht: (2026)
Enhancing MMDiT-Based Text-to-Image Models for Similar Subject Generation
von: Wei, Tianyi, et al.
Veröffentlicht: (2024)
von: Wei, Tianyi, et al.
Veröffentlicht: (2024)
Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation
von: Liu, Yifei, et al.
Veröffentlicht: (2026)
von: Liu, Yifei, et al.
Veröffentlicht: (2026)
Expressive Text-to-Image Generation with Rich Text
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
Few-Step Distillation for Text-to-Image Generation: A Practical Guide
von: Pu, Yifan, et al.
Veröffentlicht: (2025)
von: Pu, Yifan, et al.
Veröffentlicht: (2025)
DynaIP: Dynamic Image Prompt Adapter for Scalable Zero-shot Personalized Text-to-Image Generation
von: Wang, Zhizhong, et al.
Veröffentlicht: (2025)
von: Wang, Zhizhong, et al.
Veröffentlicht: (2025)
Salient Object-Aware Background Generation using Text-Guided Diffusion Models
von: Eshratifar, Amir Erfan, et al.
Veröffentlicht: (2024)
von: Eshratifar, Amir Erfan, et al.
Veröffentlicht: (2024)
VISTAR:A User-Centric and Role-Driven Benchmark for Text-to-Image Evaluation
von: Jiang, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Kaiyuan, et al.
Veröffentlicht: (2025)
PanoGen++: Domain-Adapted Text-Guided Panoramic Environment Generation for Vision-and-Language Navigation
von: Wang, Sen, et al.
Veröffentlicht: (2025)
von: Wang, Sen, et al.
Veröffentlicht: (2025)
ForCenNet: Foreground-Centric Network for Document Image Rectification
von: Cai, Peng, et al.
Veröffentlicht: (2025)
von: Cai, Peng, et al.
Veröffentlicht: (2025)
Text-Guided Semantic Image Encoder
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2025)
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2025)
Text-Conditioned Background Generation for Editable Multi-Layer Documents
von: Kang, Taewon, et al.
Veröffentlicht: (2025)
von: Kang, Taewon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MPJudge: Towards Perceptual Assessment of Music-Induced Paintings
von: Jiang, Shiqi, et al.
Veröffentlicht: (2025) -
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025) -
ExpertGen: Training-Free Expert Guidance for Controllable Text-to-Face Generation
von: Shi, Liang, et al.
Veröffentlicht: (2025) -
EmoGen: Emotional Image Content Generation with Text-to-Image Diffusion Models
von: Yang, Jingyuan, et al.
Veröffentlicht: (2024) -
StyleTextGen: Style-Conditioned Multilingual Scene Text Generation
von: Chen, Zeyu, et al.
Veröffentlicht: (2026)