Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zhennan, Li, Yajie, Wang, Haofan, Chen, Zhibo, Jiang, Zhengkai, Li, Jun, Wang, Qian, Yang, Jian, Tai, Ying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Investigating Text Insulation and Attention Mechanisms for Complex Visual Text Generation
von: Tai, Ying, et al.
Veröffentlicht: (2025)
von: Tai, Ying, et al.
Veröffentlicht: (2025)
RepText: Rendering Visual Text via Replicating
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
von: Wang, Haofan, et al.
Veröffentlicht: (2025)
Soft then Hard: Rethinking the Quantization in Neural Image Compression
von: Guo, Zongyu, et al.
Veröffentlicht: (2021)
von: Guo, Zongyu, et al.
Veröffentlicht: (2021)
Synthesis of Trifluoromethyl‐Substituted Tetrazoles via KI/TBHP‐Mediated Oxidative Annulation of Trifluoroacetimidohydrazides and Nitromethane
von: Jian Chen, et al.
Veröffentlicht: (2025)
von: Jian Chen, et al.
Veröffentlicht: (2025)
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
DiffProxy: Multi-View Human Mesh Recovery via Diffusion-Generated Dense Proxies
von: Wang, Renke, et al.
Veröffentlicht: (2026)
von: Wang, Renke, et al.
Veröffentlicht: (2026)
AFLoRA: Adaptive Federated Fine-Tuning of Large Language Models with Resource-Aware Low-Rank Adaption
von: Zhou, Yajie, et al.
Veröffentlicht: (2025)
von: Zhou, Yajie, et al.
Veröffentlicht: (2025)
InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting
von: Yu, Yue, et al.
Veröffentlicht: (2026)
von: Yu, Yue, et al.
Veröffentlicht: (2026)
AGSwap: Overcoming Category Boundaries in Object Fusion via Adaptive Group Swapping
von: Zhang, Zedong, et al.
Veröffentlicht: (2025)
von: Zhang, Zedong, et al.
Veröffentlicht: (2025)
DiverseAR: Boosting Diversity in Bitwise Autoregressive Image Generation
von: Yang, Ying, et al.
Veröffentlicht: (2025)
von: Yang, Ying, et al.
Veröffentlicht: (2025)
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
von: Gao, Zhanxin, et al.
Veröffentlicht: (2025)
von: Gao, Zhanxin, et al.
Veröffentlicht: (2025)
MambaLLIE: Implicit Retinex-Aware Low Light Enhancement with Global-then-Local State Space
von: Weng, Jiangwei, et al.
Veröffentlicht: (2024)
von: Weng, Jiangwei, et al.
Veröffentlicht: (2024)
CSGO: Content-Style Composition in Text-to-Image Generation
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
L2P: Unlocking Latent Potential for Pixel Generation
von: Chen, Zhennan, et al.
Veröffentlicht: (2026)
von: Chen, Zhennan, et al.
Veröffentlicht: (2026)
G-Refine: A General Quality Refiner for Text-to-Image Generation
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
HardSecBench: Benchmarking the Security Awareness of LLMs for Hardware Code Generation
von: Chen, Qirui, et al.
Veröffentlicht: (2026)
von: Chen, Qirui, et al.
Veröffentlicht: (2026)
TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement
von: Xie, Tianyidan, et al.
Veröffentlicht: (2026)
von: Xie, Tianyidan, et al.
Veröffentlicht: (2026)
Category-Aware 3D Object Composition with Disentangled Texture and Shape Multi-view Diffusion
von: Xiong, Zeren, et al.
Veröffentlicht: (2025)
von: Xiong, Zeren, et al.
Veröffentlicht: (2025)
From Zero to Detail: Deconstructing Ultra-High-Definition Image Restoration from Progressive Spectral Perspective
von: Zhao, Chen, et al.
Veröffentlicht: (2025)
von: Zhao, Chen, et al.
Veröffentlicht: (2025)
RH-SQL: Refined Schema and Hardness Prompt for Text-to-SQL
von: Yi, Jiawen, et al.
Veröffentlicht: (2024)
von: Yi, Jiawen, et al.
Veröffentlicht: (2024)
Adaptive Guidance Learning for Camouflaged Object Detection
von: Chen, Zhennan, et al.
Veröffentlicht: (2024)
von: Chen, Zhennan, et al.
Veröffentlicht: (2024)
Learning Generalized Residual Exchange-Correlation-Uncertain Functional for Density Functional Theory
von: Jin, Sizhuo, et al.
Veröffentlicht: (2024)
von: Jin, Sizhuo, et al.
Veröffentlicht: (2024)
Novel Object Synthesis via Adaptive Text-Image Harmony
von: Xiong, Zeren, et al.
Veröffentlicht: (2024)
von: Xiong, Zeren, et al.
Veröffentlicht: (2024)
InstantStyle-Plus: Style Transfer with Content-Preserving in Text-to-Image Generation
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
von: Wang, Haofan, et al.
Veröffentlicht: (2024)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
Iterative Prompt Refinement for Safer Text-to-Image Generation
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
von: Kou, Siqi, et al.
Veröffentlicht: (2026)
von: Kou, Siqi, et al.
Veröffentlicht: (2026)
Text-Driven Image Editing via Learnable Regions
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
TIP and Polish: Text-Image-Prototype Guided Multi-Modal Generation via Commonality-Discrepancy Modeling and Refinement
von: Ma, Zhiyong, et al.
Veröffentlicht: (2025)
von: Ma, Zhiyong, et al.
Veröffentlicht: (2025)
Pool-Select-Refine: Allocation-Aware Generative Dataset Distillation with Soft-Label-Guided Latent Refinement
von: Li, Wenmin, et al.
Veröffentlicht: (2026)
von: Li, Wenmin, et al.
Veröffentlicht: (2026)
Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regularization
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4$\times$RTX 4090s
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
DiP: Taming Diffusion Models in Pixel Space
von: Chen, Zhennan, et al.
Veröffentlicht: (2025)
von: Chen, Zhennan, et al.
Veröffentlicht: (2025)
TeViR: Text-to-Video Reward with Diffusion Models for Efficient Reinforcement Learning
von: Chen, Yuhui, et al.
Veröffentlicht: (2025)
von: Chen, Yuhui, et al.
Veröffentlicht: (2025)
Focus-N-Fix: Region-Aware Fine-Tuning for Text-to-Image Generation
von: Xing, Xiaoying, et al.
Veröffentlicht: (2025)
von: Xing, Xiaoying, et al.
Veröffentlicht: (2025)
Golden RPG: Confidence-Adaptive Region-Aware Noise for Compositional Text-to-Image Generation
von: Li, Hao
Veröffentlicht: (2026)
von: Li, Hao
Veröffentlicht: (2026)
HardSATGEN: Understanding the Difficulty of Hard SAT Formula Generation and A Strong Structure-Hardness-Aware Baseline
von: Li, Yang, et al.
Veröffentlicht: (2023)
von: Li, Yang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Investigating Text Insulation and Attention Mechanisms for Complex Visual Text Generation
von: Tai, Ying, et al.
Veröffentlicht: (2025) -
RepText: Rendering Visual Text via Replicating
von: Wang, Haofan, et al.
Veröffentlicht: (2025) -
Soft then Hard: Rethinking the Quantization in Neural Image Compression
von: Guo, Zongyu, et al.
Veröffentlicht: (2021) -
Synthesis of Trifluoromethyl‐Substituted Tetrazoles via KI/TBHP‐Mediated Oxidative Annulation of Trifluoroacetimidohydrazides and Nitromethane
von: Jian Chen, et al.
Veröffentlicht: (2025) -
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)