FRAP: Faithful and Realistic Text-to-Image Generation with Adaptive Prompt Weighting
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Liyao, Hassanpour, Negar, Salameh, Mohammad, Singamsetti, Mohan Sai, Sun, Fengyu, Lu, Wei, Niu, Di |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024)
by: Jiang, Liyao, et al.
Published: (2024)
Qua$^2$SeDiMo: Quantifiable Quantization Sensitivity of Diffusion Models
by: Mills, Keith G., et al.
Published: (2024)
by: Mills, Keith G., et al.
Published: (2024)
RAISE: Requirement-Adaptive Evolutionary Refinement for Training-Free Text-to-Image Alignment
by: Jiang, Liyao, et al.
Published: (2026)
by: Jiang, Liyao, et al.
Published: (2026)
CascadedGaze: Efficiency in Global Context Extraction for Image Restoration
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
FunEditor: Achieving Complex Image Edits via Function Aggregation with Diffusion Models
by: Samadi, Mohammadreza, et al.
Published: (2024)
by: Samadi, Mohammadreza, et al.
Published: (2024)
Griffin: Generative Reference and Layout Guided Image Composition
by: Mikaeili, Aryan, et al.
Published: (2025)
by: Mikaeili, Aryan, et al.
Published: (2025)
Building Optimal Neural Architectures using Interpretable Knowledge
by: Mills, Keith G., et al.
Published: (2024)
by: Mills, Keith G., et al.
Published: (2024)
Re-ttention: Ultra Sparse Visual Generation via Attention Statistical Reshape
by: Chen, Ruichen, et al.
Published: (2025)
by: Chen, Ruichen, et al.
Published: (2025)
Adaptive Auxiliary Prompt Blending for Target-Faithful Diffusion Generation
by: Lee, Kwanyoung, et al.
Published: (2026)
by: Lee, Kwanyoung, et al.
Published: (2026)
Learning Truncated Causal History Model for Video Restoration
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
RePack then Refine: Efficient Diffusion Transformer with Vision Foundation Model
by: Dong, Guanfang, et al.
Published: (2025)
by: Dong, Guanfang, et al.
Published: (2025)
EPIC: Efficient Prompt Interaction for Text-Image Classification
by: Yu, Xinyao, et al.
Published: (2025)
by: Yu, Xinyao, et al.
Published: (2025)
Adaptive Routing of Text-to-Image Generation Requests Between Large Cloud Model and Light-Weight Edge Model
by: Xin, Zewei, et al.
Published: (2024)
by: Xin, Zewei, et al.
Published: (2024)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
by: Tong, Lei, et al.
Published: (2025)
by: Tong, Lei, et al.
Published: (2025)
Offline Evaluation of Set-Based Text-to-Image Generation
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
by: Wu, Fan, et al.
Published: (2025)
by: Wu, Fan, et al.
Published: (2025)
Optimizing Prompts for Text-to-Image Generation
by: Hao, Yaru, et al.
Published: (2022)
by: Hao, Yaru, et al.
Published: (2022)
Adaptive Prompt Elicitation for Text-to-Image Generation
by: Wen, Xinyi, et al.
Published: (2026)
by: Wen, Xinyi, et al.
Published: (2026)
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
by: Yu, Xinyao, et al.
Published: (2024)
by: Yu, Xinyao, et al.
Published: (2024)
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
by: Manukyan, Hayk, et al.
Published: (2023)
by: Manukyan, Hayk, et al.
Published: (2023)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
by: Wang, Yuanzhi, et al.
Published: (2026)
by: Wang, Yuanzhi, et al.
Published: (2026)
Fast Prompt Alignment for Text-to-Image Generation
by: Mrini, Khalil, et al.
Published: (2024)
by: Mrini, Khalil, et al.
Published: (2024)
Noise Diffusion for Enhancing Semantic Faithfulness in Text-to-Image Synthesis
by: Miao, Boming, et al.
Published: (2024)
by: Miao, Boming, et al.
Published: (2024)
Text-guided Controllable Diffusion for Realistic Camouflage Images Generation
by: Qian, Yuhang, et al.
Published: (2025)
by: Qian, Yuhang, et al.
Published: (2025)
CritiFusion: Semantic Critique and Spectral Alignment for Faithful Text-to-Image Generation
by: Chen, ZhenQi, et al.
Published: (2025)
by: Chen, ZhenQi, et al.
Published: (2025)
HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances
by: Narasimhaswamy, Supreeth, et al.
Published: (2024)
by: Narasimhaswamy, Supreeth, et al.
Published: (2024)
Prompt Refinement with Image Pivot for Text-to-Image Generation
by: Zhan, Jingtao, et al.
Published: (2024)
by: Zhan, Jingtao, et al.
Published: (2024)
Test-Time Personalization with Meta Prompt for Gaze Estimation
by: Liu, Huan, et al.
Published: (2024)
by: Liu, Huan, et al.
Published: (2024)
ARGENT: Adaptive Hierarchical Image-Text Representations
by: Huynh, Chuong, et al.
Published: (2026)
by: Huynh, Chuong, et al.
Published: (2026)
Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis
by: Yuan, Yu, et al.
Published: (2024)
by: Yuan, Yu, et al.
Published: (2024)
TextToucher: Fine-Grained Text-to-Touch Generation
by: Tu, Jiahang, et al.
Published: (2024)
by: Tu, Jiahang, et al.
Published: (2024)
ProDehaze: Prompting Diffusion Models Toward Faithful Image Dehazing
by: Zhou, Tianwen, et al.
Published: (2025)
by: Zhou, Tianwen, et al.
Published: (2025)
Generating Faithful and Salient Text from Multimodal Data
by: Hashem, Tahsina, et al.
Published: (2024)
by: Hashem, Tahsina, et al.
Published: (2024)
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
by: Jing, Zonglei, et al.
Published: (2025)
by: Jing, Zonglei, et al.
Published: (2025)
Can Better Text Semantics in Prompt Tuning Improve VLM Generalization?
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
Applying Graph Explanation to Operator Fusion
by: Mills, Keith G., et al.
Published: (2024)
by: Mills, Keith G., et al.
Published: (2024)
Prompting Medical Vision-Language Models to Mitigate Diagnosis Bias by Generating Realistic Dermoscopic Images
by: Munia, Nusrat, et al.
Published: (2025)
by: Munia, Nusrat, et al.
Published: (2025)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
by: Wu, Mingrui, et al.
Published: (2025)
by: Wu, Mingrui, et al.
Published: (2025)
PPBoost: Progressive Prompt Boosting for Text-Driven Medical Image Segmentation
by: Li, Xuchen, et al.
Published: (2025)
by: Li, Xuchen, et al.
Published: (2025)
Similar Items
-
PixelMan: Consistent Object Editing with Diffusion Models via Pixel Manipulation and Generation
by: Jiang, Liyao, et al.
Published: (2024) -
Qua$^2$SeDiMo: Quantifiable Quantization Sensitivity of Diffusion Models
by: Mills, Keith G., et al.
Published: (2024) -
RAISE: Requirement-Adaptive Evolutionary Refinement for Training-Free Text-to-Image Alignment
by: Jiang, Liyao, et al.
Published: (2026) -
CascadedGaze: Efficiency in Global Context Extraction for Image Restoration
by: Ghasemabadi, Amirhosein, et al.
Published: (2024) -
FunEditor: Achieving Complex Image Edits via Function Aggregation with Diffusion Models
by: Samadi, Mohammadreza, et al.
Published: (2024)