Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeong, Suchae, Choi, Inseong, Yun, Youngsik, Kim, Jihie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CIC: A Framework for Culturally-Aware Image Captioning
von: Yun, Youngsik, et al.
Veröffentlicht: (2024)
von: Yun, Youngsik, et al.
Veröffentlicht: (2024)
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
von: Park, Jiho, et al.
Veröffentlicht: (2025)
von: Park, Jiho, et al.
Veröffentlicht: (2025)
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
von: Kang, Sungmin, et al.
Veröffentlicht: (2024)
von: Kang, Sungmin, et al.
Veröffentlicht: (2024)
When Cultures Meet: Multicultural Text-to-Image Generation
von: Bhalerao, Parth, et al.
Veröffentlicht: (2025)
von: Bhalerao, Parth, et al.
Veröffentlicht: (2025)
Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
von: Park, Sangha, et al.
Veröffentlicht: (2025)
von: Park, Sangha, et al.
Veröffentlicht: (2025)
Improving LLM Classification of Logical Errors by Integrating Error Relationship into Prompts
von: Lee, Yanggyu, et al.
Veröffentlicht: (2024)
von: Lee, Yanggyu, et al.
Veröffentlicht: (2024)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
von: Park, SoYoung, et al.
Veröffentlicht: (2025)
von: Park, SoYoung, et al.
Veröffentlicht: (2025)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
von: Wu, Shang, et al.
Veröffentlicht: (2026)
von: Wu, Shang, et al.
Veröffentlicht: (2026)
VisDoT : Enhancing Visual Reasoning through Human-Like Interpretation Grounding and Decomposition of Thought
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
RusCode: Russian Cultural Code Benchmark for Text-to-Image Generation
von: Vasilev, Viacheslav, et al.
Veröffentlicht: (2025)
von: Vasilev, Viacheslav, et al.
Veröffentlicht: (2025)
Seeing is Improving: Visual Feedback for Iterative Text Layout Refinement
von: Guo, Junrong, et al.
Veröffentlicht: (2026)
von: Guo, Junrong, et al.
Veröffentlicht: (2026)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
Local Representative Token Guided Merging for Text-to-Image Generation
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
von: Lee, Min-Jeong, et al.
Veröffentlicht: (2025)
DragText: Rethinking Text Embedding in Point-based Image Editing
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
von: Choi, Gayoon, et al.
Veröffentlicht: (2024)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
PhyT2V: LLM-Guided Iterative Self-Refinement for Physics-Grounded Text-to-Video Generation
von: Xue, Qiyao, et al.
Veröffentlicht: (2024)
von: Xue, Qiyao, et al.
Veröffentlicht: (2024)
Iterative Prompt Refinement for Safer Text-to-Image Generation
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
von: Jeon, Jinwoo, et al.
Veröffentlicht: (2025)
MagicMan: Generative Novel View Synthesis of Humans with 3D-Aware Diffusion and Iterative Refinement
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Retrieval, Refinement, and Ranking for Text-to-Video Generation via Prompt Optimization and Test-Time Scaling
von: Rahman, Zillur, et al.
Veröffentlicht: (2026)
von: Rahman, Zillur, et al.
Veröffentlicht: (2026)
Refining Text-to-Image Generation: Towards Accurate Training-Free Glyph-Enhanced Image Generation
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
Iterative Refinement Improves Compositional Image Generation
von: Jaiswal, Shantanu, et al.
Veröffentlicht: (2026)
von: Jaiswal, Shantanu, et al.
Veröffentlicht: (2026)
AIR: Zero-shot Generative Model Adaptation with Iterative Refinement
von: Liu, Guimeng, et al.
Veröffentlicht: (2025)
von: Liu, Guimeng, et al.
Veröffentlicht: (2025)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
A Picture is Worth a Thousand Prompts? Efficacy of Iterative Human-Driven Prompt Refinement in Image Regeneration Tasks
von: Trinh, Khoi, et al.
Veröffentlicht: (2025)
von: Trinh, Khoi, et al.
Veröffentlicht: (2025)
Reverse Prompt: Cracking the Recipe Inside Text-to-Image Generation
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
von: Ren, Zhiyao, et al.
Veröffentlicht: (2025)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
Culture In a Frame: C$^3$B as a Comic-Based Benchmark for Multimodal Culturally Awareness
von: Song, Yuchen, et al.
Veröffentlicht: (2025)
von: Song, Yuchen, et al.
Veröffentlicht: (2025)
Text-to-Image Diffusion Models Cannot Count, and Prompt Refinement Cannot Help
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
Improved Iterative Refinement for Chart-to-Code Generation via Structured Instruction
von: Xu, Chengzhi, et al.
Veröffentlicht: (2025)
von: Xu, Chengzhi, et al.
Veröffentlicht: (2025)
Face-MakeUp: Multimodal Facial Prompts for Text-to-Image Generation
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
von: Dai, Dawei, et al.
Veröffentlicht: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
von: Saichandran, Ketan Suhaas, et al.
Veröffentlicht: (2025)
FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation
von: Teo, Christopher T. H, et al.
Veröffentlicht: (2024)
von: Teo, Christopher T. H, et al.
Veröffentlicht: (2024)
SCoFT: Self-Contrastive Fine-Tuning for Equitable Image Generation
von: Liu, Zhixuan, et al.
Veröffentlicht: (2024)
von: Liu, Zhixuan, et al.
Veröffentlicht: (2024)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
von: Lee, Mingyu, et al.
Veröffentlicht: (2024)
von: Lee, Mingyu, et al.
Veröffentlicht: (2024)
Erased but Exploitable: Black-box Embedding-Aware Prompting Against Unlearned Text-to-Image Diffusion Models
von: Koma, Arian Komaei, et al.
Veröffentlicht: (2026)
von: Koma, Arian Komaei, et al.
Veröffentlicht: (2026)
Exploiting Style Latent Flows for Generalizing Deepfake Video Detection
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
von: Lim, Youngsun, et al.
Veröffentlicht: (2026)
von: Lim, Youngsun, et al.
Veröffentlicht: (2026)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
von: Kim, Sanghyun, et al.
Veröffentlicht: (2024)
Naïve PAINE: Lightweight Text-to-Image Generation Improvement with Prompt Evaluation
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CIC: A Framework for Culturally-Aware Image Captioning
von: Yun, Youngsik, et al.
Veröffentlicht: (2024) -
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
von: Park, Jiho, et al.
Veröffentlicht: (2025) -
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
von: Kang, Sungmin, et al.
Veröffentlicht: (2024) -
When Cultures Meet: Multicultural Text-to-Image Generation
von: Bhalerao, Parth, et al.
Veröffentlicht: (2025) -
Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)