When Cultures Meet: Multicultural Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhalerao, Parth, Yalamarty, Mounika, Trinh, Brian, Ignat, Oana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation
von: Li, Shuowei, et al.
Veröffentlicht: (2026)
von: Li, Shuowei, et al.
Veröffentlicht: (2026)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
von: Bai, Longju, et al.
Veröffentlicht: (2024)
von: Bai, Longju, et al.
Veröffentlicht: (2024)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
von: Ignat, Oana, et al.
Veröffentlicht: (2024)
von: Ignat, Oana, et al.
Veröffentlicht: (2024)
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)
Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
von: Bhalerao, Parth, et al.
Veröffentlicht: (2026)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
von: Jeong, Suchae, et al.
Veröffentlicht: (2025)
When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators
von: Adamkiewicz, Krzysztof, et al.
Veröffentlicht: (2026)
von: Adamkiewicz, Krzysztof, et al.
Veröffentlicht: (2026)
Where Culture Fades: Revealing the Cultural Gap in Text-to-Image Generation
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
von: Shi, Chuancheng, et al.
Veröffentlicht: (2025)
When Color-Space Decoupling Meets Diffusion for Adverse-Weather Image Restoration
von: Fang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Fang, Wenxuan, et al.
Veröffentlicht: (2025)
MDS-ViTNet: Improving saliency prediction for Eye-Tracking with Vision Transformer
von: Ignat, Polezhaev, et al.
Veröffentlicht: (2024)
von: Ignat, Polezhaev, et al.
Veröffentlicht: (2024)
Enlighten-Your-Voice: When Multimodal Meets Zero-shot Low-light Image Enhancement
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaofeng, et al.
Veröffentlicht: (2023)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
RusCode: Russian Cultural Code Benchmark for Text-to-Image Generation
von: Vasilev, Viacheslav, et al.
Veröffentlicht: (2025)
von: Vasilev, Viacheslav, et al.
Veröffentlicht: (2025)
When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
von: Luo, Junwei, et al.
Veröffentlicht: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
When Text and Images Don't Mix: Bias-Correcting Language-Image Similarity Scores for Anomaly Detection
von: Goodge, Adam, et al.
Veröffentlicht: (2024)
von: Goodge, Adam, et al.
Veröffentlicht: (2024)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
von: Nwatu, Joan, et al.
Veröffentlicht: (2024)
When Words Smile: Generating Diverse Emotional Facial Expressions from Text
von: Xu, Haidong, et al.
Veröffentlicht: (2024)
von: Xu, Haidong, et al.
Veröffentlicht: (2024)
Agentic Retoucher for Text-To-Image Generation
von: Shen, Shaocheng, et al.
Veröffentlicht: (2026)
von: Shen, Shaocheng, et al.
Veröffentlicht: (2026)
STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding
von: Kim, Junho, et al.
Veröffentlicht: (2026)
von: Kim, Junho, et al.
Veröffentlicht: (2026)
When Cars Have Stereotypes: Auditing Demographic Bias in Objects from Text-to-Image Models
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
von: Choi, Dasol, et al.
Veröffentlicht: (2025)
When VLMs Meet Image Classification: Test Sets Renovation via Missing Label Identification
von: Pang, Zirui, et al.
Veröffentlicht: (2025)
von: Pang, Zirui, et al.
Veröffentlicht: (2025)
Addressing Image Authenticity When Cameras Use Generative AI
von: Masud, Umar, et al.
Veröffentlicht: (2026)
von: Masud, Umar, et al.
Veröffentlicht: (2026)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
von: Nayak, Shravan, et al.
Veröffentlicht: (2025)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
Condition Weaving Meets Expert Modulation: Towards Universal and Controllable Image Generation
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
AI-Generated Images: What Humans and Machines See When They Look at the Same Image
von: Poletti, Silvia, et al.
Veröffentlicht: (2026)
von: Poletti, Silvia, et al.
Veröffentlicht: (2026)
Personalized Reward Modeling for Text-to-Image Generation
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
von: Lee, Jeongeun, et al.
Veröffentlicht: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
CLAReSNet: When Convolution Meets Latent Attention for Hyperspectral Image Classification
von: Bandyopadhyay, Asmit, et al.
Veröffentlicht: (2025)
von: Bandyopadhyay, Asmit, et al.
Veröffentlicht: (2025)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
von: Lim, Youngsun, et al.
Veröffentlicht: (2024)
Symbolic Rule Extraction from Attention-Guided Sparse Representations in Vision Transformers
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
TypeScore: A Text Fidelity Metric for Text-to-Image Generative Models
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
von: Sampaio, Georgia Gabriela, et al.
Veröffentlicht: (2024)
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data
von: Berman, William, et al.
Veröffentlicht: (2024)
von: Berman, William, et al.
Veröffentlicht: (2024)
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
von: Li, Meiling, et al.
Veröffentlicht: (2024)
von: Li, Meiling, et al.
Veröffentlicht: (2024)
Refining Text-to-Image Generation: Towards Accurate Training-Free Glyph-Enhanced Image Generation
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
von: Lakhanpal, Sanyam, et al.
Veröffentlicht: (2024)
HARIVO: Harnessing Text-to-Image Models for Video Generation
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
von: Kwon, Mingi, et al.
Veröffentlicht: (2024)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
Interactive Visual Assessment for Text-to-Image Generation Models
von: Mi, Xiaoyue, et al.
Veröffentlicht: (2024)
von: Mi, Xiaoyue, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation
von: Li, Shuowei, et al.
Veröffentlicht: (2026) -
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
von: Bai, Longju, et al.
Veröffentlicht: (2024) -
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
von: Zhao, Yuming, et al.
Veröffentlicht: (2026) -
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
von: Ignat, Oana, et al.
Veröffentlicht: (2024) -
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
von: Nwatu, Joan, et al.
Veröffentlicht: (2025)