Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Muxi, Liu, Yi, Yi, Jian, Xu, Changran, Lai, Qiuxia, Wang, Hongliang, Ho, Tsung-Yi, Xu, Qiang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RIGID: A Training-free and Model-Agnostic Framework for Robust AI-Generated Image Detection
by: He, Zhiyuan, et al.
Published: (2024)
by: He, Zhiyuan, et al.
Published: (2024)
Image Regeneration: Evaluating Text-to-Image Model via Generating Identical Image with Multimodal Large Language Models
by: Meng, Chutian, et al.
Published: (2024)
by: Meng, Chutian, et al.
Published: (2024)
MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
by: Zhao, Chenchen, et al.
Published: (2026)
by: Zhao, Chenchen, et al.
Published: (2026)
HiBug2: Efficient and Interpretable Error Slice Discovery for Comprehensive Model Debugging
by: Chen, Muxi, et al.
Published: (2025)
by: Chen, Muxi, et al.
Published: (2025)
Unsupervised Out-of-Distribution Detection in Medical Imaging Using Multi-Exit Class Activation Maps and Feature Masking
by: Chen, Yu-Jen, et al.
Published: (2025)
by: Chen, Yu-Jen, et al.
Published: (2025)
3DIS: Depth-Driven Decoupled Instance Synthesis for Text-to-Image Generation
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
by: Wang, Xinran, et al.
Published: (2025)
by: Wang, Xinran, et al.
Published: (2025)
MMA-Diffusion: MultiModal Attack on Diffusion Models
by: Yang, Yijun, et al.
Published: (2023)
by: Yang, Yijun, et al.
Published: (2023)
Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition
by: Lin, Tiancheng, et al.
Published: (2024)
by: Lin, Tiancheng, et al.
Published: (2024)
Temporal Consistency-Aware Text-to-Motion Generation
by: Wang, Hongsong, et al.
Published: (2026)
by: Wang, Hongsong, et al.
Published: (2026)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
by: Wu, Fan, et al.
Published: (2025)
by: Wu, Fan, et al.
Published: (2025)
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation
by: Wang, Wenhao, et al.
Published: (2024)
by: Wang, Wenhao, et al.
Published: (2024)
An Empirical Study and Analysis of Text-to-Image Generation Using Large Language Model-Powered Textual Representation
by: Tan, Zhiyu, et al.
Published: (2024)
by: Tan, Zhiyu, et al.
Published: (2024)
Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation
by: Li, Niantong, et al.
Published: (2026)
by: Li, Niantong, et al.
Published: (2026)
Text-IF: Leveraging Semantic Text Guidance for Degradation-Aware and Interactive Image Fusion
by: Yi, Xunpeng, et al.
Published: (2024)
by: Yi, Xunpeng, et al.
Published: (2024)
Fine-grained Text to Image Synthesis
by: Ouyang, Xu, et al.
Published: (2024)
by: Ouyang, Xu, et al.
Published: (2024)
Towards Understanding the Working Mechanism of Text-to-Image Diffusion Model
by: Yi, Mingyang, et al.
Published: (2024)
by: Yi, Mingyang, et al.
Published: (2024)
Vector Quantization Prompting for Continual Learning
by: Jiao, Li, et al.
Published: (2024)
by: Jiao, Li, et al.
Published: (2024)
HyperVQ: Enabling Hyperprior Entropy Modeling for VQ-Based Generative Image Compression
by: Yi, Niu, et al.
Published: (2025)
by: Yi, Niu, et al.
Published: (2025)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
by: Cui, Tianyu, et al.
Published: (2025)
by: Cui, Tianyu, et al.
Published: (2025)
Origin Identification for Text-Guided Image-to-Image Diffusion Models
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
MPDS: A Movie Posters Dataset for Image Generation with Diffusion Model
by: Xu, Meng, et al.
Published: (2024)
by: Xu, Meng, et al.
Published: (2024)
BideDPO: Conditional Image Generation with Simultaneous Text and Condition Alignment
by: Zhou, Dewei, et al.
Published: (2025)
by: Zhou, Dewei, et al.
Published: (2025)
TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models
by: Hsiao, Teng-Fang, et al.
Published: (2025)
by: Hsiao, Teng-Fang, et al.
Published: (2025)
Multi-rater Prompting for Ambiguous Medical Image Segmentation
by: Wang, Jinhong, et al.
Published: (2024)
by: Wang, Jinhong, et al.
Published: (2024)
HumanCoser: Layered 3D Human Generation via Semantic-Aware Diffusion Model
by: Wang, Yi, et al.
Published: (2024)
by: Wang, Yi, et al.
Published: (2024)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
by: Han, Jian, et al.
Published: (2024)
by: Han, Jian, et al.
Published: (2024)
EmotiCrafter: Text-to-Emotional-Image Generation based on Valence-Arousal Model
by: Dang, Shengqi, et al.
Published: (2025)
by: Dang, Shengqi, et al.
Published: (2025)
Iterative Online Image Synthesis via Diffusion Model for Imbalanced Classification
by: Li, Shuhan, et al.
Published: (2024)
by: Li, Shuhan, et al.
Published: (2024)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
by: Hu, Teng, et al.
Published: (2023)
by: Hu, Teng, et al.
Published: (2023)
TIPO: Text to Image with Text Presampling for Prompt Optimization
by: Yeh, Shih-Ying, et al.
Published: (2024)
by: Yeh, Shih-Ying, et al.
Published: (2024)
Information Bottleneck Approach to Spatial Attention Learning
by: Lai, Qiuxia, et al.
Published: (2021)
by: Lai, Qiuxia, et al.
Published: (2021)
Intriguing Properties of Diffusion Models: An Empirical Study of the Natural Attack Capability in Text-to-Image Generative Models
by: Sato, Takami, et al.
Published: (2023)
by: Sato, Takami, et al.
Published: (2023)
InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing
by: Yu, Haoran, et al.
Published: (2025)
by: Yu, Haoran, et al.
Published: (2025)
Sparsity Meets Similarity: Leveraging Long-Tail Distribution for Dynamic Optimized Token Representation in Multimodal Large Language Models
by: Yu, Gaotong, et al.
Published: (2024)
by: Yu, Gaotong, et al.
Published: (2024)
Layered 3D Human Generation via Semantic-Aware Diffusion Model
by: Wang, Yi, et al.
Published: (2023)
by: Wang, Yi, et al.
Published: (2023)
CogBlender: Towards Continuous Cognitive Intervention in Text-to-Image Generation
by: Dang, Shengqi, et al.
Published: (2026)
by: Dang, Shengqi, et al.
Published: (2026)
MIGC++: Advanced Multi-Instance Generation Controller for Image Synthesis
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
Text-Image Conditioned 3D Generation
by: Cen, Jiazhong, et al.
Published: (2026)
by: Cen, Jiazhong, et al.
Published: (2026)
Similar Items
-
RIGID: A Training-free and Model-Agnostic Framework for Robust AI-Generated Image Detection
by: He, Zhiyuan, et al.
Published: (2024) -
Image Regeneration: Evaluating Text-to-Image Model via Generating Identical Image with Multimodal Large Language Models
by: Meng, Chutian, et al.
Published: (2024) -
MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
by: Zhou, Dewei, et al.
Published: (2024) -
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
by: Zhao, Chenchen, et al.
Published: (2026) -
HiBug2: Efficient and Interpretable Error Slice Discovery for Comprehensive Model Debugging
by: Chen, Muxi, et al.
Published: (2025)