Human Image Generation: A Comprehensive Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Jia, Zhen, Zhang, Zhang, Wang, Liang, Tan, Tieniu |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts
by: Liang, Jian, et al.
Published: (2023)
by: Liang, Jian, et al.
Published: (2023)
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
by: Xie, Wulin, et al.
Published: (2025)
by: Xie, Wulin, et al.
Published: (2025)
Aligning Multimodal LLM with Human Preference: A Survey
by: Yu, Tao, et al.
Published: (2025)
by: Yu, Tao, et al.
Published: (2025)
Artifact Feature Purification for Cross-domain Detection of AI-generated Images
by: Meng, Zheling, et al.
Published: (2024)
by: Meng, Zheling, et al.
Published: (2024)
Making Images Real Again: A Comprehensive Survey on Deep Image Composition
by: Niu, Li, et al.
Published: (2021)
by: Niu, Li, et al.
Published: (2021)
How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing
by: Zhang, Huanyu, et al.
Published: (2026)
by: Zhang, Huanyu, et al.
Published: (2026)
Learning the Degradation Distribution for Blind Image Super-Resolution
by: Luo, Zhengxiong, et al.
Published: (2022)
by: Luo, Zhengxiong, et al.
Published: (2022)
CTForensics: A Comprehensive Dataset and Method for AI-Generated CT Image Detection
by: Li, Yiheng, et al.
Published: (2026)
by: Li, Yiheng, et al.
Published: (2026)
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
by: Zhang, Yi-Fan, et al.
Published: (2024)
by: Zhang, Yi-Fan, et al.
Published: (2024)
Self-Supervised Learning for Image Segmentation: A Comprehensive Survey
by: Akilan, Thangarajah, et al.
Published: (2025)
by: Akilan, Thangarajah, et al.
Published: (2025)
The Illusion of Progress? A Critical Look at Test-Time Adaptation for Vision-Language Models
by: Sheng, Lijun, et al.
Published: (2025)
by: Sheng, Lijun, et al.
Published: (2025)
DeltaEdit: Exploring Text-free Training for Text-Driven Image Manipulation
by: Lyu, Yueming, et al.
Published: (2023)
by: Lyu, Yueming, et al.
Published: (2023)
Emotion Recognition from Skeleton Data: A Comprehensive Survey
by: Lu, Haifeng, et al.
Published: (2025)
by: Lu, Haifeng, et al.
Published: (2025)
Chatting with Images for Introspective Visual Thinking
by: Wu, Junfei, et al.
Published: (2026)
by: Wu, Junfei, et al.
Published: (2026)
Realistic Unsupervised CLIP Fine-tuning with Universal Entropy Optimization
by: Liang, Jian, et al.
Published: (2023)
by: Liang, Jian, et al.
Published: (2023)
Is Sora a World Simulator? A Comprehensive Survey on General World Models and Beyond
by: Zhu, Zheng, et al.
Published: (2024)
by: Zhu, Zheng, et al.
Published: (2024)
Erasing Concepts, Steering Generations: A Comprehensive Survey of Concept Suppression
by: Xie, Yiwei, et al.
Published: (2025)
by: Xie, Yiwei, et al.
Published: (2025)
Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing
by: Wu, Sihao, et al.
Published: (2025)
by: Wu, Sihao, et al.
Published: (2025)
TEST-V: TEst-time Support-set Tuning for Zero-shot Video Classification
by: Yan, Rui, et al.
Published: (2025)
by: Yan, Rui, et al.
Published: (2025)
Enhancing End-to-End Autonomous Driving with Latent World Model
by: Li, Yingyan, et al.
Published: (2024)
by: Li, Yingyan, et al.
Published: (2024)
Generating Multimodal Images with GAN: Integrating Text, Image, and Style
by: Tan, Chaoyi, et al.
Published: (2025)
by: Tan, Chaoyi, et al.
Published: (2025)
A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights
by: Lei, Wentao, et al.
Published: (2024)
by: Lei, Wentao, et al.
Published: (2024)
Diffusion Models for Image Restoration and Enhancement: A Comprehensive Survey
by: Li, Xin, et al.
Published: (2023)
by: Li, Xin, et al.
Published: (2023)
Harmonizing Visual Text Comprehension and Generation
by: Zhao, Zhen, et al.
Published: (2024)
by: Zhao, Zhen, et al.
Published: (2024)
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing
by: Chen, Zhihong, et al.
Published: (2025)
by: Chen, Zhihong, et al.
Published: (2025)
Towards Compatible Fine-tuning for Vision-Language Model Updates
by: Wang, Zhengbo, et al.
Published: (2024)
by: Wang, Zhengbo, et al.
Published: (2024)
SAM2 for Image and Video Segmentation: A Comprehensive Survey
by: Jiaxing, Zhang, et al.
Published: (2025)
by: Jiaxing, Zhang, et al.
Published: (2025)
RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images
by: Wang, Benzhi, et al.
Published: (2024)
by: Wang, Benzhi, et al.
Published: (2024)
Vision Transformer with Super Token Sampling
by: Huang, Huaibo, et al.
Published: (2022)
by: Huang, Huaibo, et al.
Published: (2022)
MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition
by: Wei, Xinyu, et al.
Published: (2025)
by: Wei, Xinyu, et al.
Published: (2025)
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
Prompt Mechanisms in Medical Imaging: A Comprehensive Survey
by: Yang, Hao, et al.
Published: (2025)
by: Yang, Hao, et al.
Published: (2025)
Connecting the Dots: Collaborative Fine-tuning for Black-Box Vision-Language Models
by: Wang, Zhengbo, et al.
Published: (2024)
by: Wang, Zhengbo, et al.
Published: (2024)
Exploring Vacant Classes in Label-Skewed Federated Learning
by: Guo, Kuangpu, et al.
Published: (2024)
by: Guo, Kuangpu, et al.
Published: (2024)
The Role of World Models in Shaping Autonomous Driving: A Comprehensive Survey
by: Tu, Sifan, et al.
Published: (2025)
by: Tu, Sifan, et al.
Published: (2025)
Recovering 3D Human Mesh from Monocular Images: A Survey
by: Tian, Yating, et al.
Published: (2022)
by: Tian, Yating, et al.
Published: (2022)
Evaluating and Predicting Distorted Human Body Parts for Generated Images
by: Ma, Lu, et al.
Published: (2025)
by: Ma, Lu, et al.
Published: (2025)
Learning Multi-dimensional Human Preference for Text-to-Image Generation
by: Zhang, Sixian, et al.
Published: (2024)
by: Zhang, Sixian, et al.
Published: (2024)
Single-Image Shadow Removal Using Deep Learning: A Comprehensive Survey
by: Guo, Laniqng, et al.
Published: (2024)
by: Guo, Laniqng, et al.
Published: (2024)
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey
by: Wang, Yaoting, et al.
Published: (2025)
by: Wang, Yaoting, et al.
Published: (2025)
Similar Items
-
A Comprehensive Survey on Test-Time Adaptation under Distribution Shifts
by: Liang, Jian, et al.
Published: (2023) -
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
by: Xie, Wulin, et al.
Published: (2025) -
Aligning Multimodal LLM with Human Preference: A Survey
by: Yu, Tao, et al.
Published: (2025) -
Artifact Feature Purification for Cross-domain Detection of AI-generated Images
by: Meng, Zheling, et al.
Published: (2024) -
Making Images Real Again: A Comprehensive Survey on Deep Image Composition
by: Niu, Li, et al.
Published: (2021)