A Survey on Quality Metrics for Text-to-Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Hartwig, Sebastian, Engel, Dominik, Sick, Leon, Kniesel, Hannah, Payer, Tristan, Poonam, Poonam, Glöckler, Michael, Bäuerle, Alex, Ropinski, Timo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design
by: Engel, Dominik, et al.
Published: (2023)
by: Engel, Dominik, et al.
Published: (2023)
Weakly Supervised Virus Capsid Detection with Image-Level Annotations in Electron Microscopy Images
by: Kniesel, Hannah, et al.
Published: (2025)
by: Kniesel, Hannah, et al.
Published: (2025)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Attention-Guided Masked Autoencoders For Learning Image Representations
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Unsupervised Semantic Segmentation Through Depth-Guided Feature Correlation and Sampling
by: Sick, Leon, et al.
Published: (2023)
by: Sick, Leon, et al.
Published: (2023)
Evaluating Graphical Perception Capabilities of Vision Transformers
by: Poonam, Poonam, et al.
Published: (2026)
by: Poonam, Poonam, et al.
Published: (2026)
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
by: Kniesel, Hannah, et al.
Published: (2025)
by: Kniesel, Hannah, et al.
Published: (2025)
HPSCAN: Human Perception‐Based Scattered Data Clustering
by: S. Hartwig, et al.
Published: (2024)
by: S. Hartwig, et al.
Published: (2024)
S2D: Sparse-To-Dense Keymask Distillation for Unsupervised Video Instance Segmentation
by: Sick, Leon, et al.
Published: (2025)
by: Sick, Leon, et al.
Published: (2025)
EfficientMonoHair: Fast Strand-Level Reconstruction from Monocular Video via Multi-View Direction Fusion
by: Li, Da, et al.
Published: (2026)
by: Li, Da, et al.
Published: (2026)
Differentiable Electron Microscopy Simulation: Methods and Applications for Visualization
by: Nguyen, Ngan, et al.
Published: (2022)
by: Nguyen, Ngan, et al.
Published: (2022)
A Survey On Text-to-3D Contents Generation In The Wild
by: Jiang, Chenhan
Published: (2024)
by: Jiang, Chenhan
Published: (2024)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
by: Shah, Ansh, et al.
Published: (2024)
by: Shah, Ansh, et al.
Published: (2024)
CGVQM+D: Computer Graphics Video Quality Metric and Dataset
by: Jindal, Akshay, et al.
Published: (2025)
by: Jindal, Akshay, et al.
Published: (2025)
Bridging Text and Video Generation: A Survey
by: Kumar, Nilay, et al.
Published: (2025)
by: Kumar, Nilay, et al.
Published: (2025)
Multi-Layer Gaussian Splatting for Immersive Anatomy Visualization
by: Kleinbeck, Constantin, et al.
Published: (2024)
by: Kleinbeck, Constantin, et al.
Published: (2024)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
by: Zhang, Peiying, et al.
Published: (2025)
by: Zhang, Peiying, et al.
Published: (2025)
Expressive Text-to-Image Generation with Rich Text
by: Ge, Songwei, et al.
Published: (2023)
by: Ge, Songwei, et al.
Published: (2023)
Advancing Digital Twin Generation Through a Novel Simulation Framework and Quantitative Benchmarking
by: Rubinstein, Jacob, et al.
Published: (2026)
by: Rubinstein, Jacob, et al.
Published: (2026)
JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
by: Zeng, Yu, et al.
Published: (2024)
by: Zeng, Yu, et al.
Published: (2024)
Generating Human Interaction Motions in Scenes with Text Control
by: Yi, Hongwei, et al.
Published: (2024)
by: Yi, Hongwei, et al.
Published: (2024)
Generating Multi-Image Synthetic Data for Text-to-Image Customization
by: Kumari, Nupur, et al.
Published: (2025)
by: Kumari, Nupur, et al.
Published: (2025)
Improved 3D Scene Stylization via Text-Guided Generative Image Editing with Region-Based Control
by: Fujiwara, Haruo, et al.
Published: (2025)
by: Fujiwara, Haruo, et al.
Published: (2025)
SMPL-GPTexture: Dual-View 3D Human Texture Estimation using Text-to-Image Generation Models
by: Tu, Mingxiao, et al.
Published: (2025)
by: Tu, Mingxiao, et al.
Published: (2025)
GaussianDreamerPro: Text to Manipulable 3D Gaussians with Highly Enhanced Quality
by: Yi, Taoran, et al.
Published: (2024)
by: Yi, Taoran, et al.
Published: (2024)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
Controllable Video Generation: A Survey
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
Make It Count: Text-to-Image Generation with an Accurate Number of Objects
by: Binyamin, Lital, et al.
Published: (2024)
by: Binyamin, Lital, et al.
Published: (2024)
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
by: Lin, Yuanze, et al.
Published: (2024)
by: Lin, Yuanze, et al.
Published: (2024)
Text-to-Vector Generation with Neural Path Representation
by: Zhang, Peiying, et al.
Published: (2024)
by: Zhang, Peiying, et al.
Published: (2024)
Advances in 3D Generation: A Survey
by: Li, Xiaoyu, et al.
Published: (2024)
by: Li, Xiaoyu, et al.
Published: (2024)
MultiAct: Text-to-Motion Generation from Composite Text via Tailored Attention Guidance
by: Sala, Nathan, et al.
Published: (2026)
by: Sala, Nathan, et al.
Published: (2026)
Geometry Image Diffusion: Fast and Data-Efficient Text-to-3D with Image-Based Surface Representation
by: Elizarov, Slava, et al.
Published: (2024)
by: Elizarov, Slava, et al.
Published: (2024)
Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models
by: Fortes, Armando, et al.
Published: (2025)
by: Fortes, Armando, et al.
Published: (2025)
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
by: Khan, Mohammad Sadil, et al.
Published: (2024)
by: Khan, Mohammad Sadil, et al.
Published: (2024)
Text2NeRF: Text-Driven 3D Scene Generation with Neural Radiance Fields
by: Zhang, Jingbo, et al.
Published: (2023)
by: Zhang, Jingbo, et al.
Published: (2023)
LiftNav: Path Planning via Semantic Lifting in TSDF-Guided Gaussian Splatting
by: Schieber, Hannah, et al.
Published: (2026)
by: Schieber, Hannah, et al.
Published: (2026)
GenLit: Reformulating Single-Image Relighting as Video Generation
by: Bharadwaj, Shrisha, et al.
Published: (2024)
by: Bharadwaj, Shrisha, et al.
Published: (2024)
RustNeRF: Robust Neural Radiance Field with Low-Quality Images
by: Li, Mengfei, et al.
Published: (2024)
by: Li, Mengfei, et al.
Published: (2024)
Similar Items
-
Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design
by: Engel, Dominik, et al.
Published: (2023) -
Weakly Supervised Virus Capsid Detection with Image-Level Annotations in Electron Microscopy Images
by: Kniesel, Hannah, et al.
Published: (2025) -
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
by: Sick, Leon, et al.
Published: (2024) -
Attention-Guided Masked Autoencoders For Learning Image Representations
by: Sick, Leon, et al.
Published: (2024) -
Unsupervised Semantic Segmentation Through Depth-Guided Feature Correlation and Sampling
by: Sick, Leon, et al.
Published: (2023)