Towards Understanding and Quantifying Uncertainty for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Franchi, Gianni, Trong, Dat Nguyen, Belkhir, Nacim, Xia, Guoxuan, Pilzer, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NECO: NEural Collapse Based Out-of-distribution detection
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2023)
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2023)
Towards Understanding Why Label Smoothing Degrades Selective Classification and How to Fix It
von: Xia, Guoxuan, et al.
Veröffentlicht: (2024)
von: Xia, Guoxuan, et al.
Veröffentlicht: (2024)
From Local Geometry to Global Pseudo Labeling for Robust Positive Unlabeled Learning under Covariate Shift
von: Gabetni, Firas, et al.
Veröffentlicht: (2026)
von: Gabetni, Firas, et al.
Veröffentlicht: (2026)
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
von: Gabetni, Firas, et al.
Veröffentlicht: (2025)
von: Gabetni, Firas, et al.
Veröffentlicht: (2025)
Foundation Models and Transformers for Anomaly Detection: A Survey
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2025)
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2025)
A Geometric Unification of Concept Learning with Concept Cones
von: Rocchi--Henry, Alexandre, et al.
Veröffentlicht: (2025)
von: Rocchi--Henry, Alexandre, et al.
Veröffentlicht: (2025)
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
von: Chorna, Sofiia, et al.
Veröffentlicht: (2025)
von: Chorna, Sofiia, et al.
Veröffentlicht: (2025)
COOkeD: Ensemble-based OOD detection in the era of zero-shot CLIP
von: Humblot-Renaux, Galadrielle, et al.
Veröffentlicht: (2025)
von: Humblot-Renaux, Galadrielle, et al.
Veröffentlicht: (2025)
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
von: Hahn, Meera, et al.
Veröffentlicht: (2024)
Understanding Implosion in Text-to-Image Generative Models
von: Ding, Wenxin, et al.
Veröffentlicht: (2024)
von: Ding, Wenxin, et al.
Veröffentlicht: (2024)
Quantifying Deep Learning Model Uncertainty in Conformal Prediction
von: Karimi, Hamed, et al.
Veröffentlicht: (2023)
von: Karimi, Hamed, et al.
Veröffentlicht: (2023)
GRADEO: Towards Human-Like Evaluation for Text-to-Video Generation via Multi-Step Reasoning
von: Mou, Zhun, et al.
Veröffentlicht: (2025)
von: Mou, Zhun, et al.
Veröffentlicht: (2025)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
Text-To-Image with Generative Adversarial Networks
von: Momen-Tayefeh, Mehrshad
Veröffentlicht: (2024)
von: Momen-Tayefeh, Mehrshad
Veröffentlicht: (2024)
Learning to Stop Overthinking at Test Time
von: Bao, Hieu Tran, et al.
Veröffentlicht: (2025)
von: Bao, Hieu Tran, et al.
Veröffentlicht: (2025)
Towards Effective Usage of Human-Centric Priors in Diffusion Models for Text-based Human Image Generation
von: Wang, Junyan, et al.
Veröffentlicht: (2024)
von: Wang, Junyan, et al.
Veröffentlicht: (2024)
Towards Evaluating Robustness of Prompt Adherence in Text to Image Models
von: Vemishetty, Sujith, et al.
Veröffentlicht: (2025)
von: Vemishetty, Sujith, et al.
Veröffentlicht: (2025)
On the Scalability of Diffusion-based Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
EdgeFusion: On-Device Text-to-Image Generation
von: Castells, Thibault, et al.
Veröffentlicht: (2024)
von: Castells, Thibault, et al.
Veröffentlicht: (2024)
Dual Diffusion for Unified Image Generation and Understanding
von: Li, Zijie, et al.
Veröffentlicht: (2024)
von: Li, Zijie, et al.
Veröffentlicht: (2024)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
von: Mamedov, Timur, et al.
Veröffentlicht: (2026)
Evaluating Text-to-Visual Generation with Image-to-Text Generation
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2024)
von: Lin, Zhiqiu, et al.
Veröffentlicht: (2024)
Compositional Text-to-Image Generation with Dense Blob Representations
von: Nie, Weili, et al.
Veröffentlicht: (2024)
von: Nie, Weili, et al.
Veröffentlicht: (2024)
Towards Understanding Deep Learning Model in Image Recognition via Coverage Test
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
von: Li, Wenkai, et al.
Veröffentlicht: (2025)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
von: Kang, Wonjun, et al.
Veröffentlicht: (2025)
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
von: Yu, Xin, et al.
Veröffentlicht: (2025)
von: Yu, Xin, et al.
Veröffentlicht: (2025)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
JetFormer: An Autoregressive Generative Model of Raw Images and Text
von: Tschannen, Michael, et al.
Veröffentlicht: (2024)
von: Tschannen, Michael, et al.
Veröffentlicht: (2024)
Minority-Focused Text-to-Image Generation via Prompt Optimization
von: Um, Soobin, et al.
Veröffentlicht: (2024)
von: Um, Soobin, et al.
Veröffentlicht: (2024)
PreciseCam: Precise Camera Control for Text-to-Image Generation
von: Bernal-Berdun, Edurne, et al.
Veröffentlicht: (2025)
von: Bernal-Berdun, Edurne, et al.
Veröffentlicht: (2025)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Towards Resolving Optimization Conflicts Between Image- and Text-Based Person Re-Identification
von: Kvanchiani, Karina, et al.
Veröffentlicht: (2026)
von: Kvanchiani, Karina, et al.
Veröffentlicht: (2026)
Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2026)
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2026)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
Naïve PAINE: Lightweight Text-to-Image Generation Improvement with Prompt Evaluation
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
Harmonizing Generalization and Specialization: Uncertainty-Informed Collaborative Learning for Semi-supervised Medical Image Segmentation
von: Lu, Wenjing, et al.
Veröffentlicht: (2025)
von: Lu, Wenjing, et al.
Veröffentlicht: (2025)
Training-Free Consistent Text-to-Image Generation
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
von: Tewel, Yoad, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
NECO: NEural Collapse Based Out-of-distribution detection
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2023) -
Towards Understanding Why Label Smoothing Degrades Selective Classification and How to Fix It
von: Xia, Guoxuan, et al.
Veröffentlicht: (2024) -
From Local Geometry to Global Pseudo Labeling for Robust Positive Unlabeled Learning under Covariate Shift
von: Gabetni, Firas, et al.
Veröffentlicht: (2026) -
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
von: Gabetni, Firas, et al.
Veröffentlicht: (2025) -
Foundation Models and Transformers for Anomaly Detection: A Survey
von: Ammar, Mouïn Ben, et al.
Veröffentlicht: (2025)