TIAM -- A Metric for Evaluating Alignment in Text-to-Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Grimal, Paul, Borgne, Hervé Le, Ferret, Olivier, Tourille, Julien |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Text-to-Image Alignment in Denoising-Based Models through Step Selection
por: Grimal, Paul, et al.
Publicado: (2025)
por: Grimal, Paul, et al.
Publicado: (2025)
SAGA: Learning Signal-Aligned Distributions for Improved Text-to-Image Generation
por: Grimal, Paul, et al.
Publicado: (2025)
por: Grimal, Paul, et al.
Publicado: (2025)
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
por: Nayak, Shravan, et al.
Publicado: (2025)
por: Nayak, Shravan, et al.
Publicado: (2025)
Evaluating Text-to-Visual Generation with Image-to-Text Generation
por: Lin, Zhiqiu, et al.
Publicado: (2024)
por: Lin, Zhiqiu, et al.
Publicado: (2024)
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
por: Zhang, Huixuan, et al.
Publicado: (2025)
por: Zhang, Huixuan, et al.
Publicado: (2025)
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
por: Ahmadi, Saba, et al.
Publicado: (2023)
por: Ahmadi, Saba, et al.
Publicado: (2023)
Davidsonian Scene Graph: Improving Reliability in Fine-grained Evaluation for Text-to-Image Generation
por: Cho, Jaemin, et al.
Publicado: (2023)
por: Cho, Jaemin, et al.
Publicado: (2023)
Erasing with Precision: Evaluating Specific Concept Erasure from Text-to-Image Generative Models
por: Fuchi, Masane, et al.
Publicado: (2025)
por: Fuchi, Masane, et al.
Publicado: (2025)
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective
por: Huang, Hailang, et al.
Publicado: (2024)
por: Huang, Hailang, et al.
Publicado: (2024)
Holistic Evaluation for Interleaved Text-and-Image Generation
por: Liu, Minqian, et al.
Publicado: (2024)
por: Liu, Minqian, et al.
Publicado: (2024)
Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation
por: Chen, Wenting, et al.
Publicado: (2023)
por: Chen, Wenting, et al.
Publicado: (2023)
Translation-Enhanced Multilingual Text-to-Image Generation
por: Li, Yaoyiran, et al.
Publicado: (2023)
por: Li, Yaoyiran, et al.
Publicado: (2023)
Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)
por: Saxon, Michael, et al.
Publicado: (2024)
por: Saxon, Michael, et al.
Publicado: (2024)
Interleaving Reasoning for Better Text-to-Image Generation
por: Huang, Wenxuan, et al.
Publicado: (2025)
por: Huang, Wenxuan, et al.
Publicado: (2025)
Tables as Texts or Images: Evaluating the Table Reasoning Ability of LLMs and MLLMs
por: Deng, Naihao, et al.
Publicado: (2024)
por: Deng, Naihao, et al.
Publicado: (2024)
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
por: Winata, Genta Indra, et al.
Publicado: (2024)
por: Winata, Genta Indra, et al.
Publicado: (2024)
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations
por: Lan, Tian, et al.
Publicado: (2025)
por: Lan, Tian, et al.
Publicado: (2025)
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation
por: Wang, Andrew Z., et al.
Publicado: (2025)
por: Wang, Andrew Z., et al.
Publicado: (2025)
Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation
por: Yuan, Huizhuo, et al.
Publicado: (2024)
por: Yuan, Huizhuo, et al.
Publicado: (2024)
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
por: He, Yutong, et al.
Publicado: (2024)
por: He, Yutong, et al.
Publicado: (2024)
DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation
por: Jiang, Dongzhi, et al.
Publicado: (2025)
por: Jiang, Dongzhi, et al.
Publicado: (2025)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
por: Zhu, Mengdan, et al.
Publicado: (2025)
por: Zhu, Mengdan, et al.
Publicado: (2025)
EXPERT: An Explainable Image Captioning Evaluation Metric with Structured Explanations
por: Kim, Hyunjong, et al.
Publicado: (2025)
por: Kim, Hyunjong, et al.
Publicado: (2025)
MINOS: A Multimodal Evaluation Model for Bidirectional Generation Between Image and Text
por: Zhang, Junzhe, et al.
Publicado: (2025)
por: Zhang, Junzhe, et al.
Publicado: (2025)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
por: Chen, Wenting, et al.
Publicado: (2024)
por: Chen, Wenting, et al.
Publicado: (2024)
Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation
por: Zhang, Yuhui, et al.
Publicado: (2023)
por: Zhang, Yuhui, et al.
Publicado: (2023)
SELMA: Learning and Merging Skill-Specific Text-to-Image Experts with Auto-Generated Data
por: Li, Jialu, et al.
Publicado: (2024)
por: Li, Jialu, et al.
Publicado: (2024)
Guided Score identity Distillation for Data-Free One-Step Text-to-Image Generation
por: Zhou, Mingyuan, et al.
Publicado: (2024)
por: Zhou, Mingyuan, et al.
Publicado: (2024)
A Multimodal, Multitask System for Generating E Commerce Text Listings from Images
por: Singh, Nayan Kumar
Publicado: (2025)
por: Singh, Nayan Kumar
Publicado: (2025)
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning
por: Matsuda, Kazuki, et al.
Publicado: (2024)
por: Matsuda, Kazuki, et al.
Publicado: (2024)
Evaluating Semantic Variation in Text-to-Image Synthesis: A Causal Perspective
por: Zhu, Xiangru, et al.
Publicado: (2024)
por: Zhu, Xiangru, et al.
Publicado: (2024)
Explaining How Visual, Textual and Multimodal Encoders Share Concepts
por: Cornet, Clément, et al.
Publicado: (2025)
por: Cornet, Clément, et al.
Publicado: (2025)
Unified Vision-Language Modeling via Concept Space Alignment
por: Qiu, Yifu, et al.
Publicado: (2026)
por: Qiu, Yifu, et al.
Publicado: (2026)
GenAI-Bench: Evaluating and Improving Compositional Text-to-Visual Generation
por: Li, Baiqi, et al.
Publicado: (2024)
por: Li, Baiqi, et al.
Publicado: (2024)
CapsFusion: Rethinking Image-Text Data at Scale
por: Yu, Qiying, et al.
Publicado: (2023)
por: Yu, Qiying, et al.
Publicado: (2023)
Vision Learners Meet Web Image-Text Pairs
por: Zhao, Bingchen, et al.
Publicado: (2023)
por: Zhao, Bingchen, et al.
Publicado: (2023)
Multi-Modal Language Models as Text-to-Image Model Evaluators
por: Chen, Jiahui, et al.
Publicado: (2025)
por: Chen, Jiahui, et al.
Publicado: (2025)
TempViz: On the Evaluation of Temporal Knowledge in Text-to-Image Models
por: Holtermann, Carolin, et al.
Publicado: (2026)
por: Holtermann, Carolin, et al.
Publicado: (2026)
Preference Adaptive and Sequential Text-to-Image Generation
por: Nabati, Ofir, et al.
Publicado: (2024)
por: Nabati, Ofir, et al.
Publicado: (2024)
SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation
por: Qu, Leigang, et al.
Publicado: (2024)
por: Qu, Leigang, et al.
Publicado: (2024)
Ejemplares similares
-
Text-to-Image Alignment in Denoising-Based Models through Step Selection
por: Grimal, Paul, et al.
Publicado: (2025) -
SAGA: Learning Signal-Aligned Distributions for Improved Text-to-Image Generation
por: Grimal, Paul, et al.
Publicado: (2025) -
CulturalFrames: Assessing Cultural Expectation Alignment in Text-to-Image Models and Evaluation Metrics
por: Nayak, Shravan, et al.
Publicado: (2025) -
Evaluating Text-to-Visual Generation with Image-to-Text Generation
por: Lin, Zhiqiu, et al.
Publicado: (2024) -
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
por: Zhang, Huixuan, et al.
Publicado: (2025)