Rethinking FID: Towards a Better Evaluation Metric for Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Jayasumana, Sadeep, Ramalingam, Srikumar, Veit, Andreas, Glasner, Daniel, Chakrabarti, Ayan, Kumar, Sanjiv |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LatentCRF: Continuous CRF for Efficient Latent Diffusion
di: Ranasinghe, Kanchana, et al.
Pubblicazione: (2024)
di: Ranasinghe, Kanchana, et al.
Pubblicazione: (2024)
Reviewing FID and SID Metrics on Generative Adversarial Networks
di: de Deijn, Ricardo, et al.
Pubblicazione: (2024)
di: de Deijn, Ricardo, et al.
Pubblicazione: (2024)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
di: Sam, Dylan, et al.
Pubblicazione: (2025)
di: Sam, Dylan, et al.
Pubblicazione: (2025)
Rethinking FID Through the Geometry of the Reference Dataset
di: Lee, Yunghee, et al.
Pubblicazione: (2026)
di: Lee, Yunghee, et al.
Pubblicazione: (2026)
Towards A Better Metric for Text-to-Video Generation
di: Wu, Jay Zhangjie, et al.
Pubblicazione: (2024)
di: Wu, Jay Zhangjie, et al.
Pubblicazione: (2024)
Making Reconstruction FID Predictive of Diffusion Generation FID
di: Xu, Tongda, et al.
Pubblicazione: (2026)
di: Xu, Tongda, et al.
Pubblicazione: (2026)
VectorSynth: Fine-Grained Satellite Image Synthesis with Structured Semantics
di: Cher, Daniel, et al.
Pubblicazione: (2025)
di: Cher, Daniel, et al.
Pubblicazione: (2025)
SurfR: Surface Reconstruction with Multi-scale Attention
di: Ranade, Siddhant, et al.
Pubblicazione: (2025)
di: Ranade, Siddhant, et al.
Pubblicazione: (2025)
F-Bench: Rethinking Human Preference Evaluation Metrics for Benchmarking Face Generation, Customization, and Restoration
di: Liu, Lu, et al.
Pubblicazione: (2024)
di: Liu, Lu, et al.
Pubblicazione: (2024)
Towards Better & Faster Autoregressive Image Generation: From the Perspective of Entropy
di: Ma, Xiaoxiao, et al.
Pubblicazione: (2025)
di: Ma, Xiaoxiao, et al.
Pubblicazione: (2025)
Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation
di: Kasaei, Seyed Amir, et al.
Pubblicazione: (2025)
di: Kasaei, Seyed Amir, et al.
Pubblicazione: (2025)
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
di: Cheng, Jiaxin, et al.
Pubblicazione: (2024)
di: Cheng, Jiaxin, et al.
Pubblicazione: (2024)
RAG-HAR: Retrieval Augmented Generation-based Human Activity Recognition
di: Sivaroopan, Nirhoshan, et al.
Pubblicazione: (2025)
di: Sivaroopan, Nirhoshan, et al.
Pubblicazione: (2025)
Toward Better Optimization of Low-Dose CT Enhancement: A Critical Analysis of Loss Functions and Image Quality Assessment Metrics
di: Yousra, Taifour, et al.
Pubblicazione: (2025)
di: Yousra, Taifour, et al.
Pubblicazione: (2025)
Towards a Better Evaluation of Out-of-Domain Generalization
di: Hwang, Duhun, et al.
Pubblicazione: (2024)
di: Hwang, Duhun, et al.
Pubblicazione: (2024)
GeoSynth: Contextually-Aware High-Resolution Satellite Image Synthesis
di: Sastry, Srikumar, et al.
Pubblicazione: (2024)
di: Sastry, Srikumar, et al.
Pubblicazione: (2024)
Toward A Better Understanding of Monocular Depth Evaluation
di: Wu, Siyang, et al.
Pubblicazione: (2025)
di: Wu, Siyang, et al.
Pubblicazione: (2025)
Rethinking the Evaluation of Visible and Infrared Image Fusion
di: Guan, Dayan, et al.
Pubblicazione: (2024)
di: Guan, Dayan, et al.
Pubblicazione: (2024)
InsightEdit: Towards Better Instruction Following for Image Editing
di: Xu, Yingjing, et al.
Pubblicazione: (2024)
di: Xu, Yingjing, et al.
Pubblicazione: (2024)
Rethinking HTG Evaluation: Bridging Generation and Recognition
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
di: Nikolaidou, Konstantina, et al.
Pubblicazione: (2024)
Towards Better Cephalometric Landmark Detection with Diffusion Data Generation
di: Guo, Dongqian, et al.
Pubblicazione: (2025)
di: Guo, Dongqian, et al.
Pubblicazione: (2025)
Image Classification using Fuzzy Pooling in Convolutional Kolmogorov-Arnold Networks
di: Igali, Ayan, et al.
Pubblicazione: (2024)
di: Igali, Ayan, et al.
Pubblicazione: (2024)
Guidance Matters: Rethinking the Evaluation Pitfall for Text-to-Image Generation
di: Xie, Dian, et al.
Pubblicazione: (2026)
di: Xie, Dian, et al.
Pubblicazione: (2026)
HICEScore: A Hierarchical Metric for Image Captioning Evaluation
di: Zeng, Zequn, et al.
Pubblicazione: (2024)
di: Zeng, Zequn, et al.
Pubblicazione: (2024)
Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage
di: Lee, Saehyung, et al.
Pubblicazione: (2024)
di: Lee, Saehyung, et al.
Pubblicazione: (2024)
FID-Net: A Feature-Enhanced Deep Learning Network for Forest Infestation Detection
di: Zhang, Yan, et al.
Pubblicazione: (2025)
di: Zhang, Yan, et al.
Pubblicazione: (2025)
Generating customized prompts for Zero-Shot Rare Event Medical Image Classification using LLM
di: Kamboj, Payal, et al.
Pubblicazione: (2025)
di: Kamboj, Payal, et al.
Pubblicazione: (2025)
On Distributed Larger-Than-Memory Subset Selection With Pairwise Submodular Functions
di: Böther, Maximilian, et al.
Pubblicazione: (2024)
di: Böther, Maximilian, et al.
Pubblicazione: (2024)
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models
di: Hua, Hang, et al.
Pubblicazione: (2025)
di: Hua, Hang, et al.
Pubblicazione: (2025)
Towards Better Optimization For Listwise Preference in Diffusion Models
di: Bai, Jiamu, et al.
Pubblicazione: (2025)
di: Bai, Jiamu, et al.
Pubblicazione: (2025)
Evaluating Remote Sensing Image Captions Beyond Metric Biases
di: Chen, Ziyun, et al.
Pubblicazione: (2026)
di: Chen, Ziyun, et al.
Pubblicazione: (2026)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
di: Ku, Max, et al.
Pubblicazione: (2023)
di: Ku, Max, et al.
Pubblicazione: (2023)
MIA-Bench: Towards Better Instruction Following Evaluation of Multimodal LLMs
di: Qian, Yusu, et al.
Pubblicazione: (2024)
di: Qian, Yusu, et al.
Pubblicazione: (2024)
SimLBR: Learning to Detect Fake Images by Learning to Detect Real Images
di: Dhakal, Aayush, et al.
Pubblicazione: (2026)
di: Dhakal, Aayush, et al.
Pubblicazione: (2026)
Sat2Cap: Mapping Fine-Grained Textual Descriptions from Satellite Images
di: Dhakal, Aayush, et al.
Pubblicazione: (2023)
di: Dhakal, Aayush, et al.
Pubblicazione: (2023)
OpenAnimals: Revisiting Person Re-Identification for Animals Towards Better Generalization
di: Hou, Saihui, et al.
Pubblicazione: (2024)
di: Hou, Saihui, et al.
Pubblicazione: (2024)
Towards Better De-raining Generalization via Rainy Characteristics Memorization and Replay
di: Wang, Kunyu, et al.
Pubblicazione: (2025)
di: Wang, Kunyu, et al.
Pubblicazione: (2025)
BodyMetric: Evaluating the Realism of Human Bodies in Text-to-Image Generation
di: Andreou, Nefeli, et al.
Pubblicazione: (2024)
di: Andreou, Nefeli, et al.
Pubblicazione: (2024)
Rethinking Perceptual Metrics for Medical Image Translation
di: Konz, Nicholas, et al.
Pubblicazione: (2024)
di: Konz, Nicholas, et al.
Pubblicazione: (2024)
Revisiting Text-to-Image Evaluation with Gecko: On Metrics, Prompts, and Human Ratings
di: Wiles, Olivia, et al.
Pubblicazione: (2024)
di: Wiles, Olivia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
LatentCRF: Continuous CRF for Efficient Latent Diffusion
di: Ranasinghe, Kanchana, et al.
Pubblicazione: (2024) -
Reviewing FID and SID Metrics on Generative Adversarial Networks
di: de Deijn, Ricardo, et al.
Pubblicazione: (2024) -
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
di: Sam, Dylan, et al.
Pubblicazione: (2025) -
Rethinking FID Through the Geometry of the Reference Dataset
di: Lee, Yunghee, et al.
Pubblicazione: (2026) -
Towards A Better Metric for Text-to-Video Generation
di: Wu, Jay Zhangjie, et al.
Pubblicazione: (2024)