EvalGIM: A Library for Evaluating Generative Image Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hall, Melissa, Mañas, Oscar, Askari-Hemmat, Reyhane, Ibrahim, Mark, Ross, Candace, Astolfi, Pietro, Ifriqi, Tariq Berrada, Havasi, Marton, Benchetrit, Yohann, Ullrich, Karen, Braga, Carolina, Charnalia, Abhishek, Ryan, Maeve, Rabbat, Mike, Drozdzal, Michal, Verbeek, Jakob, Romero-Soriano, Adriana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On Improved Conditioning Mechanisms and Pre-training Strategies for Diffusion Models
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2024)
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2024)
Boosting Latent Diffusion with Perceptual Objectives
von: Berrada, Tariq, et al.
Veröffentlicht: (2024)
von: Berrada, Tariq, et al.
Veröffentlicht: (2024)
Entropy Rectifying Guidance for Diffusion and Flow Models
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2025)
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2025)
Improving Geo-diversity of Generated Images with Contextualized Vendi Score Guidance
von: Hemmat, Reyhane Askari, et al.
Veröffentlicht: (2024)
von: Hemmat, Reyhane Askari, et al.
Veröffentlicht: (2024)
Improving the Scaling Laws of Synthetic Data with Deliberate Practice
von: Askari-Hemmat, Reyhane, et al.
Veröffentlicht: (2025)
von: Askari-Hemmat, Reyhane, et al.
Veröffentlicht: (2025)
Multi-Modal Language Models as Text-to-Image Model Evaluators
von: Chen, Jiahui, et al.
Veröffentlicht: (2025)
von: Chen, Jiahui, et al.
Veröffentlicht: (2025)
Feedback-guided Data Synthesis for Imbalanced Classification
von: Hemmat, Reyhane Askari, et al.
Veröffentlicht: (2023)
von: Hemmat, Reyhane Askari, et al.
Veröffentlicht: (2023)
Increasing the Utility of Synthetic Images through Chamfer Guidance
von: Dall'Asen, Nicola, et al.
Veröffentlicht: (2025)
von: Dall'Asen, Nicola, et al.
Veröffentlicht: (2025)
Flowception: Temporally Expansive Flow Matching for Video Generation
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2025)
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2025)
Why Less is More (Sometimes): A Theory of Data Curation
von: Dohmatob, Elvis, et al.
Veröffentlicht: (2025)
von: Dohmatob, Elvis, et al.
Veröffentlicht: (2025)
OneFlow: Concurrent Mixed-Modal and Interleaved Generation with Edit Flows
von: Nguyen, John, et al.
Veröffentlicht: (2025)
von: Nguyen, John, et al.
Veröffentlicht: (2025)
Consistency-diversity-realism Pareto fronts of conditional image generative models
von: Astolfi, Pietro, et al.
Veröffentlicht: (2024)
von: Astolfi, Pietro, et al.
Veröffentlicht: (2024)
Improving Text-to-Image Consistency via Automatic Prompt Optimization
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Tokenization Bias in Language Models
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Inference-time Physics Alignment of Video Generative Models with Latent World Models
von: Yuan, Jianhao, et al.
Veröffentlicht: (2026)
von: Yuan, Jianhao, et al.
Veröffentlicht: (2026)
Improving the Physics of Video Generation with VJEPA-2 Reward Signal
von: Yuan, Jianhao, et al.
Veröffentlicht: (2025)
von: Yuan, Jianhao, et al.
Veröffentlicht: (2025)
Unlocking Pre-trained Image Backbones for Semantic Image Synthesis
von: Berrada, Tariq, et al.
Veröffentlicht: (2023)
von: Berrada, Tariq, et al.
Veröffentlicht: (2023)
QGen: On the Ability to Generalize in Quantization Aware Training
von: AskariHemmat, MohammadHossein, et al.
Veröffentlicht: (2024)
von: AskariHemmat, MohammadHossein, et al.
Veröffentlicht: (2024)
Unified Text-Image Generation with Weakness-Targeted Post-Training
von: Chen, Jiahui, et al.
Veröffentlicht: (2026)
von: Chen, Jiahui, et al.
Veröffentlicht: (2026)
Multimodal RewardBench 2: Evaluating Omni Reward Models for Interleaved Text and Image
von: Hu, Yushi, et al.
Veröffentlicht: (2025)
von: Hu, Yushi, et al.
Veröffentlicht: (2025)
Object-centric Binding in Contrastive Language-Image Pretraining
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
von: Assouel, Rim, et al.
Veröffentlicht: (2025)
DP-RDM: Adapting Diffusion Models to Private Domains Without Fine-Tuning
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
von: Lebensold, Jonathan, et al.
Veröffentlicht: (2024)
Dynadiff: Single-stage Decoding of Images from Continuously Evolving fMRI
von: Careil, Marlène, et al.
Veröffentlicht: (2025)
von: Careil, Marlène, et al.
Veröffentlicht: (2025)
Brain decoding: toward real-time reconstruction of visual perception
von: Benchetrit, Yohann, et al.
Veröffentlicht: (2023)
von: Benchetrit, Yohann, et al.
Veröffentlicht: (2023)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
von: Hall, Melissa, et al.
Veröffentlicht: (2023)
von: Hall, Melissa, et al.
Veröffentlicht: (2023)
Guarantee Regions for Local Explanations
von: Havasi, Marton, et al.
Veröffentlicht: (2024)
von: Havasi, Marton, et al.
Veröffentlicht: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
von: Brown, Katrina, et al.
Veröffentlicht: (2024)
von: Brown, Katrina, et al.
Veröffentlicht: (2024)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
von: Hall, Melissa, et al.
Veröffentlicht: (2024)
von: Hall, Melissa, et al.
Veröffentlicht: (2024)
Edit Flows: Flow Matching with Edit Operations
von: Havasi, Marton, et al.
Veröffentlicht: (2025)
von: Havasi, Marton, et al.
Veröffentlicht: (2025)
Controlling Multimodal LLMs via Reward-guided Decoding
von: Mañas, Oscar, et al.
Veröffentlicht: (2025)
von: Mañas, Oscar, et al.
Veröffentlicht: (2025)
Reemergencia de enfermedades prevenibles por vacunas: el caso del sarampión en las Américas
von: Andres Benchetrit
Veröffentlicht: (2019)
von: Andres Benchetrit
Veröffentlicht: (2019)
GIM: Learning Generalizable Image Matcher From Internet Videos
von: Shen, Xuelun, et al.
Veröffentlicht: (2024)
von: Shen, Xuelun, et al.
Veröffentlicht: (2024)
Scaling laws for decoding images from brain activity
von: Banville, Hubert, et al.
Veröffentlicht: (2025)
von: Banville, Hubert, et al.
Veröffentlicht: (2025)
TRIBE: TRImodal Brain Encoder for whole-brain fMRI response prediction
von: d'Ascoli, Stéphane, et al.
Veröffentlicht: (2025)
von: d'Ascoli, Stéphane, et al.
Veröffentlicht: (2025)
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
von: Chen, Zixi, et al.
Veröffentlicht: (2022)
von: Chen, Zixi, et al.
Veröffentlicht: (2022)
PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs
von: Assouel, Rim, et al.
Veröffentlicht: (2026)
von: Assouel, Rim, et al.
Veröffentlicht: (2026)
The Intricate Dance of Prompt Complexity, Quality, Diversity, and Consistency in T2I Models
von: Xiaofeng, Zhang, et al.
Veröffentlicht: (2025)
von: Xiaofeng, Zhang, et al.
Veröffentlicht: (2025)
Can repeller dynamics explain dominant pebble axis ratios?
von: Havasi-Tóth, Balázs
Veröffentlicht: (2024)
von: Havasi-Tóth, Balázs
Veröffentlicht: (2024)
GIM: Evaluating models via tasks that integrate multiple cognitive domains
von: Patel, Rohit, et al.
Veröffentlicht: (2026)
von: Patel, Rohit, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
On Improved Conditioning Mechanisms and Pre-training Strategies for Diffusion Models
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2024) -
Boosting Latent Diffusion with Perceptual Objectives
von: Berrada, Tariq, et al.
Veröffentlicht: (2024) -
Entropy Rectifying Guidance for Diffusion and Flow Models
von: Ifriqi, Tariq Berrada, et al.
Veröffentlicht: (2025) -
Improving Geo-diversity of Generated Images with Contextualized Vendi Score Guidance
von: Hemmat, Reyhane Askari, et al.
Veröffentlicht: (2024) -
Improving the Scaling Laws of Synthetic Data with Deliberate Practice
von: Askari-Hemmat, Reyhane, et al.
Veröffentlicht: (2025)