CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
Fuente:
arXiv
Saved in:
| Main Authors: | Leiter, Christoph, Asano, Yuki M., Keuper, Margret, Eger, Steffen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TikZero: Zero-Shot Text-Guided Graphics Program Synthesis
by: Belouadi, Jonas, et al.
Published: (2025)
by: Belouadi, Jonas, et al.
Published: (2025)
Corner Cases: How Size and Position of Objects Challenge ImageNet-Trained Models
by: Fatima, Mishal, et al.
Published: (2025)
by: Fatima, Mishal, et al.
Published: (2025)
Smart Eyes for Silent Threats: VLMs and In-Context Learning for THz Imaging
by: Poggi, Nicolas, et al.
Published: (2025)
by: Poggi, Nicolas, et al.
Published: (2025)
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
by: Leiter, Christoph, et al.
Published: (2024)
by: Leiter, Christoph, et al.
Published: (2024)
AutomaTikZ: Text-Guided Synthesis of Scientific Vector Graphics with TikZ
by: Belouadi, Jonas, et al.
Published: (2023)
by: Belouadi, Jonas, et al.
Published: (2023)
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
by: Greisinger, Christian, et al.
Published: (2026)
by: Greisinger, Christian, et al.
Published: (2026)
TRIX- Trading Adversarial Fairness via Mixed Adversarial Training
by: Medi, Tejaswini, et al.
Published: (2025)
by: Medi, Tejaswini, et al.
Published: (2025)
FAIR-TAT: Improving Model Fairness Using Targeted Adversarial Training
by: Medi, Tejaswini, et al.
Published: (2024)
by: Medi, Tejaswini, et al.
Published: (2024)
BMX: Boosting Natural Language Generation Metrics with Explainability
by: Leiter, Christoph, et al.
Published: (2022)
by: Leiter, Christoph, et al.
Published: (2022)
DeTikZify: Synthesizing Graphics Programs for Scientific Figures and Sketches with TikZ
by: Belouadi, Jonas, et al.
Published: (2024)
by: Belouadi, Jonas, et al.
Published: (2024)
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
by: Zhang, Leixin, et al.
Published: (2024)
by: Zhang, Leixin, et al.
Published: (2024)
Deepfakes: we need to re-think the concept of "real" images
by: Keuper, Janis, et al.
Published: (2025)
by: Keuper, Janis, et al.
Published: (2025)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
by: Leiter, Christoph, et al.
Published: (2024)
by: Leiter, Christoph, et al.
Published: (2024)
CosPGD: an efficient white-box adversarial attack for pixel-wise prediction tasks
by: Agnihotri, Shashank, et al.
Published: (2023)
by: Agnihotri, Shashank, et al.
Published: (2023)
Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics
by: Roy, Subhadeep, et al.
Published: (2026)
by: Roy, Subhadeep, et al.
Published: (2026)
Is RobustBench/AutoAttack a suitable Benchmark for Adversarial Robustness?
by: Lorenz, Peter, et al.
Published: (2021)
by: Lorenz, Peter, et al.
Published: (2021)
As large as it gets: Learning infinitely large Filters via Neural Implicit Functions in the Fourier Domain
by: Grabinski, Julia, et al.
Published: (2023)
by: Grabinski, Julia, et al.
Published: (2023)
Vision At Night: Exploring Biologically Inspired Preprocessing For Improved Robustness Via Color And Contrast Transformations
by: Stracke, Lorena, et al.
Published: (2025)
by: Stracke, Lorena, et al.
Published: (2025)
Fix your downsampling ASAP! Be natively more robust via Aliasing and Spectral Artifact free Pooling
by: Grabinski, Julia, et al.
Published: (2023)
by: Grabinski, Julia, et al.
Published: (2023)
Local Spherical Harmonics Improve Skeleton-Based Hand Action Recognition
by: Prasse, Katharina, et al.
Published: (2023)
by: Prasse, Katharina, et al.
Published: (2023)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
by: Simoncini, Walter, et al.
Published: (2024)
by: Simoncini, Walter, et al.
Published: (2024)
Divide & Bind Your Attention for Improved Generative Semantic Nursing
by: Li, Yumeng, et al.
Published: (2023)
by: Li, Yumeng, et al.
Published: (2023)
Little Data, Big Impact: Privacy-Aware Visual Language Models via Minimal Tuning
by: Samson, Laurens, et al.
Published: (2024)
by: Samson, Laurens, et al.
Published: (2024)
Towards Class-wise Robustness Analysis
by: Medi, Tejaswini, et al.
Published: (2024)
by: Medi, Tejaswini, et al.
Published: (2024)
Unfolding Local Growth Rate Estimates for (Almost) Perfect Adversarial Detection
by: Lorenz, Peter, et al.
Published: (2022)
by: Lorenz, Peter, et al.
Published: (2022)
How Do Training Methods Influence the Utilization of Vision Models?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
I Spy With My Little Eye: A Minimum Cost Multicut Investigation of Dataset Frames
by: Prasse, Katharina, et al.
Published: (2024)
by: Prasse, Katharina, et al.
Published: (2024)
Towards Explainable Evaluation Metrics for Machine Translation
by: Leiter, Christoph, et al.
Published: (2023)
by: Leiter, Christoph, et al.
Published: (2023)
From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media
by: Prasse, Katharina, et al.
Published: (2026)
by: Prasse, Katharina, et al.
Published: (2026)
LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
by: Zhang, Kaichen, et al.
Published: (2024)
by: Zhang, Kaichen, et al.
Published: (2024)
Know Yourself Better: Diverse Object-Related Features Improve Open Set Recognition
by: Xu, Jiawen, et al.
Published: (2024)
by: Xu, Jiawen, et al.
Published: (2024)
Beware of Aliases -- Signal Preservation is Crucial for Robust Image Restoration
by: Agnihotri, Shashank, et al.
Published: (2024)
by: Agnihotri, Shashank, et al.
Published: (2024)
Pseudo-Prompt Generating in Pre-trained Vision-Language Models for Multi-Label Medical Image Classification
by: Ye, Yaoqin, et al.
Published: (2024)
by: Ye, Yaoqin, et al.
Published: (2024)
Examining the Impact of Optical Aberrations to Image Classification and Object Detection Models
by: Müller, Patrick, et al.
Published: (2025)
by: Müller, Patrick, et al.
Published: (2025)
Improving Feature Stability during Upsampling -- Spectral Artifacts and the Importance of Spatial Context
by: Agnihotri, Shashank, et al.
Published: (2023)
by: Agnihotri, Shashank, et al.
Published: (2023)
Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation
by: Kasaei, Seyed Amir, et al.
Published: (2025)
by: Kasaei, Seyed Amir, et al.
Published: (2025)
ChartCheck: Explainable Fact-Checking over Real-World Chart Images
by: Akhtar, Mubashara, et al.
Published: (2023)
by: Akhtar, Mubashara, et al.
Published: (2023)
Images as Tables: In-Context Learning with TabPFN for Low-Data Detection of AI-Generated Images
by: Walter, Jan Philip, et al.
Published: (2026)
by: Walter, Jan Philip, et al.
Published: (2026)
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning
by: Matsuda, Kazuki, et al.
Published: (2024)
by: Matsuda, Kazuki, et al.
Published: (2024)
Better Language Models Exhibit Higher Visual Alignment
by: Ruthardt, Jona, et al.
Published: (2024)
by: Ruthardt, Jona, et al.
Published: (2024)
Similar Items
-
TikZero: Zero-Shot Text-Guided Graphics Program Synthesis
by: Belouadi, Jonas, et al.
Published: (2025) -
Corner Cases: How Size and Position of Objects Challenge ImageNet-Trained Models
by: Fatima, Mishal, et al.
Published: (2025) -
Smart Eyes for Silent Threats: VLMs and In-Context Learning for THz Imaging
by: Poggi, Nicolas, et al.
Published: (2025) -
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
by: Leiter, Christoph, et al.
Published: (2024) -
AutomaTikZ: Text-Guided Synthesis of Scientific Vector Graphics with TikZ
by: Belouadi, Jonas, et al.
Published: (2023)