RePOPE: Impact of Annotation Errors on the POPE Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Neuhaus, Yannic, Hein, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DASH: Detection and Assessment of Systematic Hallucinations of VLMs
by: Augustin, Maximilian, et al.
Published: (2025)
by: Augustin, Maximilian, et al.
Published: (2025)
DiG-IN: Diffusion Guidance for Investigating Networks -- Uncovering Classifier Differences Neuron Visualisations and Visual Counterfactual Explanations
by: Augustin, Maximilian, et al.
Published: (2023)
by: Augustin, Maximilian, et al.
Published: (2023)
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
by: Neuhaus, Yannic, et al.
Published: (2026)
by: Neuhaus, Yannic, et al.
Published: (2026)
POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
by: Qu, Yuxiao, et al.
Published: (2026)
by: Qu, Yuxiao, et al.
Published: (2026)
Quality Assured: Rethinking Annotation Strategies in Imaging AI
by: Rädsch, Tim, et al.
Published: (2024)
by: Rädsch, Tim, et al.
Published: (2024)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
by: Schlarmann, Christian, et al.
Published: (2024)
by: Schlarmann, Christian, et al.
Published: (2024)
Toxicity Assessment in Preclinical Histopathology via Class-Aware Mahalanobis Distance for Known and Novel Anomalies
by: Graf, Olga, et al.
Published: (2026)
by: Graf, Olga, et al.
Published: (2026)
Privacy Meets Explainability: A Comprehensive Impact Benchmark
by: Saifullah, Saifullah, et al.
Published: (2022)
by: Saifullah, Saifullah, et al.
Published: (2022)
Robustness in Both Domains: CLIP Needs a Robust Text Encoder
by: Rocamora, Elias Abad, et al.
Published: (2025)
by: Rocamora, Elias Abad, et al.
Published: (2025)
A Deep U-Net Framework for Flood Hazard Mapping Using Hydraulic Simulations of the Wupper Catchment
by: Lammers, Christian, et al.
Published: (2026)
by: Lammers, Christian, et al.
Published: (2026)
Kill it with FIRE: On Leveraging Latent Space Directions for Runtime Backdoor Mitigation in Deep Neural Networks
by: Ahlers, Enrico, et al.
Published: (2026)
by: Ahlers, Enrico, et al.
Published: (2026)
H-POPE: Hierarchical Polling-based Probing Evaluation of Hallucinations in Large Vision-Language Models
by: Pham, Nhi, et al.
Published: (2024)
by: Pham, Nhi, et al.
Published: (2024)
ReText: Text Boosts Generalization in Image-Based Person Re-identification
by: Mamedov, Timur, et al.
Published: (2026)
by: Mamedov, Timur, et al.
Published: (2026)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
by: Mamedov, Timur, et al.
Published: (2024)
by: Mamedov, Timur, et al.
Published: (2024)
Automatic Image Annotation for Mapped Features Detection
by: Noizet, Maxime, et al.
Published: (2024)
by: Noizet, Maxime, et al.
Published: (2024)
ReDepth Anything: Test-Time Depth Refinement via Self-Supervised Re-lighting
by: Bhattarai, Ananta R., et al.
Published: (2025)
by: Bhattarai, Ananta R., et al.
Published: (2025)
Evaluating Self-Supervised Learning in Medical Imaging: A Benchmark for Robustness, Generalizability, and Multi-Domain Impact
by: Bundele, Valay, et al.
Published: (2024)
by: Bundele, Valay, et al.
Published: (2024)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
by: Parmar, Mihir, et al.
Published: (2022)
by: Parmar, Mihir, et al.
Published: (2022)
Annotation-Efficient Active Test-Time Adaptation with Conformal Prediction
by: Shi, Tingyu, et al.
Published: (2025)
by: Shi, Tingyu, et al.
Published: (2025)
V-Zero: Self-Improving Multimodal Reasoning with Zero Annotation
by: Wang, Han, et al.
Published: (2026)
by: Wang, Han, et al.
Published: (2026)
Continual Error Correction on Low-Resource Devices
by: Paramonov, Kirill, et al.
Published: (2025)
by: Paramonov, Kirill, et al.
Published: (2025)
Automated Classification of Model Errors on ImageNet
by: Peychev, Momchil, et al.
Published: (2023)
by: Peychev, Momchil, et al.
Published: (2023)
MultiFloodSynth: Multi-Annotated Flood Synthetic Dataset Generation
by: Kang, YoonJe, et al.
Published: (2025)
by: Kang, YoonJe, et al.
Published: (2025)
Perceptual Quality-based Model Training under Annotator Label Uncertainty
by: Zhou, Chen, et al.
Published: (2024)
by: Zhou, Chen, et al.
Published: (2024)
Hierarchical Classification for Automated Image Annotation of Coral Reef Benthic Structures
by: Blondin, Célia, et al.
Published: (2024)
by: Blondin, Célia, et al.
Published: (2024)
BRAIxDet: Learning to Detect Malignant Breast Lesion with Incomplete Annotations
by: Chen, Yuanhong, et al.
Published: (2023)
by: Chen, Yuanhong, et al.
Published: (2023)
Exploring the Camera Bias of Person Re-identification
by: Song, Myungseo, et al.
Published: (2025)
by: Song, Myungseo, et al.
Published: (2025)
RadioActive: 3D Radiological Interactive Segmentation Benchmark
by: Ulrich, Constantin, et al.
Published: (2024)
by: Ulrich, Constantin, et al.
Published: (2024)
When Person Re-Identification Meets Event Camera: A Benchmark Dataset and An Attribute-guided Re-Identification Framework
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Human and AI Perceptual Differences in Image Classification Errors
by: Liu, Minghao, et al.
Published: (2023)
by: Liu, Minghao, et al.
Published: (2023)
Prediction Error-based Classification for Class-Incremental Learning
by: Zając, Michał, et al.
Published: (2023)
by: Zając, Michał, et al.
Published: (2023)
Multi-Task Learning with Multi-Annotation Triplet Loss for Improved Object Detection
by: Zhou, Meilun, et al.
Published: (2025)
by: Zhou, Meilun, et al.
Published: (2025)
Weakly Supervised Pretraining and Multi-Annotator Supervised Finetuning for Facial Wrinkle Detection
by: Moon, Ik Jun, et al.
Published: (2024)
by: Moon, Ik Jun, et al.
Published: (2024)
CountCLIP -- [Re] Teaching CLIP to Count to Ten
by: Mestha, Harshvardhan, et al.
Published: (2024)
by: Mestha, Harshvardhan, et al.
Published: (2024)
RepAct: The Re-parameterizable Adaptive Activation Function
by: Wu, Xian, et al.
Published: (2024)
by: Wu, Xian, et al.
Published: (2024)
LAuReL: Learned Augmented Residual Layer
by: Menghani, Gaurav, et al.
Published: (2024)
by: Menghani, Gaurav, et al.
Published: (2024)
MonoSOWA: Scalable monocular 3D Object detector Without human Annotations
by: Skvrna, Jan, et al.
Published: (2025)
by: Skvrna, Jan, et al.
Published: (2025)
What Can We Learn from Inter-Annotator Variability in Skin Lesion Segmentation?
by: Abhishek, Kumar, et al.
Published: (2025)
by: Abhishek, Kumar, et al.
Published: (2025)
CLoPA: Continual Low Parameter Adaptation of Interactive Segmentation for Medical Image Annotation
by: Esmaeili, Parhom, et al.
Published: (2026)
by: Esmaeili, Parhom, et al.
Published: (2026)
Understanding and Mitigating Human-Labelling Errors in Supervised Contrastive Learning
by: Long, Zijun, et al.
Published: (2024)
by: Long, Zijun, et al.
Published: (2024)
Similar Items
-
DASH: Detection and Assessment of Systematic Hallucinations of VLMs
by: Augustin, Maximilian, et al.
Published: (2025) -
DiG-IN: Diffusion Guidance for Investigating Networks -- Uncovering Classifier Differences Neuron Visualisations and Visual Counterfactual Explanations
by: Augustin, Maximilian, et al.
Published: (2023) -
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
by: Neuhaus, Yannic, et al.
Published: (2026) -
POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
by: Qu, Yuxiao, et al.
Published: (2026) -
Quality Assured: Rethinking Annotation Strategies in Imaging AI
by: Rädsch, Tim, et al.
Published: (2024)