Dataset Diversity Metrics and Impact on Classification Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sourget, Théo, Claßen, Niclas, Xu, Jack Junchi, van der Goot, Rob, Cheplygina, Veronika |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mask of truth: model sensitivity to unexpected regions of medical images
by: Sourget, Théo, et al.
Published: (2024)
by: Sourget, Théo, et al.
Published: (2024)
[Citation needed] Data usage and citation practices in medical imaging conferences
by: Sourget, Théo, et al.
Published: (2024)
by: Sourget, Théo, et al.
Published: (2024)
Source Matters: Source Dataset Impact on Model Robustness in Medical Imaging
by: Juodelyte, Dovile, et al.
Published: (2024)
by: Juodelyte, Dovile, et al.
Published: (2024)
Copycats: the many lives of a publicly available medical imaging dataset
by: Jiménez-Sánchez, Amelia, et al.
Published: (2024)
by: Jiménez-Sánchez, Amelia, et al.
Published: (2024)
Augmenting Chest X-ray Datasets with Non-Expert Annotations
by: Cheplygina, Veronika, et al.
Published: (2023)
by: Cheplygina, Veronika, et al.
Published: (2023)
Fairness and Robustness of CLIP-Based Models for Chest X-rays
by: Sourget, Théo, et al.
Published: (2025)
by: Sourget, Théo, et al.
Published: (2025)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
by: Lu, Yucheng, et al.
Published: (2025)
by: Lu, Yucheng, et al.
Published: (2025)
Detecting Shortcuts in Medical Images -- A Case Study in Chest X-rays
by: Jiménez-Sánchez, Amelia, et al.
Published: (2022)
by: Jiménez-Sánchez, Amelia, et al.
Published: (2022)
SeagrassFinder: Deep Learning for Eelgrass Detection and Coverage Estimation in the Wild
by: Elsäßer, Jannik, et al.
Published: (2024)
by: Elsäßer, Jannik, et al.
Published: (2024)
Exploring connections of spectral analysis and transfer learning in medical imaging
by: Lu, Yucheng, et al.
Published: (2024)
by: Lu, Yucheng, et al.
Published: (2024)
Detection Transformer for Teeth Detection, Segmentation, and Numbering in Oral Rare Diseases: Focus on Data Augmentation and Inpainting Techniques
by: Kadi, Hocine, et al.
Published: (2024)
by: Kadi, Hocine, et al.
Published: (2024)
Learning to Harmonize Cross-vendor X-ray Images by Non-linear Image Dynamics Correction
by: Lu, Yucheng, et al.
Published: (2025)
by: Lu, Yucheng, et al.
Published: (2025)
On dataset transferability in medical image classification
by: Juodelyte, Dovile, et al.
Published: (2024)
by: Juodelyte, Dovile, et al.
Published: (2024)
Robustness and sex differences in skin cancer detection: logistic regression vs CNNs
by: Pedersen, Nikolette, et al.
Published: (2025)
by: Pedersen, Nikolette, et al.
Published: (2025)
GeoDE: a Geographically Diverse Evaluation Dataset for Object Recognition
by: Ramaswamy, Vikram V., et al.
Published: (2023)
by: Ramaswamy, Vikram V., et al.
Published: (2023)
Towards More Diverse and Challenging Pre-training for Point Cloud Learning: Self-Supervised Cross Reconstruction with Decoupled Views
by: Zhang, Xiangdong, et al.
Published: (2025)
by: Zhang, Xiangdong, et al.
Published: (2025)
LiFMCR: Dataset and Benchmark for Light Field Multi-Camera Registration
by: Fleith, Aymeric, et al.
Published: (2025)
by: Fleith, Aymeric, et al.
Published: (2025)
Jigsaw Regularization in Whole-Slide Image Classification
by: Jeong, So Won, et al.
Published: (2026)
by: Jeong, So Won, et al.
Published: (2026)
In the Picture: Medical Imaging Datasets, Artifacts, and their Living Review
by: Jiménez-Sánchez, Amelia, et al.
Published: (2025)
by: Jiménez-Sánchez, Amelia, et al.
Published: (2025)
Unveiling the Potential: Harnessing Deep Metric Learning to Circumvent Video Streaming Encryption
by: Gansekoele, Arwin, et al.
Published: (2024)
by: Gansekoele, Arwin, et al.
Published: (2024)
Exploring the Effect of Dataset Diversity in Self-Supervised Learning for Surgical Computer Vision
by: Jaspers, Tim J. M., et al.
Published: (2024)
by: Jaspers, Tim J. M., et al.
Published: (2024)
Multi-label Scene Classification for Autonomous Vehicles: Acquiring and Accumulating Knowledge from Diverse Datasets
by: Li, Ke, et al.
Published: (2025)
by: Li, Ke, et al.
Published: (2025)
UAV (Unmanned Aerial Vehicles): Diverse Applications of UAV Datasets in Segmentation, Classification, Detection, and Tracking
by: Rahman, Md. Mahfuzur, et al.
Published: (2024)
by: Rahman, Md. Mahfuzur, et al.
Published: (2024)
On the Evaluation and Refinement of Vision-Language Instruction Tuning Datasets
by: Liao, Ning, et al.
Published: (2023)
by: Liao, Ning, et al.
Published: (2023)
Evaluating the Impact of Adversarial Attacks on Traffic Sign Classification using the LISA Dataset
by: Tadessa, Nabeyou, et al.
Published: (2025)
by: Tadessa, Nabeyou, et al.
Published: (2025)
Diverse and Lifespan Facial Age Transformation Synthesis with Identity Variation Rationality Metric
by: Xie, Jiu-Cheng, et al.
Published: (2024)
by: Xie, Jiu-Cheng, et al.
Published: (2024)
FSOCO: The Formula Student Objects in Context Dataset
by: Vödisch, Niclas, et al.
Published: (2020)
by: Vödisch, Niclas, et al.
Published: (2020)
StructChart: On the Schema, Metric, and Augmentation for Visual Chart Understanding
by: Xia, Renqiu, et al.
Published: (2023)
by: Xia, Renqiu, et al.
Published: (2023)
No-Reference Rendered Video Quality Assessment: Dataset and Metrics
by: Yang, Sipeng, et al.
Published: (2025)
by: Yang, Sipeng, et al.
Published: (2025)
Imitating Radiological Scrolling: A Global-Local Attention Model for 3D Chest CT Volumes Multi-Label Anomaly Classification
by: Di Piazza, Theo, et al.
Published: (2025)
by: Di Piazza, Theo, et al.
Published: (2025)
Grounding and Enhancing Grid-based Models for Neural Fields
by: Zhao, Zelin, et al.
Published: (2024)
by: Zhao, Zelin, et al.
Published: (2024)
LLM-Seg: Bridging Image Segmentation and Large Language Model Reasoning
by: Wang, Junchi, et al.
Published: (2024)
by: Wang, Junchi, et al.
Published: (2024)
Dark Side Augmentation: Generating Diverse Night Examples for Metric Learning
by: Mohwald, Albert, et al.
Published: (2023)
by: Mohwald, Albert, et al.
Published: (2023)
EMMA: Concept Erasure Benchmark with Comprehensive Semantic Metrics and Diverse Categories
by: Wei, Lu, et al.
Published: (2025)
by: Wei, Lu, et al.
Published: (2025)
Subjective-Aligned Dataset and Metric for Text-to-Video Quality Assessment
by: Kou, Tengchuan, et al.
Published: (2024)
by: Kou, Tengchuan, et al.
Published: (2024)
Democratising Pathology Co-Pilots: An Open Pipeline and Dataset for Whole-Slide Vision-Language Modelling
by: Moonemans, Sander, et al.
Published: (2025)
by: Moonemans, Sander, et al.
Published: (2025)
Explaining Human Preferences via Metrics for Structured 3D Reconstruction
by: Langerman, Jack, et al.
Published: (2025)
by: Langerman, Jack, et al.
Published: (2025)
SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking
by: Zhao, Weiguang, et al.
Published: (2026)
by: Zhao, Weiguang, et al.
Published: (2026)
Perceptual Quality Assessment of 3D Gaussian Splatting: A Subjective Dataset and Prediction Metric
by: Wan, Zhaolin, et al.
Published: (2025)
by: Wan, Zhaolin, et al.
Published: (2025)
Dataset Distillation for Histopathology Image Classification
by: Cong, Cong, et al.
Published: (2024)
by: Cong, Cong, et al.
Published: (2024)
Similar Items
-
Mask of truth: model sensitivity to unexpected regions of medical images
by: Sourget, Théo, et al.
Published: (2024) -
[Citation needed] Data usage and citation practices in medical imaging conferences
by: Sourget, Théo, et al.
Published: (2024) -
Source Matters: Source Dataset Impact on Model Robustness in Medical Imaging
by: Juodelyte, Dovile, et al.
Published: (2024) -
Copycats: the many lives of a publicly available medical imaging dataset
by: Jiménez-Sánchez, Amelia, et al.
Published: (2024) -
Augmenting Chest X-ray Datasets with Non-Expert Annotations
by: Cheplygina, Veronika, et al.
Published: (2023)