Confidence Intervals for Error Rates in 1:1 Matching Tasks: Critical Statistical Analysis and Recommendations
Fuente:
arXiv
Saved in:
| Main Authors: | Fogliato, Riccardo, Patil, Pratik, Perona, Pietro |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Framework for Efficient Model Evaluation through Stratification, Sampling, and Estimation
by: Fogliato, Riccardo, et al.
Published: (2024)
by: Fogliato, Riccardo, et al.
Published: (2024)
Precise Model Benchmarking with Only a Few Observations
by: Fogliato, Riccardo, et al.
Published: (2024)
by: Fogliato, Riccardo, et al.
Published: (2024)
Statistical analysis of multivariate planar curves and applications to X-ray classification
by: Moindjié, Issam-Ali, et al.
Published: (2025)
by: Moindjié, Issam-Ali, et al.
Published: (2025)
A False Discovery Rate Control Method Using a Fully Connected Hidden Markov Random Field for Neuroimaging Data
by: Kim, Taehyo, et al.
Published: (2025)
by: Kim, Taehyo, et al.
Published: (2025)
Single View Seafloor Recovery from Imaging Sonar via Differentiable Rendering
by: Brodjian, Sevan, et al.
Published: (2026)
by: Brodjian, Sevan, et al.
Published: (2026)
Unsupervised Representation Learning from Sparse Transformation Analysis
by: Song, Yue, et al.
Published: (2024)
by: Song, Yue, et al.
Published: (2024)
Is CLIP ideal? No. Can we fix it? Yes!
by: Kang, Raphi, et al.
Published: (2025)
by: Kang, Raphi, et al.
Published: (2025)
Evaluating AI systems under uncertain ground truth: a case study in dermatology
by: Stutz, David, et al.
Published: (2023)
by: Stutz, David, et al.
Published: (2023)
Conformal Prediction for Long-Tailed Classification
by: Ding, Tiffany, et al.
Published: (2025)
by: Ding, Tiffany, et al.
Published: (2025)
Conformal Prediction Sets for Instance Segmentation
by: Lu, Kerri, et al.
Published: (2026)
by: Lu, Kerri, et al.
Published: (2026)
Optimizing Diffusion Priors in Image Reconstruction from a Single Observation
by: Wang, Frederic, et al.
Published: (2026)
by: Wang, Frederic, et al.
Published: (2026)
Registration-Free Monitoring of Unstructured Point Cloud Data via Intrinsic Geometrical Properties
by: Patalano, Mariafrancesca, et al.
Published: (2025)
by: Patalano, Mariafrancesca, et al.
Published: (2025)
Sample-efficient evidence estimation of score based priors for model selection
by: Wang, Frederic, et al.
Published: (2026)
by: Wang, Frederic, et al.
Published: (2026)
Mixstyle-Entropy: Domain Generalization with Causal Intervention and Perturbation
by: Tang, Luyao, et al.
Published: (2024)
by: Tang, Luyao, et al.
Published: (2024)
A Fourier Space Perspective on Diffusion Models
by: Falck, Fabian, et al.
Published: (2025)
by: Falck, Fabian, et al.
Published: (2025)
Overcoming Common Flaws in the Evaluation of Selective Classification Systems
by: Traub, Jeremias, et al.
Published: (2024)
by: Traub, Jeremias, et al.
Published: (2024)
Non-Negative Stiefel Approximating Flow: Orthogonalish Matrix Optimization for Interpretable Embeddings
by: Avants, Brian B., et al.
Published: (2025)
by: Avants, Brian B., et al.
Published: (2025)
CaRiNG: Learning Temporal Causal Representation under Non-Invertible Generation Process
by: Chen, Guangyi, et al.
Published: (2024)
by: Chen, Guangyi, et al.
Published: (2024)
CAD-VAE: Leveraging Correlation-Aware Latents for Comprehensive Fair Disentanglement
by: Ma, Chenrui, et al.
Published: (2025)
by: Ma, Chenrui, et al.
Published: (2025)
Efficient Neural Network Training via Subset Pretraining
by: Spörer, Jan, et al.
Published: (2024)
by: Spörer, Jan, et al.
Published: (2024)
Automatic Scoring of Cognition Drawings: Assessing the Quality of Machine-Based Scores Against a Gold Standard
by: Bethmann, Arne, et al.
Published: (2023)
by: Bethmann, Arne, et al.
Published: (2023)
Evaluating Reliability in Medical DNNs: A Critical Analysis of Feature and Confidence-Based OOD Detection
by: Anthony, Harry, et al.
Published: (2024)
by: Anthony, Harry, et al.
Published: (2024)
Representational Difference Explanations
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
Confidence Intervals for Performance Estimates in Brain MRI Segmentation
by: Jurdi, R. El, et al.
Published: (2023)
by: Jurdi, R. El, et al.
Published: (2023)
On Pitfalls of $\textit{RemOve-And-Retrain}$: Data Processing Inequality Perspective
by: Song, Junhwa, et al.
Published: (2023)
by: Song, Junhwa, et al.
Published: (2023)
Deep Neural-network Prior for Orbit Recovery from Method of Moments
by: Khoo, Yuehaw, et al.
Published: (2023)
by: Khoo, Yuehaw, et al.
Published: (2023)
Understanding Model Calibration -- A gentle introduction and visual exploration of calibration and the expected calibration error (ECE)
by: Pavlovic, Maja
Published: (2025)
by: Pavlovic, Maja
Published: (2025)
Regularizing Attention Scores with Bootstrapping
by: Chung, Neo Christopher, et al.
Published: (2026)
by: Chung, Neo Christopher, et al.
Published: (2026)
ECG-IMN: Interpretable Mesomorphic Neural Networks for 12-Lead Electrocardiogram Interpretation
by: Thambawita, Vajira, et al.
Published: (2026)
by: Thambawita, Vajira, et al.
Published: (2026)
Attacking Bayes: On the Adversarial Robustness of Bayesian Neural Networks
by: Feng, Yunzhen, et al.
Published: (2024)
by: Feng, Yunzhen, et al.
Published: (2024)
Social Perception of Faces in a Vision-Language Model
by: Hausladen, Carina I., et al.
Published: (2024)
by: Hausladen, Carina I., et al.
Published: (2024)
Confidence-Based Task Prediction in Continual Disease Classification Using Probability Distribution
by: Verma, Tanvi, et al.
Published: (2024)
by: Verma, Tanvi, et al.
Published: (2024)
ZipIt! Merging Models from Different Tasks without Training
by: Stoica, George, et al.
Published: (2023)
by: Stoica, George, et al.
Published: (2023)
AGFS-Tractometry: A Novel Atlas-Guided Fine-Scale Tractometry Approach for Enhanced Along-Tract Group Statistical Comparison Using Diffusion MRI Tractography
by: Zheng, Ruixi, et al.
Published: (2025)
by: Zheng, Ruixi, et al.
Published: (2025)
Causal Attribution of Model Performance Gaps in Medical Imaging Under Distribution Shifts
by: Gordaliza, Pedro M., et al.
Published: (2025)
by: Gordaliza, Pedro M., et al.
Published: (2025)
Kuramoto Orientation Diffusion Models
by: Song, Yue, et al.
Published: (2025)
by: Song, Yue, et al.
Published: (2025)
Statistical Opportunities in Neuroimaging
by: Kang, Jian, et al.
Published: (2026)
by: Kang, Jian, et al.
Published: (2026)
Many Perception Tasks are Highly Redundant Functions of their Input Data
by: Ramesh, Rahul, et al.
Published: (2024)
by: Ramesh, Rahul, et al.
Published: (2024)
Soft Dice Confidence: A Near-Optimal Confidence Estimator for Selective Prediction in Semantic Segmentation
by: Borges, Bruno Laboissiere Camargos, et al.
Published: (2024)
by: Borges, Bruno Laboissiere Camargos, et al.
Published: (2024)
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model
by: Wang, Xiyao, et al.
Published: (2025)
by: Wang, Xiyao, et al.
Published: (2025)
Similar Items
-
A Framework for Efficient Model Evaluation through Stratification, Sampling, and Estimation
by: Fogliato, Riccardo, et al.
Published: (2024) -
Precise Model Benchmarking with Only a Few Observations
by: Fogliato, Riccardo, et al.
Published: (2024) -
Statistical analysis of multivariate planar curves and applications to X-ray classification
by: Moindjié, Issam-Ali, et al.
Published: (2025) -
A False Discovery Rate Control Method Using a Fully Connected Hidden Markov Random Field for Neuroimaging Data
by: Kim, Taehyo, et al.
Published: (2025) -
Single View Seafloor Recovery from Imaging Sonar via Differentiable Rendering
by: Brodjian, Sevan, et al.
Published: (2026)