A Confidence Interval for the $\ell_2$ Expected Calibration Error
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Yan, Chaudhari, Pratik, Barnett, Ian J., Dobriban, Edgar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Watermarking Language Models with Error Correcting Codes
by: Chao, Patrick, et al.
Published: (2024)
by: Chao, Patrick, et al.
Published: (2024)
Singleton-Optimized Conformal Prediction
by: Wang, Tao, et al.
Published: (2025)
by: Wang, Tao, et al.
Published: (2025)
Confidence Intervals for Error Rates in 1:1 Matching Tasks: Critical Statistical Analysis and Recommendations
by: Fogliato, Riccardo, et al.
Published: (2023)
by: Fogliato, Riccardo, et al.
Published: (2023)
Statistical Methods in Generative AI
by: Dobriban, Edgar
Published: (2025)
by: Dobriban, Edgar
Published: (2025)
Solving a Research Problem in Mathematical Statistics with AI Assistance
by: Dobriban, Edgar
Published: (2025)
by: Dobriban, Edgar
Published: (2025)
MultiRisk: Multiple Risk Control via Iterative Score Thresholding
by: Joshi, Sunay, et al.
Published: (2025)
by: Joshi, Sunay, et al.
Published: (2025)
Optimal Decision-Making Based on Prediction Sets
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Fair Classification by Direct Intervention on Operating Characteristics
by: Jiang, Kevin, et al.
Published: (2025)
by: Jiang, Kevin, et al.
Published: (2025)
Soft Mean Expected Calibration Error (SMECE): A Calibration Metric for Probabilistic Labels
by: Leznik, Michael
Published: (2026)
by: Leznik, Michael
Published: (2026)
Model Zoo: A Growing "Brain" That Learns Continually
by: Ramesh, Rahul, et al.
Published: (2021)
by: Ramesh, Rahul, et al.
Published: (2021)
Information-theoretic Generalization Analysis for Expected Calibration Error
by: Futami, Futoshi, et al.
Published: (2024)
by: Futami, Futoshi, et al.
Published: (2024)
Minimax Statistical Estimation under Wasserstein Contamination
by: Chao, Patrick, et al.
Published: (2023)
by: Chao, Patrick, et al.
Published: (2023)
Constructing Confidence Intervals for 'the' Generalization Error -- a Comprehensive Benchmark Study
by: Schulz-Kümpel, Hannah, et al.
Published: (2024)
by: Schulz-Kümpel, Hannah, et al.
Published: (2024)
An Effective Gram Matrix Characterizes Generalization in Deep Networks
by: Yang, Rubing, et al.
Published: (2025)
by: Yang, Rubing, et al.
Published: (2025)
An Information-Geometric Distance on the Space of Tasks
by: Gao, Yansong, et al.
Published: (2020)
by: Gao, Yansong, et al.
Published: (2020)
SymmPI: Predictive Inference for Data with Group Symmetries
by: Dobriban, Edgar, et al.
Published: (2023)
by: Dobriban, Edgar, et al.
Published: (2023)
Where to Spend Rollouts: Hit-Utility Optimal Rollout Allocation for Group-Based RLVR
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Online Conformal Prediction via Universal Portfolio Algorithms
by: Liu, Tuo, et al.
Published: (2026)
by: Liu, Tuo, et al.
Published: (2026)
Bayes-Optimal Classifiers under Group Fairness
by: Zeng, Xianli, et al.
Published: (2022)
by: Zeng, Xianli, et al.
Published: (2022)
LLMs are Overconfident: Evaluating Confidence Interval Calibration with FermiEval
by: Epstein, Elliot L., et al.
Published: (2025)
by: Epstein, Elliot L., et al.
Published: (2025)
Sample-efficient Multiclass Calibration under $\ell_{p}$ Error
by: Bairaktari, Konstantina, et al.
Published: (2025)
by: Bairaktari, Konstantina, et al.
Published: (2025)
On Nonasymptotic Confidence Intervals for Treatment Effects in Randomized Experiments
by: Sandoval, Ricardo J., et al.
Published: (2026)
by: Sandoval, Ricardo J., et al.
Published: (2026)
Confidence Intervals and Simultaneous Confidence Bands Based on Deep Learning
by: Arie, Asaf Ben, et al.
Published: (2024)
by: Arie, Asaf Ben, et al.
Published: (2024)
Confidence Intervals for Evaluation of Data Mining
by: Yuan, Zheng, et al.
Published: (2025)
by: Yuan, Zheng, et al.
Published: (2025)
A Theory of Non-Linear Feature Learning with One Gradient Step in Two-Layer Neural Networks
by: Moniri, Behrad, et al.
Published: (2023)
by: Moniri, Behrad, et al.
Published: (2023)
PERRY: Policy Evaluation with Confidence Intervals using Auxiliary Data
by: Mandyam, Aishwarya, et al.
Published: (2025)
by: Mandyam, Aishwarya, et al.
Published: (2025)
Provable tradeoffs in adversarially robust classification
by: Dobriban, Edgar, et al.
Published: (2020)
by: Dobriban, Edgar, et al.
Published: (2020)
Inference in Randomized Least Squares and PCA via Normality of Quadratic Forms
by: Wang, Leda, et al.
Published: (2024)
by: Wang, Leda, et al.
Published: (2024)
Minimax Optimal Fair Classification with Bounded Demographic Disparity
by: Zeng, Xianli, et al.
Published: (2024)
by: Zeng, Xianli, et al.
Published: (2024)
Evaluating the Performance of Large Language Models via Debates
by: Moniri, Behrad, et al.
Published: (2024)
by: Moniri, Behrad, et al.
Published: (2024)
CONTINA: Confidence Interval for Traffic Demand Prediction with Coverage Guarantee
by: Yang, Chao, et al.
Published: (2025)
by: Yang, Chao, et al.
Published: (2025)
Budgeting Counterfactual for Offline RL
by: Liu, Yao, et al.
Published: (2023)
by: Liu, Yao, et al.
Published: (2023)
Synthetic-Powered Multiple Testing with FDR Control
by: Lee, Yonghoon, et al.
Published: (2026)
by: Lee, Yonghoon, et al.
Published: (2026)
Learning Capacity: A Measure of the Effective Dimensionality of a Model
by: Chen, Daiwei, et al.
Published: (2023)
by: Chen, Daiwei, et al.
Published: (2023)
Confidence-Aware Multi-Field Model Calibration
by: Zhao, Yuang, et al.
Published: (2024)
by: Zhao, Yuang, et al.
Published: (2024)
Efficient and Multiply Robust Risk Estimation under General Forms of Dataset Shift
by: Qiu, Hongxiang, et al.
Published: (2023)
by: Qiu, Hongxiang, et al.
Published: (2023)
Conformalized Interval Arithmetic with Symmetric Calibration
by: Luo, Rui, et al.
Published: (2024)
by: Luo, Rui, et al.
Published: (2024)
Uncertainty in Language Models: Assessment through Rank-Calibration
by: Huang, Xinmeng, et al.
Published: (2024)
by: Huang, Xinmeng, et al.
Published: (2024)
Bayes-Optimal Fair Classification with Linear Disparity Constraints via Pre-, In-, and Post-processing
by: Zeng, Xianli, et al.
Published: (2024)
by: Zeng, Xianli, et al.
Published: (2024)
Risk-Controlled Post-Processing of Decision Policies
by: Joshi, Sunay, et al.
Published: (2026)
by: Joshi, Sunay, et al.
Published: (2026)
Similar Items
-
Watermarking Language Models with Error Correcting Codes
by: Chao, Patrick, et al.
Published: (2024) -
Singleton-Optimized Conformal Prediction
by: Wang, Tao, et al.
Published: (2025) -
Confidence Intervals for Error Rates in 1:1 Matching Tasks: Critical Statistical Analysis and Recommendations
by: Fogliato, Riccardo, et al.
Published: (2023) -
Statistical Methods in Generative AI
by: Dobriban, Edgar
Published: (2025) -
Solving a Research Problem in Mathematical Statistics with AI Assistance
by: Dobriban, Edgar
Published: (2025)