When +1% Is Not Enough: A Paired Bootstrap Protocol for Evaluating Small Improvements
Fuente:
arXiv
Guardado en:
| Autor principal: | Du, Wenzhang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Non-Heuristic Selection via Hybrid Regularized and Machine Learning Models for Insurance
por: Galvão, Luciano Ribeiro, et al.
Publicado: (2025)
por: Galvão, Luciano Ribeiro, et al.
Publicado: (2025)
Calibrated Counterfactual Conformal Fairness ($C^3F$): Post-hoc, Shift-Aware Coverage Parity via Conformal Prediction and Counterfactual Regularization
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
Enhancing Visual Interpretability and Explainability in Functional Survival Trees and Forests
por: Loffredo, Giuseppe, et al.
Publicado: (2025)
por: Loffredo, Giuseppe, et al.
Publicado: (2025)
An Imbalance-Robust Evaluation Framework for Extreme Risk Forecasts
por: Nikolopoulos, Sotirios D.
Publicado: (2025)
por: Nikolopoulos, Sotirios D.
Publicado: (2025)
Closed-Form Beta Distribution Estimation from Sparse Statistics with Random Forest Implicit Regularization
por: Landers, Jonathan R.
Publicado: (2025)
por: Landers, Jonathan R.
Publicado: (2025)
ML-driven detection and reduction of ballast information in multi-modal datasets
por: Solovko, Yaroslav
Publicado: (2026)
por: Solovko, Yaroslav
Publicado: (2026)
ICE: Intervention-Consistent Explanation Evaluation with Statistical Grounding for LLMs
por: Basu, Abhinaba, et al.
Publicado: (2026)
por: Basu, Abhinaba, et al.
Publicado: (2026)
Evaluation of the number of clusters in a data set using $p$-values from Multiple Tests of Hypotheses
por: Modak, Soumita
Publicado: (2026)
por: Modak, Soumita
Publicado: (2026)
Augmented Functional Random Forests: Classifier Construction and Unbiased Functional Principal Components Importance through Ad-Hoc Conditional Permutations
por: Maturo, Fabrizio, et al.
Publicado: (2024)
por: Maturo, Fabrizio, et al.
Publicado: (2024)
Conservative Decisions with Risk Scores
por: Wei, Yishu, et al.
Publicado: (2025)
por: Wei, Yishu, et al.
Publicado: (2025)
Censored lifespans in a double-truncated sample: Maximum likelihood inference for the exponential distribution
por: Sieg, Fiete, et al.
Publicado: (2025)
por: Sieg, Fiete, et al.
Publicado: (2025)
Optimised Support Vector Regression for California Housing Price Prediction: The Critical Role of Feature Engineering and Hyperparameter Tuning
por: Adutwum, Emmanuel
Publicado: (2026)
por: Adutwum, Emmanuel
Publicado: (2026)
The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction
por: Wan, Shu, et al.
Publicado: (2026)
por: Wan, Shu, et al.
Publicado: (2026)
Bootstrapping not under the null?
por: Derumigny, Alexis, et al.
Publicado: (2025)
por: Derumigny, Alexis, et al.
Publicado: (2025)
Exploring Hierarchical Classification Performance for Time Series Data: Dissimilarity Measures and Classifier Comparisons
por: Alagoz, Celal
Publicado: (2024)
por: Alagoz, Celal
Publicado: (2024)
Forecasting intermittent time series with Gaussian Processes and Tweedie likelihood
por: Damato, Stefano, et al.
Publicado: (2025)
por: Damato, Stefano, et al.
Publicado: (2025)
Ostrom-Weighted Bootstrap: A Theoretically Optimal and Provably Complete Framework for Hierarchical Imputation in Multi-Agent Systems
por: Wakimoto, Hirofumi
Publicado: (2025)
por: Wakimoto, Hirofumi
Publicado: (2025)
SplitWise Regression: Stepwise Modeling with Adaptive Dummy Encoding
por: Kurbucz, Marcell T., et al.
Publicado: (2025)
por: Kurbucz, Marcell T., et al.
Publicado: (2025)
Scalable Bayesian Clustering for Integrative Analysis of Multi-View Data
por: Cabral, Rafael, et al.
Publicado: (2024)
por: Cabral, Rafael, et al.
Publicado: (2024)
Integration of Structural Equation Modeling and Bayesian Networks in the Context of Causal Inference: A Case Study on Personal Positive Youth Development
por: Benitez, Edgar, et al.
Publicado: (2024)
por: Benitez, Edgar, et al.
Publicado: (2024)
Dimension-independent rates for structured neural density estimation
por: Vandermeulen, Robert A., et al.
Publicado: (2024)
por: Vandermeulen, Robert A., et al.
Publicado: (2024)
Ranking of Multi-Response Experiment Treatments
por: Pebes-Trujillo, Miguel R., et al.
Publicado: (2024)
por: Pebes-Trujillo, Miguel R., et al.
Publicado: (2024)
CSTS: A Benchmark for the Discovery of Correlation Structures in Time Series Clustering
por: Degen, Isabella, et al.
Publicado: (2025)
por: Degen, Isabella, et al.
Publicado: (2025)
Statistical Measures for Explainable Aspect-Based Sentiment Analysis: A Case Study on Environmental Discourse in Reddit
por: Stracqualursi, Luisa, et al.
Publicado: (2026)
por: Stracqualursi, Luisa, et al.
Publicado: (2026)
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
por: Manokhin, Valery, et al.
Publicado: (2026)
por: Manokhin, Valery, et al.
Publicado: (2026)
When to Trust Confidence Thresholding: Calibration Diagnostics for Pseudo-Labelled Regression
por: Kurbucz, Marcell T.
Publicado: (2026)
por: Kurbucz, Marcell T.
Publicado: (2026)
Geometric Mixture Classifier (GMC): A Discriminative Per-Class Mixture of Hyperplanes
por: K, Prasanth K, et al.
Publicado: (2025)
por: K, Prasanth K, et al.
Publicado: (2025)
Asymptotic Consistency and Generalization in Hybrid Models of Regularized Selection and Nonlinear Learning
por: Galvão, Luciano Ribeiro, et al.
Publicado: (2025)
por: Galvão, Luciano Ribeiro, et al.
Publicado: (2025)
Evaluating the Quality of the Quantified Uncertainty for (Re)Calibration of Data-Driven Regression Models
por: Wibbeke, Jelke, et al.
Publicado: (2025)
por: Wibbeke, Jelke, et al.
Publicado: (2025)
Factorizable joint shift revisited
por: Tasche, Dirk
Publicado: (2026)
por: Tasche, Dirk
Publicado: (2026)
Confidence Sets for Multidimensional Scaling
por: Vishwanath, Siddharth, et al.
Publicado: (2025)
por: Vishwanath, Siddharth, et al.
Publicado: (2025)
A Bayesian Bootstrap for Mixture Models
por: Cui, Fuheng, et al.
Publicado: (2023)
por: Cui, Fuheng, et al.
Publicado: (2023)
A Hybrid Mixture Approach for Clustering and Characterizing Cancer Data
por: Kareem, Kazeem, et al.
Publicado: (2025)
por: Kareem, Kazeem, et al.
Publicado: (2025)
Pair Correlation Factor and the Sample Complexity of Gaussian Mixtures
por: Aryan, Farzad
Publicado: (2025)
por: Aryan, Farzad
Publicado: (2025)
PRIM-cipal components analysis
por: Liu, Tianhao, et al.
Publicado: (2026)
por: Liu, Tianhao, et al.
Publicado: (2026)
5G-DIL: Domain Incremental Learning with Similarity-Aware Sampling for Dynamic 5G Indoor Localization
por: Raichur, Nisha Lakshmana, et al.
Publicado: (2025)
por: Raichur, Nisha Lakshmana, et al.
Publicado: (2025)
Causal Deep Learning
por: Vasilescu, M. Alex O.
Publicado: (2023)
por: Vasilescu, M. Alex O.
Publicado: (2023)
Multi-layered model-based characterisation of the local-Universe galaxy data from the GAMA survey
por: Dai, Fan, et al.
Publicado: (2026)
por: Dai, Fan, et al.
Publicado: (2026)
Statistical Multicriteria Benchmarking via the GSD-Front
por: Jansen, Christoph, et al.
Publicado: (2024)
por: Jansen, Christoph, et al.
Publicado: (2024)
Dense Video Understanding with Gated Residual Tokenization
por: Zhang, Haichao, et al.
Publicado: (2025)
por: Zhang, Haichao, et al.
Publicado: (2025)
Ejemplares similares
-
Non-Heuristic Selection via Hybrid Regularized and Machine Learning Models for Insurance
por: Galvão, Luciano Ribeiro, et al.
Publicado: (2025) -
Calibrated Counterfactual Conformal Fairness ($C^3F$): Post-hoc, Shift-Aware Coverage Parity via Conformal Prediction and Counterfactual Regularization
por: Alpay, Faruk, et al.
Publicado: (2025) -
Enhancing Visual Interpretability and Explainability in Functional Survival Trees and Forests
por: Loffredo, Giuseppe, et al.
Publicado: (2025) -
An Imbalance-Robust Evaluation Framework for Extreme Risk Forecasts
por: Nikolopoulos, Sotirios D.
Publicado: (2025) -
Closed-Form Beta Distribution Estimation from Sparse Statistics with Random Forest Implicit Regularization
por: Landers, Jonathan R.
Publicado: (2025)