Position: Why We Must Rethink Empirical Research in Machine Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Herrmann, Moritz, Lange, F. Julian D., Eggensperger, Katharina, Casalicchio, Giuseppe, Wever, Marcel, Feurer, Matthias, Rügamer, David, Hüllermeier, Eyke, Boulesteix, Anne-Laure, Bischl, Bernd |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CASHomon Sets: Efficient Rashomon Sets Across Multiple Model Classes and their Hyperparameters
por: Ewald, Fiona Katharina, et al.
Publicado: (2026)
por: Ewald, Fiona Katharina, et al.
Publicado: (2026)
Over-optimism in benchmark studies and the multiplicity of design and analysis options when interpreting their results
por: Nießl, Christina, et al.
Publicado: (2021)
por: Nießl, Christina, et al.
Publicado: (2021)
Information Leakage Detection through Approximate Bayes-optimal Prediction
por: Gupta, Pritha, et al.
Publicado: (2024)
por: Gupta, Pritha, et al.
Publicado: (2024)
Reliable Part-of-Speech Tagging of Historical Corpora through Set-Valued Prediction
por: Heid, Stefan, et al.
Publicado: (2020)
por: Heid, Stefan, et al.
Publicado: (2020)
Mind the Gap: Measuring Generalization Performance Across Multiple Objectives
por: Feurer, Matthias, et al.
Publicado: (2022)
por: Feurer, Matthias, et al.
Publicado: (2022)
Structured Credal Learning
por: Venkatesh, Varun, et al.
Publicado: (2026)
por: Venkatesh, Varun, et al.
Publicado: (2026)
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration
por: Rodemann, Julian, et al.
Publicado: (2024)
por: Rodemann, Julian, et al.
Publicado: (2024)
Analyzing Error Sources in Global Feature Effect Estimation
por: Heiß, Timo, et al.
Publicado: (2026)
por: Heiß, Timo, et al.
Publicado: (2026)
On the Robustness of Global Feature Effect Explanations
por: Baniecki, Hubert, et al.
Publicado: (2024)
por: Baniecki, Hubert, et al.
Publicado: (2024)
Efficient and Accurate Explanation Estimation with Distribution Compression
por: Baniecki, Hubert, et al.
Publicado: (2024)
por: Baniecki, Hubert, et al.
Publicado: (2024)
Overtuning in Hyperparameter Optimization
por: Schneider, Lennart, et al.
Publicado: (2025)
por: Schneider, Lennart, et al.
Publicado: (2025)
Mitigating Label Noise through Data Ambiguation
por: Lienen, Julian, et al.
Publicado: (2023)
por: Lienen, Julian, et al.
Publicado: (2023)
Rethinking the handling of method failure in comparison studies
por: Wünsch, Milena, et al.
Publicado: (2024)
por: Wünsch, Milena, et al.
Publicado: (2024)
xplainfi: Feature Importance and Statistical Inference for Machine Learning in R
por: Burk, Lukas, et al.
Publicado: (2026)
por: Burk, Lukas, et al.
Publicado: (2026)
Position Paper: Bridging the Gap Between Machine Learning and Sensitivity Analysis
por: Scholbeck, Christian A., et al.
Publicado: (2023)
por: Scholbeck, Christian A., et al.
Publicado: (2023)
mlr3summary: Concise and interpretable summaries for machine learning models
por: Dandl, Susanne, et al.
Publicado: (2024)
por: Dandl, Susanne, et al.
Publicado: (2024)
Semi-Implicit Variational Inference via Kernelized Path Gradient Descent
por: Pielok, Tobias, et al.
Publicado: (2025)
por: Pielok, Tobias, et al.
Publicado: (2025)
Revisiting Unbiased Implicit Variational Inference
por: Pielok, Tobias, et al.
Publicado: (2025)
por: Pielok, Tobias, et al.
Publicado: (2025)
Evolutionary Mapping of Neural Networks to Spatial Accelerators
por: Pierro, Alessandro, et al.
Publicado: (2026)
por: Pierro, Alessandro, et al.
Publicado: (2026)
Reshuffling Resampling Splits Can Improve Generalization of Hyperparameter Optimization
por: Nagler, Thomas, et al.
Publicado: (2024)
por: Nagler, Thomas, et al.
Publicado: (2024)
fmeffects: An R Package for Forward Marginal Effects
por: Löwe, Holger, et al.
Publicado: (2023)
por: Löwe, Holger, et al.
Publicado: (2023)
Optimal Transport Group Counterfactual Explanations
por: Valero-Leal, Enrique, et al.
Publicado: (2026)
por: Valero-Leal, Enrique, et al.
Publicado: (2026)
Decomposing Global Feature Effects Based on Feature Interactions
por: Herbinger, Julia, et al.
Publicado: (2023)
por: Herbinger, Julia, et al.
Publicado: (2023)
A Guide to Feature Importance Methods for Scientific Inference
por: Ewald, Fiona Katharina, et al.
Publicado: (2024)
por: Ewald, Fiona Katharina, et al.
Publicado: (2024)
Can Fairness be Automated? Guidelines and Opportunities for Fairness-aware AutoML
por: Weerts, Hilde, et al.
Publicado: (2023)
por: Weerts, Hilde, et al.
Publicado: (2023)
Constructing Confidence Intervals for 'the' Generalization Error -- a Comprehensive Benchmark Study
por: Schulz-Kümpel, Hannah, et al.
Publicado: (2024)
por: Schulz-Kümpel, Hannah, et al.
Publicado: (2024)
ALPBench: A Benchmark for Active Learning Pipelines on Tabular Data
por: Margraf, Valentin, et al.
Publicado: (2024)
por: Margraf, Valentin, et al.
Publicado: (2024)
On "Confirmatory" Methodological Research in Statistics and Related Fields
por: Lange, F. J. D., et al.
Publicado: (2025)
por: Lange, F. J. D., et al.
Publicado: (2025)
On “Confirmatory” Methodological Research in Statistics and Related Fields
por: F. J. D. Lange, et al.
Publicado: (2025)
por: F. J. D. Lange, et al.
Publicado: (2025)
Linear Opinion Pooling for Uncertainty Quantification on Graphs
por: Damke, Clemens, et al.
Publicado: (2024)
por: Damke, Clemens, et al.
Publicado: (2024)
Distribution Matching for Graph Quantification Under Structural Covariate Shift
por: Damke, Clemens, et al.
Publicado: (2025)
por: Damke, Clemens, et al.
Publicado: (2025)
Adjusted Count Quantification Learning on Graphs
por: Damke, Clemens, et al.
Publicado: (2025)
por: Damke, Clemens, et al.
Publicado: (2025)
CUQ-GNN: Committee-based Graph Uncertainty Quantification using Posterior Networks
por: Damke, Clemens, et al.
Publicado: (2024)
por: Damke, Clemens, et al.
Publicado: (2024)
Aleatoric and Epistemic Uncertainty Measures for Ordinal Classification through Binary Reduction
por: Haas, Stefan, et al.
Publicado: (2025)
por: Haas, Stefan, et al.
Publicado: (2025)
To tweak or not to tweak. How exploiting flexibilities in gene set analysis leads to over-optimism
por: Wünsch, Milena, et al.
Publicado: (2024)
por: Wünsch, Milena, et al.
Publicado: (2024)
To Tweak or Not to Tweak. How Exploiting Flexibilities in Gene Set Analysis Leads to Overoptimism
por: Milena Wünsch, et al.
Publicado: (2024)
por: Milena Wünsch, et al.
Publicado: (2024)
Differentiable Sparsity via $D$-Gating: Simple and Versatile Structured Penalization
por: Kolb, Chris, et al.
Publicado: (2025)
por: Kolb, Chris, et al.
Publicado: (2025)
Calibrating LLMs with Information-Theoretic Evidential Deep Learning
por: Li, Yawei, et al.
Publicado: (2025)
por: Li, Yawei, et al.
Publicado: (2025)
Deep Weight Factorization: Sparse Learning Through the Lens of Artificial Symmetries
por: Kolb, Chris, et al.
Publicado: (2025)
por: Kolb, Chris, et al.
Publicado: (2025)
Post-hoc Orthogonalization for Mitigation of Protected Feature Bias in CXR Embeddings
por: Weber, Tobias, et al.
Publicado: (2023)
por: Weber, Tobias, et al.
Publicado: (2023)
Ejemplares similares
-
CASHomon Sets: Efficient Rashomon Sets Across Multiple Model Classes and their Hyperparameters
por: Ewald, Fiona Katharina, et al.
Publicado: (2026) -
Over-optimism in benchmark studies and the multiplicity of design and analysis options when interpreting their results
por: Nießl, Christina, et al.
Publicado: (2021) -
Information Leakage Detection through Approximate Bayes-optimal Prediction
por: Gupta, Pritha, et al.
Publicado: (2024) -
Reliable Part-of-Speech Tagging of Historical Corpora through Set-Valued Prediction
por: Heid, Stefan, et al.
Publicado: (2020) -
Mind the Gap: Measuring Generalization Performance Across Multiple Objectives
por: Feurer, Matthias, et al.
Publicado: (2022)