Which Leakage Types Matter? A Quantitative Landscape Across 2,047 Benchmark Datasets
Fuente:
arXiv
Salvato in:
| Autore principale: | Roth, Simon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Grammar of Machine Learning Workflows: Rejecting Data Leakage at Call Time
di: Roth, Simon
Pubblicazione: (2026)
di: Roth, Simon
Pubblicazione: (2026)
The Optimization Landscape of SGD Across the Feature Learning Strength
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
di: Atanasov, Alexander, et al.
Pubblicazione: (2024)
EuroCropsML: A Time Series Benchmark Dataset For Few-Shot Crop Type Classification
di: Reuss, Joana, et al.
Pubblicazione: (2024)
di: Reuss, Joana, et al.
Pubblicazione: (2024)
Bayesian Optimisation: Which Constraints Matter?
di: Lin, Xietao Wang, et al.
Pubblicazione: (2025)
di: Lin, Xietao Wang, et al.
Pubblicazione: (2025)
Impact of Leakage on Data Harmonization in Machine Learning Pipelines in Class Imbalance Across Sites
di: Nieto, Nicolás, et al.
Pubblicazione: (2024)
di: Nieto, Nicolás, et al.
Pubblicazione: (2024)
Benchmarking Benchmark Leakage in Large Language Models
di: Xu, Ruijie, et al.
Pubblicazione: (2024)
di: Xu, Ruijie, et al.
Pubblicazione: (2024)
Benchmarking Deep Learning Models for Raman Spectroscopy Across Open-Source Datasets
di: Sineesh, Adithya, et al.
Pubblicazione: (2026)
di: Sineesh, Adithya, et al.
Pubblicazione: (2026)
Data Leakage and Redundancy in the LIT-PCBA Benchmark
di: Huang, Amber, et al.
Pubblicazione: (2025)
di: Huang, Amber, et al.
Pubblicazione: (2025)
How Contaminated Is Your Benchmark? Quantifying Dataset Leakage in Large Language Models with Kernel Divergence
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2025)
A Comprehensive Benchmark of Machine and Deep Learning Across Diverse Tabular Datasets
di: Shmuel, Assaf, et al.
Pubblicazione: (2024)
di: Shmuel, Assaf, et al.
Pubblicazione: (2024)
Which Attention Heads Matter for In-Context Learning?
di: Yin, Kayo, et al.
Pubblicazione: (2025)
di: Yin, Kayo, et al.
Pubblicazione: (2025)
LiveClin: A Live Clinical Benchmark without Leakage
di: Wang, Xidong, et al.
Pubblicazione: (2026)
di: Wang, Xidong, et al.
Pubblicazione: (2026)
Benchmarking Anomaly Detection Across Heterogeneous Cloud Telemetry Datasets
di: Islam, Mohammad Saiful, et al.
Pubblicazione: (2026)
di: Islam, Mohammad Saiful, et al.
Pubblicazione: (2026)
Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?
di: Zhang, Mingqiao, et al.
Pubblicazione: (2026)
di: Zhang, Mingqiao, et al.
Pubblicazione: (2026)
Thought Anchors: Which LLM Reasoning Steps Matter?
di: Bogdan, Paul C., et al.
Pubblicazione: (2025)
di: Bogdan, Paul C., et al.
Pubblicazione: (2025)
Shapley Neuron Values for Continual Learning: Which Neurons Matter Most?
di: Vahedifar, Mohammad Ali, et al.
Pubblicazione: (2026)
di: Vahedifar, Mohammad Ali, et al.
Pubblicazione: (2026)
EMBER2024 -- A Benchmark Dataset for Holistic Evaluation of Malware Classifiers
di: Joyce, Robert J., et al.
Pubblicazione: (2025)
di: Joyce, Robert J., et al.
Pubblicazione: (2025)
Risk In Context: Benchmarking Privacy Leakage of Foundation Models in Synthetic Tabular Data Generation
di: Byun, Jessup, et al.
Pubblicazione: (2025)
di: Byun, Jessup, et al.
Pubblicazione: (2025)
Which Rewards Matter? Reward Selection for Reinforcement Learning under Limited Feedback
di: Chaudhari, Shreyas, et al.
Pubblicazione: (2025)
di: Chaudhari, Shreyas, et al.
Pubblicazione: (2025)
Age model of sediment core M77/2_047-2
di: Erdem, Zeynep
Pubblicazione: (2019)
di: Erdem, Zeynep
Pubblicazione: (2019)
Quantum Kernel Methods under Scrutiny: A Benchmarking Study
di: Schnabel, Jan, et al.
Pubblicazione: (2024)
di: Schnabel, Jan, et al.
Pubblicazione: (2024)
MacrOData: New Benchmarks of Thousands of Datasets for Tabular Outlier Detection
di: Ding, Xueying, et al.
Pubblicazione: (2026)
di: Ding, Xueying, et al.
Pubblicazione: (2026)
Benchmarking for Practice: Few-Shot Time-Series Crop-Type Classification on the EuroCropsML Dataset
di: Reuss, Joana, et al.
Pubblicazione: (2025)
di: Reuss, Joana, et al.
Pubblicazione: (2025)
Unveiling Client Privacy Leakage from Public Dataset Usage in Federated Distillation
di: Shi, Haonan, et al.
Pubblicazione: (2025)
di: Shi, Haonan, et al.
Pubblicazione: (2025)
Porewater geochemistry of sediment core M77/2_047-3
di: Sommer, Stefan, et al.
Pubblicazione: (2019)
di: Sommer, Stefan, et al.
Pubblicazione: (2019)
Particulate geochemistry of sediment core M77/2_047-3
di: Sommer, Stefan, et al.
Pubblicazione: (2019)
di: Sommer, Stefan, et al.
Pubblicazione: (2019)
Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't
di: Svete, Anej, et al.
Pubblicazione: (2026)
di: Svete, Anej, et al.
Pubblicazione: (2026)
Leakage-Aware Bandgap Prediction on the JARVIS-DFT Dataset: A Phase-Wise Feature Analysis
di: Sharma, Gaurav Kumar
Pubblicazione: (2025)
di: Sharma, Gaurav Kumar
Pubblicazione: (2025)
On Leakage in Machine Learning Pipelines
di: Sasse, Leonard, et al.
Pubblicazione: (2023)
di: Sasse, Leonard, et al.
Pubblicazione: (2023)
Which Side Are You On? A Multi-task Dataset for End-to-End Argument Summarisation and Evaluation
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Which Company Adjustment Matter? Insights from Uplift Modeling on Financial Health
di: Wang, Xinlin, et al.
Pubblicazione: (2025)
di: Wang, Xinlin, et al.
Pubblicazione: (2025)
What Does it Take to Generalize SER Model Across Datasets? A Comprehensive Benchmark
di: Ibrahim, Adham, et al.
Pubblicazione: (2024)
di: Ibrahim, Adham, et al.
Pubblicazione: (2024)
Which Imputation Fits Which Feature Selection Method? A Survey-Based Simulation Study
di: Schwerter, Jakob, et al.
Pubblicazione: (2024)
di: Schwerter, Jakob, et al.
Pubblicazione: (2024)
InSpaceType: Dataset and Benchmark for Reconsidering Cross-Space Type Performance in Indoor Monocular Depth
di: Wu, Cho-Ying, et al.
Pubblicazione: (2024)
di: Wu, Cho-Ying, et al.
Pubblicazione: (2024)
EngineAD: A Real-World Vehicle Engine Anomaly Detection Dataset
di: Hojjati, Hadi, et al.
Pubblicazione: (2026)
di: Hojjati, Hadi, et al.
Pubblicazione: (2026)
Augmenting Biological Fitness Prediction Benchmarks with Landscapes Features from GraphFLA
di: Huang, Mingyu, et al.
Pubblicazione: (2025)
di: Huang, Mingyu, et al.
Pubblicazione: (2025)
Ubiquitous Symmetry at Critical Points Across Diverse Optimization Landscapes
di: Schneider, Irmi
Pubblicazione: (2025)
di: Schneider, Irmi
Pubblicazione: (2025)
Benchmark Dataset for Catalysis on 2D MXenes
di: Melnyk, Pavlo, et al.
Pubblicazione: (2026)
di: Melnyk, Pavlo, et al.
Pubblicazione: (2026)
Spurious Privacy Leakage in Neural Networks
di: Zhang, Chenxiang, et al.
Pubblicazione: (2025)
di: Zhang, Chenxiang, et al.
Pubblicazione: (2025)
Hidden Leaks in Time Series Forecasting: How Data Leakage Affects LSTM Evaluation Across Configurations and Validation Strategies
di: Albelali, Salma, et al.
Pubblicazione: (2025)
di: Albelali, Salma, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Grammar of Machine Learning Workflows: Rejecting Data Leakage at Call Time
di: Roth, Simon
Pubblicazione: (2026) -
The Optimization Landscape of SGD Across the Feature Learning Strength
di: Atanasov, Alexander, et al.
Pubblicazione: (2024) -
EuroCropsML: A Time Series Benchmark Dataset For Few-Shot Crop Type Classification
di: Reuss, Joana, et al.
Pubblicazione: (2024) -
Bayesian Optimisation: Which Constraints Matter?
di: Lin, Xietao Wang, et al.
Pubblicazione: (2025) -
Impact of Leakage on Data Harmonization in Machine Learning Pipelines in Class Imbalance Across Sites
di: Nieto, Nicolás, et al.
Pubblicazione: (2024)