STADE: Standard Deviation as a Pruning Metric
Fuente:
arXiv
Guardado en:
| Autores principales: | Mecke, Diego Coello de Portugal, Alyoussef, Haya, Stubbemann, Maximilian, Koloiarov, Ilia, Hanika, Tom, Schmidt-Thieme, Lars |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prune, Update and Trim: Robust Structured Pruning for Large Language Models
por: Mecke, Diego Coello de Portugal, et al.
Publicado: (2026)
por: Mecke, Diego Coello de Portugal, et al.
Publicado: (2026)
HexFormer: Hyperbolic Vision Transformer with Exponential Map Aggregation
por: Alyoussef, Haya, et al.
Publicado: (2026)
por: Alyoussef, Haya, et al.
Publicado: (2026)
Are EEG Sequences Time Series? EEG Classification with Time Series Models and Joint Subject Training
por: Burchert, Johannes, et al.
Publicado: (2024)
por: Burchert, Johannes, et al.
Publicado: (2024)
Reproducibility and Geometric Intrinsic Dimensionality: An Investigation on Graph Neural Network Research
por: Hille, Tobias, et al.
Publicado: (2024)
por: Hille, Tobias, et al.
Publicado: (2024)
A Cross-Domain Benchmark for Active Learning
por: Werner, Thorben, et al.
Publicado: (2024)
por: Werner, Thorben, et al.
Publicado: (2024)
ProbSAINT: Probabilistic Tabular Regression for Used Car Pricing
por: Madhusudhanan, Kiran, et al.
Publicado: (2024)
por: Madhusudhanan, Kiran, et al.
Publicado: (2024)
Functional Latent Dynamics for Irregularly Sampled Time Series Forecasting
por: Klötergens, Christian, et al.
Publicado: (2024)
por: Klötergens, Christian, et al.
Publicado: (2024)
Valid and Expressive Copulas for Irregular Multivariate Time Series
por: Klötergens, Christian, et al.
Publicado: (2026)
por: Klötergens, Christian, et al.
Publicado: (2026)
Moco: A Learnable Meta Optimizer for Combinatorial Optimization
por: Dernedde, Tim, et al.
Publicado: (2024)
por: Dernedde, Tim, et al.
Publicado: (2024)
LAtte: Hyperbolic Lorentz Attention for Cross-Subject EEG Classification
por: Bdeir, Ahmad, et al.
Publicado: (2026)
por: Bdeir, Ahmad, et al.
Publicado: (2026)
TabResFlow: A Normalizing Spline Flow Model for Probabilistic Univariate Tabular Regression
por: Madhusudhanan, Kiran, et al.
Publicado: (2025)
por: Madhusudhanan, Kiran, et al.
Publicado: (2025)
Physiome-ODE: A Benchmark for Irregularly Sampled Multivariate Time Series Forecasting Based on Biological ODEs
por: Klötergens, Christian, et al.
Publicado: (2025)
por: Klötergens, Christian, et al.
Publicado: (2025)
Channel Dependence, Limited Lookback Windows, and the Simplicity of Datasets: How Biased is Time Series Forecasting?
por: Abdelmalak, Ibram, et al.
Publicado: (2025)
por: Abdelmalak, Ibram, et al.
Publicado: (2025)
Bayesian Active Learning By Distribution Disagreement
por: Werner, Thorben, et al.
Publicado: (2025)
por: Werner, Thorben, et al.
Publicado: (2025)
Recurrent State Encoders for Efficient Neural Combinatorial Optimization
por: Dernedde, Tim, et al.
Publicado: (2025)
por: Dernedde, Tim, et al.
Publicado: (2025)
Towards Comparable Active Learning
por: Werner, Thorben, et al.
Publicado: (2023)
por: Werner, Thorben, et al.
Publicado: (2023)
Hyperparameter Tuning MLPs for Probabilistic Time Series Forecasting
por: Madhusudhanan, Kiran, et al.
Publicado: (2024)
por: Madhusudhanan, Kiran, et al.
Publicado: (2024)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
por: Hanika, Tom, et al.
Publicado: (2024)
por: Hanika, Tom, et al.
Publicado: (2024)
The Role of Active Learning in Modern Machine Learning
por: Werner, Thorben, et al.
Publicado: (2025)
por: Werner, Thorben, et al.
Publicado: (2025)
Probabilistic Circuits for Irregular Multivariate Time Series Forecasting
por: Klötergens, Christian, et al.
Publicado: (2026)
por: Klötergens, Christian, et al.
Publicado: (2026)
HMAR: Hierarchical Masked Attention for Multi-Behaviour Recommendation
por: Elsayed, Shereen, et al.
Publicado: (2024)
por: Elsayed, Shereen, et al.
Publicado: (2024)
Conceptual Views of Neural Networks: A Framework for Neuro-Symbolic Analysis
por: Hirth, Johannes, et al.
Publicado: (2022)
por: Hirth, Johannes, et al.
Publicado: (2022)
On Distributional Dependent Performance of Classical and Neural Routing Solvers
por: Thyssens, Daniela, et al.
Publicado: (2025)
por: Thyssens, Daniela, et al.
Publicado: (2025)
HPMixer: Hierarchical Patching for Multivariate Time Series Forecasting
por: Choi, Jung Min, et al.
Publicado: (2026)
por: Choi, Jung Min, et al.
Publicado: (2026)
Temporal Patch Shuffle (TPS): Leveraging Patch-Level Shuffling to Boost Generalization and Robustness in Time Series Forecasting
por: Bakhshaliyev, Jafar, et al.
Publicado: (2026)
por: Bakhshaliyev, Jafar, et al.
Publicado: (2026)
NPMixer: Hierarchical Neighboring Patch Mixing for Time Series Forecasting
por: Choi, Jung Min, et al.
Publicado: (2026)
por: Choi, Jung Min, et al.
Publicado: (2026)
Probabilistic Forecasting of Irregular Time Series via Conditional Flows
por: Yalavarthi, Vijaya Krishna, et al.
Publicado: (2024)
por: Yalavarthi, Vijaya Krishna, et al.
Publicado: (2024)
Mixing It Up: Exploring Mixer Networks for Irregular Multivariate Time Series Forecasting
por: Klötergens, Christian, et al.
Publicado: (2025)
por: Klötergens, Christian, et al.
Publicado: (2025)
Rethinking Convolutional Networks for Attribute-Aware Sequential Recommendation
por: Elsayed, Shereen, et al.
Publicado: (2026)
por: Elsayed, Shereen, et al.
Publicado: (2026)
Marginalization Consistent Probabilistic Forecasting of Irregular Time Series via Mixture of Separable flows
por: Yalavarthi, Vijaya Krishna, et al.
Publicado: (2024)
por: Yalavarthi, Vijaya Krishna, et al.
Publicado: (2024)
Effective Layer Pruning Through Similarity Metric Perspective
por: Pons, Ian, et al.
Publicado: (2024)
por: Pons, Ian, et al.
Publicado: (2024)
Evaluation of Neural Networks Defenses and Attacks using NDCG and Reciprocal Rank Metrics
por: Brama, Haya, et al.
Publicado: (2022)
por: Brama, Haya, et al.
Publicado: (2022)
Standard-Deviation-Inspired Regularization for Improving Adversarial Robustness
por: Fakorede, Olukorede, et al.
Publicado: (2024)
por: Fakorede, Olukorede, et al.
Publicado: (2024)
Computing the Distance between unbalanced Distributions -- The flat Metric
por: Schmidt, Henri, et al.
Publicado: (2023)
por: Schmidt, Henri, et al.
Publicado: (2023)
Distribution-free Deviation Bounds and The Role of Domain Knowledge in Learning via Model Selection with Cross-validation Risk Estimation
por: Marcondes, Diego, et al.
Publicado: (2023)
por: Marcondes, Diego, et al.
Publicado: (2023)
Machine Learning needs Better Randomness Standards: Randomised Smoothing and PRNG-based attacks
por: Dahiya, Pranav, et al.
Publicado: (2023)
por: Dahiya, Pranav, et al.
Publicado: (2023)
Effective Data Pruning through Score Extrapolation
por: Schmidt, Sebastian, et al.
Publicado: (2025)
por: Schmidt, Sebastian, et al.
Publicado: (2025)
Standardizing Structural Causal Models
por: Ormaniec, Weronika, et al.
Publicado: (2024)
por: Ormaniec, Weronika, et al.
Publicado: (2024)
UniPruning: Unifying Local Metric and Global Feedback for Scalable Sparse LLMs
por: Ding, Yizhuo, et al.
Publicado: (2025)
por: Ding, Yizhuo, et al.
Publicado: (2025)
KVzap: Fast, Adaptive, and Faithful KV Cache Pruning
por: Jegou, Simon, et al.
Publicado: (2026)
por: Jegou, Simon, et al.
Publicado: (2026)
Ejemplares similares
-
Prune, Update and Trim: Robust Structured Pruning for Large Language Models
por: Mecke, Diego Coello de Portugal, et al.
Publicado: (2026) -
HexFormer: Hyperbolic Vision Transformer with Exponential Map Aggregation
por: Alyoussef, Haya, et al.
Publicado: (2026) -
Are EEG Sequences Time Series? EEG Classification with Time Series Models and Joint Subject Training
por: Burchert, Johannes, et al.
Publicado: (2024) -
Reproducibility and Geometric Intrinsic Dimensionality: An Investigation on Graph Neural Network Research
por: Hille, Tobias, et al.
Publicado: (2024) -
A Cross-Domain Benchmark for Active Learning
por: Werner, Thorben, et al.
Publicado: (2024)