What Scalable Second-Order Information Knows for Pruning at Initialization
Fuente:
arXiv
Saved in:
| Main Authors: | Navarrete, Ivo Gollini, Ávila, Nicolás Mauricio Cuadrado, Takáč, Martin, Horváth, Samuel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hard Samples, Bad Labels: Robust Loss Functions That Know When to Back Off
by: Pellegrino, Nicholas, et al.
Published: (2025)
by: Pellegrino, Nicholas, et al.
Published: (2025)
Effects of Initialization Biases on Deep Neural Network Training Dynamics
by: Pellegrino, Nicholas, et al.
Published: (2025)
by: Pellegrino, Nicholas, et al.
Published: (2025)
Distinguished In Uniform: Self Attention Vs. Virtual Nodes
by: Rosenbluth, Eran, et al.
Published: (2024)
by: Rosenbluth, Eran, et al.
Published: (2024)
Contrastive Mutual Information Learning: Toward Robust Representations without Positive-Pair Augmentations
by: Livne, Micha
Published: (2025)
by: Livne, Micha
Published: (2025)
Contrastive MIM: A Contrastive Mutual Information Framework for Unified Generative and Discriminative Representation Learning
by: Livne, Micha
Published: (2025)
by: Livne, Micha
Published: (2025)
Epistemology of Generative AI: The Geometry of Knowing
by: Levin, Ilya
Published: (2026)
by: Levin, Ilya
Published: (2026)
Order-Robust Class Incremental Learning: Graph-Driven Dynamic Similarity Grouping
by: Lai, Guannan, et al.
Published: (2025)
by: Lai, Guannan, et al.
Published: (2025)
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
by: Hussonnois, Maxence, et al.
Published: (2025)
by: Hussonnois, Maxence, et al.
Published: (2025)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
by: Atamuradov, Sanjar
Published: (2025)
by: Atamuradov, Sanjar
Published: (2025)
Adaptive Weighting in Knowledge Distillation: An Axiomatic Framework for Multi-Scale Teacher Ensemble Optimization
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
When Redundancy Matters: Machine Teaching of Representations
by: Ferri, Cèsar, et al.
Published: (2024)
by: Ferri, Cèsar, et al.
Published: (2024)
Data structure > labels? Unsupervised heuristics for SVM hyperparameter estimation
by: Cholewa, Michał, et al.
Published: (2021)
by: Cholewa, Michał, et al.
Published: (2021)
Generative Pretrained Embedding and Hierarchical Irregular Time Series Representation for Daily Living Activity Recognition
by: Bouchabou, Damien, et al.
Published: (2024)
by: Bouchabou, Damien, et al.
Published: (2024)
Reinforcement Learning for Stock Transactions
by: Zhou, Ziyi, et al.
Published: (2025)
by: Zhou, Ziyi, et al.
Published: (2025)
Understanding the Limits of Deep Tabular Methods with Temporal Shift
by: Cai, Hao-Run, et al.
Published: (2025)
by: Cai, Hao-Run, et al.
Published: (2025)
Multi-Teacher Ensemble Distillation: A Mathematical Framework for Probability-Domain Knowledge Aggregation
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
RVI-SAC: Average Reward Off-Policy Deep Reinforcement Learning
by: Hisaki, Yukinari, et al.
Published: (2024)
by: Hisaki, Yukinari, et al.
Published: (2024)
Grammar-based evolutionary approach for automated workflow composition with domain-specific operators and ensemble diversity
by: Barbudo, Rafael, et al.
Published: (2024)
by: Barbudo, Rafael, et al.
Published: (2024)
Feature-aware Modulation for Learning from Temporal Tabular Data
by: Cai, Hao-Run, et al.
Published: (2025)
by: Cai, Hao-Run, et al.
Published: (2025)
Temporal Taskification in Streaming Continual Learning: A Source of Evaluation Instability
by: Filat, Nicolae, et al.
Published: (2026)
by: Filat, Nicolae, et al.
Published: (2026)
Fine-Tuning Regimes Define Distinct Continual Learning Problems
by: Iordache, Paul-Tiberiu, et al.
Published: (2026)
by: Iordache, Paul-Tiberiu, et al.
Published: (2026)
Adaptive Exploration for Latent-State Bandits
by: Jin, Jikai, et al.
Published: (2026)
by: Jin, Jikai, et al.
Published: (2026)
Data-Driven Preference Sampling for Pareto Front Learning
by: Ye, Rongguang, et al.
Published: (2024)
by: Ye, Rongguang, et al.
Published: (2024)
Quantum Machine Learning for Predicting Anastomotic Leak: A Clinical Study
by: Novák, Vojtěch, et al.
Published: (2025)
by: Novák, Vojtěch, et al.
Published: (2025)
Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis
by: Suetake, Yamato, et al.
Published: (2026)
by: Suetake, Yamato, et al.
Published: (2026)
Adaptive Bernstein Change Detector for High-Dimensional Data Streams
by: Heyden, Marco, et al.
Published: (2023)
by: Heyden, Marco, et al.
Published: (2023)
I Know What I Don't Know: Latent Posterior Factor Models for Multi-Evidence Probabilistic Reasoning
by: Alege, Aliyu Agboola
Published: (2026)
by: Alege, Aliyu Agboola
Published: (2026)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols
by: Reitich, Fernando
Published: (2026)
by: Reitich, Fernando
Published: (2026)
PACT: Reducing Alert Fatigue in Low-Prevalence SOC Streams with Triggered Active Learning
by: Ndichu, Samuel, et al.
Published: (2026)
by: Ndichu, Samuel, et al.
Published: (2026)
Repetition Makes Perfect: Recurrent Graph Neural Networks Match Message-Passing Limit
by: Rosenbluth, Eran, et al.
Published: (2025)
by: Rosenbluth, Eran, et al.
Published: (2025)
Formulation and Therapeutic Assessment of a Zinc Oxide, Silver, and Cerium Oxide Enriched Ointment for Accelerated Wound Healing in Aged Models
by: Yousaf, Iqra, et al.
Published: (2025)
by: Yousaf, Iqra, et al.
Published: (2025)
How much do LLMs learn from negative examples?
by: Hamdan, Shadi, et al.
Published: (2025)
by: Hamdan, Shadi, et al.
Published: (2025)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
by: Saini, Saurabh, et al.
Published: (2026)
by: Saini, Saurabh, et al.
Published: (2026)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
by: Kim, Heejun, et al.
Published: (2026)
by: Kim, Heejun, et al.
Published: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
by: Singh, Diyansha
Published: (2026)
by: Singh, Diyansha
Published: (2026)
Adaptive Mixture Importance Sampling for Automated Ads Auction Tuning
by: Jia, Yimeng, et al.
Published: (2024)
by: Jia, Yimeng, et al.
Published: (2024)
Post-Training Probability Manifold Correction via Structured SVD Pruning and Self-Referential Distillation
by: Flouro, Aaron R., et al.
Published: (2026)
by: Flouro, Aaron R., et al.
Published: (2026)
Harnessing non-adversarial robustness in large language models
by: Zhou, Qinghua, et al.
Published: (2026)
by: Zhou, Qinghua, et al.
Published: (2026)
Similar Items
-
Hard Samples, Bad Labels: Robust Loss Functions That Know When to Back Off
by: Pellegrino, Nicholas, et al.
Published: (2025) -
Effects of Initialization Biases on Deep Neural Network Training Dynamics
by: Pellegrino, Nicholas, et al.
Published: (2025) -
Distinguished In Uniform: Self Attention Vs. Virtual Nodes
by: Rosenbluth, Eran, et al.
Published: (2024) -
Contrastive Mutual Information Learning: Toward Robust Representations without Positive-Pair Augmentations
by: Livne, Micha
Published: (2025) -
Contrastive MIM: A Contrastive Mutual Information Framework for Unified Generative and Discriminative Representation Learning
by: Livne, Micha
Published: (2025)