It's Not a Lottery, It's a Race: Understanding How Gradient Descent Adapts the Network's Capacity to the Task
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Pinson, Hannah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Winning the Lottery by Preserving Network Training Dynamics with Concrete Ticket Search
von: Arora, Tanay, et al.
Veröffentlicht: (2025)
von: Arora, Tanay, et al.
Veröffentlicht: (2025)
Arc Gradient Descent: A Geometrically Motivated Gradient Descent-based Optimiser with Phase-Aware, User-Controlled Step Dynamics (proof-of-concept)
von: Verma, Nikhil, et al.
Veröffentlicht: (2025)
von: Verma, Nikhil, et al.
Veröffentlicht: (2025)
Stochastic Gradient Sampling for Enhancing Neural Networks Training
von: Yun, Juyoung
Veröffentlicht: (2023)
von: Yun, Juyoung
Veröffentlicht: (2023)
CODE-CL: Conceptor-Based Gradient Projection for Deep Continual Learning
von: Apolinario, Marco Paul E., et al.
Veröffentlicht: (2024)
von: Apolinario, Marco Paul E., et al.
Veröffentlicht: (2024)
Understanding Transformer-based Vision Models through Inversion
von: Rathjens, Jan, et al.
Veröffentlicht: (2024)
von: Rathjens, Jan, et al.
Veröffentlicht: (2024)
Breaking Free: How to Hack Safety Guardrails in Black-Box Diffusion Models!
von: Kotyan, Shashank, et al.
Veröffentlicht: (2024)
von: Kotyan, Shashank, et al.
Veröffentlicht: (2024)
The Cooperative Network Architecture: Learning Structured Networks as Representation of Sensory Patterns
von: Sager, Pascal J., et al.
Veröffentlicht: (2024)
von: Sager, Pascal J., et al.
Veröffentlicht: (2024)
B-DENSE: Branching For Dense Ensemble Network Supervision Efficiency
von: Puniani, Cherish, et al.
Veröffentlicht: (2026)
von: Puniani, Cherish, et al.
Veröffentlicht: (2026)
Benchmarking Spiking Neural Network Learning Methods with Varying Locality
von: Lin, Jiaqi, et al.
Veröffentlicht: (2024)
von: Lin, Jiaqi, et al.
Veröffentlicht: (2024)
Brain-Inspired Efficient Pruning: Exploiting Criticality in Spiking Neural Networks
von: Chen, Shuo, et al.
Veröffentlicht: (2023)
von: Chen, Shuo, et al.
Veröffentlicht: (2023)
Threshold Modulation for Online Test-Time Adaptation of Spiking Neural Networks
von: Zhao, Kejie, et al.
Veröffentlicht: (2025)
von: Zhao, Kejie, et al.
Veröffentlicht: (2025)
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
von: Yao, Xingting, et al.
Veröffentlicht: (2025)
von: Yao, Xingting, et al.
Veröffentlicht: (2025)
Temporal Flexibility in Spiking Neural Networks: Towards Generalization Across Time Steps and Deployment Friendliness
von: Du, Kangrui, et al.
Veröffentlicht: (2025)
von: Du, Kangrui, et al.
Veröffentlicht: (2025)
Guiding a Diffusion Model with a Bad Version of Itself
von: Karras, Tero, et al.
Veröffentlicht: (2024)
von: Karras, Tero, et al.
Veröffentlicht: (2024)
Applying Guidance in a Limited Interval Improves Sample and Distribution Quality in Diffusion Models
von: Kynkäänniemi, Tuomas, et al.
Veröffentlicht: (2024)
von: Kynkäänniemi, Tuomas, et al.
Veröffentlicht: (2024)
Forward-Cooperation-Backward (FCB) learning in a Multi-Encoding Uni-Decoding neural network architecture
von: Dutta, Prasun, et al.
Veröffentlicht: (2025)
von: Dutta, Prasun, et al.
Veröffentlicht: (2025)
SMART: Scalable Mesh-free Aerodynamic Simulations from Raw Geometries using a Transformer-based Surrogate Model
von: Hagnberger, Jan, et al.
Veröffentlicht: (2026)
von: Hagnberger, Jan, et al.
Veröffentlicht: (2026)
Sharpness-Aware Minimization with Z-Score Gradient Filtering
von: Yun, Vincent-Daniel
Veröffentlicht: (2025)
von: Yun, Vincent-Daniel
Veröffentlicht: (2025)
Accuracy-Robustness Trade Off via Spiking Neural Network Gradient Sparsity Trail
von: Nhan, Luu Trong, et al.
Veröffentlicht: (2025)
von: Nhan, Luu Trong, et al.
Veröffentlicht: (2025)
Revisiting the Platonic Representation Hypothesis: An Aristotelian View
von: Gröger, Fabian, et al.
Veröffentlicht: (2026)
von: Gröger, Fabian, et al.
Veröffentlicht: (2026)
Socrates Loss: Unifying Confidence Calibration and Classification by Leveraging the Unknown
von: Gómez-Gálvez, Sandra, et al.
Veröffentlicht: (2026)
von: Gómez-Gálvez, Sandra, et al.
Veröffentlicht: (2026)
Neural Sentinel: Unified Vision Language Model (VLM) for License Plate Recognition with Human-in-the-Loop Continual Learning
von: Sivakoti, Karthik
Veröffentlicht: (2026)
von: Sivakoti, Karthik
Veröffentlicht: (2026)
Retinal Cyst Detection from Optical Coherence Tomography Images
von: Dharmaratnakar, Abhishek, et al.
Veröffentlicht: (2026)
von: Dharmaratnakar, Abhishek, et al.
Veröffentlicht: (2026)
Trust-SSL: Additive-Residual Selective Invariance for Robust Aerial Self-Supervised Learning
von: Boulila, Wadii, et al.
Veröffentlicht: (2026)
von: Boulila, Wadii, et al.
Veröffentlicht: (2026)
Planktonzilla: Multimodal dataset and models for understanding plankton ecosystems
von: Montanares, Alan Gerson Contreras, et al.
Veröffentlicht: (2026)
von: Montanares, Alan Gerson Contreras, et al.
Veröffentlicht: (2026)
Updating the standard neuron model in artificial neural networks
von: Mohedano, Raul, et al.
Veröffentlicht: (2026)
von: Mohedano, Raul, et al.
Veröffentlicht: (2026)
Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models
von: McAllister, David, et al.
Veröffentlicht: (2026)
von: McAllister, David, et al.
Veröffentlicht: (2026)
NeuRN: Neuro-inspired Domain Generalization for Image Classification
von: Jalil, Hamd, et al.
Veröffentlicht: (2025)
von: Jalil, Hamd, et al.
Veröffentlicht: (2025)
Mice to Machines: Neural Representations from Visual Cortex for Domain Generalization
von: Qazi, Ahmed, et al.
Veröffentlicht: (2025)
von: Qazi, Ahmed, et al.
Veröffentlicht: (2025)
Concept Probing: Where to Find Human-Defined Concepts (Extended Version)
von: Ribeiro, Manuel de Sousa, et al.
Veröffentlicht: (2025)
von: Ribeiro, Manuel de Sousa, et al.
Veröffentlicht: (2025)
APTx Neuron: A Unified Trainable Neuron Architecture Integrating Activation and Computation
von: Kumar, Ravin
Veröffentlicht: (2025)
von: Kumar, Ravin
Veröffentlicht: (2025)
The Platonic Representation Hypothesis
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
Neuromimetic metaplasticity for adaptive continual learning
von: Cho, Suhee, et al.
Veröffentlicht: (2024)
von: Cho, Suhee, et al.
Veröffentlicht: (2024)
CoNO: Complex Neural Operator for Continous Dynamical Physical Systems
von: Tiwari, Karn, et al.
Veröffentlicht: (2024)
von: Tiwari, Karn, et al.
Veröffentlicht: (2024)
Learning Evolution via Optimization Knowledge Adaptation
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
When Person Re-Identification Meets Event Camera: A Benchmark Dataset and An Attribute-guided Re-Identification Framework
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
An extended asymmetric sigmoid with Perceptron (SIGTRON) for imbalanced linear classification
von: Woo, Hyenkyun
Veröffentlicht: (2023)
von: Woo, Hyenkyun
Veröffentlicht: (2023)
APTx: better activation function than MISH, SWISH, and ReLU's variants used in deep learning
von: Kumar, Ravin
Veröffentlicht: (2022)
von: Kumar, Ravin
Veröffentlicht: (2022)
Data augmentation with automated machine learning: approaches and performance comparison with classical data augmentation methods
von: Mumuni, Alhassan, et al.
Veröffentlicht: (2024)
von: Mumuni, Alhassan, et al.
Veröffentlicht: (2024)
Fooling Contrastive Language-Image Pre-trained Models with CLIPMasterPrints
von: Freiberger, Matthias, et al.
Veröffentlicht: (2023)
von: Freiberger, Matthias, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Winning the Lottery by Preserving Network Training Dynamics with Concrete Ticket Search
von: Arora, Tanay, et al.
Veröffentlicht: (2025) -
Arc Gradient Descent: A Geometrically Motivated Gradient Descent-based Optimiser with Phase-Aware, User-Controlled Step Dynamics (proof-of-concept)
von: Verma, Nikhil, et al.
Veröffentlicht: (2025) -
Stochastic Gradient Sampling for Enhancing Neural Networks Training
von: Yun, Juyoung
Veröffentlicht: (2023) -
CODE-CL: Conceptor-Based Gradient Projection for Deep Continual Learning
von: Apolinario, Marco Paul E., et al.
Veröffentlicht: (2024) -
Understanding Transformer-based Vision Models through Inversion
von: Rathjens, Jan, et al.
Veröffentlicht: (2024)