Sharpness-Aware Minimization with Z-Score Gradient Filtering
Fuente:
arXiv
Saved in:
| Main Author: | Yun, Vincent-Daniel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast Fourier Transform-Based Spectral and Temporal Gradient Filtering for Differential Privacy
by: Shin, Hyeju, et al.
Published: (2025)
by: Shin, Hyeju, et al.
Published: (2025)
Stochastic Gradient Sampling for Enhancing Neural Networks Training
by: Yun, Juyoung
Published: (2023)
by: Yun, Juyoung
Published: (2023)
Recursive KL Divergence Optimization: A Dynamic Framework for Representation Learning
by: Martin, Anthony D
Published: (2025)
by: Martin, Anthony D
Published: (2025)
CODE-CL: Conceptor-Based Gradient Projection for Deep Continual Learning
by: Apolinario, Marco Paul E., et al.
Published: (2024)
by: Apolinario, Marco Paul E., et al.
Published: (2024)
Arc Gradient Descent: A Geometrically Motivated Gradient Descent-based Optimiser with Phase-Aware, User-Controlled Step Dynamics (proof-of-concept)
by: Verma, Nikhil, et al.
Published: (2025)
by: Verma, Nikhil, et al.
Published: (2025)
It's Not a Lottery, It's a Race: Understanding How Gradient Descent Adapts the Network's Capacity to the Task
by: Pinson, Hannah
Published: (2026)
by: Pinson, Hannah
Published: (2026)
Sharpness-Aware Surrogate Training for On-Sensor Spiking Neural Networks
by: Nicholson, Maximilian
Published: (2026)
by: Nicholson, Maximilian
Published: (2026)
DUA-D2C: Dynamic Uncertainty Aware Method for Overfitting Remediation in Deep Learning
by: Siddiqui, Md. Saiful Bari, et al.
Published: (2024)
by: Siddiqui, Md. Saiful Bari, et al.
Published: (2024)
Modern Hopfield Networks meet Encoded Neural Representations -- Addressing Practical Considerations
by: Kashyap, Satyananda, et al.
Published: (2024)
by: Kashyap, Satyananda, et al.
Published: (2024)
ExTTNet: A Deep Learning Algorithm for Extracting Table Texts from Invoice Images
by: Akdoğan, Adem, et al.
Published: (2024)
by: Akdoğan, Adem, et al.
Published: (2024)
Decomposer Networks: Deep Component Analysis and Synthesis
by: Joneidi, Mohsen
Published: (2025)
by: Joneidi, Mohsen
Published: (2025)
A Unified Measure-Theoretic View of Diffusion, Score-Based, and Flow Matching Generative Models
by: Ranganath, Aditya, et al.
Published: (2026)
by: Ranganath, Aditya, et al.
Published: (2026)
NeuRN: Neuro-inspired Domain Generalization for Image Classification
by: Jalil, Hamd, et al.
Published: (2025)
by: Jalil, Hamd, et al.
Published: (2025)
Mice to Machines: Neural Representations from Visual Cortex for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
Concept Probing: Where to Find Human-Defined Concepts (Extended Version)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
APTx Neuron: A Unified Trainable Neuron Architecture Integrating Activation and Computation
by: Kumar, Ravin
Published: (2025)
by: Kumar, Ravin
Published: (2025)
Winning the Lottery by Preserving Network Training Dynamics with Concrete Ticket Search
by: Arora, Tanay, et al.
Published: (2025)
by: Arora, Tanay, et al.
Published: (2025)
Learning Evolution via Optimization Knowledge Adaptation
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
When Person Re-Identification Meets Event Camera: A Benchmark Dataset and An Attribute-guided Re-Identification Framework
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Temporal Flexibility in Spiking Neural Networks: Towards Generalization Across Time Steps and Deployment Friendliness
by: Du, Kangrui, et al.
Published: (2025)
by: Du, Kangrui, et al.
Published: (2025)
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
by: Yao, Xingting, et al.
Published: (2025)
by: Yao, Xingting, et al.
Published: (2025)
A Neural Architecture Search Method using Auxiliary Evaluation Metric based on ResNet Architecture
by: Wang, Shang, et al.
Published: (2025)
by: Wang, Shang, et al.
Published: (2025)
Noise-Tolerant Coreset-Based Class Incremental Continual Learning
by: Mucllari, Edison, et al.
Published: (2025)
by: Mucllari, Edison, et al.
Published: (2025)
On the Performance of Concept Probing: The Influence of the Data (Extended Version)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
Forward-Cooperation-Backward (FCB) learning in a Multi-Encoding Uni-Decoding neural network architecture
by: Dutta, Prasun, et al.
Published: (2025)
by: Dutta, Prasun, et al.
Published: (2025)
NeuGen: Amplifying the 'Neural' in Neural Radiance Fields for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
Dynamic Activation with Knowledge Distillation for Energy-Efficient Spiking NN Ensembles
by: Konstantaropoulos, Orestis, et al.
Published: (2025)
by: Konstantaropoulos, Orestis, et al.
Published: (2025)
STAS: Spatio-Temporal Adaptive Computation Time for Spiking Transformers
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
Training Frozen Feature Pyramid DINOv2 for Eyelid Measurements with Infinite Encoding and Orthogonal Regularization
by: Chen, Chun-Hung
Published: (2025)
by: Chen, Chun-Hung
Published: (2025)
Foundry: Distilling 3D Foundation Models for the Edge
by: Letellier, Guillaume, et al.
Published: (2025)
by: Letellier, Guillaume, et al.
Published: (2025)
Conditional Morphogenesis: Emergent Generation of Structural Digits via Neural Cellular Automata
by: Sakour, Ali
Published: (2025)
by: Sakour, Ali
Published: (2025)
Generative Classifiers Avoid Shortcut Solutions
by: Li, Alexander C., et al.
Published: (2025)
by: Li, Alexander C., et al.
Published: (2025)
Optimized Spectral Fault Receptive Fields for Diagnosis-Informed Prognosis
by: Gutiérrez, Stan Muñoz, et al.
Published: (2025)
by: Gutiérrez, Stan Muñoz, et al.
Published: (2025)
Binary Quadratic Quantization: Beyond First-Order Quantization for Real-Valued Matrix Compression
by: Kuroki, Kyo, et al.
Published: (2025)
by: Kuroki, Kyo, et al.
Published: (2025)
Neural Architecture Search with Mixed Bio-inspired Learning Rules
by: Hamzaoui, Imane, et al.
Published: (2025)
by: Hamzaoui, Imane, et al.
Published: (2025)
Threshold Modulation for Online Test-Time Adaptation of Spiking Neural Networks
by: Zhao, Kejie, et al.
Published: (2025)
by: Zhao, Kejie, et al.
Published: (2025)
DRiVE: Dynamic Recognition in VEhicles using snnTorch
by: Vora, Heerak, et al.
Published: (2025)
by: Vora, Heerak, et al.
Published: (2025)
The Platonic Representation Hypothesis
by: Huh, Minyoung, et al.
Published: (2024)
by: Huh, Minyoung, et al.
Published: (2024)
Neuromimetic metaplasticity for adaptive continual learning
by: Cho, Suhee, et al.
Published: (2024)
by: Cho, Suhee, et al.
Published: (2024)
CoNO: Complex Neural Operator for Continous Dynamical Physical Systems
by: Tiwari, Karn, et al.
Published: (2024)
by: Tiwari, Karn, et al.
Published: (2024)
Similar Items
-
Fast Fourier Transform-Based Spectral and Temporal Gradient Filtering for Differential Privacy
by: Shin, Hyeju, et al.
Published: (2025) -
Stochastic Gradient Sampling for Enhancing Neural Networks Training
by: Yun, Juyoung
Published: (2023) -
Recursive KL Divergence Optimization: A Dynamic Framework for Representation Learning
by: Martin, Anthony D
Published: (2025) -
CODE-CL: Conceptor-Based Gradient Projection for Deep Continual Learning
by: Apolinario, Marco Paul E., et al.
Published: (2024) -
Arc Gradient Descent: A Geometrically Motivated Gradient Descent-based Optimiser with Phase-Aware, User-Controlled Step Dynamics (proof-of-concept)
by: Verma, Nikhil, et al.
Published: (2025)