Feather: An Elegant Solution to Effective DNN Sparsification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Georgoulakis, Athanasios Glentis, Retsinas, George, Maragos, Petros |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime
von: Wraight, Petros Georgoulas, et al.
Veröffentlicht: (2025)
von: Wraight, Petros Georgoulas, et al.
Veröffentlicht: (2025)
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
von: Svyezhentsev, Davyd, et al.
Veröffentlicht: (2024)
von: Svyezhentsev, Davyd, et al.
Veröffentlicht: (2024)
Mushroom Segmentation and 3D Pose Estimation from Point Clouds using Fully Convolutional Geometric Features and Implicit Pose Encoding
von: Retsinas, George, et al.
Veröffentlicht: (2024)
von: Retsinas, George, et al.
Veröffentlicht: (2024)
Training Deep Morphological Neural Networks as Universal Approximators
von: Fotopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
von: Fotopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates
von: Glentis, Athanasios, et al.
Veröffentlicht: (2026)
von: Glentis, Athanasios, et al.
Veröffentlicht: (2026)
A Transformer-Based Framework for Greek Sign Language Production using Extended Skeletal Motion Representations
von: Pratikaki, Chrysa, et al.
Veröffentlicht: (2025)
von: Pratikaki, Chrysa, et al.
Veröffentlicht: (2025)
Sparse Hybrid Linear-Morphological Networks
von: Fotopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
von: Fotopoulos, Konstantinos, et al.
Veröffentlicht: (2025)
Proactive tactile exploration for object-agnostic shape reconstruction from minimal visual priors
von: Oikonomou, Paris, et al.
Veröffentlicht: (2025)
von: Oikonomou, Paris, et al.
Veröffentlicht: (2025)
Only-Style: Stylistic Consistency in Image Generation without Content Leakage
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2025)
von: Aravanis, Tilemachos, et al.
Veröffentlicht: (2025)
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
Shared-Weights Extender and Gradient Voting for Neural Network Expansion
von: Chatzis, Nikolas, et al.
Veröffentlicht: (2025)
von: Chatzis, Nikolas, et al.
Veröffentlicht: (2025)
Category-Level 6D Object Pose Estimation in Agricultural Settings Using a Lattice-Deformation Framework and Diffusion-Augmented Synthetic Data
von: Glytsos, Marios, et al.
Veröffentlicht: (2025)
von: Glytsos, Marios, et al.
Veröffentlicht: (2025)
Uncertainty-Driven Anomaly Detection for Psychotic Relapse Using Smartwatches: Forecasting and Multi-Task Learning Fusion
von: Tsalkitzis, Nikolaos, et al.
Veröffentlicht: (2026)
von: Tsalkitzis, Nikolaos, et al.
Veröffentlicht: (2026)
Structure-Aware Spectral Sparsification via Uniform Edge Sampling
von: He, Kaiwen, et al.
Veröffentlicht: (2025)
von: He, Kaiwen, et al.
Veröffentlicht: (2025)
Diffusion-Based Scenario Tree Generation for Multivariate Time Series Prediction and Multistage Stochastic Optimization
von: Zarifis, Stelios, et al.
Veröffentlicht: (2025)
von: Zarifis, Stelios, et al.
Veröffentlicht: (2025)
Diffusion-Based Forecasting for Uncertainty-Aware Model Predictive Control
von: Zarifis, Stelios, et al.
Veröffentlicht: (2025)
von: Zarifis, Stelios, et al.
Veröffentlicht: (2025)
EMA-Nesterov: Stabilizing Nesterov's Lookahead for Accelerated Deep Learning Optimization
von: Yau, Chung-Yiu, et al.
Veröffentlicht: (2026)
von: Yau, Chung-Yiu, et al.
Veröffentlicht: (2026)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
von: Giannakakis, Nikos, et al.
Veröffentlicht: (2025)
von: Giannakakis, Nikos, et al.
Veröffentlicht: (2025)
Scalable Parameter and Memory Efficient Pretraining for LLM: Recent Algorithmic Advances and Benchmarking
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
MiCRO: Near-Zero Cost Gradient Sparsification for Scaling and Accelerating Distributed DNN Training
von: Yoon, Daegun, et al.
Veröffentlicht: (2023)
von: Yoon, Daegun, et al.
Veröffentlicht: (2023)
Registration-Free Learnable Multi-View Capture of Faces in Dense Semantic Correspondence
von: Filntisis, Panagiotis P., et al.
Veröffentlicht: (2026)
von: Filntisis, Panagiotis P., et al.
Veröffentlicht: (2026)
Mask in the Mirror: Implicit Sparsification
von: Jacobs, Tom, et al.
Veröffentlicht: (2024)
von: Jacobs, Tom, et al.
Veröffentlicht: (2024)
Spectral Neural Graph Sparsification
von: Liguori, Angelica, et al.
Veröffentlicht: (2025)
von: Liguori, Angelica, et al.
Veröffentlicht: (2025)
Synthetic Data is an Elegant GIFT for Continual Vision-Language Models
von: Wu, Bin, et al.
Veröffentlicht: (2025)
von: Wu, Bin, et al.
Veröffentlicht: (2025)
Secure Aggregation Meets Sparsification in Decentralized Learning
von: Biswas, Sayan, et al.
Veröffentlicht: (2024)
von: Biswas, Sayan, et al.
Veröffentlicht: (2024)
MSQ: Memory-Efficient Bit Sparsification Quantization
von: Han, Seokho, et al.
Veröffentlicht: (2025)
von: Han, Seokho, et al.
Veröffentlicht: (2025)
Efficient Unbiased Sparsification
von: Barnes, Leighton, et al.
Veröffentlicht: (2024)
von: Barnes, Leighton, et al.
Veröffentlicht: (2024)
Joint Model and Data Sparsification via the Marginal Likelihood
von: Timans, Alexander, et al.
Veröffentlicht: (2026)
von: Timans, Alexander, et al.
Veröffentlicht: (2026)
BitSnap: Checkpoint Sparsification and Quantization in LLM Training
von: Peng, Yanxin, et al.
Veröffentlicht: (2025)
von: Peng, Yanxin, et al.
Veröffentlicht: (2025)
Mobility-Aware Asynchronous Federated Learning with Dynamic Sparsification
von: Yan, Jintao, et al.
Veröffentlicht: (2025)
von: Yan, Jintao, et al.
Veröffentlicht: (2025)
Automatic and Structure-Aware Sparsification of Hybrid Neural ODEs
von: Zou, Bob Junyi, et al.
Veröffentlicht: (2025)
von: Zou, Bob Junyi, et al.
Veröffentlicht: (2025)
Graph Sparsification via Mixture of Graphs
von: Zhang, Guibin, et al.
Veröffentlicht: (2024)
von: Zhang, Guibin, et al.
Veröffentlicht: (2024)
Graph Sparsification for Enhanced Conformal Prediction in Graph Neural Networks
von: He, Yuntian, et al.
Veröffentlicht: (2024)
von: He, Yuntian, et al.
Veröffentlicht: (2024)
Stochastic Matching via Local Sparsification
von: Ahmadian, Sara, et al.
Veröffentlicht: (2026)
von: Ahmadian, Sara, et al.
Veröffentlicht: (2026)
Empirical Error Estimates for Graph Sparsification
von: Wang, Siyao, et al.
Veröffentlicht: (2025)
von: Wang, Siyao, et al.
Veröffentlicht: (2025)
Learning Effective Dynamics across Spatio-Temporal Scales of Complex Flows
von: Gao, Han, et al.
Veröffentlicht: (2025)
von: Gao, Han, et al.
Veröffentlicht: (2025)
Extending Input Contexts of Language Models through Training on Segmented Sequences
von: Karypis, Petros, et al.
Veröffentlicht: (2023)
von: Karypis, Petros, et al.
Veröffentlicht: (2023)
Importance Sparsification for Sinkhorn Algorithm
von: Li, Mengyu, et al.
Veröffentlicht: (2023)
von: Li, Mengyu, et al.
Veröffentlicht: (2023)
Requests of a Feather Must Flock Together: Batch Size vs. Prefix Homogeneity in LLM Inference
von: Rathi, Saksham, et al.
Veröffentlicht: (2026)
von: Rathi, Saksham, et al.
Veröffentlicht: (2026)
Mitigating Over-Squashing in Graph Neural Networks by Spectrum-Preserving Sparsification
von: Liang, Langzhang, et al.
Veröffentlicht: (2025)
von: Liang, Langzhang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime
von: Wraight, Petros Georgoulas, et al.
Veröffentlicht: (2025) -
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
von: Svyezhentsev, Davyd, et al.
Veröffentlicht: (2024) -
Mushroom Segmentation and 3D Pose Estimation from Point Clouds using Fully Convolutional Geometric Features and Implicit Pose Encoding
von: Retsinas, George, et al.
Veröffentlicht: (2024) -
Training Deep Morphological Neural Networks as Universal Approximators
von: Fotopoulos, Konstantinos, et al.
Veröffentlicht: (2025) -
Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates
von: Glentis, Athanasios, et al.
Veröffentlicht: (2026)