SparseDM: Toward Sparse Efficient Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Kafeng, Chen, Jianfei, Li, He, Mi, Zhenpeng, Zhu, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RCR-AF: Enhancing Model Generalization via Rademacher Complexity Reduction Activation Function
von: Yu, Yunrui, et al.
Veröffentlicht: (2025)
von: Yu, Yunrui, et al.
Veröffentlicht: (2025)
Sparsely Supervised Diffusion
von: Zhao, Wenshuai, et al.
Veröffentlicht: (2026)
von: Zhao, Wenshuai, et al.
Veröffentlicht: (2026)
Sparse Autoencoders, Again?
von: Lu, Yin, et al.
Veröffentlicht: (2025)
von: Lu, Yin, et al.
Veröffentlicht: (2025)
SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
Topology-Aware Revival for Efficient Sparse Training
von: Jin, Meiling, et al.
Veröffentlicht: (2026)
von: Jin, Meiling, et al.
Veröffentlicht: (2026)
Enhancing the Resilience of Graph Neural Networks to Topological Perturbations in Sparse Graphs
von: He, Shuqi, et al.
Veröffentlicht: (2024)
von: He, Shuqi, et al.
Veröffentlicht: (2024)
SpargeAttention: Accurate and Training-free Sparse Attention Accelerating Any Model Inference
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
Sparse Training of Discrete Diffusion Models for Graph Generation
von: Qin, Yiming, et al.
Veröffentlicht: (2023)
von: Qin, Yiming, et al.
Veröffentlicht: (2023)
Generative modeling of Sparse Approximate Inverse Preconditioners
von: Li, Mou, et al.
Veröffentlicht: (2024)
von: Li, Mou, et al.
Veröffentlicht: (2024)
Towards Faster Training of Diffusion Models: An Inspiration of A Consistency Phenomenon
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
SLA2: Sparse-Linear Attention with Learnable Routing and QAT
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
EsaCL: Efficient Continual Learning of Sparse Models
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
von: Ren, Weijieying, et al.
Veröffentlicht: (2024)
ReLU$^2$ Wins: Discovering Efficient Activation Functions for Sparse LLMs
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengyan, et al.
Veröffentlicht: (2024)
SLaB: Sparse-Lowrank-Binary Decomposition for Efficient Large Language Models
von: Li, Ziwei, et al.
Veröffentlicht: (2026)
von: Li, Ziwei, et al.
Veröffentlicht: (2026)
SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders
von: Cywiński, Bartosz, et al.
Veröffentlicht: (2025)
von: Cywiński, Bartosz, et al.
Veröffentlicht: (2025)
Sparse-Aware Neural Networks for Nonlinear Functionals: Mitigating the Exponential Dependence on Dimension
von: Li, Jianfei, et al.
Veröffentlicht: (2026)
von: Li, Jianfei, et al.
Veröffentlicht: (2026)
Towards Robust Knowledge Tracing Models via k-Sparse Attention
von: Huang, Shuyan, et al.
Veröffentlicht: (2024)
von: Huang, Shuyan, et al.
Veröffentlicht: (2024)
Towards Interpretable Adversarial Examples via Sparse Adversarial Attack
von: Lin, Fudong, et al.
Veröffentlicht: (2025)
von: Lin, Fudong, et al.
Veröffentlicht: (2025)
FlashOmni: A Unified Sparse Attention Engine for Diffusion Transformers
von: Qiao, Liang, et al.
Veröffentlicht: (2025)
von: Qiao, Liang, et al.
Veröffentlicht: (2025)
Uncertainty-Calibrated Spatiotemporal Field Diffusion with Sparse Supervision
von: Valencia, Kevin, et al.
Veröffentlicht: (2026)
von: Valencia, Kevin, et al.
Veröffentlicht: (2026)
HISA: Efficient Hierarchical Indexing for Fine-Grained Sparse Attention
von: Xu, Yufei, et al.
Veröffentlicht: (2026)
von: Xu, Yufei, et al.
Veröffentlicht: (2026)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
von: Xu, Hongtao, et al.
Veröffentlicht: (2026)
von: Xu, Hongtao, et al.
Veröffentlicht: (2026)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
von: Christopoulou, Fenia, et al.
Veröffentlicht: (2024)
von: Christopoulou, Fenia, et al.
Veröffentlicht: (2024)
Unified Sparse-Matrix Representations for Diverse Neural Architectures
von: Zhu, Yuzhou
Veröffentlicht: (2025)
von: Zhu, Yuzhou
Veröffentlicht: (2025)
Sparse, Efficient and Explainable Data Attribution with DualXDA
von: Yolcu, Galip Ümit, et al.
Veröffentlicht: (2024)
von: Yolcu, Galip Ümit, et al.
Veröffentlicht: (2024)
Backbone-Equated Diffusion OOD via Sparse Internal Snapshots
von: Rouzoumka, Yadang Alexis, et al.
Veröffentlicht: (2026)
von: Rouzoumka, Yadang Alexis, et al.
Veröffentlicht: (2026)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
CurvZO: Adaptive Curvature-Guided Sparse Zeroth-Order Optimization for Efficient LLM Fine-Tuning
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
von: Wang, Shuo, et al.
Veröffentlicht: (2026)
MonoSparse-CAM: Efficient Tree Model Processing via Monotonicity and Sparsity in CAMs
von: Molom-Ochir, Tergel, et al.
Veröffentlicht: (2024)
von: Molom-Ochir, Tergel, et al.
Veröffentlicht: (2024)
Are Sparse Autoencoders Useful? A Case Study in Sparse Probing
von: Kantamneni, Subhash, et al.
Veröffentlicht: (2025)
von: Kantamneni, Subhash, et al.
Veröffentlicht: (2025)
MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
von: MiniCPM Team, et al.
Veröffentlicht: (2026)
von: MiniCPM Team, et al.
Veröffentlicht: (2026)
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
Sparse Inducing Points in Deep Gaussian Processes: Enhancing Modeling with Denoising Diffusion Variational Inference
von: Xu, Jian, et al.
Veröffentlicht: (2024)
von: Xu, Jian, et al.
Veröffentlicht: (2024)
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
von: Chen, Lei, et al.
Veröffentlicht: (2026)
von: Chen, Lei, et al.
Veröffentlicht: (2026)
Meta Additive Model: Interpretable Sparse Learning With Auto Weighting
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
Displacement-Sparse Neural Optimal Transport
von: Chen, Peter, et al.
Veröffentlicht: (2025)
von: Chen, Peter, et al.
Veröffentlicht: (2025)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
von: Ayonrinde, Kola
Veröffentlicht: (2024)
von: Ayonrinde, Kola
Veröffentlicht: (2024)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
von: Hartman, Max, et al.
Veröffentlicht: (2025)
von: Hartman, Max, et al.
Veröffentlicht: (2025)
Towards Interpretable Protein Structure Prediction with Sparse Autoencoders
von: Parsan, Nithin, et al.
Veröffentlicht: (2025)
von: Parsan, Nithin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RCR-AF: Enhancing Model Generalization via Rademacher Complexity Reduction Activation Function
von: Yu, Yunrui, et al.
Veröffentlicht: (2025) -
Sparsely Supervised Diffusion
von: Zhao, Wenshuai, et al.
Veröffentlicht: (2026) -
Sparse Autoencoders, Again?
von: Lu, Yin, et al.
Veröffentlicht: (2025) -
SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention
von: Zhang, Jintao, et al.
Veröffentlicht: (2025) -
Topology-Aware Revival for Efficient Sparse Training
von: Jin, Meiling, et al.
Veröffentlicht: (2026)