Topology-Aware Revival for Efficient Sparse Training
Fuente:
arXiv
Guardado en:
| Autores principales: | Jin, Meiling, Wang, Fei, Yuan, Xiaoyun, Qian, Chen, Cheng, Yuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
por: Chen, Lei, et al.
Publicado: (2026)
por: Chen, Lei, et al.
Publicado: (2026)
Topology-Aware and Highly Generalizable Deep Reinforcement Learning for Efficient Retrieval in Multi-Deep Storage Systems
por: Li, Funing, et al.
Publicado: (2025)
por: Li, Funing, et al.
Publicado: (2025)
BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding
por: He, Liang, et al.
Publicado: (2026)
por: He, Liang, et al.
Publicado: (2026)
Towards Heterogeneity-Aware and Energy-Efficient Topology Optimization for Decentralized Federated Learning in Edge Environment
por: Liu, Yuze, et al.
Publicado: (2025)
por: Liu, Yuze, et al.
Publicado: (2025)
TED: Training-Free Experience Distillation for Multimodal Reasoning
por: Yuan, Shuozhi, et al.
Publicado: (2026)
por: Yuan, Shuozhi, et al.
Publicado: (2026)
SparseDM: Toward Sparse Efficient Diffusion Models
por: Wang, Kafeng, et al.
Publicado: (2024)
por: Wang, Kafeng, et al.
Publicado: (2024)
HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces
por: Ullah, Nasib, et al.
Publicado: (2026)
por: Ullah, Nasib, et al.
Publicado: (2026)
On Effectiveness and Efficiency of Agentic Tool-calling and RL Training
por: Liu, Tong, et al.
Publicado: (2026)
por: Liu, Tong, et al.
Publicado: (2026)
Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training
por: Lin, Jinhong, et al.
Publicado: (2026)
por: Lin, Jinhong, et al.
Publicado: (2026)
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
por: Tan, Qitao, et al.
Publicado: (2025)
por: Tan, Qitao, et al.
Publicado: (2025)
DistrAttention: An Efficient and Flexible Self-Attention Mechanism on Modern GPUs
por: Jin, Haolin, et al.
Publicado: (2025)
por: Jin, Haolin, et al.
Publicado: (2025)
Efficient Network Automatic Relevance Determination
por: Zhang, Hongwei, et al.
Publicado: (2025)
por: Zhang, Hongwei, et al.
Publicado: (2025)
EcoSpa: Efficient Transformer Training with Coupled Sparsity
por: Xiao, Jinqi, et al.
Publicado: (2025)
por: Xiao, Jinqi, et al.
Publicado: (2025)
HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference
por: Gong, Ping, et al.
Publicado: (2025)
por: Gong, Ping, et al.
Publicado: (2025)
EfficientQAT: Efficient Quantization-Aware Training for Large Language Models
por: Chen, Mengzhao, et al.
Publicado: (2024)
por: Chen, Mengzhao, et al.
Publicado: (2024)
MOSS: Efficient and Accurate FP8 LLM Training with Microscaling and Automatic Scaling
por: Zhang, Yu, et al.
Publicado: (2025)
por: Zhang, Yu, et al.
Publicado: (2025)
Magnitude-Modulated Equivariant Adapter for Parameter-Efficient Fine-Tuning of Equivariant Graph Neural Networks
por: Jin, Dian, et al.
Publicado: (2025)
por: Jin, Dian, et al.
Publicado: (2025)
SlimPipe: Memory-Thrifty and Efficient Pipeline Parallelism for Long-Context LLM Training
por: Li, Zhouyang, et al.
Publicado: (2025)
por: Li, Zhouyang, et al.
Publicado: (2025)
Efficient Equivariant High-Order Crystal Tensor Prediction via Cartesian Local-Environment Many-Body Coupling
por: Jin, Dian, et al.
Publicado: (2026)
por: Jin, Dian, et al.
Publicado: (2026)
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
por: Zhang, Tuo, et al.
Publicado: (2025)
por: Zhang, Tuo, et al.
Publicado: (2025)
HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime
por: Sana, Mohamed, et al.
Publicado: (2026)
por: Sana, Mohamed, et al.
Publicado: (2026)
Gradient-Congruity Guided Federated Sparse Training
por: Tian, Chris Xing, et al.
Publicado: (2024)
por: Tian, Chris Xing, et al.
Publicado: (2024)
Towards Interpretable Adversarial Examples via Sparse Adversarial Attack
por: Lin, Fudong, et al.
Publicado: (2025)
por: Lin, Fudong, et al.
Publicado: (2025)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
por: Xu, Hongtao, et al.
Publicado: (2026)
por: Xu, Hongtao, et al.
Publicado: (2026)
Data-Efficient Training by Evolved Sampling
por: Cheng, Ziheng, et al.
Publicado: (2025)
por: Cheng, Ziheng, et al.
Publicado: (2025)
Optimal Corpus Aware Training for Neural Machine Translation
por: Liao, Yi-Hsiu, et al.
Publicado: (2025)
por: Liao, Yi-Hsiu, et al.
Publicado: (2025)
Enhancing the Resilience of Graph Neural Networks to Topological Perturbations in Sparse Graphs
por: He, Shuqi, et al.
Publicado: (2024)
por: He, Shuqi, et al.
Publicado: (2024)
Advancing On-Device Neural Network Training with TinyPropv2: Dynamic, Sparse, and Efficient Backpropagation
por: Rüb, Marcus, et al.
Publicado: (2024)
por: Rüb, Marcus, et al.
Publicado: (2024)
ssProp: Energy-Efficient Training for Convolutional Neural Networks with Scheduled Sparse Back Propagation
por: Zhong, Lujia, et al.
Publicado: (2024)
por: Zhong, Lujia, et al.
Publicado: (2024)
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
por: Yuan, Mingqi, et al.
Publicado: (2025)
por: Yuan, Mingqi, et al.
Publicado: (2025)
Sparse-VQ Transformer: An FFN-Free Framework with Vector Quantization for Enhanced Time Series Forecasting
por: Zhao, Yanjun, et al.
Publicado: (2024)
por: Zhao, Yanjun, et al.
Publicado: (2024)
FAST: Topology-Aware Frequency-Domain Distribution Matching for Coreset Selection
por: Cui, Jin, et al.
Publicado: (2025)
por: Cui, Jin, et al.
Publicado: (2025)
Efficient GNN Training Through Structure-Aware Randomized Mini-Batching
por: Balaji, Vignesh, et al.
Publicado: (2025)
por: Balaji, Vignesh, et al.
Publicado: (2025)
Conflict-Aware Adversarial Training
por: Xue, Zhiyu, et al.
Publicado: (2024)
por: Xue, Zhiyu, et al.
Publicado: (2024)
Fast Adversarial Training against Sparse Attacks Requires Loss Smoothing
por: Zhong, Xuyang, et al.
Publicado: (2025)
por: Zhong, Xuyang, et al.
Publicado: (2025)
BigMac: A Communication-Efficient Mixture-of-Experts Model Structure for Fast Training and Inference
por: Jin, Zewen, et al.
Publicado: (2025)
por: Jin, Zewen, et al.
Publicado: (2025)
An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference
por: Yao, Feiyu, et al.
Publicado: (2026)
por: Yao, Feiyu, et al.
Publicado: (2026)
SynergyKGC: Reconciling Topological Heterogeneity in Knowledge Graph Completion via Topology-Aware Synergy
por: Zou, Xuecheng, et al.
Publicado: (2026)
por: Zou, Xuecheng, et al.
Publicado: (2026)
Reasoning-Aware Training for Time Series Forecasting
por: Ahamed, Md Atik, et al.
Publicado: (2026)
por: Ahamed, Md Atik, et al.
Publicado: (2026)
Online Prompt Pricing based on Combinatorial Multi-Armed Bandit and Hierarchical Stackelberg Game
por: Li, Meiling, et al.
Publicado: (2024)
por: Li, Meiling, et al.
Publicado: (2024)
Ejemplares similares
-
WiSparse: Boosting LLM Inference Efficiency with Weight-Aware Mixed Activation Sparsity
por: Chen, Lei, et al.
Publicado: (2026) -
Topology-Aware and Highly Generalizable Deep Reinforcement Learning for Efficient Retrieval in Multi-Deep Storage Systems
por: Li, Funing, et al.
Publicado: (2025) -
BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding
por: He, Liang, et al.
Publicado: (2026) -
Towards Heterogeneity-Aware and Energy-Efficient Topology Optimization for Decentralized Federated Learning in Edge Environment
por: Liu, Yuze, et al.
Publicado: (2025) -
TED: Training-Free Experience Distillation for Multimodal Reasoning
por: Yuan, Shuozhi, et al.
Publicado: (2026)