SEAP: Training-free Sparse Expert Activation Pruning Unlock the Brainpower of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Xun, Wang, Hanyu, Lai, Huayi, Niu, Simin, Song, Shichao, Yang, Jiawei, Zhao, Jihao, Xiong, Feiyu, Tang, Bo, Li, Zhiyu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Empowering Large Language Models to Set up a Knowledge Retrieval Indexer via Self-Learning
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents
by: Lai, Huayi, et al.
Published: (2026)
by: Lai, Huayi, et al.
Published: (2026)
MoM: Mixtures of Scenario-Aware Document Memories for Retrieval-Augmented Generation Systems
by: Zhao, Jihao, et al.
Published: (2025)
by: Zhao, Jihao, et al.
Published: (2025)
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model
by: Liang, Xun, et al.
Published: (2025)
by: Liang, Xun, et al.
Published: (2025)
MoC: Mixtures of Text Chunking Learners for Retrieval-Augmented Generation System
by: Zhao, Jihao, et al.
Published: (2025)
by: Zhao, Jihao, et al.
Published: (2025)
HRDE: Retrieval-Augmented Large Language Models for Chinese Health Rumor Detection and Explainability
by: Chen, Yanfang, et al.
Published: (2024)
by: Chen, Yanfang, et al.
Published: (2024)
Controlled Text Generation for Large Language Model with Dynamic Attribute Graphs
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
Controllable Text Generation for Large Language Models: A Survey
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
SurveyX: Academic Survey Automation via Large Language Models
by: Liang, Xun, et al.
Published: (2025)
by: Liang, Xun, et al.
Published: (2025)
Dissecting Role Cognition in Medical LLMs via Neuronal Ablation
by: Liang, Xun, et al.
Published: (2025)
by: Liang, Xun, et al.
Published: (2025)
Meta-Chunking: Learning Text Segmentation and Semantic Completion via Logical Perception
by: Zhao, Jihao, et al.
Published: (2024)
by: Zhao, Jihao, et al.
Published: (2024)
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation
by: Liang, Xun, et al.
Published: (2023)
by: Liang, Xun, et al.
Published: (2023)
GuessArena: Guess Who I Am? A Self-Adaptive Framework for Evaluating LLMs in Domain-Specific Knowledge and Reasoning
by: Yu, Qingchen, et al.
Published: (2025)
by: Yu, Qingchen, et al.
Published: (2025)
MemOS: An Operating System for Memory-Augmented Generation (MAG) in Large Language Models
by: Li, Zhiyu, et al.
Published: (2025)
by: Li, Zhiyu, et al.
Published: (2025)
Grimoire is All You Need for Enhancing Large Language Models
by: Chen, Ding, et al.
Published: (2024)
by: Chen, Ding, et al.
Published: (2024)
xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation
by: Yu, Qingchen, et al.
Published: (2024)
by: Yu, Qingchen, et al.
Published: (2024)
Using Brainpower in the Classroom
by: Garnett, Steve
Published: (2025)
by: Garnett, Steve
Published: (2025)
TurtleBench: Evaluating Top Language Models via Real-World Yes/No Puzzles
by: Yu, Qingchen, et al.
Published: (2024)
by: Yu, Qingchen, et al.
Published: (2024)
MemFactory: Unified Inference & Training Framework for Agent Memory
by: Guo, Ziliang, et al.
Published: (2026)
by: Guo, Ziliang, et al.
Published: (2026)
Internal Consistency and Self-Feedback in Large Language Models: A Survey
by: Liang, Xun, et al.
Published: (2024)
by: Liang, Xun, et al.
Published: (2024)
Attention Heads of Large Language Models: A Survey
by: Zheng, Zifan, et al.
Published: (2024)
by: Zheng, Zifan, et al.
Published: (2024)
HaluMem: Evaluating Hallucinations in Memory Systems of Agents
by: Chen, Ding, et al.
Published: (2025)
by: Chen, Ding, et al.
Published: (2025)
MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents
by: Chen, Yining, et al.
Published: (2026)
by: Chen, Yining, et al.
Published: (2026)
Inside Out: Evolving User-Centric Core Memory Trees for Long-Term Personalized Dialogue Systems
by: Zhao, Jihao, et al.
Published: (2026)
by: Zhao, Jihao, et al.
Published: (2026)
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models
by: Lyu, Yuanjie, et al.
Published: (2024)
by: Lyu, Yuanjie, et al.
Published: (2024)
QAEncoder: Towards Aligned Representation Learning in Question Answering Systems
by: Wang, Zhengren, et al.
Published: (2024)
by: Wang, Zhengren, et al.
Published: (2024)
Fake Artificial Intelligence Generated Contents (FAIGC): A Survey of Theories, Detection Methods, and Opportunities
by: Yu, Xiaomin, et al.
Published: (2024)
by: Yu, Xiaomin, et al.
Published: (2024)
EvoESAP: Non-Uniform Expert Pruning for Sparse MoE
by: Liu, Zongfang, et al.
Published: (2026)
by: Liu, Zongfang, et al.
Published: (2026)
LNPT: Label-free Network Pruning and Training
by: Xiao, Jinying, et al.
Published: (2024)
by: Xiao, Jinying, et al.
Published: (2024)
MemReader: From Passive to Active Extraction for Long-Term Agent Memory
by: Kang, Jingyi, et al.
Published: (2026)
by: Kang, Jingyi, et al.
Published: (2026)
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
by: Liu, Hongyi, et al.
Published: (2026)
by: Liu, Hongyi, et al.
Published: (2026)
Mosaic Pruning: A Hierarchical Framework for Generalizable Pruning of Mixture-of-Experts Models
by: Hu, Wentao, et al.
Published: (2025)
by: Hu, Wentao, et al.
Published: (2025)
Diversifying the Expert Knowledge for Task-Agnostic Pruning in Sparse Mixture-of-Experts
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
TAdaRAG: Task Adaptive Retrieval-Augmented Generation via On-the-Fly Knowledge Graph Construction
by: Zhang, Jie, et al.
Published: (2025)
by: Zhang, Jie, et al.
Published: (2025)
GRAPHMOE: Amplifying Cognitive Depth of Mixture-of-Experts Network via Introducing Self-Rethinking Mechanism
by: Lv, Bo, et al.
Published: (2025)
by: Lv, Bo, et al.
Published: (2025)
Brawn and Brainpower: Acute Resistance Exercise Improves Behavioral and Neuroelectric Measures of Executive Function
by: Nicholas W. Baumgartner, et al.
Published: (2025)
by: Nicholas W. Baumgartner, et al.
Published: (2025)
DuoGPT: Training-free Dual Sparsity through Activation-aware Pruning in LLMs
by: Yin, Ruokai, et al.
Published: (2025)
by: Yin, Ruokai, et al.
Published: (2025)
An Efficient Hybrid Sparse Attention with CPU-GPU Parallelism for Long-Context Inference
by: Yao, Feiyu, et al.
Published: (2026)
by: Yao, Feiyu, et al.
Published: (2026)
SparseX: Efficient Segment-Level KV Cache Sharing for Interleaved LLM Serving
by: Zhang, Quqing, et al.
Published: (2026)
by: Zhang, Quqing, et al.
Published: (2026)
Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs
by: Zhou, Yixiao, et al.
Published: (2025)
by: Zhou, Yixiao, et al.
Published: (2025)
Similar Items
-
Empowering Large Language Models to Set up a Knowledge Retrieval Indexer via Self-Learning
by: Liang, Xun, et al.
Published: (2024) -
RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents
by: Lai, Huayi, et al.
Published: (2026) -
MoM: Mixtures of Scenario-Aware Document Memories for Retrieval-Augmented Generation Systems
by: Zhao, Jihao, et al.
Published: (2025) -
SafeRAG: Benchmarking Security in Retrieval-Augmented Generation of Large Language Model
by: Liang, Xun, et al.
Published: (2025) -
MoC: Mixtures of Text Chunking Learners for Retrieval-Augmented Generation System
by: Zhao, Jihao, et al.
Published: (2025)