Resting Neurons, Active Insights: Robustifying Activation Sparsity in LLMs via Spontaneity
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Haotian, Yang, Jiannan, Gao, Tian, Weng, Tsui-Wei, Ma, Tengfei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph Concept Bottleneck Models
by: Xu, Haotian, et al.
Published: (2025)
by: Xu, Haotian, et al.
Published: (2025)
Linear Explanations for Individual Neurons
by: Oikarinen, Tuomas, et al.
Published: (2024)
by: Oikarinen, Tuomas, et al.
Published: (2024)
Evaluating Neuron Explanations: A Unified Framework with Sanity Checks
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
Faithful and Stable Neuron Explanations for Trustworthy Mechanistic Interpretability
by: Yan, Ge, et al.
Published: (2025)
by: Yan, Ge, et al.
Published: (2025)
When Structure Doesn't Help: LLMs Do Not Read Text-Attributed Graphs as Effectively as We Expected
by: Xu, Haotian, et al.
Published: (2025)
by: Xu, Haotian, et al.
Published: (2025)
Joint Training Across Multiple Activation Sparsity Regimes
by: Wang, Haotian
Published: (2026)
by: Wang, Haotian
Published: (2026)
Self-Supervised Learning on Molecular Graphs: A Systematic Investigation of Masking Design
by: Yang, Jiannan, et al.
Published: (2025)
by: Yang, Jiannan, et al.
Published: (2025)
Abstracted Shapes as Tokens -- A Generalizable and Interpretable Model for Time-series Classification
by: Wen, Yunshi, et al.
Published: (2024)
by: Wen, Yunshi, et al.
Published: (2024)
Interpreting Neurons in Deep Vision Networks with Language Models
by: Bai, Nicholas, et al.
Published: (2024)
by: Bai, Nicholas, et al.
Published: (2024)
Efficient Imputation for Patch-based Missing Single-cell Data via Cluster-regularized Optimal Transport
by: Liu, Yuyu, et al.
Published: (2026)
by: Liu, Yuyu, et al.
Published: (2026)
Beyond Top Activations: Efficient and Reliable Crowdsourced Evaluation of Automated Interpretability
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
by: Madhyastha, Meghana, et al.
Published: (2026)
by: Madhyastha, Meghana, et al.
Published: (2026)
Breaking the Barrier: Enhanced Utility and Robustness in Smoothed DRL Agents
by: Sun, Chung-En, et al.
Published: (2024)
by: Sun, Chung-En, et al.
Published: (2024)
Neuron Activation Coverage: Rethinking Out-of-distribution Detection and Generalization
by: Liu, Yibing, et al.
Published: (2023)
by: Liu, Yibing, et al.
Published: (2023)
Towards the Connection between Activation Sparsity and Flat Minima
by: Peng, Ze, et al.
Published: (2026)
by: Peng, Ze, et al.
Published: (2026)
Distillation Robustifies Unlearning
by: Lee, Bruce W., et al.
Published: (2025)
by: Lee, Bruce W., et al.
Published: (2025)
Wasserstein Distances, Neuronal Entanglement, and Sparsity
by: Sawmya, Shashata, et al.
Published: (2024)
by: Sawmya, Shashata, et al.
Published: (2024)
ProTransformer: Robustify Transformers via Plug-and-Play Paradigm
by: Hou, Zhichao, et al.
Published: (2024)
by: Hou, Zhichao, et al.
Published: (2024)
Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities
by: Sun, Chung-En, et al.
Published: (2024)
by: Sun, Chung-En, et al.
Published: (2024)
RAT: Boosting Misclassification Detection Ability without Extra Data
by: Yan, Ge, et al.
Published: (2025)
by: Yan, Ge, et al.
Published: (2025)
Interpretability-Guided Test-Time Adversarial Defense
by: Kulkarni, Akshay, et al.
Published: (2024)
by: Kulkarni, Akshay, et al.
Published: (2024)
DuoGPT: Training-free Dual Sparsity through Activation-aware Pruning in LLMs
by: Yin, Ruokai, et al.
Published: (2025)
by: Yin, Ruokai, et al.
Published: (2025)
Understanding Fixed Predictions via Confined Regions
by: Lawless, Connor, et al.
Published: (2025)
by: Lawless, Connor, et al.
Published: (2025)
R-Sparse: Rank-Aware Activation Sparsity for Efficient LLM Inference
by: Zhang, Zhenyu, et al.
Published: (2025)
by: Zhang, Zhenyu, et al.
Published: (2025)
Týr-the-Pruner: Structural Pruning LLMs via Global Sparsity Distribution Optimization
by: Li, Guanchen, et al.
Published: (2025)
by: Li, Guanchen, et al.
Published: (2025)
Explore Activation Sparsity in Recurrent LLMs for Energy-Efficient Neuromorphic Computing
by: Knunyants, Ivan, et al.
Published: (2025)
by: Knunyants, Ivan, et al.
Published: (2025)
ThinkEdit: Interpretable Weight Editing to Mitigate Overly Short Thinking in Reasoning Models
by: Sun, Chung-En, et al.
Published: (2025)
by: Sun, Chung-En, et al.
Published: (2025)
Effective Skill Unlearning through Intervention and Abstention
by: Li, Yongce, et al.
Published: (2025)
by: Li, Yongce, et al.
Published: (2025)
Crafting Large Language Models for Enhanced Interpretability
by: Sun, Chung-En, et al.
Published: (2024)
by: Sun, Chung-En, et al.
Published: (2024)
Reinforcement Learning With Sparse-Executing Actions via Sparsity Regularization
by: Pang, Jing-Cheng, et al.
Published: (2021)
by: Pang, Jing-Cheng, et al.
Published: (2021)
Robustifying Conditional Portfolio Decisions via Optimal Transport
by: Nguyen, Viet Anh, et al.
Published: (2021)
by: Nguyen, Viet Anh, et al.
Published: (2021)
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
by: Srivastava, Divyansh, et al.
Published: (2024)
by: Srivastava, Divyansh, et al.
Published: (2024)
Task-Specific Data Selection for Instruction Tuning via Monosemantic Neuronal Activations
by: Ma, Da, et al.
Published: (2025)
by: Ma, Da, et al.
Published: (2025)
Robustifying and Boosting Training-Free Neural Architecture Search
by: He, Zhenfeng, et al.
Published: (2024)
by: He, Zhenfeng, et al.
Published: (2024)
Deep Neural Network Initialization with Sparsity Inducing Activations
by: Price, Ilan, et al.
Published: (2024)
by: Price, Ilan, et al.
Published: (2024)
Prediction without Preclusion: Recourse Verification with Reachable Sets
by: Kothari, Avni, et al.
Published: (2023)
by: Kothari, Avni, et al.
Published: (2023)
Provably Robust Conformal Prediction with Improved Efficiency
by: Yan, Ge, et al.
Published: (2024)
by: Yan, Ge, et al.
Published: (2024)
Robustifying Point Cloud Networks by Refocusing
by: Levi, Meir Yossef, et al.
Published: (2023)
by: Levi, Meir Yossef, et al.
Published: (2023)
Zeroth-Order Fine-Tuning of LLMs with Extreme Sparsity
by: Guo, Wentao, et al.
Published: (2024)
by: Guo, Wentao, et al.
Published: (2024)
SAU: Sparsity-Aware Unlearning for LLMs via Gradient Masking and Importance Redistribution
by: Wang, Yuze, et al.
Published: (2026)
by: Wang, Yuze, et al.
Published: (2026)
Similar Items
-
Graph Concept Bottleneck Models
by: Xu, Haotian, et al.
Published: (2025) -
Linear Explanations for Individual Neurons
by: Oikarinen, Tuomas, et al.
Published: (2024) -
Evaluating Neuron Explanations: A Unified Framework with Sanity Checks
by: Oikarinen, Tuomas, et al.
Published: (2025) -
Faithful and Stable Neuron Explanations for Trustworthy Mechanistic Interpretability
by: Yan, Ge, et al.
Published: (2025) -
When Structure Doesn't Help: LLMs Do Not Read Text-Attributed Graphs as Effectively as We Expected
by: Xu, Haotian, et al.
Published: (2025)