Sparse Brains are Also Adaptive Brains: Cognitive-Load-Aware Dynamic Activation for LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yiheng, Wang, Yujie, Ma, Chi, Yu, Lei, Chersoni, Emmanuele, Huang, Chu-Ren |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
von: Wang, Yu, et al.
Veröffentlicht: (2026)
von: Wang, Yu, et al.
Veröffentlicht: (2026)
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
von: Prange, Jakob, et al.
Veröffentlicht: (2023)
von: Prange, Jakob, et al.
Veröffentlicht: (2023)
From BERT to LLMs: Comparing and Understanding Chinese Classifier Prediction in Language Models
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025)
Learning to Look at the Other Side: A Semantic Probing Study of Word Embeddings in LLMs with Enabled Bidirectional Attention
von: Feng, Zhaoxin, et al.
Veröffentlicht: (2025)
von: Feng, Zhaoxin, et al.
Veröffentlicht: (2025)
Is Length Really A Liability? An Evaluation of Multi-turn LLM Conversations using BoolQ
von: Neergaard, Karl, et al.
Veröffentlicht: (2026)
von: Neergaard, Karl, et al.
Veröffentlicht: (2026)
Composing or Not Composing? Towards Distributional Construction Grammars
von: Blache, Philippe, et al.
Veröffentlicht: (2024)
von: Blache, Philippe, et al.
Veröffentlicht: (2024)
First Activations Matter: Training-Free Methods for Dynamic Activation in Large Language Models
von: Ma, Chi, et al.
Veröffentlicht: (2024)
von: Ma, Chi, et al.
Veröffentlicht: (2024)
Good Arguments Against the People Pleasers: How Reasoning Mitigates (Yet Masks) LLM Sycophancy
von: Feng, Zhaoxin, et al.
Veröffentlicht: (2026)
von: Feng, Zhaoxin, et al.
Veröffentlicht: (2026)
StockGenChaR: A Study on the Evaluation of Large Vision-Language Models on Stock Chart Captioning
von: Qiu, Le, et al.
Veröffentlicht: (2024)
von: Qiu, Le, et al.
Veröffentlicht: (2024)
Log Probabilities Are a Reliable Estimate of Semantic Plausibility in Base and Instruction-Tuned Language Models
von: Kauf, Carina, et al.
Veröffentlicht: (2024)
von: Kauf, Carina, et al.
Veröffentlicht: (2024)
LLMs are Also Effective Embedding Models: An In-depth Overview
von: Tao, Chongyang, et al.
Veröffentlicht: (2024)
von: Tao, Chongyang, et al.
Veröffentlicht: (2024)
Semantics-Adaptive Activation Intervention for LLMs via Dynamic Steering Vectors
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
Instruction-tuning Aligns LLMs to the Human Brain
von: Aw, Khai Loong, et al.
Veröffentlicht: (2023)
von: Aw, Khai Loong, et al.
Veröffentlicht: (2023)
Linguistics and Human Brain: A Perspective of Computational Neuroscience
von: Zhang, Fudong, et al.
Veröffentlicht: (2026)
von: Zhang, Fudong, et al.
Veröffentlicht: (2026)
Optimal Brain Restoration for Joint Quantization and Sparsification of LLMs
von: Guo, Hang, et al.
Veröffentlicht: (2025)
von: Guo, Hang, et al.
Veröffentlicht: (2025)
Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility
von: Wang, Sheng-Fu, et al.
Veröffentlicht: (2025)
von: Wang, Sheng-Fu, et al.
Veröffentlicht: (2025)
Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step
von: Li, Liunian Harold, et al.
Veröffentlicht: (2023)
von: Li, Liunian Harold, et al.
Veröffentlicht: (2023)
From Human Cognition to Neural Activations: Probing the Computational Primitives of Spatial Reasoning in LLMs
von: An, Jiyuan, et al.
Veröffentlicht: (2026)
von: An, Jiyuan, et al.
Veröffentlicht: (2026)
LLMs Can Also Do Well! Breaking Barriers in Semantic Role Labeling via Large Language Models
von: Li, Xinxin, et al.
Veröffentlicht: (2025)
von: Li, Xinxin, et al.
Veröffentlicht: (2025)
Contextualized Automatic Speech Recognition with Dynamic Vocabulary Prediction and Activation
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
von: Lin, Zhennan, et al.
Veröffentlicht: (2025)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
Brain in a Vat: On Missing Pieces Towards Artificial General Intelligence in Large Language Models
von: Ma, Yuxi, et al.
Veröffentlicht: (2023)
von: Ma, Yuxi, et al.
Veröffentlicht: (2023)
Do Large Language Models Think Like the Brain? Sentence-Level Evidences from Layer-Wise Embeddings and fMRI
von: Lei, Yu, et al.
Veröffentlicht: (2025)
von: Lei, Yu, et al.
Veröffentlicht: (2025)
AI Meets Brain: Memory Systems from Cognitive Neuroscience to Autonomous Agents
von: Liang, Jiafeng, et al.
Veröffentlicht: (2025)
von: Liang, Jiafeng, et al.
Veröffentlicht: (2025)
Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
von: Chi, Jinwei, et al.
Veröffentlicht: (2025)
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
von: Zhou, Xiabin, et al.
Veröffentlicht: (2024)
Q-Sparse: All Large Language Models can be Fully Sparsely-Activated
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
SASFT: Sparse Autoencoder-guided Supervised Finetuning to Mitigate Unexpected Code-Switching in LLMs
von: Deng, Boyi, et al.
Veröffentlicht: (2025)
von: Deng, Boyi, et al.
Veröffentlicht: (2025)
Time-Aware Feature Selection: Adaptive Temporal Masking for Stable Sparse Autoencoder Training
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
BrainLLM: Generative Language Decoding from Brain Recordings
von: Ye, Ziyi, et al.
Veröffentlicht: (2023)
von: Ye, Ziyi, et al.
Veröffentlicht: (2023)
Enhancing Multiple Dimensions of Trustworthiness in LLMs via Sparse Activation Control
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2024)
LLMs Can Get "Brain Rot": A Pilot Study on Twitter/X
von: Xing, Shuo, et al.
Veröffentlicht: (2025)
von: Xing, Shuo, et al.
Veröffentlicht: (2025)
Sparse-dLLM: Accelerating Diffusion LLMs with Dynamic Cache Eviction
von: Song, Yuerong, et al.
Veröffentlicht: (2025)
von: Song, Yuerong, et al.
Veröffentlicht: (2025)
Weaker LLMs' Opinions Also Matter: Mixture of Opinions Enhances LLM's Mathematical Reasoning
von: Chen, Yanan, et al.
Veröffentlicht: (2025)
von: Chen, Yanan, et al.
Veröffentlicht: (2025)
Achieving Sparse Activation in Small Language Models
von: Song, Jifeng, et al.
Veröffentlicht: (2024)
von: Song, Jifeng, et al.
Veröffentlicht: (2024)
Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech Models
von: Moussa, Omer, et al.
Veröffentlicht: (2025)
von: Moussa, Omer, et al.
Veröffentlicht: (2025)
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
von: Ye, Chunyu, et al.
Veröffentlicht: (2025)
von: Ye, Chunyu, et al.
Veröffentlicht: (2025)
One Brain, Omni Modalities: Towards Unified Non-Invasive Brain Decoding with Large Language Models
von: Tang, Changli, et al.
Veröffentlicht: (2026)
von: Tang, Changli, et al.
Veröffentlicht: (2026)
Self-Adaptive Cognitive Debiasing for Large Language Models in Decision-Making
von: Lyu, Yougang, et al.
Veröffentlicht: (2025)
von: Lyu, Yougang, et al.
Veröffentlicht: (2025)
BrainWavLM: Fine-tuning Speech Representations with Brain Responses to Language
von: Vattikonda, Nishitha, et al.
Veröffentlicht: (2025)
von: Vattikonda, Nishitha, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
von: Wang, Yu, et al.
Veröffentlicht: (2026) -
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
von: Prange, Jakob, et al.
Veröffentlicht: (2023) -
From BERT to LLMs: Comparing and Understanding Chinese Classifier Prediction in Language Models
von: Zhang, Ziqi, et al.
Veröffentlicht: (2025) -
Learning to Look at the Other Side: A Semantic Probing Study of Word Embeddings in LLMs with Enabled Bidirectional Attention
von: Feng, Zhaoxin, et al.
Veröffentlicht: (2025) -
Is Length Really A Liability? An Evaluation of Multi-turn LLM Conversations using BoolQ
von: Neergaard, Karl, et al.
Veröffentlicht: (2026)