Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Junyan, Gao, Yubo, Yan, Yibo, Li, Jungang, Hou, Zhaorui, Tao, Sicheng, Liu, Shuliang, Dai, Song, Hei, Yonghua, Li, Junzhuo, Hu, Xuming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unlocking Speech Instruction Data Potential with Query Rewriting
von: Hei, Yonghua, et al.
Veröffentlicht: (2025)
von: Hei, Yonghua, et al.
Veröffentlicht: (2025)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
von: Dai, Song, et al.
Veröffentlicht: (2025)
von: Dai, Song, et al.
Veröffentlicht: (2025)
CoheMark: A Novel Sentence-Level Watermark for Enhanced Text Quality
von: Zhang, Junyan, et al.
Veröffentlicht: (2025)
von: Zhang, Junyan, et al.
Veröffentlicht: (2025)
Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?
von: Zhang, Junyan, et al.
Veröffentlicht: (2025)
von: Zhang, Junyan, et al.
Veröffentlicht: (2025)
MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning
von: Tao, Sicheng, et al.
Veröffentlicht: (2025)
von: Tao, Sicheng, et al.
Veröffentlicht: (2025)
EffiReason-Bench: A Unified Benchmark for Evaluating and Advancing Efficient Reasoning in Large Language Models
von: Huang, Junquan, et al.
Veröffentlicht: (2025)
von: Huang, Junquan, et al.
Veröffentlicht: (2025)
Self-Review Framework for Enhancing Instruction Following Capability of LLM
von: Park, Sihyun
Veröffentlicht: (2025)
von: Park, Sihyun
Veröffentlicht: (2025)
Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs
von: Gao, Yubo, et al.
Veröffentlicht: (2026)
von: Gao, Yubo, et al.
Veröffentlicht: (2026)
Dynamic Expert Specialization: Towards Catastrophic Forgetting-Free Multi-Domain MoE Adaptation
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
When Instructions Multiply: Measuring and Estimating LLM Capabilities of Multiple Instructions Following
von: Harada, Keno, et al.
Veröffentlicht: (2025)
von: Harada, Keno, et al.
Veröffentlicht: (2025)
Temporal Gains, Spatial Costs: Revisiting Video Fine-Tuning in Multimodal Large Language Models
von: Zhang, Linghao, et al.
Veröffentlicht: (2026)
von: Zhang, Linghao, et al.
Veröffentlicht: (2026)
Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation
von: Zheng, Kening, et al.
Veröffentlicht: (2026)
von: Zheng, Kening, et al.
Veröffentlicht: (2026)
Decoding Knowledge Attribution in Mixture-of-Experts: A Framework of Basic-Refinement Collaboration and Efficiency Analysis
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
von: Li, Junzhuo, et al.
Veröffentlicht: (2025)
VideoMark: A Distortion-Free Robust Watermarking Framework for Video Diffusion Models
von: Hu, Xuming, et al.
Veröffentlicht: (2025)
von: Hu, Xuming, et al.
Veröffentlicht: (2025)
Optimal Expert-Attention Allocation in Mixture-of-Experts: A Scalable Law for Dynamic Model Design
von: Li, Junzhuo, et al.
Veröffentlicht: (2026)
von: Li, Junzhuo, et al.
Veröffentlicht: (2026)
SAVEn-Vid: Synergistic Audio-Visual Integration for Enhanced Understanding in Long Video Context
von: Li, Jungang, et al.
Veröffentlicht: (2024)
von: Li, Jungang, et al.
Veröffentlicht: (2024)
Quantifying LLM Biases Across Instruction Boundary in Mixed Question Forms
von: Ling, Zipeng, et al.
Veröffentlicht: (2025)
von: Ling, Zipeng, et al.
Veröffentlicht: (2025)
MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
von: Huo, Jiahao, et al.
Veröffentlicht: (2024)
von: Huo, Jiahao, et al.
Veröffentlicht: (2024)
PMark: Towards Robust and Distortion-free Semantic-level Watermarking with Channel Constraints
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
DocPruner: A Storage-Efficient Framework for Multi-Vector Visual Document Retrieval via Adaptive Patch-Level Embedding Pruning
von: Yan, Yibo, et al.
Veröffentlicht: (2025)
von: Yan, Yibo, et al.
Veröffentlicht: (2025)
MINER: Mining the Underlying Pattern of Modality-Specific Neurons in Multimodal Large Language Models
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
Enhancing LLM Instruction Following: An Evaluation-Driven Multi-Agentic Workflow for Prompt Instructions Optimization
von: Purpura, Alberto, et al.
Veröffentlicht: (2026)
von: Purpura, Alberto, et al.
Veröffentlicht: (2026)
Internal Chain-of-Thought: Empirical Evidence for Layer-wise Subtask Scheduling in LLMs
von: Yang, Zhipeng, et al.
Veröffentlicht: (2025)
von: Yang, Zhipeng, et al.
Veröffentlicht: (2025)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
von: Yan, Kaiwen, et al.
Veröffentlicht: (2025)
von: Yan, Kaiwen, et al.
Veröffentlicht: (2025)
A Visual Semantic Adaptive Watermark grounded by Prefix-Tuning for Large Vision-Language Model
von: Zheng, Qi, et al.
Veröffentlicht: (2026)
von: Zheng, Qi, et al.
Veröffentlicht: (2026)
Marco-Bench-MIF: On Multilingual Instruction-Following Capability of Large Language Models
von: Zeng, Bo, et al.
Veröffentlicht: (2025)
von: Zeng, Bo, et al.
Veröffentlicht: (2025)
Non-instructional Fine-tuning: Enabling Instruction-Following Capabilities in Pre-trained Language Models without Instruction-Following Data
von: Xie, Juncheng, et al.
Veröffentlicht: (2024)
von: Xie, Juncheng, et al.
Veröffentlicht: (2024)
Instruction-Following Pruning for Large Language Models
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
von: Hou, Bairu, et al.
Veröffentlicht: (2025)
CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
von: Wang, Peiding, et al.
Veröffentlicht: (2025)
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy
von: Yang, Te, et al.
Veröffentlicht: (2024)
von: Yang, Te, et al.
Veröffentlicht: (2024)
VerIF: Verification Engineering for Reinforcement Learning in Instruction Following
von: Peng, Hao, et al.
Veröffentlicht: (2025)
von: Peng, Hao, et al.
Veröffentlicht: (2025)
Natural Language based Specification and Verification
von: Li, Zhaorui, et al.
Veröffentlicht: (2026)
von: Li, Zhaorui, et al.
Veröffentlicht: (2026)
Visual Instruction Pretraining for Domain-Specific Foundation Models
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
Towards Better Instruction Following Retrieval Models
von: Zhuang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhuang, Yuchen, et al.
Veröffentlicht: (2025)
Instruction Following without Instruction Tuning
von: Hewitt, John, et al.
Veröffentlicht: (2024)
von: Hewitt, John, et al.
Veröffentlicht: (2024)
Enhancing and Assessing Instruction-Following with Fine-Grained Instruction Variants
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
CATMark: A Context-Aware Thresholding Framework for Robust Cross-Task Watermarking in Large Language Models
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation
von: Gao, Ning, et al.
Veröffentlicht: (2025)
von: Gao, Ning, et al.
Veröffentlicht: (2025)
Task-Specific Data Selection for Instruction Tuning via Monosemantic Neuronal Activations
von: Ma, Da, et al.
Veröffentlicht: (2025)
von: Ma, Da, et al.
Veröffentlicht: (2025)
On the Paradoxical Interference between Instruction-Following and Task Solving
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Unlocking Speech Instruction Data Potential with Query Rewriting
von: Hei, Yonghua, et al.
Veröffentlicht: (2025) -
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
von: Dai, Song, et al.
Veröffentlicht: (2025) -
CoheMark: A Novel Sentence-Level Watermark for Enhanced Text Quality
von: Zhang, Junyan, et al.
Veröffentlicht: (2025) -
Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?
von: Zhang, Junyan, et al.
Veröffentlicht: (2025) -
MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning
von: Tao, Sicheng, et al.
Veröffentlicht: (2025)