ARM: Role-Conditioned Neuron Transplantation for Training-Free Generalist LLM Agent Merging
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Feng, Zhuoka, Chen, Kang, Zhao, Sihan, Xiong, Kai, Wang, Yaoning, Yu, Minshen, Nian, Junjie, Xiao, Changyi, Cao, Yixin, Jiang, Yugang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NEX: Neuron Explore-Exploit Scoring for Label-Free Chain-of-Thought Selection and Model Ranking
von: Chen, Kang, et al.
Veröffentlicht: (2026)
von: Chen, Kang, et al.
Veröffentlicht: (2026)
Do LLMs Signal When They're Right? Evidence from Neuron Agreement
von: Chen, Kang, et al.
Veröffentlicht: (2025)
von: Chen, Kang, et al.
Veröffentlicht: (2025)
Thinking Traps in Long Chain-of-Thought: A Measurable Study and Trap-Aware Adaptive Restart
von: Chen, Kang, et al.
Veröffentlicht: (2026)
von: Chen, Kang, et al.
Veröffentlicht: (2026)
TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories
von: Nian, Junjie, et al.
Veröffentlicht: (2026)
von: Nian, Junjie, et al.
Veröffentlicht: (2026)
SliceGraph: Mapping Process Isomers in Multi-Run Chain-of-Thought Reasoning
von: Chen, Kang, et al.
Veröffentlicht: (2026)
von: Chen, Kang, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Conditional Expectation Reward
von: Xiao, Changyi, et al.
Veröffentlicht: (2026)
von: Xiao, Changyi, et al.
Veröffentlicht: (2026)
EffiEval: Efficient and Generalizable Model Evaluation via Capability Coverage Maximization
von: Wang, Yaoning, et al.
Veröffentlicht: (2025)
von: Wang, Yaoning, et al.
Veröffentlicht: (2025)
Complex Logical Query Answering by Calibrating Knowledge Graph Completion Models
von: Xiao, Changyi, et al.
Veröffentlicht: (2024)
von: Xiao, Changyi, et al.
Veröffentlicht: (2024)
Knowledge Graph Completion by Intermediate Variables Regularization
von: Xiao, Changyi, et al.
Veröffentlicht: (2025)
von: Xiao, Changyi, et al.
Veröffentlicht: (2025)
Model Utility Law: Evaluating LLMs beyond Performance through Mechanism Interpretable Metric
von: Cao, Yixin, et al.
Veröffentlicht: (2025)
von: Cao, Yixin, et al.
Veröffentlicht: (2025)
Long or short CoT? Investigating Instance-level Switch of Large Reasoning Models
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiqi, et al.
Veröffentlicht: (2025)
BNPO: Beta Normalization Policy Optimization
von: Xiao, Changyi, et al.
Veröffentlicht: (2025)
von: Xiao, Changyi, et al.
Veröffentlicht: (2025)
Knowledge Graph Embedding by Normalizing Flows
von: Xiao, Changyi, et al.
Veröffentlicht: (2024)
von: Xiao, Changyi, et al.
Veröffentlicht: (2024)
A Hybrid Vectorized Merge Sort on ARM NEON
von: Zhou, Jincheng, et al.
Veröffentlicht: (2024)
von: Zhou, Jincheng, et al.
Veröffentlicht: (2024)
Optimal Kernel Learning for Gaussian Process Models with High-Dimensional Input
von: Kang, Lulu, et al.
Veröffentlicht: (2025)
von: Kang, Lulu, et al.
Veröffentlicht: (2025)
SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning
von: Xu, Caijun, et al.
Veröffentlicht: (2026)
von: Xu, Caijun, et al.
Veröffentlicht: (2026)
MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent
von: Fu, Yuxia, et al.
Veröffentlicht: (2025)
von: Fu, Yuxia, et al.
Veröffentlicht: (2025)
Finding and Editing Multi-Modal Neurons in Pre-Trained Transformers
von: Pan, Haowen, et al.
Veröffentlicht: (2023)
von: Pan, Haowen, et al.
Veröffentlicht: (2023)
Massively Multiagent Minigames for Training Generalist Agents
von: Choe, Kyoung Whan, et al.
Veröffentlicht: (2024)
von: Choe, Kyoung Whan, et al.
Veröffentlicht: (2024)
Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation
von: Yu, Ye, et al.
Veröffentlicht: (2026)
von: Yu, Ye, et al.
Veröffentlicht: (2026)
MIN-Merging: Merge the Important Neurons for Model Merging
von: Liang, Yunfei
Veröffentlicht: (2025)
von: Liang, Yunfei
Veröffentlicht: (2025)
Bayesian Bridge Gaussian Process Regression
von: Xu, Minshen, et al.
Veröffentlicht: (2025)
von: Xu, Minshen, et al.
Veröffentlicht: (2025)
Orthogonal Model Merging
von: Yang, Sihan, et al.
Veröffentlicht: (2026)
von: Yang, Sihan, et al.
Veröffentlicht: (2026)
ARM: Adaptive Reasoning Model
von: Wu, Siye, et al.
Veröffentlicht: (2025)
von: Wu, Siye, et al.
Veröffentlicht: (2025)
Unlocking Green Growth: The Role of Business Environment in Enhancing Environmental Sustainability
von: Yugang He
Veröffentlicht: (2026)
von: Yugang He
Veröffentlicht: (2026)
CureAgent: A Training-Free Executor-Analyst Framework for Clinical Reasoning
von: Xie, Ting-Ting, et al.
Veröffentlicht: (2025)
von: Xie, Ting-Ting, et al.
Veröffentlicht: (2025)
Training-Free Pretrained Model Merging
von: Xu, Zhengqi, et al.
Veröffentlicht: (2024)
von: Xu, Zhengqi, et al.
Veröffentlicht: (2024)
Mora: Enabling Generalist Video Generation via A Multi-Agent Framework
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2024)
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2024)
CAT Merging: A Training-Free Approach for Resolving Conflicts in Model Merging
von: Sun, Wenju, et al.
Veröffentlicht: (2025)
von: Sun, Wenju, et al.
Veröffentlicht: (2025)
Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
von: Hu, Taihang, et al.
Veröffentlicht: (2024)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
von: Xiong, Kai, et al.
Veröffentlicht: (2023)
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2025)
CoDiQ: Test-Time Scaling for Controllable Difficult Question Generation
von: Peng, Zhongyuan, et al.
Veröffentlicht: (2026)
von: Peng, Zhongyuan, et al.
Veröffentlicht: (2026)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
OmniText: A Training-Free Generalist for Controllable Text-Image Manipulation
von: Gunawan, Agus, et al.
Veröffentlicht: (2025)
von: Gunawan, Agus, et al.
Veröffentlicht: (2025)
Skill-as-Pseudocode: Refactoring Skill Libraries to Pseudocode for LLM Agents
von: Li, Xinze, et al.
Veröffentlicht: (2026)
von: Li, Xinze, et al.
Veröffentlicht: (2026)
DOTResize: Reducing LLM Width via Discrete Optimal Transport-based Neuron Merging
von: Verma, Neha, et al.
Veröffentlicht: (2025)
von: Verma, Neha, et al.
Veröffentlicht: (2025)
Arquitectura Neuronal con Aprendizaje Incremental y Creación de Mapas: el Modelo ARM
von: S. Domínguez
Veröffentlicht: (2000)
von: S. Domínguez
Veröffentlicht: (2000)
Governance by Construction for Generalist Agents
von: Shlomov, Segev, et al.
Veröffentlicht: (2026)
von: Shlomov, Segev, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
NEX: Neuron Explore-Exploit Scoring for Label-Free Chain-of-Thought Selection and Model Ranking
von: Chen, Kang, et al.
Veröffentlicht: (2026) -
Do LLMs Signal When They're Right? Evidence from Neuron Agreement
von: Chen, Kang, et al.
Veröffentlicht: (2025) -
Thinking Traps in Long Chain-of-Thought: A Measurable Study and Trap-Aware Adaptive Restart
von: Chen, Kang, et al.
Veröffentlicht: (2026) -
TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories
von: Nian, Junjie, et al.
Veröffentlicht: (2026) -
SliceGraph: Mapping Process Isomers in Multi-Run Chain-of-Thought Reasoning
von: Chen, Kang, et al.
Veröffentlicht: (2026)