Improving Model Fusion by Training-time Neuron Alignment with Fixed Neuron Anchors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zexi, Li, Zhiqi, Lin, Jie, Shen, Tao, Xiao, Jun, Guo, Yike, Lin, Tao, Wu, Chao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FedGuCci: Making Local Models More Connected in Landscape for Federated Learning
von: Li, Zexi, et al.
Veröffentlicht: (2024)
von: Li, Zexi, et al.
Veröffentlicht: (2024)
SafeNeuron: Neuron-Level Safety Alignment for Large Language Models
von: Wang, Zhaoxin, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoxin, et al.
Veröffentlicht: (2026)
Light Alignment Improves LLM Safety via Model Self-Reflection with a Single Neuron
von: Shen, Sicheng, et al.
Veröffentlicht: (2026)
von: Shen, Sicheng, et al.
Veröffentlicht: (2026)
Text-to-Model: Text-Conditioned Neural Network Diffusion for Train-Once-for-All Personalization
von: Li, Zexi, et al.
Veröffentlicht: (2024)
von: Li, Zexi, et al.
Veröffentlicht: (2024)
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026)
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026)
Model Fusion via Neuron Transplantation
von: Öz, Muhammed, et al.
Veröffentlicht: (2025)
von: Öz, Muhammed, et al.
Veröffentlicht: (2025)
Optimal Parameter and Neuron Pruning for Out-of-Distribution Detection
von: Chen, Chao, et al.
Veröffentlicht: (2024)
von: Chen, Chao, et al.
Veröffentlicht: (2024)
Controllable Value Alignment in Large Language Models through Neuron-Level Editing
von: Yang, Yonghui, et al.
Veröffentlicht: (2026)
von: Yang, Yonghui, et al.
Veröffentlicht: (2026)
Neuron-level Balance between Stability and Plasticity in Deep Reinforcement Learning
von: Lan, Jiahua, et al.
Veröffentlicht: (2025)
von: Lan, Jiahua, et al.
Veröffentlicht: (2025)
TS-LIF: A Temporal Segment Spiking Neuron Network for Time Series Forecasting
von: Feng, Shibo, et al.
Veröffentlicht: (2025)
von: Feng, Shibo, et al.
Veröffentlicht: (2025)
GRAFT: Grid-Aware Load Forecasting with Multi-Source Textual Alignment and Fusion
von: Lin, Fangzhou, et al.
Veröffentlicht: (2025)
von: Lin, Fangzhou, et al.
Veröffentlicht: (2025)
Trainable Weight Averaging: Accelerating Training and Improving Generalization
von: Li, Tao, et al.
Veröffentlicht: (2022)
von: Li, Tao, et al.
Veröffentlicht: (2022)
FedEve: On Bridging the Client Drift and Period Drift for Cross-device Federated Learning
von: Shen, Tao, et al.
Veröffentlicht: (2025)
von: Shen, Tao, et al.
Veröffentlicht: (2025)
NeuronSeek: On Stability and Expressivity of Task-driven Neurons
von: Pei, Hanyu, et al.
Veröffentlicht: (2025)
von: Pei, Hanyu, et al.
Veröffentlicht: (2025)
Mirror-Neuron Patterns in AI Alignment
von: Wyrick, Robyn
Veröffentlicht: (2025)
von: Wyrick, Robyn
Veröffentlicht: (2025)
Training Neural Networks by Optimizing Neuron Positions
von: Erb, Laura, et al.
Veröffentlicht: (2025)
von: Erb, Laura, et al.
Veröffentlicht: (2025)
Threshold Neuron: A Brain-inspired Artificial Neuron for Efficient On-device Inference
von: Zheng, Zihao, et al.
Veröffentlicht: (2024)
von: Zheng, Zihao, et al.
Veröffentlicht: (2024)
NeuronTune: Fine-Grained Neuron Modulation for Balanced Safety-Utility Alignment in LLMs
von: Pan, Birong, et al.
Veröffentlicht: (2025)
von: Pan, Birong, et al.
Veröffentlicht: (2025)
Zeroth-Order Adaptive Neuron Alignment Based Pruning without Re-Training
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
von: Cunegatti, Elia, et al.
Veröffentlicht: (2024)
Neuronal Fluctuations: Learning Rates vs Participating Neurons
von: Pareek, Darsh, et al.
Veröffentlicht: (2025)
von: Pareek, Darsh, et al.
Veröffentlicht: (2025)
NoRA: Nested Low-Rank Adaptation for Efficient Fine-Tuning Large Models
von: Lin, Cheng, et al.
Veröffentlicht: (2024)
von: Lin, Cheng, et al.
Veröffentlicht: (2024)
Gated Parametric Neuron for Spike-based Audio Recognition
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
Cross-Modal Unlearning via Influential Neuron Path Editing in Multimodal Large Language Models
von: Li, Kunhao, et al.
Veröffentlicht: (2025)
von: Li, Kunhao, et al.
Veröffentlicht: (2025)
Accelerating Training with Neuron Interaction and Nowcasting Networks
von: Knyazev, Boris, et al.
Veröffentlicht: (2024)
von: Knyazev, Boris, et al.
Veröffentlicht: (2024)
Improving Neuron-level Interpretability with White-box Language Models
von: Bai, Hao, et al.
Veröffentlicht: (2024)
von: Bai, Hao, et al.
Veröffentlicht: (2024)
Harnessing Neuron Stability to Improve DNN Verification
von: Duong, Hai, et al.
Veröffentlicht: (2024)
von: Duong, Hai, et al.
Veröffentlicht: (2024)
Training Verification-Friendly Neural Networks via Neuron Behavior Consistency
von: Liu, Zongxin, et al.
Veröffentlicht: (2024)
von: Liu, Zongxin, et al.
Veröffentlicht: (2024)
Variational Neurons in Transformers for Language Modeling
von: Ruffenach, Yves
Veröffentlicht: (2026)
von: Ruffenach, Yves
Veröffentlicht: (2026)
Enhancing Clustered Federated Learning: Integration of Strategies and Improved Methodologies
von: Guo, Yongxin, et al.
Veröffentlicht: (2023)
von: Guo, Yongxin, et al.
Veröffentlicht: (2023)
Open Vocabulary Compositional Explanations for Neuron Alignment
von: La Rosa, Biagio, et al.
Veröffentlicht: (2025)
von: La Rosa, Biagio, et al.
Veröffentlicht: (2025)
Property Neurons in Self-Supervised Speech Transformers
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2024)
von: Lin, Tzu-Quan, et al.
Veröffentlicht: (2024)
Variational Distributional Neuron
von: Ruffenach, Yves
Veröffentlicht: (2026)
von: Ruffenach, Yves
Veröffentlicht: (2026)
Expand Neurons, Not Parameters
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
Early Neuron Alignment in Two-layer ReLU Networks with Small Initialization
von: Min, Hancheng, et al.
Veröffentlicht: (2023)
von: Min, Hancheng, et al.
Veröffentlicht: (2023)
Client2Vec: Improving Federated Learning by Distribution Shifts Aware Client Indexing
von: Guo, Yongxin, et al.
Veröffentlicht: (2024)
von: Guo, Yongxin, et al.
Veröffentlicht: (2024)
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
NeuronScope: A Multi-Agent Framework for Explaining Polysemantic Neurons in Language Models
von: Liu, Weiqi, et al.
Veröffentlicht: (2026)
von: Liu, Weiqi, et al.
Veröffentlicht: (2026)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models
von: Zhou, Yi, et al.
Veröffentlicht: (2025)
von: Zhou, Yi, et al.
Veröffentlicht: (2025)
Confidence Regulation Neurons in Language Models
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FedGuCci: Making Local Models More Connected in Landscape for Federated Learning
von: Li, Zexi, et al.
Veröffentlicht: (2024) -
SafeNeuron: Neuron-Level Safety Alignment for Large Language Models
von: Wang, Zhaoxin, et al.
Veröffentlicht: (2026) -
Light Alignment Improves LLM Safety via Model Self-Reflection with a Single Neuron
von: Shen, Sicheng, et al.
Veröffentlicht: (2026) -
Text-to-Model: Text-Conditioned Neural Network Diffusion for Train-Once-for-All Personalization
von: Li, Zexi, et al.
Veröffentlicht: (2024) -
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
von: Madhyastha, Meghana, et al.
Veröffentlicht: (2026)