Multi-Modality Expansion and Retention for LLMs through Parameter Merging and Decoupling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Junlin, DU, Guodong, Li, Jing, Goh, Sim Kuan, Wang, Wenya, Wang, Yequan, Liu, Fangming, Tang, Ho-Kin, Alharbi, Saleh, He, Daojing, Zhang, Min |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parameter Competition Balancing for Model Merging
von: Du, Guodong, et al.
Veröffentlicht: (2024)
von: Du, Guodong, et al.
Veröffentlicht: (2024)
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
Neural Parameter Search for Slimmer Fine-Tuned Models and Better Transfer
von: Du, Guodong, et al.
Veröffentlicht: (2025)
von: Du, Guodong, et al.
Veröffentlicht: (2025)
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
von: Guo, Weiyang, et al.
Veröffentlicht: (2025)
von: Guo, Weiyang, et al.
Veröffentlicht: (2025)
Knowledge Editing with Dynamic Knowledge Graphs for Multi-Hop Question Answering
von: Lu, Yifan, et al.
Veröffentlicht: (2024)
von: Lu, Yifan, et al.
Veröffentlicht: (2024)
Knowledge Fusion of Large Language Models Via Modular SkillPacks
von: Du, Guodong, et al.
Veröffentlicht: (2025)
von: Du, Guodong, et al.
Veröffentlicht: (2025)
Evolutionary Neural Architecture Search for 3D Point Cloud Analysis
von: Yang, Yisheng, et al.
Veröffentlicht: (2024)
von: Yang, Yisheng, et al.
Veröffentlicht: (2024)
Modality-Decoupled Online Recursive Editing
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
Multimodal Reasoning with Multimodal Knowledge Graph
von: Lee, Junlin, et al.
Veröffentlicht: (2024)
von: Lee, Junlin, et al.
Veröffentlicht: (2024)
Multi-objective Large Language Model Alignment with Hierarchical Experts
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming
von: Guo, Weiyang, et al.
Veröffentlicht: (2025)
von: Guo, Weiyang, et al.
Veröffentlicht: (2025)
CADE: Cosine Annealing Differential Evolution for Spiking Neural Network
von: Jiang, Runhua, et al.
Veröffentlicht: (2024)
von: Jiang, Runhua, et al.
Veröffentlicht: (2024)
Knowledge Fusion By Evolving Weights of Language Models
von: Du, Guodong, et al.
Veröffentlicht: (2024)
von: Du, Guodong, et al.
Veröffentlicht: (2024)
Impacts of Darwinian Evolution on Pre-trained Deep Neural Networks
von: Du, Guodong, et al.
Veröffentlicht: (2024)
von: Du, Guodong, et al.
Veröffentlicht: (2024)
Mitigating Context-Memory Conflicts in LLMs through Dynamic Cognitive Reconciliation Decoding
von: Zhou, Yigeng, et al.
Veröffentlicht: (2026)
von: Zhou, Yigeng, et al.
Veröffentlicht: (2026)
Meta-heuristic Optimizer Inspired by the Philosophy of Yi Jing
von: Yang, Yisheng, et al.
Veröffentlicht: (2024)
von: Yang, Yisheng, et al.
Veröffentlicht: (2024)
Vocabulary Hijacking in LVLMs: Unveiling Critical Attention Heads by Excluding Inert Tokens to Mitigate Hallucination
von: Chen, Yangneng, et al.
Veröffentlicht: (2026)
von: Chen, Yangneng, et al.
Veröffentlicht: (2026)
Adaptive Detoxification: Safeguarding General Capabilities of LLMs through Toxicity-Aware Knowledge Editing
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
von: Lu, Yifan, et al.
Veröffentlicht: (2025)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
von: Li, Wu, et al.
Veröffentlicht: (2026)
von: Li, Wu, et al.
Veröffentlicht: (2026)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
von: Li, Linbao, et al.
Veröffentlicht: (2025)
von: Li, Linbao, et al.
Veröffentlicht: (2025)
AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
Learning to Watermark: A Selective Watermarking Framework for Large Language Models via Multi-Objective Optimization
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
von: Wang, Chenrui, et al.
Veröffentlicht: (2025)
NeurIPT: Foundation Model for Neural Interfaces
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
von: Fang, Zitao, et al.
Veröffentlicht: (2025)
Unveiling Modality Bias: Automated Sample-Specific Analysis for Multimodal Misinformation Benchmarks
von: Lin, Hehai, et al.
Veröffentlicht: (2025)
von: Lin, Hehai, et al.
Veröffentlicht: (2025)
Few-Shot Learner Generalizes Across AI-Generated Image Detection
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
FedMerge: Federated Personalization via Model Merging
von: Chen, Shutong, et al.
Veröffentlicht: (2025)
von: Chen, Shutong, et al.
Veröffentlicht: (2025)
OmniDFA: A Unified Framework for Open Set Synthesis Image Detection and Few-Shot Attribution
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
von: Wu, Shiyu, et al.
Veröffentlicht: (2025)
LocateEdit-Bench: A Benchmark for Instruction-Based Editing Localization
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
Merging Parameter Estimation and Classification Using LASSO
von: Wang, Le, et al.
Veröffentlicht: (2024)
von: Wang, Le, et al.
Veröffentlicht: (2024)
Comparison of Assessment Methods
von: DU, ZHIQIANG
Veröffentlicht: (2025)
von: DU, ZHIQIANG
Veröffentlicht: (2025)
Masked Structural Growth for 2x Faster Language Model Pre-training
von: Yao, Yiqun, et al.
Veröffentlicht: (2023)
von: Yao, Yiqun, et al.
Veröffentlicht: (2023)
Survive the economic downturn: Operating flexibility, productivity, and stock crash
von: Yang Li, et al.
Veröffentlicht: (2024)
von: Yang Li, et al.
Veröffentlicht: (2024)
Personalizing LLMs with Binary Feedback: A Preference-Corrected Optimization Framework
von: Ma, Xilai, et al.
Veröffentlicht: (2026)
von: Ma, Xilai, et al.
Veröffentlicht: (2026)
Segment Anything in Pathology Images with Natural Language
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
STARE at the Structure: Steering ICL Exemplar Selection with Structural Alignment
von: Li, Jiaqian, et al.
Veröffentlicht: (2025)
von: Li, Jiaqian, et al.
Veröffentlicht: (2025)
Merge then Realign: Simple and Effective Modality-Incremental Continual Learning for Multimodal LLMs
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2025)
Uni-MDTrack: Learning Decoupled Memory and Dynamic States for Parameter-Efficient Visual Tracking in All Modality
von: Cai, Wenrui, et al.
Veröffentlicht: (2026)
von: Cai, Wenrui, et al.
Veröffentlicht: (2026)
Commonsense Knowledge Editing Based on Free-Text in LLMs
von: Huang, Xiusheng, et al.
Veröffentlicht: (2024)
von: Huang, Xiusheng, et al.
Veröffentlicht: (2024)
MergePipe: A Budget-Aware Parameter Management System for Scalable LLM Merging
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
von: Wang, Yuanyi, et al.
Veröffentlicht: (2026)
STEPS: A Temporal Smooth Error Propagation Solver on the Manifolds for Test-Time Adaptation in Time Series Forecasting
von: Liu, Jiaqi, et al.
Veröffentlicht: (2026)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Parameter Competition Balancing for Model Merging
von: Du, Guodong, et al.
Veröffentlicht: (2024) -
To See a World in a Spark of Neuron: Disentangling Multi-task Interference for Training-free Model Merging
von: Fang, Zitao, et al.
Veröffentlicht: (2025) -
Neural Parameter Search for Slimmer Fine-Tuned Models and Better Transfer
von: Du, Guodong, et al.
Veröffentlicht: (2025) -
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
von: Guo, Weiyang, et al.
Veröffentlicht: (2025) -
Knowledge Editing with Dynamic Knowledge Graphs for Multi-Hop Question Answering
von: Lu, Yifan, et al.
Veröffentlicht: (2024)