Weight Spectra Induced Efficient Model Adaptation
Fuente:
arXiv
Saved in:
| Main Authors: | Si, Chongjie, Yang, Xuankun, Liu, Muqing, Wang, Yadao, Yang, Xiaokang, Su, Wenbo, Zheng, Bo, Shen, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MAP: Revisiting Weight Decomposition for Low-Rank Adaptation
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
NAN: A Training-Free Solution to Coefficient Estimation in Model Merging
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
See Further for Parameter Efficient Fine-tuning by Standing on the Shoulders of Decomposition
by: Si, Chongjie, et al.
Published: (2024)
by: Si, Chongjie, et al.
Published: (2024)
FlexLoRA: Entropy-Guided Flexible Low-Rank Adaptation
by: Liu, Muqing, et al.
Published: (2026)
by: Liu, Muqing, et al.
Published: (2026)
Why Can Accurate Models Be Learned from Inaccurate Annotations?
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
Appeal: Allow Mislabeled Samples the Chance to be Rectified in Partial Label Learning
by: Si, Chongjie, et al.
Published: (2023)
by: Si, Chongjie, et al.
Published: (2023)
Revisiting Sparsity Constraint Under High-Rank Property in Partial Multi-Label Learning
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
Generalized Tensor-based Parameter-Efficient Fine-Tuning via Lie Group Transformations
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
Unveiling the Mystery of Weight in Large Foundation Models: Gaussian Distribution Never Fades
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
Task-Specific Directions: Definition, Exploration, and Utilization in Parameter Efficient Fine-Tuning
by: Si, Chongjie, et al.
Published: (2024)
by: Si, Chongjie, et al.
Published: (2024)
AdaMuon: Adaptive Muon Optimizer
by: Si, Chongjie, et al.
Published: (2025)
by: Si, Chongjie, et al.
Published: (2025)
Expert Divergence Learning for MoE-based Language Models
by: Li, Jiaang, et al.
Published: (2026)
by: Li, Jiaang, et al.
Published: (2026)
SpiralFormer: Looped Transformers Can Learn Hierarchical Dependencies via Multi-Resolution Recursion
by: Yu, Chengting, et al.
Published: (2026)
by: Yu, Chengting, et al.
Published: (2026)
Tendency-driven Mutual Exclusivity for Weakly Supervised Incremental Semantic Segmentation
by: Si, Chongjie, et al.
Published: (2024)
by: Si, Chongjie, et al.
Published: (2024)
MeSH: Memory-as-State-Highways for Recursive Transformers
by: Yu, Chengting, et al.
Published: (2025)
by: Yu, Chengting, et al.
Published: (2025)
Low Rank Adaptation for Adversarial Perturbation
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
Dynamic Weight Adjusting Deep Q-Networks for Real-Time Environmental Adaptation
by: Zhang, Xinhao, et al.
Published: (2024)
by: Zhang, Xinhao, et al.
Published: (2024)
DoTA: Weight-Decomposed Tensor Adaptation for Large Language Models
by: Hu, Xiaolin, et al.
Published: (2024)
by: Hu, Xiaolin, et al.
Published: (2024)
Model-Based Reinforcement Learning with Multi-Task Offline Pretraining
by: Pan, Minting, et al.
Published: (2023)
by: Pan, Minting, et al.
Published: (2023)
DiffNMR: Diffusion Models for Nuclear Magnetic Resonance Spectra Elucidation
by: Yang, Qingsong, et al.
Published: (2025)
by: Yang, Qingsong, et al.
Published: (2025)
Efficient Multi-agent Reinforcement Learning by Planning
by: Liu, Qihan, et al.
Published: (2024)
by: Liu, Qihan, et al.
Published: (2024)
PEANuT: Parameter-Efficient Adaptation with Weight-aware Neural Tweakers
by: Zhong, Yibo, et al.
Published: (2024)
by: Zhong, Yibo, et al.
Published: (2024)
Enhancing Parameter Efficiency and Generalization in Large-Scale Models: A Regularized and Masked Low-Rank Adaptation Approach
by: Mao, Yuzhu, et al.
Published: (2024)
by: Mao, Yuzhu, et al.
Published: (2024)
ProgCo: Program Helps Self-Correction of Large Language Models
by: Song, Xiaoshuai, et al.
Published: (2025)
by: Song, Xiaoshuai, et al.
Published: (2025)
Tackling Interference Induced by Data Training Loops in A/B Tests: A Weighted Training Approach
by: Si, Nian
Published: (2023)
by: Si, Nian
Published: (2023)
Model Predictive Task Sampling for Efficient and Robust Adaptation
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Noisy Test-Time Adaptation in Vision-Language Models
by: Cao, Chentao, et al.
Published: (2025)
by: Cao, Chentao, et al.
Published: (2025)
Words & Weights: Streamlining Multi-Turn Interactions via Co-Adaptation
by: Wei, Chenxing, et al.
Published: (2026)
by: Wei, Chenxing, et al.
Published: (2026)
FinLoRA: Finetuning Quantized Financial Large Language Models Using Low-Rank Adaptation
by: Wang, Dannong, et al.
Published: (2024)
by: Wang, Dannong, et al.
Published: (2024)
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging
by: Shen, Li, et al.
Published: (2024)
by: Shen, Li, et al.
Published: (2024)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
by: Wang, Qi, et al.
Published: (2023)
by: Wang, Qi, et al.
Published: (2023)
CoMeT: Collaborative Memory Transformer for Efficient Long Context Modeling
by: Zhao, Runsong, et al.
Published: (2026)
by: Zhao, Runsong, et al.
Published: (2026)
Video-Enhanced Offline Reinforcement Learning: A Model-Based Approach
by: Pan, Minting, et al.
Published: (2025)
by: Pan, Minting, et al.
Published: (2025)
Bidirectional Time-Frequency Pyramid Network for Enhanced Robust EEG Classification
by: Hong, Jiahui, et al.
Published: (2025)
by: Hong, Jiahui, et al.
Published: (2025)
Translating Flow to Policy via Hindsight Online Imitation
by: Zheng, Yitian, et al.
Published: (2025)
by: Zheng, Yitian, et al.
Published: (2025)
ReAugment: Model Zoo-Guided RL for Few-Shot Time Series Augmentation and Forecasting
by: Yuan, Haochen, et al.
Published: (2024)
by: Yuan, Haochen, et al.
Published: (2024)
MS-BART: Unified Modeling of Mass Spectra and Molecules for Structure Elucidation
by: Han, Yang, et al.
Published: (2025)
by: Han, Yang, et al.
Published: (2025)
MetaGS: A Meta-Learned Gaussian-Phong Model for Out-of-Distribution 3D Scene Relighting
by: He, Yumeng, et al.
Published: (2024)
by: He, Yumeng, et al.
Published: (2024)
Weight Space Representation Learning via Neural Field Adaptation
by: Yang, Zhuoqian, et al.
Published: (2025)
by: Yang, Zhuoqian, et al.
Published: (2025)
Spectra-to-Structure and Structure-to-Spectra Inference Across the Periodic Table
by: Wang, Yufeng, et al.
Published: (2025)
by: Wang, Yufeng, et al.
Published: (2025)
Similar Items
-
MAP: Revisiting Weight Decomposition for Low-Rank Adaptation
by: Si, Chongjie, et al.
Published: (2025) -
NAN: A Training-Free Solution to Coefficient Estimation in Model Merging
by: Si, Chongjie, et al.
Published: (2025) -
See Further for Parameter Efficient Fine-tuning by Standing on the Shoulders of Decomposition
by: Si, Chongjie, et al.
Published: (2024) -
FlexLoRA: Entropy-Guided Flexible Low-Rank Adaptation
by: Liu, Muqing, et al.
Published: (2026) -
Why Can Accurate Models Be Learned from Inaccurate Annotations?
by: Si, Chongjie, et al.
Published: (2025)