Forget the Data and Fine-Tuning! Just Fold the Network to Compress
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Dong, Šikić, Haris, Thiele, Lothar, Saukh, Olga |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
by: Saukh, Olga, et al.
Published: (2026)
by: Saukh, Olga, et al.
Published: (2026)
GRAIL: Post-hoc Compensation by Linear Reconstruction for Compressed Networks
by: Tang, Wenwu, et al.
Published: (2026)
by: Tang, Wenwu, et al.
Published: (2026)
From LLMs to Edge: Parameter-Efficient Fine-Tuning on Edge Devices
by: Slamanig, Georg, et al.
Published: (2025)
by: Slamanig, Georg, et al.
Published: (2025)
Subspace-Configurable Networks
by: Wang, Dong, et al.
Published: (2023)
by: Wang, Dong, et al.
Published: (2023)
PCDCNet: A Surrogate Model for Air Quality Forecasting with Physical-Chemical Dynamics and Constraints
by: Wang, Shuo, et al.
Published: (2025)
by: Wang, Shuo, et al.
Published: (2025)
Physics-Guided Inductive Spatiotemporal Kriging for PM2.5 with Satellite Gradient Constraints
by: Wang, Shuo, et al.
Published: (2025)
by: Wang, Shuo, et al.
Published: (2025)
RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting
by: Li, Yuduo, et al.
Published: (2026)
by: Li, Yuduo, et al.
Published: (2026)
Upweighting Easy Samples in Fine-Tuning Mitigates Forgetting
by: Sanyal, Sunny, et al.
Published: (2025)
by: Sanyal, Sunny, et al.
Published: (2025)
RL Fine-Tuning Heals OOD Forgetting in SFT
by: Jin, Hangzhan, et al.
Published: (2025)
by: Jin, Hangzhan, et al.
Published: (2025)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
by: Schweighofer, Kajetan, et al.
Published: (2026)
by: Schweighofer, Kajetan, et al.
Published: (2026)
MoFO: Momentum-Filtered Optimizer for Mitigating Forgetting in LLM Fine-Tuning
by: Chen, Yupeng, et al.
Published: (2024)
by: Chen, Yupeng, et al.
Published: (2024)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
by: Lai, Song, et al.
Published: (2025)
by: Lai, Song, et al.
Published: (2025)
Not All Forgetting Is Equal: Architecture-Dependent Retention Dynamics in Fine-Tuned Image Classifiers
by: Daga, Miit, et al.
Published: (2026)
by: Daga, Miit, et al.
Published: (2026)
TuneComp: Joint Fine-tuning and Compression for Large Foundation Models
by: Chen, Xiangyu, et al.
Published: (2025)
by: Chen, Xiangyu, et al.
Published: (2025)
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
by: Wang, Changsheng, et al.
Published: (2025)
by: Wang, Changsheng, et al.
Published: (2025)
Entropy-Adaptive Fine-Tuning: Resolving Confident Conflicts to Mitigate Forgetting
by: Diao, Muxi, et al.
Published: (2026)
by: Diao, Muxi, et al.
Published: (2026)
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
by: Dabas, Mahavir, et al.
Published: (2025)
by: Dabas, Mahavir, et al.
Published: (2025)
CURLoRA: Stable LLM Continual Fine-Tuning and Catastrophic Forgetting Mitigation
by: Fawi, Muhammad
Published: (2024)
by: Fawi, Muhammad
Published: (2024)
Memory-Efficient Fine-Tuning via Low-Rank Activation Compression
by: Shi, Jiang-Xin, et al.
Published: (2025)
by: Shi, Jiang-Xin, et al.
Published: (2025)
Alignment Dynamics in LLM Fine-Tuning
by: Huang, Yuhan, et al.
Published: (2026)
by: Huang, Yuhan, et al.
Published: (2026)
Catastrophic Forgetting in Kolmogorov-Arnold Networks
by: Rahman, Mohammad Marufur, et al.
Published: (2025)
by: Rahman, Mohammad Marufur, et al.
Published: (2025)
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
by: Liu, Siyuan, et al.
Published: (2026)
by: Liu, Siyuan, et al.
Published: (2026)
Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning
by: Ren, Weijieying, et al.
Published: (2024)
by: Ren, Weijieying, et al.
Published: (2024)
Goal-Conditioned Supervised Learning for LLM Fine-Tuning
by: Li, Shijun, et al.
Published: (2026)
by: Li, Shijun, et al.
Published: (2026)
SEFE: Superficial and Essential Forgetting Eliminator for Multimodal Continual Instruction Tuning
by: Chen, Jinpeng, et al.
Published: (2025)
by: Chen, Jinpeng, et al.
Published: (2025)
H2Tune: Federated Foundation Model Fine-Tuning with Hybrid Heterogeneity
by: Guo, Wei, et al.
Published: (2025)
by: Guo, Wei, et al.
Published: (2025)
Fine-Tuning without Performance Degradation
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
ImitAL: Learned Active Learning Strategy on Synthetic Data
by: Gonsior, Julius, et al.
Published: (2022)
by: Gonsior, Julius, et al.
Published: (2022)
Delayed Bottlenecking: Alleviating Forgetting in Pre-trained Graph Neural Networks
by: Zhao, Zhe, et al.
Published: (2024)
by: Zhao, Zhe, et al.
Published: (2024)
Efficient Inference Using Large Language Models with Limited Human Data: Fine-Tuning then Rectification
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved)
by: Qin, Chongli, et al.
Published: (2025)
by: Qin, Chongli, et al.
Published: (2025)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
by: Agarwal, Ishika, et al.
Published: (2025)
by: Agarwal, Ishika, et al.
Published: (2025)
Unlocking the Power of Function Vectors for Characterizing and Mitigating Catastrophic Forgetting in Continual Instruction Tuning
by: Jiang, Gangwei, et al.
Published: (2025)
by: Jiang, Gangwei, et al.
Published: (2025)
RoRA: Efficient Fine-Tuning of LLM with Reliability Optimization for Rank Adaptation
by: Liu, Jun, et al.
Published: (2025)
by: Liu, Jun, et al.
Published: (2025)
Understanding Forgetting in LLM Supervised Fine-Tuning and Preference Learning -- A Convex Optimization Perspective
by: Fernando, Heshan, et al.
Published: (2024)
by: Fernando, Heshan, et al.
Published: (2024)
From Theory to Application: Fine-Tuning Large EEG Model with Real-World Stress Data
by: Wang, Siwen, et al.
Published: (2025)
by: Wang, Siwen, et al.
Published: (2025)
Parameter-Efficient Fine-Tuning for Foundation Models
by: Zhang, Dan, et al.
Published: (2025)
by: Zhang, Dan, et al.
Published: (2025)
DeltaDQ: Ultra-High Delta Compression for Fine-Tuned LLMs via Group-wise Dropout and Separate Quantization
by: Jiang, Yanfeng, et al.
Published: (2024)
by: Jiang, Yanfeng, et al.
Published: (2024)
Silent Inconsistency in Data-Parallel Full Fine-Tuning: Diagnosing Worker-Level Optimization Misalignment
by: Li, Hong, et al.
Published: (2026)
by: Li, Hong, et al.
Published: (2026)
Human-like Forgetting Curves in Deep Neural Networks
by: Kline, Dylan
Published: (2025)
by: Kline, Dylan
Published: (2025)
Similar Items
-
Cut Less, Fold More: Model Compression through the Lens of Projection Geometry
by: Saukh, Olga, et al.
Published: (2026) -
GRAIL: Post-hoc Compensation by Linear Reconstruction for Compressed Networks
by: Tang, Wenwu, et al.
Published: (2026) -
From LLMs to Edge: Parameter-Efficient Fine-Tuning on Edge Devices
by: Slamanig, Georg, et al.
Published: (2025) -
Subspace-Configurable Networks
by: Wang, Dong, et al.
Published: (2023) -
PCDCNet: A Surrogate Model for Air Quality Forecasting with Physical-Chemical Dynamics and Constraints
by: Wang, Shuo, et al.
Published: (2025)