Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yi-Chen, Zhang, Fuxiang, Qiu, Wenjie, Yuan, Lei, Jia, Chengxing, Zhang, Zongzhang, Yu, Yang, An, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary Dynamics
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
by: Jia, Chengxing, et al.
Published: (2024)
by: Jia, Chengxing, et al.
Published: (2024)
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference
by: Qiu, Wenjie, et al.
Published: (2025)
by: Qiu, Wenjie, et al.
Published: (2025)
Improving Sample Efficiency of Reinforcement Learning with Background Knowledge from Large Language Models
by: Zhang, Fuxiang, et al.
Published: (2024)
by: Zhang, Fuxiang, et al.
Published: (2024)
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
by: Gao, Chen-Xiao, et al.
Published: (2024)
by: Gao, Chen-Xiao, et al.
Published: (2024)
Utilization of Pre-trained Language Model for Adapter-based Knowledge Transfer in Software Engineering
by: Saberi, Iman, et al.
Published: (2023)
by: Saberi, Iman, et al.
Published: (2023)
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
Robust Multi-agent Communication via Multi-view Message Certification
by: Yuan, Lei, et al.
Published: (2023)
by: Yuan, Lei, et al.
Published: (2023)
Stable Continual Reinforcement Learning via Diffusion-based Trajectory Replay
by: Chen, Feng, et al.
Published: (2024)
by: Chen, Feng, et al.
Published: (2024)
Delayed Bottlenecking: Alleviating Forgetting in Pre-trained Graph Neural Networks
by: Zhao, Zhe, et al.
Published: (2024)
by: Zhao, Zhe, et al.
Published: (2024)
Practical Continual Forgetting for Pre-trained Vision Models
by: Zhao, Hongbo, et al.
Published: (2025)
by: Zhao, Hongbo, et al.
Published: (2025)
The Stability of Singular Distribution: A Spectral Perspective on the Two-Phase Dynamics of Language Model Pre-training
by: Zhang, Hongtao, et al.
Published: (2026)
by: Zhang, Hongtao, et al.
Published: (2026)
Can LLMs Learn New Concepts Incrementally without Forgetting?
by: Zheng, Junhao, et al.
Published: (2024)
by: Zheng, Junhao, et al.
Published: (2024)
Incentivizing LLMs to Self-Verify Their Answers
by: Zhang, Fuxiang, et al.
Published: (2025)
by: Zhang, Fuxiang, et al.
Published: (2025)
Wings: Learning Multimodal LLMs without Text-only Forgetting
by: Zhang, Yi-Kai, et al.
Published: (2024)
by: Zhang, Yi-Kai, et al.
Published: (2024)
ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression
by: Yu, Wneya, et al.
Published: (2026)
by: Yu, Wneya, et al.
Published: (2026)
Instruction Backdoor Attacks Against Customized LLMs
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Can Distillation Mitigate Backdoor Attacks in Pre-trained Encoders?
by: Han, TIngxu, et al.
Published: (2024)
by: Han, TIngxu, et al.
Published: (2024)
An Empirical Analysis of Forgetting in Pre-trained Models with Incremental Low-Rank Updates
by: Soutif--Cormerais, Albin, et al.
Published: (2024)
by: Soutif--Cormerais, Albin, et al.
Published: (2024)
Multi-agent In-context Coordination via Decentralized Memory Retrieval
by: Jiang, Tao, et al.
Published: (2025)
by: Jiang, Tao, et al.
Published: (2025)
Gradient-based Fine-Tuning through Pre-trained Model Regularization
by: Liu, Xuanbo, et al.
Published: (2025)
by: Liu, Xuanbo, et al.
Published: (2025)
Mutual Information Guided Backdoor Mitigation for Pre-trained Encoders
by: Han, Tingxu, et al.
Published: (2024)
by: Han, Tingxu, et al.
Published: (2024)
BWArea Model: Learning World Model, Inverse Dynamics, and Policy for Controllable Language Generation
by: Jia, Chengxing, et al.
Published: (2024)
by: Jia, Chengxing, et al.
Published: (2024)
PreLoRA: Hybrid Pre-training of Vision Transformers with Full Training and Low-Rank Adapters
by: Thapa, Krishu K, et al.
Published: (2025)
by: Thapa, Krishu K, et al.
Published: (2025)
Self-Expansion of Pre-trained Models with Mixture of Adapters for Continual Learning
by: Wang, Huiyi, et al.
Published: (2024)
by: Wang, Huiyi, et al.
Published: (2024)
Free Lunch in Medical Image Foundation Model Pre-training via Randomized Synthesis and Disentanglement
by: Wei, Yuhan, et al.
Published: (2026)
by: Wei, Yuhan, et al.
Published: (2026)
Bias Mitigation in Fine-tuning Pre-trained Models for Enhanced Fairness and Efficiency
by: Zhang, Yixuan, et al.
Published: (2024)
by: Zhang, Yixuan, et al.
Published: (2024)
HG-Adapter: Improving Pre-Trained Heterogeneous Graph Neural Networks with Dual Adapters
by: Mo, Yujie, et al.
Published: (2024)
by: Mo, Yujie, et al.
Published: (2024)
Beyond Reasoning Gains: Mitigating General Capabilities Forgetting in Large Reasoning Models
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
Efficient Adapter Tuning of Pre-trained Speech Models for Automatic Speaker Verification
by: Sang, Mufan, et al.
Published: (2024)
by: Sang, Mufan, et al.
Published: (2024)
MEGA: Second-Order Gradient Alignment for Catastrophic Forgetting Mitigation in GFSCIL
by: Pang, Jinhui, et al.
Published: (2025)
by: Pang, Jinhui, et al.
Published: (2025)
Mitigating Mismatch within Reference-based Preference Optimization
by: Yuan, Suqin, et al.
Published: (2026)
by: Yuan, Suqin, et al.
Published: (2026)
Any-step Dynamics Model Improves Future Predictions for Online and Offline Reinforcement Learning
by: Lin, Haoxin, et al.
Published: (2024)
by: Lin, Haoxin, et al.
Published: (2024)
Bootstrapping LLMs via Preference-Based Policy Optimization
by: Jia, Chen
Published: (2025)
by: Jia, Chen
Published: (2025)
MolGA: Molecular Graph Adaptation with Pre-trained 2D Graph Encoder
by: Yu, Xingtong, et al.
Published: (2025)
by: Yu, Xingtong, et al.
Published: (2025)
Text-Free Multi-domain Graph Pre-training: Toward Graph Foundation Models
by: Yu, Xingtong, et al.
Published: (2024)
by: Yu, Xingtong, et al.
Published: (2024)
Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
by: Huang, Wei, et al.
Published: (2025)
by: Huang, Wei, et al.
Published: (2025)
Fine-Grained Gradient Restriction: A Simple Approach for Mitigating Catastrophic Forgetting
by: Liu, Bo, et al.
Published: (2024)
by: Liu, Bo, et al.
Published: (2024)
ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning
by: Gao, Chen-Xiao, et al.
Published: (2023)
by: Gao, Chen-Xiao, et al.
Published: (2023)
Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning
by: Gao, Chen-Xiao, et al.
Published: (2025)
by: Gao, Chen-Xiao, et al.
Published: (2025)
Similar Items
-
Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary Dynamics
by: Zhang, Xinyu, et al.
Published: (2024) -
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
by: Jia, Chengxing, et al.
Published: (2024) -
Sentence-level Reward Model can Generalize Better for Aligning LLM from Human Preference
by: Qiu, Wenjie, et al.
Published: (2025) -
Improving Sample Efficiency of Reinforcement Learning with Background Knowledge from Large Language Models
by: Zhang, Fuxiang, et al.
Published: (2024) -
Hindsight Preference Learning for Offline Preference-based Reinforcement Learning
by: Gao, Chen-Xiao, et al.
Published: (2024)