Adaptive Budget Allocation for Orthogonal-Subspace Adapter Tuning in LLMs Continual Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Wan, Zhiyi, Du, Wanrou, Li, Liang, Pan, Miao, Qin, Xiaoqi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Zeroth-Order Fine-Tuning of LLMs in Random Subspaces
por: Yu, Ziming, et al.
Publicado: (2024)
por: Yu, Ziming, et al.
Publicado: (2024)
Do Protein Transformers Have Biological Intelligence?
por: Lin, Fudong, et al.
Publicado: (2025)
por: Lin, Fudong, et al.
Publicado: (2025)
CoBA-RL: Capability-Oriented Budget Allocation for Reinforcement Learning in LLMs
por: Yao, Zhiyuan, et al.
Publicado: (2026)
por: Yao, Zhiyuan, et al.
Publicado: (2026)
Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA
por: Rahulamathavan, Yogachandran, et al.
Publicado: (2026)
por: Rahulamathavan, Yogachandran, et al.
Publicado: (2026)
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
por: Li, Ziniu, et al.
Publicado: (2025)
por: Li, Ziniu, et al.
Publicado: (2025)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
por: Bozkurt, Alper Kamil, et al.
Publicado: (2026)
por: Bozkurt, Alper Kamil, et al.
Publicado: (2026)
Sparse Adapter Fusion for Continual Learning in NLP
por: Zeng, Min, et al.
Publicado: (2026)
por: Zeng, Min, et al.
Publicado: (2026)
Context-Aware Adapter Tuning for Few-Shot Relation Learning in Knowledge Graphs
por: Liu, Ran, et al.
Publicado: (2024)
por: Liu, Ran, et al.
Publicado: (2024)
DUET: Optimize Token-Budget Allocation for Reinforcement Learning with Verifiable Rewards
por: Hu, Haoyu, et al.
Publicado: (2026)
por: Hu, Haoyu, et al.
Publicado: (2026)
SPARC: Subspace-Aware Prompt Adaptation for Robust Continual Learning in LLMs
por: Jayasuriya, Dinithi, et al.
Publicado: (2025)
por: Jayasuriya, Dinithi, et al.
Publicado: (2025)
Lotus: Efficient LLM Training by Randomized Low-Rank Gradient Projection with Adaptive Subspace Switching
por: Miao, Tianhao, et al.
Publicado: (2026)
por: Miao, Tianhao, et al.
Publicado: (2026)
Continual Adapter Tuning with Semantic Shift Compensation for Class-Incremental Learning
por: Zhou, Qinhao, et al.
Publicado: (2024)
por: Zhou, Qinhao, et al.
Publicado: (2024)
MGAA: Multi-Granular Adaptive Allocation fof Low-Rank Compression of LLMs
por: Li, Guangyan, et al.
Publicado: (2025)
por: Li, Guangyan, et al.
Publicado: (2025)
Spectral Adapter: Fine-Tuning in Spectral Space
por: Zhang, Fangzhao, et al.
Publicado: (2024)
por: Zhang, Fangzhao, et al.
Publicado: (2024)
DARA: Few-shot Budget Allocation in Online Advertising via In-Context Decision Making with RL-Finetuned LLMs
por: Song, Mingxuan, et al.
Publicado: (2026)
por: Song, Mingxuan, et al.
Publicado: (2026)
Predict-then-Diffuse: Adaptive Response Length for Compute-Budgeted Inference in Diffusion LLMs
por: Rottoli, Michael, et al.
Publicado: (2026)
por: Rottoli, Michael, et al.
Publicado: (2026)
Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs
por: Alomrani, Mohammad Ali, et al.
Publicado: (2025)
por: Alomrani, Mohammad Ali, et al.
Publicado: (2025)
Robust and Efficient Zeroth-Order LLM Fine-Tuning via Adaptive Bayesian Subspace Optimizer
por: Feng, Jian, et al.
Publicado: (2026)
por: Feng, Jian, et al.
Publicado: (2026)
Plasticity vs. Rigidity: The Impact of Low-Rank Adapters on Reasoning on a Micro-Budget
por: Khan, Zohaib, et al.
Publicado: (2026)
por: Khan, Zohaib, et al.
Publicado: (2026)
EigenLoRAx: Recycling Adapters to Find Principal Subspaces for Resource-Efficient Adaptation and Inference
por: Kaushik, Prakhar, et al.
Publicado: (2025)
por: Kaushik, Prakhar, et al.
Publicado: (2025)
Sculpting Subspaces: Constrained Full Fine-Tuning in LLMs for Continual Learning
por: Nayak, Nikhil Shivakumar, et al.
Publicado: (2025)
por: Nayak, Nikhil Shivakumar, et al.
Publicado: (2025)
PLATE: Plasticity-Tunable Efficient Adapters for Geometry-Aware Continual Learning
por: Cosentino, Romain
Publicado: (2026)
por: Cosentino, Romain
Publicado: (2026)
How to Allocate, How to Learn? Dynamic Rollout Allocation and Advantage Modulation for Policy Optimization
por: Fang, Yangyi, et al.
Publicado: (2026)
por: Fang, Yangyi, et al.
Publicado: (2026)
Self-Evolving LLMs via Continual Instruction Tuning
por: Kang, Jiazheng, et al.
Publicado: (2025)
por: Kang, Jiazheng, et al.
Publicado: (2025)
Projecting Out the Malice: A Global Subspace Approach to LLM Detoxification
por: Duan, Zenghao, et al.
Publicado: (2026)
por: Duan, Zenghao, et al.
Publicado: (2026)
LAVa: Layer-wise KV Cache Eviction with Dynamic Budget Allocation
por: Shen, Yiqun, et al.
Publicado: (2025)
por: Shen, Yiqun, et al.
Publicado: (2025)
Budgeted Attention Allocation: Cost-Conditioned Compute Control for Efficient Transformers
por: Nidhi, Amrit
Publicado: (2026)
por: Nidhi, Amrit
Publicado: (2026)
Fisher-Orthogonal Projected Natural Gradient Descent for Continual Learning
por: Garg, Ishir, et al.
Publicado: (2026)
por: Garg, Ishir, et al.
Publicado: (2026)
AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
por: Fleshman, William, et al.
Publicado: (2024)
por: Fleshman, William, et al.
Publicado: (2024)
Elastic Multi-Gradient Descent for Parallel Continual Learning
por: Lyu, Fan, et al.
Publicado: (2024)
por: Lyu, Fan, et al.
Publicado: (2024)
BaKlaVa -- Budgeted Allocation of KV cache for Long-context Inference
por: Gulhan, Ahmed Burak, et al.
Publicado: (2025)
por: Gulhan, Ahmed Burak, et al.
Publicado: (2025)
GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets
por: Yao, Zhangyang, et al.
Publicado: (2026)
por: Yao, Zhangyang, et al.
Publicado: (2026)
FedSVD: Adaptive Orthogonalization for Private Federated Learning with LoRA
por: Lee, Seanie, et al.
Publicado: (2025)
por: Lee, Seanie, et al.
Publicado: (2025)
PLAN: Proactive Low-Rank Allocation for Continual Learning
por: Wang, Xiequn, et al.
Publicado: (2025)
por: Wang, Xiequn, et al.
Publicado: (2025)
Policy Gradient with Adaptive Entropy Annealing for Continual Fine-Tuning
por: Zhang, Yaqian, et al.
Publicado: (2026)
por: Zhang, Yaqian, et al.
Publicado: (2026)
Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging
por: Zhang, Haobo, et al.
Publicado: (2025)
por: Zhang, Haobo, et al.
Publicado: (2025)
A Survey of Continual Reinforcement Learning
por: Pan, Chaofan, et al.
Publicado: (2025)
por: Pan, Chaofan, et al.
Publicado: (2025)
Scalable Variational Bayesian Fine-Tuning of LLMs via Orthogonalized Low-Rank Adapters
por: Xiang, Haotian, et al.
Publicado: (2026)
por: Xiang, Haotian, et al.
Publicado: (2026)
Preventing Dimensional Collapse in Self-Supervised Learning via Orthogonality Regularization
por: He, Junlin, et al.
Publicado: (2024)
por: He, Junlin, et al.
Publicado: (2024)
Reinforced Interactive Continual Learning via Real-time Noisy Human Feedback
por: Yang, Yutao, et al.
Publicado: (2025)
por: Yang, Yutao, et al.
Publicado: (2025)
Ejemplares similares
-
Zeroth-Order Fine-Tuning of LLMs in Random Subspaces
por: Yu, Ziming, et al.
Publicado: (2024) -
Do Protein Transformers Have Biological Intelligence?
por: Lin, Fudong, et al.
Publicado: (2025) -
CoBA-RL: Capability-Oriented Budget Allocation for Reinforcement Learning in LLMs
por: Yao, Zhiyuan, et al.
Publicado: (2026) -
Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA
por: Rahulamathavan, Yogachandran, et al.
Publicado: (2026) -
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
por: Li, Ziniu, et al.
Publicado: (2025)