Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Araujo, Vladimir, Moens, Marie-Francine, Tuytelaars, Tinne |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023)
by: Araujo, Vladimir, et al.
Published: (2023)
Learning to Plan Long-Term for Language Modeling
by: Mai, Florian, et al.
Published: (2024)
by: Mai, Florian, et al.
Published: (2024)
Learning to Plan for Language Modeling from Unlabeled Data
by: Cornille, Nathan, et al.
Published: (2024)
by: Cornille, Nathan, et al.
Published: (2024)
Reduction of Supervision for Biomedical Knowledge Discovery
by: Theodoropoulos, Christos, et al.
Published: (2025)
by: Theodoropoulos, Christos, et al.
Published: (2025)
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
by: Dragomir, Alexandra, et al.
Published: (2026)
by: Dragomir, Alexandra, et al.
Published: (2026)
Continual Learning of Diffusion Models with Generative Distillation
by: Masip, Sergi, et al.
Published: (2023)
by: Masip, Sergi, et al.
Published: (2023)
Multiple Choice Learning of Low-Rank Adapters for Language Modeling
by: Letzelter, Victor, et al.
Published: (2025)
by: Letzelter, Victor, et al.
Published: (2025)
Two Complementary Perspectives to Continual Learning: Ask Not Only What to Optimize, But Also How
by: Hess, Timm, et al.
Published: (2023)
by: Hess, Timm, et al.
Published: (2023)
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
by: Qu, Tingyu, et al.
Published: (2024)
by: Qu, Tingyu, et al.
Published: (2024)
DAM: Dynamic Adapter Merging for Continual Video QA Learning
by: Cheng, Feng, et al.
Published: (2024)
by: Cheng, Feng, et al.
Published: (2024)
NewsRECON: News article REtrieval for image CONtextualization
by: Tonglet, Jonathan, et al.
Published: (2026)
by: Tonglet, Jonathan, et al.
Published: (2026)
Protecting multimodal large language models against misleading visualizations
by: Tonglet, Jonathan, et al.
Published: (2025)
by: Tonglet, Jonathan, et al.
Published: (2025)
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
by: Shenaj, Donald, et al.
Published: (2025)
by: Shenaj, Donald, et al.
Published: (2025)
Anti-Overestimation Dialogue Policy Learning for Task-Completion Dialogue System
by: Tian, Chang, et al.
Published: (2022)
by: Tian, Chang, et al.
Published: (2022)
Learning Dynamics in Continual Pre-Training for Large Language Models
by: Wang, Xingjin, et al.
Published: (2025)
by: Wang, Xingjin, et al.
Published: (2025)
Analytic Subspace Routing: How Recursive Least Squares Works in Continual Learning of Large Language Model
by: Tong, Kai, et al.
Published: (2025)
by: Tong, Kai, et al.
Published: (2025)
Routoo: Learning to Route to Large Language Models Effectively
by: Mohammadshahi, Alireza, et al.
Published: (2024)
by: Mohammadshahi, Alireza, et al.
Published: (2024)
MedualTime: A Dual-Adapter Language Model for Medical Time Series-Text Multimodal Learning
by: Ye, Jiexia, et al.
Published: (2024)
by: Ye, Jiexia, et al.
Published: (2024)
Prediction Error-based Classification for Class-Incremental Learning
by: Zając, Michał, et al.
Published: (2023)
by: Zając, Michał, et al.
Published: (2023)
Efficient Information Extraction in Few-Shot Relation Classification through Contrastive Representation Learning
by: Borchert, Philipp, et al.
Published: (2024)
by: Borchert, Philipp, et al.
Published: (2024)
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
by: Son, Hyegang, et al.
Published: (2024)
by: Son, Hyegang, et al.
Published: (2024)
Learning Adapter Rank via Symmetry Breaking
by: Doyle, Cooper, et al.
Published: (2025)
by: Doyle, Cooper, et al.
Published: (2025)
Bridging Language Gaps: Enhancing Few-Shot Language Adaptation
by: Borchert, Philipp, et al.
Published: (2025)
by: Borchert, Philipp, et al.
Published: (2025)
Unlocking Continual Learning Abilities in Language Models
by: Du, Wenyu, et al.
Published: (2024)
by: Du, Wenyu, et al.
Published: (2024)
End-to-end Planner Training for Language Modeling
by: Cornille, Nathan, et al.
Published: (2024)
by: Cornille, Nathan, et al.
Published: (2024)
RouteLLM: Learning to Route LLMs with Preference Data
by: Ong, Isaac, et al.
Published: (2024)
by: Ong, Isaac, et al.
Published: (2024)
Continual Learning of Large Language Models: A Comprehensive Survey
by: Shi, Haizhou, et al.
Published: (2024)
by: Shi, Haizhou, et al.
Published: (2024)
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2025)
by: Feng, Yujie, et al.
Published: (2025)
Training-Free Bayesianization for Low-Rank Adapters of Large Language Models
by: Shi, Haizhou, et al.
Published: (2024)
by: Shi, Haizhou, et al.
Published: (2024)
BBox-Adapter: Lightweight Adapting for Black-Box Large Language Models
by: Sun, Haotian, et al.
Published: (2024)
by: Sun, Haotian, et al.
Published: (2024)
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2026)
by: Bohdal, Ondrej, et al.
Published: (2026)
Generative Adapter: Contextualizing Language Models in Parameters with A Single Forward Pass
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
OrchMoE: Efficient Multi-Adapter Learning with Task-Skill Synergy
by: Wang, Haowen, et al.
Published: (2024)
by: Wang, Haowen, et al.
Published: (2024)
Learning to Route LLMs with Confidence Tokens
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
Polynomial Composition Activations: Unleashing the Dynamics of Large Language Models
by: Zhuo, Zhijian, et al.
Published: (2024)
by: Zhuo, Zhijian, et al.
Published: (2024)
AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
by: Fleshman, William, et al.
Published: (2024)
by: Fleshman, William, et al.
Published: (2024)
Filter-then-Generate: Large Language Models with Structure-Text Adapter for Knowledge Graph Completion
by: Liu, Ben, et al.
Published: (2024)
by: Liu, Ben, et al.
Published: (2024)
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2026)
by: Feng, Yujie, et al.
Published: (2026)
Task-Aware LoRA Adapter Composition via Similarity Retrieval in Vector Databases
by: Adsul, Riya, et al.
Published: (2026)
by: Adsul, Riya, et al.
Published: (2026)
Dynamic Latent Routing
by: Yu, Fangyuan, et al.
Published: (2026)
by: Yu, Fangyuan, et al.
Published: (2026)
Similar Items
-
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023) -
Learning to Plan Long-Term for Language Modeling
by: Mai, Florian, et al.
Published: (2024) -
Learning to Plan for Language Modeling from Unlabeled Data
by: Cornille, Nathan, et al.
Published: (2024) -
Reduction of Supervision for Biomedical Knowledge Discovery
by: Theodoropoulos, Christos, et al.
Published: (2025) -
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
by: Dragomir, Alexandra, et al.
Published: (2026)