Effective LoRA Adapter Routing using Task Representations
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Dhasade, Akash, Kermarrec, Anne-Marie, Pavlovic, Igor, Petrescu, Diana, Pires, Rafael, Randl, Mathis, de Vos, Martijn |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Efficient Federated Search for Retrieval-Augmented Generation using Lightweight Routing
par: Dhasade, Akash, et autres
Publié: (2025)
par: Dhasade, Akash, et autres
Publié: (2025)
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation
par: Bergman, Shai, et autres
Publié: (2025)
par: Bergman, Shai, et autres
Publié: (2025)
Catapults to the Rescue: Accelerating Vector Search by Exploiting Query Locality
par: Abuzakuk, Sami, et autres
Publié: (2026)
par: Abuzakuk, Sami, et autres
Publié: (2026)
Harnessing Increased Client Participation with Cohort-Parallel Federated Learning
par: Dhasade, Akash, et autres
Publié: (2024)
par: Dhasade, Akash, et autres
Publié: (2024)
QuickDrop: Efficient Federated Unlearning by Integrated Dataset Distillation
par: Dhasade, Akash, et autres
Publié: (2023)
par: Dhasade, Akash, et autres
Publié: (2023)
S-LoRA: Serving Thousands of Concurrent LoRA Adapters
par: Sheng, Ying, et autres
Publié: (2023)
par: Sheng, Ying, et autres
Publié: (2023)
RIVA: Leveraging LLM Agents for Reliable Configuration Drift Detection
par: Abuzakuk, Sami, et autres
Publié: (2026)
par: Abuzakuk, Sami, et autres
Publié: (2026)
Decentralized Learning Made Easy with DecentralizePy
par: Dhasade, Akash, et autres
Publié: (2023)
par: Dhasade, Akash, et autres
Publié: (2023)
Kron-LoRA: Hybrid Kronecker-LoRA Adapters for Scalable, Sustainable Fine-tuning
par: Shen, Yixin
Publié: (2025)
par: Shen, Yixin
Publié: (2025)
Practical Federated Learning without a Server
par: Dhasade, Akash, et autres
Publié: (2025)
par: Dhasade, Akash, et autres
Publié: (2025)
Decentralized Learning Made Practical with Client Sampling
par: de Vos, Martijn, et autres
Publié: (2023)
par: de Vos, Martijn, et autres
Publié: (2023)
Energy-Aware Decentralized Learning with Intermittent Model Training
par: Dhasade, Akash, et autres
Publié: (2024)
par: Dhasade, Akash, et autres
Publié: (2024)
POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving
par: Li, Shaoang, et autres
Publié: (2026)
par: Li, Shaoang, et autres
Publié: (2026)
Boosting Asynchronous Decentralized Learning with Model Fragmentation
par: Biswas, Sayan, et autres
Publié: (2024)
par: Biswas, Sayan, et autres
Publié: (2024)
LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing
par: Li, Wenbing, et autres
Publié: (2025)
par: Li, Wenbing, et autres
Publié: (2025)
Weight space Detection of Backdoors in LoRA Adapters
par: Merenciano, David Puertolas, et autres
Publié: (2026)
par: Merenciano, David Puertolas, et autres
Publié: (2026)
Task-Aware LoRA Adapter Composition via Similarity Retrieval in Vector Databases
par: Adsul, Riya, et autres
Publié: (2026)
par: Adsul, Riya, et autres
Publié: (2026)
Get More for Less in Decentralized Learning Systems
par: Dhasade, Akash, et autres
Publié: (2023)
par: Dhasade, Akash, et autres
Publié: (2023)
LoRA-Pro: Are Low-Rank Adapters Properly Optimized?
par: Wang, Zhengbo, et autres
Publié: (2024)
par: Wang, Zhengbo, et autres
Publié: (2024)
R-LoRA: Randomized Multi-Head LoRA for Efficient Multi-Task Learning
par: Liu, Jinda, et autres
Publié: (2025)
par: Liu, Jinda, et autres
Publié: (2025)
LoRA-Squeeze: Simple and Effective Post-Tuning and In-Tuning Compression of LoRA Modules
par: Vulić, Ivan, et autres
Publié: (2026)
par: Vulić, Ivan, et autres
Publié: (2026)
Optimizing Agentic Workflows using Meta-tools
par: Abuzakuk, Sami, et autres
Publié: (2026)
par: Abuzakuk, Sami, et autres
Publié: (2026)
HypeLoRA: Hyper-Network-Generated LoRA Adapters for Calibrated Language Model Fine-Tuning
par: Trojan, Bartosz, et autres
Publié: (2026)
par: Trojan, Bartosz, et autres
Publié: (2026)
HiLoRA: Adaptive Hierarchical LoRA Routing for Training-Free Domain Generalization
par: Han, Ziyi, et autres
Publié: (2025)
par: Han, Ziyi, et autres
Publié: (2025)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
par: Zou, Junyi
Publié: (2026)
par: Zou, Junyi
Publié: (2026)
Accelerating MoE Model Inference with Expert Sharding
par: Balmau, Oana, et autres
Publié: (2025)
par: Balmau, Oana, et autres
Publié: (2025)
Boosting Resource-Constrained Federated Learning Systems with Guessed Updates
par: Boukhari, Mohamed Yassine, et autres
Publié: (2021)
par: Boukhari, Mohamed Yassine, et autres
Publié: (2021)
Serving Heterogeneous LoRA Adapters in Distributed LLM Inference Systems
par: Jaiswal, Shashwat, et autres
Publié: (2025)
par: Jaiswal, Shashwat, et autres
Publié: (2025)
Parametric Retrieval-Augmented Generation using Latent Routing of LoRA Adapters
par: Su, Zhan, et autres
Publié: (2025)
par: Su, Zhan, et autres
Publié: (2025)
mLoRA: Fine-Tuning LoRA Adapters via Highly-Efficient Pipeline Parallelism in Multiple GPUs
par: Ye, Zhengmao, et autres
Publié: (2023)
par: Ye, Zhengmao, et autres
Publié: (2023)
Noiseless Privacy-Preserving Decentralized Learning
par: Biswas, Sayan, et autres
Publié: (2024)
par: Biswas, Sayan, et autres
Publié: (2024)
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
par: Brüel-Gabrielsson, Rickard, et autres
Publié: (2024)
par: Brüel-Gabrielsson, Rickard, et autres
Publié: (2024)
LoRA as Oracle
par: Arazzi, Marco, et autres
Publié: (2026)
par: Arazzi, Marco, et autres
Publié: (2026)
SEQR: Secure and Efficient QR-based LoRA Routing
par: Fleshman, William, et autres
Publié: (2025)
par: Fleshman, William, et autres
Publié: (2025)
Fairness Auditing with Multi-Agent Collaboration
par: de Vos, Martijn, et autres
Publié: (2024)
par: de Vos, Martijn, et autres
Publié: (2024)
Block-wise LoRA: Revisiting Fine-grained LoRA for Effective Personalization and Stylization in Text-to-Image Generation
par: Li, Likun, et autres
Publié: (2024)
par: Li, Likun, et autres
Publié: (2024)
AutoRAG-LoRA: Hallucination-Triggered Knowledge Retuning via Lightweight Adapters
par: Dwivedi, Kaushik, et autres
Publié: (2025)
par: Dwivedi, Kaushik, et autres
Publié: (2025)
SLAD : Shared LoRA Adapters for Task Specific Distillation
par: Bensaid, Reda, et autres
Publié: (2026)
par: Bensaid, Reda, et autres
Publié: (2026)
Collaborative Agentic AI Needs Interoperability Across Ecosystems
par: Sharma, Rishi, et autres
Publié: (2025)
par: Sharma, Rishi, et autres
Publié: (2025)
LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging
par: Lee, Seungeon, et autres
Publié: (2025)
par: Lee, Seungeon, et autres
Publié: (2025)
Documents similaires
-
Efficient Federated Search for Retrieval-Augmented Generation using Lightweight Routing
par: Dhasade, Akash, et autres
Publié: (2025) -
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation
par: Bergman, Shai, et autres
Publié: (2025) -
Catapults to the Rescue: Accelerating Vector Search by Exploiting Query Locality
par: Abuzakuk, Sami, et autres
Publié: (2026) -
Harnessing Increased Client Participation with Cohort-Parallel Federated Learning
par: Dhasade, Akash, et autres
Publié: (2024) -
QuickDrop: Efficient Federated Unlearning by Integrated Dataset Distillation
par: Dhasade, Akash, et autres
Publié: (2023)