Efficient Compositional Multi-tasking for On-device Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bohdal, Ondrej, Ozay, Mete, Moon, Jijoong, Lee, Kyeng-Hun, Ko, Hyeonmok, Michieli, Umberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
On-device System of Compositional Multi-tasking in Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025)
HydraOpt: Navigating the Efficiency-Performance Trade-off of Adapter Merging
von: Ceritli, Taha, et al.
Veröffentlicht: (2025)
von: Ceritli, Taha, et al.
Veröffentlicht: (2025)
Clustering-driven Memory Compression for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
von: Shenaj, Donald, et al.
Veröffentlicht: (2025)
von: Shenaj, Donald, et al.
Veröffentlicht: (2025)
HOP to the Next Tasks and Domains for Continual Learning in NLP
von: Michieli, Umberto, et al.
Veröffentlicht: (2024)
von: Michieli, Umberto, et al.
Veröffentlicht: (2024)
Swiss DINO: Efficient and Versatile Vision Framework for On-device Personal Object Search
von: Paramonov, Kirill, et al.
Veröffentlicht: (2024)
von: Paramonov, Kirill, et al.
Veröffentlicht: (2024)
LoRA.rar: Learning to Merge LoRAs via Hypernetworks for Subject-Style Conditioned Image Generation
von: Shenaj, Donald, et al.
Veröffentlicht: (2024)
von: Shenaj, Donald, et al.
Veröffentlicht: (2024)
Controllable Forgetting Mechanism for Few-Shot Class-Incremental Learning
von: Paramonov, Kirill, et al.
Veröffentlicht: (2025)
von: Paramonov, Kirill, et al.
Veröffentlicht: (2025)
Object-conditioned Bag of Instances for Few-Shot Personalized Instance Recognition
von: Michieli, Umberto, et al.
Veröffentlicht: (2024)
von: Michieli, Umberto, et al.
Veröffentlicht: (2024)
MeKi: Memory-based Expert Knowledge Injection for Efficient LLM Scaling
von: Ding, Ning, et al.
Veröffentlicht: (2026)
von: Ding, Ning, et al.
Veröffentlicht: (2026)
CG-TTRL: Context-Guided Test-Time Reinforcement Learning for On-Device Large Language Models
von: Hosseini, Peyman, et al.
Veröffentlicht: (2025)
von: Hosseini, Peyman, et al.
Veröffentlicht: (2025)
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
Feature-Space Generative Models for One-Shot Class-Incremental Learning
von: Foster, Jack, et al.
Veröffentlicht: (2026)
von: Foster, Jack, et al.
Veröffentlicht: (2026)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
MOCHA: Multi-modal Objects-aware Cross-arcHitecture Alignment
von: Camuffo, Elena, et al.
Veröffentlicht: (2025)
von: Camuffo, Elena, et al.
Veröffentlicht: (2025)
LoRA-Guard: Parameter-Efficient Guardrail Adaptation for Content Moderation of Large Language Models
von: Elesedy, Hayder, et al.
Veröffentlicht: (2024)
von: Elesedy, Hayder, et al.
Veröffentlicht: (2024)
Cross-Architecture Auxiliary Feature Space Translation for Efficient Few-Shot Personalized Object Detection
von: Barbato, Francesco, et al.
Veröffentlicht: (2024)
von: Barbato, Francesco, et al.
Veröffentlicht: (2024)
DreamCache: Finetuning-Free Lightweight Personalized Image Generation via Feature Caching
von: Aiello, Emanuele, et al.
Veröffentlicht: (2024)
von: Aiello, Emanuele, et al.
Veröffentlicht: (2024)
Enhanced Model Robustness to Input Corruptions by Per-corruption Adaptation of Normalization Statistics
von: Camuffo, Elena, et al.
Veröffentlicht: (2024)
von: Camuffo, Elena, et al.
Veröffentlicht: (2024)
A Model for Every User and Budget: Label-Free and Personalized Mixed-Precision Quantization
von: Fish, Edward, et al.
Veröffentlicht: (2023)
von: Fish, Edward, et al.
Veröffentlicht: (2023)
FFT-based Selection and Optimization of Statistics for Robust Recognition of Severely Corrupted Images
von: Camuffo, Elena, et al.
Veröffentlicht: (2024)
von: Camuffo, Elena, et al.
Veröffentlicht: (2024)
Hansel: Output Length Controlling Framework for Large Language Models
von: Song, Seoha, et al.
Veröffentlicht: (2024)
von: Song, Seoha, et al.
Veröffentlicht: (2024)
Continual Error Correction on Low-Resource Devices
von: Paramonov, Kirill, et al.
Veröffentlicht: (2025)
von: Paramonov, Kirill, et al.
Veröffentlicht: (2025)
Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2024)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2024)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
Multi-property Steering of Large Language Models with Dynamic Activation Composition
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
Deep Neural Network Models Trained With A Fixed Random Classifier Transfer Better Across Domains
von: Ali, Hafiz Tiomoko, et al.
Veröffentlicht: (2024)
von: Ali, Hafiz Tiomoko, et al.
Veröffentlicht: (2024)
Efficient Real-time Refinement of Language Model Text Generation
von: Ko, Joonho, et al.
Veröffentlicht: (2025)
von: Ko, Joonho, et al.
Veröffentlicht: (2025)
CALM : A Multi-task Benchmark for Comprehensive Assessment of Language Model Bias
von: Gupta, Vipul, et al.
Veröffentlicht: (2023)
von: Gupta, Vipul, et al.
Veröffentlicht: (2023)
BRIDO: Bringing Democratic Order to Abstractive Summarization
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
von: Lee, Junhyun, et al.
Veröffentlicht: (2025)
Efficient Large Language Model Inference with Neural Block Linearization
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
MetaTool: Facilitating Large Language Models to Master Tools with Meta-task Augmentation
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
Enhancing Chemical Reaction and Retrosynthesis Prediction with Large Language Model and Dual-task Learning
von: Lin, Xuan, et al.
Veröffentlicht: (2025)
von: Lin, Xuan, et al.
Veröffentlicht: (2025)
Multi-Task Pre-Finetuning of Lightweight Transformer Encoders for Text Classification and NER
von: Zhu, Junyi, et al.
Veröffentlicht: (2025)
von: Zhu, Junyi, et al.
Veröffentlicht: (2025)
RDP LoRA: Geometry-Driven Identification for Parameter-Efficient Adaptation in Large Language Models
von: Çelebi, Yusuf, et al.
Veröffentlicht: (2026)
von: Çelebi, Yusuf, et al.
Veröffentlicht: (2026)
Understanding Subword Compositionality of Large Language Models
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
The Compositional Architecture of Regret in Large Language Models
von: Cui, Xiangxiang, et al.
Veröffentlicht: (2025)
von: Cui, Xiangxiang, et al.
Veröffentlicht: (2025)
DistiLLM: Towards Streamlined Distillation for Large Language Models
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
von: Ko, Jongwoo, et al.
Veröffentlicht: (2024)
Modeling Layered Consciousness with Multi-Agent Large Language Models
von: Kim, Sang Hun, et al.
Veröffentlicht: (2025)
von: Kim, Sang Hun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026) -
On-device System of Compositional Multi-tasking in Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2025) -
HydraOpt: Navigating the Efficiency-Performance Trade-off of Adapter Merging
von: Ceritli, Taha, et al.
Veröffentlicht: (2025) -
Clustering-driven Memory Compression for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026) -
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
von: Shenaj, Donald, et al.
Veröffentlicht: (2025)