Rapid Switching and Multi-Adapter Fusion via Sparse High Rank Adapters
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhardwaj, Kartikeya, Pandey, Nilesh Prasad, Priyadarshi, Sweta, Ganapathy, Viswanath, Esteves, Rafael, Kadambi, Shreya, Borse, Shubhankar, Whatmough, Paul, Garrepalli, Risheek, Van Baalen, Mart, Teague, Harris, Nagel, Markus |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sparse High Rank Adapters
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024)
FouRA: Fourier Low Rank Adaptation
von: Borse, Shubhankar, et al.
Veröffentlicht: (2024)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2024)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
von: Kadambi, Shreya, et al.
Veröffentlicht: (2025)
von: Kadambi, Shreya, et al.
Veröffentlicht: (2025)
Oh! We Freeze: Improving Quantized Knowledge Distillation via Signal Propagation Analysis for Large Language Models
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
DuoLoRA : Cycle-consistent and Rank-disentangled Content-Style Personalization
von: Roy, Aniket, et al.
Veröffentlicht: (2025)
von: Roy, Aniket, et al.
Veröffentlicht: (2025)
Leech Lattice Vector Quantization for Efficient LLM Compression
von: van der Ouderaa, Tycho F. A., et al.
Veröffentlicht: (2026)
von: van der Ouderaa, Tycho F. A., et al.
Veröffentlicht: (2026)
Efficient LLM Inference using Dynamic Input Pruning and Cache-Aware Masking
von: Federici, Marco, et al.
Veröffentlicht: (2024)
von: Federici, Marco, et al.
Veröffentlicht: (2024)
Mixture of Cache-Conditional Experts for Efficient Mobile Device Inference
von: Skliar, Andrii, et al.
Veröffentlicht: (2024)
von: Skliar, Andrii, et al.
Veröffentlicht: (2024)
GPTVQ: The Blessing of Dimensionality for LLM Quantization
von: van Baalen, Mart, et al.
Veröffentlicht: (2024)
von: van Baalen, Mart, et al.
Veröffentlicht: (2024)
PipeFlow: Pipelined Processing and Motion-Aware Frame Selection for Long-Form Video Editing
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
Pruning vs Quantization: Which is Better?
von: Kuzmin, Andrey, et al.
Veröffentlicht: (2023)
von: Kuzmin, Andrey, et al.
Veröffentlicht: (2023)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
von: Garrepalli, Risheek, et al.
Veröffentlicht: (2024)
von: Garrepalli, Risheek, et al.
Veröffentlicht: (2024)
Video Reasoning without Training
von: Sridhar, Deepak, et al.
Veröffentlicht: (2025)
von: Sridhar, Deepak, et al.
Veröffentlicht: (2025)
Shears: Unstructured Sparsity with Neural Low-rank Adapter Search
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2024)
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2024)
FP8 Quantization: The Power of the Exponent
von: Kuzmin, Andrey, et al.
Veröffentlicht: (2022)
von: Kuzmin, Andrey, et al.
Veröffentlicht: (2022)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
Adapter-dependent Adapter Methylation Assay
von: Zhang, Jia, et al.
Veröffentlicht: (2024)
von: Zhang, Jia, et al.
Veröffentlicht: (2024)
Improving Code Switching with Supervised Fine Tuning and GELU Adapters
von: Pham, Linh
Veröffentlicht: (2025)
von: Pham, Linh
Veröffentlicht: (2025)
The LLM Surgeon
von: van der Ouderaa, Tycho F. A., et al.
Veröffentlicht: (2023)
von: van der Ouderaa, Tycho F. A., et al.
Veröffentlicht: (2023)
Resolving the Identity Crisis in Text-to-Image Generation
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
Zero-Shot Adaptation of Parameter-Efficient Fine-Tuning in Diffusion Models
von: Farhadzadeh, Farzad, et al.
Veröffentlicht: (2025)
von: Farhadzadeh, Farzad, et al.
Veröffentlicht: (2025)
LoRA-X: Bridging Foundation Models with Training-Free Cross-Model Adaptation
von: Farhadzadeh, Farzad, et al.
Veröffentlicht: (2025)
von: Farhadzadeh, Farzad, et al.
Veröffentlicht: (2025)
ResAdapter: Domain Consistent Resolution Adapter for Diffusion Models
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Cheng, Jiaxiang, et al.
Veröffentlicht: (2024)
CLIP-Adapter: Better Vision-Language Models with Feature Adapters
von: Gao, Peng, et al.
Veröffentlicht: (2021)
von: Gao, Peng, et al.
Veröffentlicht: (2021)
A Switch Protein Adapter for Anti‐LILRB4 CAR‐T Cells
von: Ryan Huang, et al.
Veröffentlicht: (2024)
von: Ryan Huang, et al.
Veröffentlicht: (2024)
Spiffy: Multiplying Diffusion LLM Acceleration via Lossless Speculative Decoding
von: Agrawal, Sudhanshu, et al.
Veröffentlicht: (2025)
von: Agrawal, Sudhanshu, et al.
Veröffentlicht: (2025)
ConFu: Contemplate the Future for Better Speculative Sampling
von: Qin, Zongyue, et al.
Veröffentlicht: (2026)
von: Qin, Zongyue, et al.
Veröffentlicht: (2026)
Masks Can Be Distracting: On Context Comprehension in Diffusion Language Models
von: Piskorz, Julianna, et al.
Veröffentlicht: (2025)
von: Piskorz, Julianna, et al.
Veröffentlicht: (2025)
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2023)
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2023)
A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection
von: Shao, Rui, et al.
Veröffentlicht: (2023)
von: Shao, Rui, et al.
Veröffentlicht: (2023)
Systems Toxicology of Bisphenol A: Mechanistic Overlap in Metabolic and Reproductive Disruption
von: Sweta Bhardwaj, et al.
Veröffentlicht: (2026)
von: Sweta Bhardwaj, et al.
Veröffentlicht: (2026)
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023)
von: Meister, Clara, et al.
Veröffentlicht: (2023)
Adapters Strike Back
von: Steitz, Jan-Martin O., et al.
Veröffentlicht: (2024)
von: Steitz, Jan-Martin O., et al.
Veröffentlicht: (2024)
ObjectAlign: Neuro-Symbolic Object Consistency Verification and Correction
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
Inv-Adapter: ID Customization Generation via Image Inversion and Lightweight Adapter
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
ELP-Adapters: Parameter Efficient Adapter Tuning for Various Speech Processing Tasks
von: Inoue, Nakamasa, et al.
Veröffentlicht: (2024)
von: Inoue, Nakamasa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Sparse High Rank Adapters
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024) -
FouRA: Fourier Low Rank Adaptation
von: Borse, Shubhankar, et al.
Veröffentlicht: (2024) -
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
von: Kadambi, Shreya, et al.
Veröffentlicht: (2025) -
Oh! We Freeze: Improving Quantized Knowledge Distillation via Signal Propagation Analysis for Large Language Models
von: Bhardwaj, Kartikeya, et al.
Veröffentlicht: (2024) -
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)