Towards Optimal Adapter Placement for Efficient Transfer Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Nowak, Aleksandra I., Mercea, Otniel-Bogdan, Arnab, Anurag, Pfeiffer, Jonas, Dauphin, Yann, Evci, Utku |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Time-, Memory- and Parameter-Efficient Visual Adaptation
by: Mercea, Otniel-Bogdan, et al.
Published: (2024)
by: Mercea, Otniel-Bogdan, et al.
Published: (2024)
Learning Fine-grained Parameter Sharing via Sparse Tensor Decomposition
by: Üyük, Cem, et al.
Published: (2024)
by: Üyük, Cem, et al.
Published: (2024)
Audio-Visual Generalized Zero-Shot Learning using Pre-Trained Large Multi-Modal Models
by: Kurzendörfer, David, et al.
Published: (2024)
by: Kurzendörfer, David, et al.
Published: (2024)
Compression Scaling Laws:Unifying Sparsity and Quantization
by: Frantar, Elias, et al.
Published: (2025)
by: Frantar, Elias, et al.
Published: (2025)
Dynamic Sparse Training with Structured Sparsity
by: Lasby, Mike, et al.
Published: (2023)
by: Lasby, Mike, et al.
Published: (2023)
Neglected Hessian component explains mysteries in Sharpness regularization
by: Dauphin, Yann N., et al.
Published: (2024)
by: Dauphin, Yann N., et al.
Published: (2024)
Robustmix: Improving Robustness by Regularizing the Frequency Bias of Deep Nets
by: Ngnawe, Jonas, et al.
Published: (2023)
by: Ngnawe, Jonas, et al.
Published: (2023)
Progressive Gradient Flow for Robust N:M Sparsity Training in Transformers
by: Bambhaniya, Abhimanyu Rajeshkumar, et al.
Published: (2024)
by: Bambhaniya, Abhimanyu Rajeshkumar, et al.
Published: (2024)
The Journey Matters: Average Parameter Count over Pre-training Unifies Sparse and Dense Scaling Laws
by: Jin, Tian, et al.
Published: (2025)
by: Jin, Tian, et al.
Published: (2025)
Avoiding spurious sharpness minimization broadens applicability of SAM
by: Singh, Sidak Pal, et al.
Published: (2025)
by: Singh, Sidak Pal, et al.
Published: (2025)
Efficient Active Learning with Abstention
by: Zhu, Yinglun, et al.
Published: (2022)
by: Zhu, Yinglun, et al.
Published: (2022)
Rethinking Adapter Placement: A Dominant Adaptation Module Perspective
by: Zhang, Suoxin, et al.
Published: (2026)
by: Zhang, Suoxin, et al.
Published: (2026)
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
by: Yang, Juncheng, et al.
Published: (2024)
by: Yang, Juncheng, et al.
Published: (2024)
ELLA: Efficient Lifelong Learning for Adapters in Large Language Models
by: Biswas, Shristi Das, et al.
Published: (2026)
by: Biswas, Shristi Das, et al.
Published: (2026)
Modular Deep Learning
by: Pfeiffer, Jonas, et al.
Published: (2023)
by: Pfeiffer, Jonas, et al.
Published: (2023)
Introduction to speech recognition
by: Dauphin, Gabriel
Published: (2024)
by: Dauphin, Gabriel
Published: (2024)
Pareto Low-Rank Adapters: Efficient Multi-Task Learning with Preferences
by: Dimitriadis, Nikolaos, et al.
Published: (2024)
by: Dimitriadis, Nikolaos, et al.
Published: (2024)
Shifting the Paradigm: A Diffeomorphism Between Time Series Data Manifolds for Achieving Shift-Invariancy in Deep Learning
by: Demirel, Berken Utku, et al.
Published: (2025)
by: Demirel, Berken Utku, et al.
Published: (2025)
A Physics Informed Machine Learning Framework for Optimal Sensor Placement and Parameter Estimation
by: Venianakis, Georgios, et al.
Published: (2025)
by: Venianakis, Georgios, et al.
Published: (2025)
Temporal Chain of Thought: Long-Video Understanding by Thinking in Frames
by: Arnab, Anurag, et al.
Published: (2025)
by: Arnab, Anurag, et al.
Published: (2025)
Towards Symmetric Low-Rank Adapters
by: Panoutsos, Tales, et al.
Published: (2025)
by: Panoutsos, Tales, et al.
Published: (2025)
A Wander Through the Multimodal Landscape: Efficient Transfer Learning via Low-rank Sequence Multimodal Adapter
by: Guo, Zirun, et al.
Published: (2024)
by: Guo, Zirun, et al.
Published: (2024)
Learning Without Augmenting: Unsupervised Time Series Representation Learning via Frame Projections
by: Demirel, Berken Utku, et al.
Published: (2025)
by: Demirel, Berken Utku, et al.
Published: (2025)
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
by: Duan, Shukai, et al.
Published: (2024)
by: Duan, Shukai, et al.
Published: (2024)
Transfer learning via Regularized Linear Discriminant Analysis
by: Zhang, Hongzhe, et al.
Published: (2025)
by: Zhang, Hongzhe, et al.
Published: (2025)
Traffic-Aware Optimal Taxi Placement Using Graph Neural Network-Based Reinforcement Learning
by: Khetarpaul, Sonia, et al.
Published: (2026)
by: Khetarpaul, Sonia, et al.
Published: (2026)
F-Adapter: Frequency-Adaptive Parameter-Efficient Fine-Tuning in Scientific Machine Learning
by: Zhang, Hangwei, et al.
Published: (2025)
by: Zhang, Hangwei, et al.
Published: (2025)
Differentiable Particle Filtering using Optimal Placement Resampling
by: Csuzdi, Domonkos, et al.
Published: (2024)
by: Csuzdi, Domonkos, et al.
Published: (2024)
PLATE: Plasticity-Tunable Efficient Adapters for Geometry-Aware Continual Learning
by: Cosentino, Romain
Published: (2026)
by: Cosentino, Romain
Published: (2026)
A density estimation perspective on learning from pairwise human preferences
by: Dumoulin, Vincent, et al.
Published: (2023)
by: Dumoulin, Vincent, et al.
Published: (2023)
A Conservative Approach for Few-Shot Transfer in Off-Dynamics Reinforcement Learning
by: Daoudi, Paul, et al.
Published: (2023)
by: Daoudi, Paul, et al.
Published: (2023)
Advancing GDP Forecasting: The Potential of Machine Learning Techniques in Economic Predictions
by: Oancea, Bogdan
Published: (2025)
by: Oancea, Bogdan
Published: (2025)
Unsupervised Machine Learning for Detecting Structural Anomalies in European Regional Statistics
by: Oancea, Bogdan
Published: (2026)
by: Oancea, Bogdan
Published: (2026)
Good Enough to Learn: LLM-based Anomaly Detection in ECU Logs without Reliable Labels
by: Bogdan, Bogdan, et al.
Published: (2025)
by: Bogdan, Bogdan, et al.
Published: (2025)
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
Exploring Sparse Adapters for Scalable Merging of Parameter Efficient Experts
by: Arnob, Samin Yeasar, et al.
Published: (2025)
by: Arnob, Samin Yeasar, et al.
Published: (2025)
Deep Optimal Sensor Placement for Black Box Stochastic Simulations
by: Cordero-Encinar, Paula, et al.
Published: (2024)
by: Cordero-Encinar, Paula, et al.
Published: (2024)
On Optimal Hyperparameters for Differentially Private Deep Transfer Learning
by: Rehn, Aki, et al.
Published: (2025)
by: Rehn, Aki, et al.
Published: (2025)
Optimal Transfer Learning for Missing Not-at-Random Matrix Completion
by: Jalan, Akhil, et al.
Published: (2025)
by: Jalan, Akhil, et al.
Published: (2025)
Efficient Optimal PAC Learning
by: Høgsgaard, Mikael Møller
Published: (2025)
by: Høgsgaard, Mikael Møller
Published: (2025)
Similar Items
-
Time-, Memory- and Parameter-Efficient Visual Adaptation
by: Mercea, Otniel-Bogdan, et al.
Published: (2024) -
Learning Fine-grained Parameter Sharing via Sparse Tensor Decomposition
by: Üyük, Cem, et al.
Published: (2024) -
Audio-Visual Generalized Zero-Shot Learning using Pre-Trained Large Multi-Modal Models
by: Kurzendörfer, David, et al.
Published: (2024) -
Compression Scaling Laws:Unifying Sparsity and Quantization
by: Frantar, Elias, et al.
Published: (2025) -
Dynamic Sparse Training with Structured Sparsity
by: Lasby, Mike, et al.
Published: (2023)