Diving into Kronecker Adapters: Component Design Matters
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Jiayu, Yu, Danchen, Liao, Zhenyu, Hou, TianQi, Zhou, Feng, Qiu, Robert C., Ling, Zenan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Interpolation Error of Nonlinear Attention versus Linear Regression
von: Liao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Liao, Zhenyu, et al.
Veröffentlicht: (2025)
Adaptive Discretization for Consistency Models
von: Bai, Jiayu, et al.
Veröffentlicht: (2025)
von: Bai, Jiayu, et al.
Veröffentlicht: (2025)
IGNN-Solver: A Graph Neural Solver for Implicit Graph Neural Networks
von: Lin, Junchao, et al.
Veröffentlicht: (2024)
von: Lin, Junchao, et al.
Veröffentlicht: (2024)
Deep Equilibrium Models are Almost Equivalent to Not-so-deep Explicit Models for High-dimensional Gaussian Mixtures
von: Ling, Zenan, et al.
Veröffentlicht: (2024)
von: Ling, Zenan, et al.
Veröffentlicht: (2024)
MoKA: Mixture of Kronecker Adapters
von: Sadeghi, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Sadeghi, Mohammadreza, et al.
Veröffentlicht: (2025)
Series-to-Series Diffusion Bridge Model
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
Analysis and Approximate Inference of Large Random Kronecker Graphs
von: Liao, Zhenyu, et al.
Veröffentlicht: (2023)
von: Liao, Zhenyu, et al.
Veröffentlicht: (2023)
Fundamental Bias in Inverting Random Sampling Matrices with Application to Sub-sampled Newton
von: Niu, Chengmei, et al.
Veröffentlicht: (2025)
von: Niu, Chengmei, et al.
Veröffentlicht: (2025)
Consistency Deep Equilibrium Models
von: Lin, Junchao, et al.
Veröffentlicht: (2026)
von: Lin, Junchao, et al.
Veröffentlicht: (2026)
Revisiting Logistic-softmax Likelihood in Bayesian Meta-Learning for Few-Shot Classification
von: Ke, Tianjun, et al.
Veröffentlicht: (2023)
von: Ke, Tianjun, et al.
Veröffentlicht: (2023)
Nonstationary Sparse Spectral Permanental Process
von: Sun, Zicheng, et al.
Veröffentlicht: (2024)
von: Sun, Zicheng, et al.
Veröffentlicht: (2024)
Kron-LoRA: Hybrid Kronecker-LoRA Adapters for Scalable, Sustainable Fine-tuning
von: Shen, Yixin
Veröffentlicht: (2025)
von: Shen, Yixin
Veröffentlicht: (2025)
Deep Kronecker Network
von: Feng, Long, et al.
Veröffentlicht: (2022)
von: Feng, Long, et al.
Veröffentlicht: (2022)
Rethinking Bregman Divergences in Kronecker-Factored Optimizers
von: Liu, Bing, et al.
Veröffentlicht: (2026)
von: Liu, Bing, et al.
Veröffentlicht: (2026)
What Information Matters? Graph Out-of-Distribution Detection via Tri-Component Information Decomposition
von: Wang, Danny, et al.
Veröffentlicht: (2026)
von: Wang, Danny, et al.
Veröffentlicht: (2026)
Structural Priors and Modular Adapters in the Composable Fine-Tuning Algorithm of Large-Scale Models
von: Wang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Wang, Yuxiao, et al.
Veröffentlicht: (2025)
Latent Iterative Refinement Flow: A Geometric Constrained Approach for Few-Shot Generation
von: Li, Songtao, et al.
Veröffentlicht: (2025)
von: Li, Songtao, et al.
Veröffentlicht: (2025)
Understanding the Role of Textual Prompts in LLM for Time Series Forecasting: an Adapter View
von: Niu, Peisong, et al.
Veröffentlicht: (2023)
von: Niu, Peisong, et al.
Veröffentlicht: (2023)
Rank Also Matters: Hierarchical Configuration for Mixture of Adapter Experts in LLM Fine-Tuning
von: Cong, Peizhuang, et al.
Veröffentlicht: (2025)
von: Cong, Peizhuang, et al.
Veröffentlicht: (2025)
Fast Inference with Kronecker-Sparse Matrices
von: Gonon, Antoine, et al.
Veröffentlicht: (2024)
von: Gonon, Antoine, et al.
Veröffentlicht: (2024)
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
von: Son, Hyegang, et al.
Veröffentlicht: (2024)
von: Son, Hyegang, et al.
Veröffentlicht: (2024)
"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach
von: Gu, Lingyu, et al.
Veröffentlicht: (2024)
von: Gu, Lingyu, et al.
Veröffentlicht: (2024)
HG-Adapter: Improving Pre-Trained Heterogeneous Graph Neural Networks with Dual Adapters
von: Mo, Yujie, et al.
Veröffentlicht: (2024)
von: Mo, Yujie, et al.
Veröffentlicht: (2024)
Block Circulant Adapter for Large Language Models
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
Kronecker-Structured Nonparametric Spatiotemporal Point Processes
von: Xu, Zhitong, et al.
Veröffentlicht: (2026)
von: Xu, Zhitong, et al.
Veröffentlicht: (2026)
Structured and Fast Optimization: The Kronecker SGD Algorithm
von: Song, Zhao, et al.
Veröffentlicht: (2023)
von: Song, Zhao, et al.
Veröffentlicht: (2023)
Scalable Gaussian Processes with Latent Kronecker Structure
von: Lin, Jihao Andreas, et al.
Veröffentlicht: (2025)
von: Lin, Jihao Andreas, et al.
Veröffentlicht: (2025)
Efficient High-Resolution Time Series Classification via Attention Kronecker Decomposition
von: Feng, Aosong, et al.
Veröffentlicht: (2024)
von: Feng, Aosong, et al.
Veröffentlicht: (2024)
Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation
von: Li, Yi-Chen, et al.
Veröffentlicht: (2024)
von: Li, Yi-Chen, et al.
Veröffentlicht: (2024)
Order Matters in Retrosynthesis: Structure-aware Generation via Reaction-Center-Guided Discrete Flow Matching
von: Wang, Chenguang, et al.
Veröffentlicht: (2026)
von: Wang, Chenguang, et al.
Veröffentlicht: (2026)
Adapter-Enhanced Semantic Prompting for Continual Learning
von: Yin, Baocai, et al.
Veröffentlicht: (2024)
von: Yin, Baocai, et al.
Veröffentlicht: (2024)
Inorganic Catalyst Efficiency Prediction Based on EAPCR Model: A Deep Learning Solution for Multi-Source Heterogeneous Data
von: Liu, Zhangdi, et al.
Veröffentlicht: (2025)
von: Liu, Zhangdi, et al.
Veröffentlicht: (2025)
Disentangling Interpretable Factors with Supervised Independent Subspace Principal Component Analysis
von: Su, Jiayu, et al.
Veröffentlicht: (2024)
von: Su, Jiayu, et al.
Veröffentlicht: (2024)
Position: AI for Science Should Treat Measurement-to-Dataset Pipelines as Inference Components
von: Zhan, Ling, et al.
Veröffentlicht: (2026)
von: Zhan, Ling, et al.
Veröffentlicht: (2026)
Kronecker-factored Approximate Curvature (KFAC) From Scratch
von: Dangel, Felix, et al.
Veröffentlicht: (2025)
von: Dangel, Felix, et al.
Veröffentlicht: (2025)
LADA: Scalable Label-Specific CLIP Adapter for Continual Learning
von: Luo, Mao-Lin, et al.
Veröffentlicht: (2025)
von: Luo, Mao-Lin, et al.
Veröffentlicht: (2025)
Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models
von: Liao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Liao, Zhenyu, et al.
Veröffentlicht: (2025)
KromHC: Manifold-Constrained Hyper-Connections with Kronecker-Product Residual Matrices
von: Zhou, Wuyang, et al.
Veröffentlicht: (2026)
von: Zhou, Wuyang, et al.
Veröffentlicht: (2026)
ONNXPruner: ONNX-Based General Model Pruning Adapter
von: Ren, Dongdong, et al.
Veröffentlicht: (2024)
von: Ren, Dongdong, et al.
Veröffentlicht: (2024)
Kronecker-Factored Approximate Curvature for Modern Neural Network Architectures
von: Eschenhagen, Runa, et al.
Veröffentlicht: (2023)
von: Eschenhagen, Runa, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
On the Interpolation Error of Nonlinear Attention versus Linear Regression
von: Liao, Zhenyu, et al.
Veröffentlicht: (2025) -
Adaptive Discretization for Consistency Models
von: Bai, Jiayu, et al.
Veröffentlicht: (2025) -
IGNN-Solver: A Graph Neural Solver for Implicit Graph Neural Networks
von: Lin, Junchao, et al.
Veröffentlicht: (2024) -
Deep Equilibrium Models are Almost Equivalent to Not-so-deep Explicit Models for High-dimensional Gaussian Mixtures
von: Ling, Zenan, et al.
Veröffentlicht: (2024) -
MoKA: Mixture of Kronecker Adapters
von: Sadeghi, Mohammadreza, et al.
Veröffentlicht: (2025)