Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among Prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Minh, Nguyen, Chau, Nguyen, Huy, Tran, Quyen, Le, Trung, Ho, Nhat |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RepLoRA: Reparameterizing Low-Rank Adaptation via the Perspective of Mixture of Experts
by: Truong, Tuan, et al.
Published: (2025)
by: Truong, Tuan, et al.
Published: (2025)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Leveraging Hierarchical Taxonomies in Prompt-based Continual Learning
by: Tran, Quyen, et al.
Published: (2024)
by: Tran, Quyen, et al.
Published: (2024)
On Zero-Initialized Attention: Optimal Prompt and Gating Factor Estimation
by: Diep, Nghiem T., et al.
Published: (2025)
by: Diep, Nghiem T., et al.
Published: (2025)
Mixture of Experts Meets Prompt-Based Continual Learning
by: Le, Minh, et al.
Published: (2024)
by: Le, Minh, et al.
Published: (2024)
Towards Convergence Rates for Parameter Estimation in Gaussian-gated Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
Promoting Ensemble Diversity with Interactive Bayesian Distributional Robustness for Fine-tuning Foundation Models
by: Pham, Ngoc-Quan, et al.
Published: (2025)
by: Pham, Ngoc-Quan, et al.
Published: (2025)
A General Theory for Softmax Gating Multinomial Logistic Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
Attack On Prompt: Backdoor Attack in Prompt-Based Continual Learning
by: Nguyen, Trang, et al.
Published: (2024)
by: Nguyen, Trang, et al.
Published: (2024)
Enhancing Domain Adaptation through Prompt Gradient Alignment
by: Phan, Hoang, et al.
Published: (2024)
by: Phan, Hoang, et al.
Published: (2024)
Improving Generalization with Flat Hilbert Bayesian Inference
by: Truong, Tuan, et al.
Published: (2024)
by: Truong, Tuan, et al.
Published: (2024)
On Parameter Estimation in Deviated Gaussian Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Quadratic Gating Mixture of Experts: Statistical Insights into Self-Attention
by: Akbarian, Pedram, et al.
Published: (2024)
by: Akbarian, Pedram, et al.
Published: (2024)
Statistical Perspective of Top-K Sparse Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
On DeepSeekMoE: Statistical Benefits of Shared Experts and Normalized Sigmoid Gating
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
A Statistical Theory of Gated Attention through the Lens of Hierarchical Mixture of Experts
by: Nguyen, Viet, et al.
Published: (2026)
by: Nguyen, Viet, et al.
Published: (2026)
Class-Prototype Conditional Diffusion Model with Gradient Projection for Continual Learning
by: Doan, Khanh, et al.
Published: (2023)
by: Doan, Khanh, et al.
Published: (2023)
Statistical Advantages of Perturbing Cosine Router in Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
On Least Square Estimation in Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Is Temperature Sample Efficient for Softmax Gaussian Mixture of Experts?
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Sigmoid Gating is More Sample Efficient than Softmax Gating in Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Convergence Rates for Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
Revisiting the Dataset Bias Problem from a Statistical Perspective
by: Do, Kien, et al.
Published: (2024)
by: Do, Kien, et al.
Published: (2024)
Understanding Expert Structures on Minimax Parameter Estimation in Contaminated Mixture of Experts
by: Yan, Fanqi, et al.
Published: (2024)
by: Yan, Fanqi, et al.
Published: (2024)
Improving Minimax Estimation Rates for Contaminated Mixture of Multinomial Logistic Experts via Expert Heterogeneity
by: Yan, Fanqi, et al.
Published: (2026)
by: Yan, Fanqi, et al.
Published: (2026)
Rethinking Multinomial Logistic Mixture of Experts with Sigmoid Gating Function
by: Pham, Tuan Minh, et al.
Published: (2026)
by: Pham, Tuan Minh, et al.
Published: (2026)
Beyond Losses Reweighting: Empowering Multi-Task Learning via the Generalization Perspective
by: Phan, Hoang, et al.
Published: (2022)
by: Phan, Hoang, et al.
Published: (2022)
Agnostic Sharpness-Aware Minimization
by: Nguyen, Van-Anh, et al.
Published: (2024)
by: Nguyen, Van-Anh, et al.
Published: (2024)
Hypernetwork-Driven Low-Rank Adaptation Across Attention Heads
by: Diep, Nghiem T., et al.
Published: (2025)
by: Diep, Nghiem T., et al.
Published: (2025)
On Minimax Estimation of Parameters in Softmax-Contaminated Mixture of Experts
by: Yan, Fanqi, et al.
Published: (2025)
by: Yan, Fanqi, et al.
Published: (2025)
KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All
by: Tran, Quyen, et al.
Published: (2023)
by: Tran, Quyen, et al.
Published: (2023)
Gap Safe Screening Rules for Fast Training of Robust Support Vector Machines under Feature Noise
by: Nguyen, Tan-Hau, et al.
Published: (2026)
by: Nguyen, Tan-Hau, et al.
Published: (2026)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
by: Nguyen, Hieu, et al.
Published: (2024)
by: Nguyen, Hieu, et al.
Published: (2024)
BSO: Safety Alignment Is Density Ratio Matching
by: Nguyen, Tien-Phat, et al.
Published: (2026)
by: Nguyen, Tien-Phat, et al.
Published: (2026)
BRIDGE: Budget-aware Reasoning via Intermediate Distillation with Guided Examples
by: Le, Xuan-An, et al.
Published: (2025)
by: Le, Xuan-An, et al.
Published: (2025)
Statistical Inference for Clustering-based Anomaly Detection
by: Phu, Nguyen Thi Minh, et al.
Published: (2025)
by: Phu, Nguyen Thi Minh, et al.
Published: (2025)
PRE: Vision-Language Prompt Learning with Reparameterization Encoder
by: Pham, Thi Minh Anh, et al.
Published: (2023)
by: Pham, Thi Minh Anh, et al.
Published: (2023)
Empowering Contrastive Federated Sequential Recommendation with LLMs
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
by: Nguyen, Thi Minh Chau, et al.
Published: (2026)
Importance Weighted Variational Inference without the Reparameterization Trick
by: Daudel, Kamélia, et al.
Published: (2026)
by: Daudel, Kamélia, et al.
Published: (2026)
Similar Items
-
RepLoRA: Reparameterizing Low-Rank Adaptation via the Perspective of Mixture of Experts
by: Truong, Tuan, et al.
Published: (2025) -
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025) -
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025) -
Leveraging Hierarchical Taxonomies in Prompt-based Continual Learning
by: Tran, Quyen, et al.
Published: (2024) -
On Zero-Initialized Attention: Optimal Prompt and Gating Factor Estimation
by: Diep, Nghiem T., et al.
Published: (2025)