CAMEx: Curvature-aware Merging of Experts
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Dung V., Nguyen, Minh H., Nguyen, Luc Q., Teo, Rachel S. Y., Nguyen, Tan M., Tran, Linh Duy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
by: Teo, Rachel S. Y., et al.
Published: (2025)
by: Teo, Rachel S. Y., et al.
Published: (2025)
MomentumSMoE: Integrating Momentum into Sparse Mixture of Experts
by: Teo, Rachel S. Y., et al.
Published: (2024)
by: Teo, Rachel S. Y., et al.
Published: (2024)
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
by: Bach, Thong, et al.
Published: (2025)
by: Bach, Thong, et al.
Published: (2025)
Accelerating Transformers with Spectrum-Preserving Token Merging
by: Tran, Hoai-Chau, et al.
Published: (2024)
by: Tran, Hoai-Chau, et al.
Published: (2024)
Mixture of Experts Meets Prompt-Based Continual Learning
by: Le, Minh, et al.
Published: (2024)
by: Le, Minh, et al.
Published: (2024)
Tight Clusters Make Specialized Experts
by: Nielsen, Stefan K., et al.
Published: (2025)
by: Nielsen, Stefan K., et al.
Published: (2025)
Unveiling the Hidden Structure of Self-Attention via Kernel Principal Component Analysis
by: Teo, Rachel S. Y., et al.
Published: (2024)
by: Teo, Rachel S. Y., et al.
Published: (2024)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Revisiting the Disequilibrium Issues in Tackling Heart Disease Classification Tasks
by: Hoang, Thao, et al.
Published: (2024)
by: Hoang, Thao, et al.
Published: (2024)
Equivariant Polynomial Functional Networks
by: Vo, Thieu N., et al.
Published: (2024)
by: Vo, Thieu N., et al.
Published: (2024)
Equivariant Neural Functional Networks for Transformers
by: Tran, Viet-Hoang, et al.
Published: (2024)
by: Tran, Viet-Hoang, et al.
Published: (2024)
FairFinGAN: Fairness-aware Synthetic Financial Data Generation
by: Quy, Tai Le, et al.
Published: (2026)
by: Quy, Tai Le, et al.
Published: (2026)
Spectral Text Fusion: A Frequency-Aware Approach to Multimodal Time-Series Forecasting
by: Nguyen, Huu Hiep, et al.
Published: (2026)
by: Nguyen, Huu Hiep, et al.
Published: (2026)
Fake Advertisements Detection Using Automated Multimodal Learning: A Case Study for Vietnamese Real Estate Data
by: Nguyen, Duy, et al.
Published: (2025)
by: Nguyen, Duy, et al.
Published: (2025)
BRIDGE: Budget-aware Reasoning via Intermediate Distillation with Guided Examples
by: Le, Xuan-An, et al.
Published: (2025)
by: Le, Xuan-An, et al.
Published: (2025)
MP-MoE: Matrix Profile-Guided Mixture of Experts for Precipitation Forecasting
by: Tran, Huyen Ngoc, et al.
Published: (2026)
by: Tran, Huyen Ngoc, et al.
Published: (2026)
Statistical Inference for Clustering-based Anomaly Detection
by: Phu, Nguyen Thi Minh, et al.
Published: (2025)
by: Phu, Nguyen Thi Minh, et al.
Published: (2025)
Optimizing Electric Vehicle Charging Station Placement Using Reinforcement Learning and Agent-Based Simulations
by: Nguyen, Minh-Duc, et al.
Published: (2025)
by: Nguyen, Minh-Duc, et al.
Published: (2025)
The Blessing and Curse of Dimensionality in Safety Alignment
by: Teo, Rachel S. Y., et al.
Published: (2025)
by: Teo, Rachel S. Y., et al.
Published: (2025)
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
by: Le, Dong, et al.
Published: (2026)
by: Le, Dong, et al.
Published: (2026)
RegMean++: Enhancing Effectiveness and Generalization of Regression Mean for Model Merging
by: Nguyen, The-Hai, et al.
Published: (2025)
by: Nguyen, The-Hai, et al.
Published: (2025)
Cost-Adaptive Recourse Recommendation by Adaptive Preference Elicitation
by: Nguyen, Duy, et al.
Published: (2024)
by: Nguyen, Duy, et al.
Published: (2024)
MP-PINN: A Multi-Phase Physics-Informed Neural Network for Epidemic Forecasting
by: Nguyen, Thang, et al.
Published: (2024)
by: Nguyen, Thang, et al.
Published: (2024)
A Framework for Controllable Multi-objective Learning with Annealed Stein Variational Hypernetworks
by: Nguyen, Minh-Duc, et al.
Published: (2025)
by: Nguyen, Minh-Duc, et al.
Published: (2025)
Rethinking Deep Alignment Through The Lens Of Incomplete Learning
by: Bach, Thong, et al.
Published: (2025)
by: Bach, Thong, et al.
Published: (2025)
Continual Safety Alignment via Gradient-Based Sample Selection
by: Bach, Thong, et al.
Published: (2026)
by: Bach, Thong, et al.
Published: (2026)
Identifying Causal Direction via Variational Bayesian Compression
by: Tran, Quang-Duy, et al.
Published: (2025)
by: Tran, Quang-Duy, et al.
Published: (2025)
CASUAL: Conditional Support Alignment for Domain Adaptation with Label Shift
by: Nguyen, Anh T, et al.
Published: (2023)
by: Nguyen, Anh T, et al.
Published: (2023)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
iMoT: Inertial Motion Transformer for Inertial Navigation
by: Nguyen, Son Minh, et al.
Published: (2024)
by: Nguyen, Son Minh, et al.
Published: (2024)
Improving Time Series Encoding with Noise-Aware Self-Supervised Learning and an Efficient Encoder
by: Nguyen, Duy A., et al.
Published: (2023)
by: Nguyen, Duy A., et al.
Published: (2023)
Activation Steering with a Feedback Controller
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
Selective Sinkhorn Routing for Improved Sparse Mixture of Experts
by: Nguyen, Duc Anh, et al.
Published: (2025)
by: Nguyen, Duc Anh, et al.
Published: (2025)
Reviving Error Correction in Modern Deep Time-Series Forecasting
by: Nguyen, Minh Hoang, et al.
Published: (2026)
by: Nguyen, Minh Hoang, et al.
Published: (2026)
Adaptive multi-gradient methods for quasiconvex vector optimization and applications to multi-task learning
by: Minh, Nguyen Anh, et al.
Published: (2024)
by: Minh, Nguyen Anh, et al.
Published: (2024)
Statistical Inference for Autoencoder-based Anomaly Detection after Representation Learning-based Domain Adaptation
by: Kiet, Tran Tuan, et al.
Published: (2025)
by: Kiet, Tran Tuan, et al.
Published: (2025)
Improving Routing in Sparse Mixture of Experts with Graph of Tokens
by: Nguyen, Tam, et al.
Published: (2025)
by: Nguyen, Tam, et al.
Published: (2025)
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network
by: Van Thanh, Nguyen, et al.
Published: (2025)
by: Van Thanh, Nguyen, et al.
Published: (2025)
Similar Items
-
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
by: Nguyen, Dung V., et al.
Published: (2025) -
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
by: Nguyen-Nhat, Minh-Khoi, et al.
Published: (2025) -
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
by: Teo, Rachel S. Y., et al.
Published: (2025) -
MomentumSMoE: Integrating Momentum into Sparse Mixture of Experts
by: Teo, Rachel S. Y., et al.
Published: (2024) -
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
by: Bach, Thong, et al.
Published: (2025)