FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Han, Xing, Chaudhari, Shravan, Ranade, Tanvi, Chellappa, Rama, Saria, Suchi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Open-Set Domain Adaptation Under Background Distribution Shift: Challenges and A Provably Efficient Solution
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
di: Nguyen, Huy, et al.
Pubblicazione: (2024)
di: Nguyen, Huy, et al.
Pubblicazione: (2024)
FuseMoE: Mixture-of-Experts Transformers for Fleximodal Fusion
di: Han, Xing, et al.
Pubblicazione: (2024)
di: Han, Xing, et al.
Pubblicazione: (2024)
Massively Multimodal Foundation Models: A Framework for Capturing Interactions with Specialized Mixture-of-Experts
di: Han, Xing, et al.
Pubblicazione: (2025)
di: Han, Xing, et al.
Pubblicazione: (2025)
Between Linear and Sinusoidal: Rethinking the Time Encoder in Dynamic Graph Learning
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2025)
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2025)
Improving Coverage in Combined Prediction Sets with Weighted p-values
di: Wong, Gina, et al.
Pubblicazione: (2025)
di: Wong, Gina, et al.
Pubblicazione: (2025)
Shared LoRA Subspaces for almost Strict Continual Learning
di: Kaushik, Prakhar, et al.
Pubblicazione: (2026)
di: Kaushik, Prakhar, et al.
Pubblicazione: (2026)
WATCH: Adaptive Monitoring for AI Deployments via Weighted-Conformal Martingales
di: Prinster, Drew, et al.
Pubblicazione: (2025)
di: Prinster, Drew, et al.
Pubblicazione: (2025)
On the Invariance and Generality of Neural Scaling Laws
di: Han, Xing, et al.
Pubblicazione: (2026)
di: Han, Xing, et al.
Pubblicazione: (2026)
The Universal Weight Subspace Hypothesis
di: Kaushik, Prakhar, et al.
Pubblicazione: (2025)
di: Kaushik, Prakhar, et al.
Pubblicazione: (2025)
MILM: Large Language Models for Multimodal Irregular Time Series with Informative Sampling
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2026)
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2026)
Conformal Validity Guarantees Exist for Any Data Distribution (and How to Find Them)
di: Prinster, Drew, et al.
Pubblicazione: (2024)
di: Prinster, Drew, et al.
Pubblicazione: (2024)
MoRE: A Mixture of Low-Rank Experts for Adaptive Multi-Task Learning
di: Zhang, Dacao, et al.
Pubblicazione: (2025)
di: Zhang, Dacao, et al.
Pubblicazione: (2025)
Sparsity and Superposition in Mixture of Experts
di: Chaudhari, Marmik, et al.
Pubblicazione: (2025)
di: Chaudhari, Marmik, et al.
Pubblicazione: (2025)
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
di: Hendawy, Ahmed, et al.
Pubblicazione: (2023)
di: Hendawy, Ahmed, et al.
Pubblicazione: (2023)
Data Augmentations for Improved (Large) Language Model Generalization
di: Feder, Amir, et al.
Pubblicazione: (2023)
di: Feder, Amir, et al.
Pubblicazione: (2023)
Continual Learning for Adaptive AI Systems
di: Amin, Md Hasibul, et al.
Pubblicazione: (2025)
di: Amin, Md Hasibul, et al.
Pubblicazione: (2025)
Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning
di: Siddika, Fatema, et al.
Pubblicazione: (2026)
di: Siddika, Fatema, et al.
Pubblicazione: (2026)
Adaptive Shared Experts with LoRA-Based Mixture of Experts for Multi-Task Learning
di: Yang, Minghao, et al.
Pubblicazione: (2025)
di: Yang, Minghao, et al.
Pubblicazione: (2025)
FLAME-MoE: A Transparent End-to-End Research Platform for Mixture-of-Experts Language Models
di: Kang, Hao, et al.
Pubblicazione: (2025)
di: Kang, Hao, et al.
Pubblicazione: (2025)
SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning
di: Xie, Zhen-Hao, et al.
Pubblicazione: (2026)
di: Xie, Zhen-Hao, et al.
Pubblicazione: (2026)
Learning to Prompt Your Domain for Vision-Language Models
di: Wei, Guoyizhe, et al.
Pubblicazione: (2023)
di: Wei, Guoyizhe, et al.
Pubblicazione: (2023)
Scene-Adaptive Continual Learning for CSI-based Human Activity Recognition with Mixture of Experts
di: Zheng, Wenhan, et al.
Pubblicazione: (2026)
di: Zheng, Wenhan, et al.
Pubblicazione: (2026)
PEMT: Multi-Task Correlation Guided Mixture-of-Experts Enables Parameter-Efficient Transfer Learning
di: Lin, Zhisheng, et al.
Pubblicazione: (2024)
di: Lin, Zhisheng, et al.
Pubblicazione: (2024)
Theory on Mixture-of-Experts in Continual Learning
di: Li, Hongbo, et al.
Pubblicazione: (2024)
di: Li, Hongbo, et al.
Pubblicazione: (2024)
Novel Node Category Detection Under Subpopulation Shift
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2024)
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2024)
FLAME: Adaptive and Reactive Concept Drift Mitigation for Federated Learning Deployments
di: Mavromatis, Ioannis, et al.
Pubblicazione: (2024)
di: Mavromatis, Ioannis, et al.
Pubblicazione: (2024)
MixTTE: Multi-Level Mixture-of-Experts for Scalable and Adaptive Travel Time Estimation
di: Jiang, Wenzhao, et al.
Pubblicazione: (2026)
di: Jiang, Wenzhao, et al.
Pubblicazione: (2026)
Mixture of Experts Meets Prompt-Based Continual Learning
di: Le, Minh, et al.
Pubblicazione: (2024)
di: Le, Minh, et al.
Pubblicazione: (2024)
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
di: Lou, Meng, et al.
Pubblicazione: (2026)
di: Lou, Meng, et al.
Pubblicazione: (2026)
Each Rank Could be an Expert: Single-Ranked Mixture of Experts LoRA for Multi-Task Learning
di: Zhao, Ziyu, et al.
Pubblicazione: (2025)
di: Zhao, Ziyu, et al.
Pubblicazione: (2025)
Learning with Expert Abstractions for Efficient Multi-Task Continuous Control
di: Jewett, Jeff, et al.
Pubblicazione: (2025)
di: Jewett, Jeff, et al.
Pubblicazione: (2025)
Quadratic Gating Mixture of Experts: Statistical Insights into Self-Attention
di: Akbarian, Pedram, et al.
Pubblicazione: (2024)
di: Akbarian, Pedram, et al.
Pubblicazione: (2024)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
Dynamic Adaptive Shared Experts with Grouped Multi-Head Attention Mixture of Experts
di: Li, Cheng, et al.
Pubblicazione: (2025)
di: Li, Cheng, et al.
Pubblicazione: (2025)
Similarity-Aware Mixture-of-Experts for Data-Efficient Continual Learning
di: Mclaughlin, Connor, et al.
Pubblicazione: (2026)
di: Mclaughlin, Connor, et al.
Pubblicazione: (2026)
Task-Agnostic Experts Composition for Continual Learning
di: Quarantiello, Luigi, et al.
Pubblicazione: (2025)
di: Quarantiello, Luigi, et al.
Pubblicazione: (2025)
Separation and Collaboration: Two-Level Routing Grouped Mixture-of-Experts for Multi-Domain Continual Learning
di: Zhou, Jialu, et al.
Pubblicazione: (2025)
di: Zhou, Jialu, et al.
Pubblicazione: (2025)
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer
di: Kong, Yilun, et al.
Pubblicazione: (2025)
di: Kong, Yilun, et al.
Pubblicazione: (2025)
EEG-Based Multimodal Learning via Hyperbolic Mixture-of-Curvature Experts
di: Zhou, Runhe, et al.
Pubblicazione: (2026)
di: Zhou, Runhe, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Open-Set Domain Adaptation Under Background Distribution Shift: Challenges and A Provably Efficient Solution
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025) -
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions
di: Nguyen, Huy, et al.
Pubblicazione: (2024) -
FuseMoE: Mixture-of-Experts Transformers for Fleximodal Fusion
di: Han, Xing, et al.
Pubblicazione: (2024) -
Massively Multimodal Foundation Models: A Framework for Capturing Interactions with Specialized Mixture-of-Experts
di: Han, Xing, et al.
Pubblicazione: (2025) -
Between Linear and Sinusoidal: Rethinking the Time Encoder in Dynamic Graph Learning
di: Chung, Hsing-Huan, et al.
Pubblicazione: (2025)