LPT++: Efficient Training on Mixture of Long-tailed Experts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Bowen, Zhou, Pan, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AMEND: A Mixture of Experts Framework for Long-tailed Trajectory Prediction
von: Mercurius, Ray Coden, et al.
Veröffentlicht: (2024)
von: Mercurius, Ray Coden, et al.
Veröffentlicht: (2024)
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling
von: Wang, Xinze, et al.
Veröffentlicht: (2025)
von: Wang, Xinze, et al.
Veröffentlicht: (2025)
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
von: Dong, Bowen, et al.
Veröffentlicht: (2025)
von: Dong, Bowen, et al.
Veröffentlicht: (2025)
Robust Experts: the Effect of Adversarial Training on CNNs with Sparse Mixture-of-Experts Layers
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2025)
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2025)
Tool-R1: Sample-Efficient Reinforcement Learning for Agentic Tool Use
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
von: Zhang, Yabo, et al.
Veröffentlicht: (2025)
Efficient Long-Horizon GUI Agents via Training-Free KV Cache Compression
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
Stable Routing for Mixture-of-Experts in Class-Incremental Learning
von: Guo, Zirui, et al.
Veröffentlicht: (2026)
von: Guo, Zirui, et al.
Veröffentlicht: (2026)
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging
von: Shen, Li, et al.
Veröffentlicht: (2024)
von: Shen, Li, et al.
Veröffentlicht: (2024)
Multilinear Mixture of Experts: Scalable Expert Specialization through Factorization
von: Oldfield, James, et al.
Veröffentlicht: (2024)
von: Oldfield, James, et al.
Veröffentlicht: (2024)
TrACT: A Training Dynamics Aware Contrastive Learning Framework for Long-tail Trajectory Prediction
von: Zhang, Junrui, et al.
Veröffentlicht: (2024)
von: Zhang, Junrui, et al.
Veröffentlicht: (2024)
Generalized Categories Discovery for Long-tailed Recognition
von: Li, Ziyun, et al.
Veröffentlicht: (2023)
von: Li, Ziyun, et al.
Veröffentlicht: (2023)
Video Relationship Detection Using Mixture of Experts
von: Shaabana, Ala, et al.
Veröffentlicht: (2024)
von: Shaabana, Ala, et al.
Veröffentlicht: (2024)
LiME: Lightweight Mixture of Experts for Efficient Multimodal Multi-task Learning
von: Kowsher, Md, et al.
Veröffentlicht: (2026)
von: Kowsher, Md, et al.
Veröffentlicht: (2026)
Mixture of Experts in Image Classification: What's the Sweet Spot?
von: Videau, Mathurin, et al.
Veröffentlicht: (2024)
von: Videau, Mathurin, et al.
Veröffentlicht: (2024)
EMoE: Eigenbasis-Guided Routing for Mixture-of-Experts
von: Cheng, Anzhe, et al.
Veröffentlicht: (2026)
von: Cheng, Anzhe, et al.
Veröffentlicht: (2026)
Mixture-of-Experts Models in Vision: Routing, Optimization, and Generalization
von: Rokah, Adam, et al.
Veröffentlicht: (2026)
von: Rokah, Adam, et al.
Veröffentlicht: (2026)
LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios
von: Huang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Huang, Zhiyuan, et al.
Veröffentlicht: (2025)
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
von: Hong, Feng, et al.
Veröffentlicht: (2025)
von: Hong, Feng, et al.
Veröffentlicht: (2025)
SMCL: Saliency Masked Contrastive Learning for Long-tailed Recognition
von: Park, Sanglee, et al.
Veröffentlicht: (2024)
von: Park, Sanglee, et al.
Veröffentlicht: (2024)
PA-Net: Precipitation-Adaptive Mixture-of-Experts for Long-Tail Rainfall Nowcasting
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
Rethinking the Bias of Foundation Model under Long-tailed Distribution
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
Contrastive Conditional-Unconditional Alignment for Long-tailed Diffusion Model
von: Chen, Fang, et al.
Veröffentlicht: (2025)
von: Chen, Fang, et al.
Veröffentlicht: (2025)
MR-GDINO: Efficient Open-World Continual Object Detection
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
Extracting Uncertainty Estimates from Mixtures of Experts for Semantic Segmentation
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2025)
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2025)
Mixture of Group Experts for Learning Invariant Representations
von: Kang, Lei, et al.
Veröffentlicht: (2025)
von: Kang, Lei, et al.
Veröffentlicht: (2025)
Merging Multi-Task Models via Weight-Ensembling Mixture of Experts
von: Tang, Anke, et al.
Veröffentlicht: (2024)
von: Tang, Anke, et al.
Veröffentlicht: (2024)
Domain-Specialized Object Detection via Model-Level Mixtures of Experts
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2026)
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2026)
Lightweight Metadata-Aware Mixture-of-Experts Masked Autoencoder for Earth Observation
von: Albughdadi, Mohanad
Veröffentlicht: (2025)
von: Albughdadi, Mohanad
Veröffentlicht: (2025)
CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering
von: Huai, Tianyu, et al.
Veröffentlicht: (2025)
von: Huai, Tianyu, et al.
Veröffentlicht: (2025)
Expert Race: A Flexible Routing Strategy for Scaling Diffusion Transformer with Mixture of Experts
von: Yuan, Yike, et al.
Veröffentlicht: (2025)
von: Yuan, Yike, et al.
Veröffentlicht: (2025)
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
Towards Adversarial Robustness of Model-Level Mixture-of-Experts Architectures for Semantic Segmentation
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2024)
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2024)
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
von: Lou, Meng, et al.
Veröffentlicht: (2026)
von: Lou, Meng, et al.
Veröffentlicht: (2026)
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
Design and Behavior of Sparse Mixture-of-Experts Layers in CNN-based Semantic Segmentation
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2026)
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2026)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
From Sparse to Soft Mixtures of Experts
von: Puigcerver, Joan, et al.
Veröffentlicht: (2023)
von: Puigcerver, Joan, et al.
Veröffentlicht: (2023)
LLM as a Complementary Optimizer to Gradient Descent: A Case Study in Prompt Tuning
von: Guo, Zixian, et al.
Veröffentlicht: (2024)
von: Guo, Zixian, et al.
Veröffentlicht: (2024)
Reweighted Flow Matching via Unbalanced OT for Label-free Long-tailed Generation
von: Song, Hyunsoo, et al.
Veröffentlicht: (2025)
von: Song, Hyunsoo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AMEND: A Mixture of Experts Framework for Long-tailed Trajectory Prediction
von: Mercurius, Ray Coden, et al.
Veröffentlicht: (2024) -
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
von: Liu, Yahui, et al.
Veröffentlicht: (2025) -
CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling
von: Wang, Xinze, et al.
Veröffentlicht: (2025) -
MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM
von: Dong, Bowen, et al.
Veröffentlicht: (2025) -
Robust Experts: the Effect of Adversarial Training on CNNs with Sparse Mixture-of-Experts Layers
von: Pavlitska, Svetlana, et al.
Veröffentlicht: (2025)