Guardado en:
| Autores principales: | Petridis, Savvas, Wedin, Ben, Yuan, Ann, Wexler, James, Thain, Nithum |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.04894 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Choose Your Agent: Tradeoffs in Adopting AI Advisors, Coaches, and Delegates in Multi-Party Negotiation
por: Zhu, Kehang, et al.
Publicado: (2026)
por: Zhu, Kehang, et al.
Publicado: (2026)
Strategic Tradeoffs Between Humans and AI in Multi-Agent Bargaining
por: Qian, Crystal, et al.
Publicado: (2025)
por: Qian, Crystal, et al.
Publicado: (2025)
Thinking Like a Scientist: Can Interactive Simulations Foster Critical AI Literacy?
por: Zhao, Yiling, et al.
Publicado: (2025)
por: Zhao, Yiling, et al.
Publicado: (2025)
Improving Neutral Point-of-View Generation with Data- and Parameter-Efficient RL
por: Hoffmann, Jessica, et al.
Publicado: (2025)
por: Hoffmann, Jessica, et al.
Publicado: (2025)
Branch-Train-MiX: Mixing Expert LLMs into a Mixture-of-Experts LLM
por: Sukhbaatar, Sainbayar, et al.
Publicado: (2024)
por: Sukhbaatar, Sainbayar, et al.
Publicado: (2024)
Superposition in Transformers: A Novel Way of Building Mixture of Experts
por: Chaliah, Ayoub Ben, et al.
Publicado: (2024)
por: Chaliah, Ayoub Ben, et al.
Publicado: (2024)
MEPT: Mixture of Expert Prompt Tuning as a Manifold Mapper
por: Zeng, Runjia, et al.
Publicado: (2025)
por: Zeng, Runjia, et al.
Publicado: (2025)
Inverse Constitutional AI: Compressing Preferences into Principles
por: Findeis, Arduin, et al.
Publicado: (2024)
por: Findeis, Arduin, et al.
Publicado: (2024)
One Prompt is not Enough: Automated Construction of a Mixture-of-Expert Prompts
por: Wang, Ruochen, et al.
Publicado: (2024)
por: Wang, Ruochen, et al.
Publicado: (2024)
Stealing User Prompts from Mixture of Experts
por: Yona, Itay, et al.
Publicado: (2024)
por: Yona, Itay, et al.
Publicado: (2024)
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
por: Liu, Zhili, et al.
Publicado: (2024)
por: Liu, Zhili, et al.
Publicado: (2024)
Training Sparse Mixture Of Experts Text Embedding Models
por: Nussbaum, Zach, et al.
Publicado: (2025)
por: Nussbaum, Zach, et al.
Publicado: (2025)
CompeteSMoE -- Statistically Guaranteed Mixture of Experts Training via Competition
por: Nguyen, Nam V., et al.
Publicado: (2025)
por: Nguyen, Nam V., et al.
Publicado: (2025)
Deliberate Lab: A Platform for Real-Time Human-AI Social Experiments
por: Qian, Crystal, et al.
Publicado: (2025)
por: Qian, Crystal, et al.
Publicado: (2025)
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
por: Pan, Bowen, et al.
Publicado: (2024)
por: Pan, Bowen, et al.
Publicado: (2024)
Nexus: Specialization meets Adaptability for Efficiently Training Mixture of Experts
por: Gritsch, Nikolas, et al.
Publicado: (2024)
por: Gritsch, Nikolas, et al.
Publicado: (2024)
Automatic Histograms: Leveraging Language Models for Text Dataset Exploration
por: Reif, Emily, et al.
Publicado: (2024)
por: Reif, Emily, et al.
Publicado: (2024)
Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
por: Wei, Tianwen, et al.
Publicado: (2024)
por: Wei, Tianwen, et al.
Publicado: (2024)
Drop-Upcycling: Training Sparse Mixture of Experts with Partial Re-initialization
por: Nakamura, Taishi, et al.
Publicado: (2025)
por: Nakamura, Taishi, et al.
Publicado: (2025)
Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts
por: Lyu, Boxuan, et al.
Publicado: (2026)
por: Lyu, Boxuan, et al.
Publicado: (2026)
ExpertFlow: Efficient Mixture-of-Experts Inference via Predictive Expert Caching and Token Scheduling
por: He, Xin, et al.
Publicado: (2024)
por: He, Xin, et al.
Publicado: (2024)
LD-MoLE: Learnable Dynamic Routing for Mixture of LoRA Experts
por: Zhuang, Yuan, et al.
Publicado: (2025)
por: Zhuang, Yuan, et al.
Publicado: (2025)
QuantMoE-Bench: Examining Post-Training Quantization for Mixture-of-Experts
por: Li, Pingzhi, et al.
Publicado: (2024)
por: Li, Pingzhi, et al.
Publicado: (2024)
Pre-Attention Expert Prediction and Prefetching for Mixture-of-Experts Large Language Models
por: Zhu, Shien, et al.
Publicado: (2025)
por: Zhu, Shien, et al.
Publicado: (2025)
ExpertPrompting: Instructing Large Language Models to be Distinguished Experts
por: Xu, Benfeng, et al.
Publicado: (2023)
por: Xu, Benfeng, et al.
Publicado: (2023)
Multi-Head Mixture-of-Experts
por: Wu, Xun, et al.
Publicado: (2024)
por: Wu, Xun, et al.
Publicado: (2024)
Routing-Free Mixture-of-Experts
por: Liu, Yilun, et al.
Publicado: (2026)
por: Liu, Yilun, et al.
Publicado: (2026)
Multilingual Routing in Mixture-of-Experts
por: Bandarkar, Lucas, et al.
Publicado: (2025)
por: Bandarkar, Lucas, et al.
Publicado: (2025)
MoPE: Mixture of Prompt Experts for Parameter-Efficient and Scalable Multimodal Fusion
por: Jiang, Ruixiang, et al.
Publicado: (2024)
por: Jiang, Ruixiang, et al.
Publicado: (2024)
SEUF: Is Unlearning One Expert Enough for Mixture-of-Experts LLMs?
por: Zhuang, Haomin, et al.
Publicado: (2024)
por: Zhuang, Haomin, et al.
Publicado: (2024)
Diversifying the Mixture-of-Experts Representation for Language Models with Orthogonal Optimizer
por: Liu, Boan, et al.
Publicado: (2023)
por: Liu, Boan, et al.
Publicado: (2023)
Mixture-of-Experts with Intermediate CTC Supervision for Accented Speech Recognition
por: Lee, Wonjun, et al.
Publicado: (2026)
por: Lee, Wonjun, et al.
Publicado: (2026)
Group then Scale: Dynamic Mixture-of-Experts Multilingual Language Model
por: Li, Chong, et al.
Publicado: (2025)
por: Li, Chong, et al.
Publicado: (2025)
On the Spatial Structure of Mixture-of-Experts in Transformers
por: Bershatsky, Daniel, et al.
Publicado: (2025)
por: Bershatsky, Daniel, et al.
Publicado: (2025)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
por: Thiombiano, Abdoul Majid O., et al.
Publicado: (2025)
por: Thiombiano, Abdoul Majid O., et al.
Publicado: (2025)
Flexible and Effective Mixing of Large Language Models into a Mixture of Domain Experts
por: Lee, Rhui Dih, et al.
Publicado: (2024)
por: Lee, Rhui Dih, et al.
Publicado: (2024)
C3AI: Crafting and Evaluating Constitutions for Constitutional AI
por: Kyrychenko, Yara, et al.
Publicado: (2025)
por: Kyrychenko, Yara, et al.
Publicado: (2025)
Towards a Comprehensive Scaling Law of Mixture-of-Experts
por: Zhao, Guoliang, et al.
Publicado: (2025)
por: Zhao, Guoliang, et al.
Publicado: (2025)
Supervisory Prompt Training
por: Billa, Jean Ghislain, et al.
Publicado: (2024)
por: Billa, Jean Ghislain, et al.
Publicado: (2024)
Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
por: Lu, Xudong, et al.
Publicado: (2024)
por: Lu, Xudong, et al.
Publicado: (2024)
Ejemplares similares
-
Choose Your Agent: Tradeoffs in Adopting AI Advisors, Coaches, and Delegates in Multi-Party Negotiation
por: Zhu, Kehang, et al.
Publicado: (2026) -
Strategic Tradeoffs Between Humans and AI in Multi-Agent Bargaining
por: Qian, Crystal, et al.
Publicado: (2025) -
Thinking Like a Scientist: Can Interactive Simulations Foster Critical AI Literacy?
por: Zhao, Yiling, et al.
Publicado: (2025) -
Improving Neutral Point-of-View Generation with Data- and Parameter-Efficient RL
por: Hoffmann, Jessica, et al.
Publicado: (2025) -
Branch-Train-MiX: Mixing Expert LLMs into a Mixture-of-Experts LLM
por: Sukhbaatar, Sainbayar, et al.
Publicado: (2024)