Neural Additive Experts: Context-Gated Experts for Controllable Model Additivity
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiong, Guangzhi, Sinha, Sanchit, Zhang, Aidong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ProtoNAM: Prototypical Neural Additive Models for Interpretable Deep Tabular Learning
por: Xiong, Guangzhi, et al.
Publicado: (2024)
por: Xiong, Guangzhi, et al.
Publicado: (2024)
CoLiDR: Concept Learning using Aggregated Disentangled Representations
por: Sinha, Sanchit, et al.
Publicado: (2024)
por: Sinha, Sanchit, et al.
Publicado: (2024)
A Comprehensive Survey on the Risks and Limitations of Concept-based Models
por: Sinha, Sanchit, et al.
Publicado: (2025)
por: Sinha, Sanchit, et al.
Publicado: (2025)
A Self-explaining Neural Architecture for Generalizable Concept Learning
por: Sinha, Sanchit, et al.
Publicado: (2024)
por: Sinha, Sanchit, et al.
Publicado: (2024)
Structural Causality-based Generalizable Concept Discovery Models
por: Sinha, Sanchit, et al.
Publicado: (2024)
por: Sinha, Sanchit, et al.
Publicado: (2024)
Retrieving Counterfactuals Improves Visual In-Context Learning
por: Xiong, Guangzhi, et al.
Publicado: (2026)
por: Xiong, Guangzhi, et al.
Publicado: (2026)
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
por: Sinha, Sanchit, et al.
Publicado: (2025)
por: Sinha, Sanchit, et al.
Publicado: (2025)
Concept-RuleNet: Grounded Multi-Agent Neurosymbolic Reasoning in Vision Language Models
por: Sinha, Sanchit, et al.
Publicado: (2025)
por: Sinha, Sanchit, et al.
Publicado: (2025)
Reasoning Beyond Chain-of-Thought: A Latent Computational Mode in Large Language Models
por: He, Zhenghao, et al.
Publicado: (2026)
por: He, Zhenghao, et al.
Publicado: (2026)
Toward Faithful Retrieval-Augmented Generation with Sparse Autoencoders
por: Xiong, Guangzhi, et al.
Publicado: (2025)
por: Xiong, Guangzhi, et al.
Publicado: (2025)
Gaussian Process Neural Additive Models
por: Zhang, Wei, et al.
Publicado: (2024)
por: Zhang, Wei, et al.
Publicado: (2024)
CASL: Concept-Aligned Sparse Latents for Interpreting Diffusion Models
por: He, Zhenghao, et al.
Publicado: (2026)
por: He, Zhenghao, et al.
Publicado: (2026)
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models
por: Xiong, Guangzhi, et al.
Publicado: (2026)
por: Xiong, Guangzhi, et al.
Publicado: (2026)
Hierarchically Gated Experts for Efficient Online Continual Learning
por: Luong, Kevin, et al.
Publicado: (2024)
por: Luong, Kevin, et al.
Publicado: (2024)
Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts
por: Nguyen, Xuan-Phi, et al.
Publicado: (2026)
por: Nguyen, Xuan-Phi, et al.
Publicado: (2026)
The Interpretable and Effective Graph Neural Additive Networks
por: Bechler-Speicher, Maya, et al.
Publicado: (2024)
por: Bechler-Speicher, Maya, et al.
Publicado: (2024)
AnyExperts: On-Demand Expert Allocation for Multimodal Language Models with Mixture of Expert
por: Gao, Yuting, et al.
Publicado: (2025)
por: Gao, Yuting, et al.
Publicado: (2025)
Tensor Polynomial Additive Model
por: Chen, Yang, et al.
Publicado: (2024)
por: Chen, Yang, et al.
Publicado: (2024)
Mixture-of-Experts Meets In-Context Reinforcement Learning
por: Wu, Wenhao, et al.
Publicado: (2025)
por: Wu, Wenhao, et al.
Publicado: (2025)
Fate: Fast Edge Inference of Mixture-of-Experts Models via Cross-Layer Gate
por: Fang, Zhiyuan, et al.
Publicado: (2025)
por: Fang, Zhiyuan, et al.
Publicado: (2025)
Efficiently Editing Mixture-of-Experts Models with Compressed Experts
por: He, Yifei, et al.
Publicado: (2025)
por: He, Yifei, et al.
Publicado: (2025)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
por: Xie, Zhitian, et al.
Publicado: (2024)
por: Xie, Zhitian, et al.
Publicado: (2024)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
por: Zou, Will Y., et al.
Publicado: (2025)
por: Zou, Will Y., et al.
Publicado: (2025)
COCO-Tree: Compositional Hierarchical Concept Trees for Enhanced Reasoning in Vision Language Models
por: Sinha, Sanchit, et al.
Publicado: (2025)
por: Sinha, Sanchit, et al.
Publicado: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
Multi-Task Learning with Additive U-Net for Image Denoising and Classification
por: Lakkavalli, Vikram, et al.
Publicado: (2026)
por: Lakkavalli, Vikram, et al.
Publicado: (2026)
Wavelet Mixture of Experts for Time Series Forecasting
por: Zhou, Zheng, et al.
Publicado: (2025)
por: Zhou, Zheng, et al.
Publicado: (2025)
Modeling Spatio-temporal Dynamical Systems with Neural Discrete Learning and Levels-of-Experts
por: Wang, Kun, et al.
Publicado: (2024)
por: Wang, Kun, et al.
Publicado: (2024)
Optimal Expert-Attention Allocation in Mixture-of-Experts: A Scalable Law for Dynamic Model Design
por: Li, Junzhuo, et al.
Publicado: (2026)
por: Li, Junzhuo, et al.
Publicado: (2026)
Not All Models Suit Expert Offloading: On Local Routing Consistency of Mixture-of-Expert Models
por: Liang, Jingcong, et al.
Publicado: (2025)
por: Liang, Jingcong, et al.
Publicado: (2025)
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
por: Wei, Jia, et al.
Publicado: (2026)
por: Wei, Jia, et al.
Publicado: (2026)
Meta Additive Model: Interpretable Sparse Learning With Auto Weighting
por: Zhang, Xuelin, et al.
Publicado: (2026)
por: Zhang, Xuelin, et al.
Publicado: (2026)
Speculating Experts Accelerates Inference for Mixture-of-Experts
por: Madan, Vivan, et al.
Publicado: (2026)
por: Madan, Vivan, et al.
Publicado: (2026)
Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
por: Lu, Xudong, et al.
Publicado: (2024)
por: Lu, Xudong, et al.
Publicado: (2024)
Mixture of Experts in Large Language Models
por: Zhang, Danyang, et al.
Publicado: (2025)
por: Zhang, Danyang, et al.
Publicado: (2025)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
por: Yang, Cheng, et al.
Publicado: (2024)
por: Yang, Cheng, et al.
Publicado: (2024)
Graph Mixing Additive Networks
por: Bechler-Speicher, Maya, et al.
Publicado: (2025)
por: Bechler-Speicher, Maya, et al.
Publicado: (2025)
Dynamic Expert Quantization for Scalable Mixture-of-Experts Inference
por: Chu, Kexin, et al.
Publicado: (2025)
por: Chu, Kexin, et al.
Publicado: (2025)
Discrete-Choice Model with Generalized Additive Utility Network
por: Nishi, Tomoki, et al.
Publicado: (2023)
por: Nishi, Tomoki, et al.
Publicado: (2023)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
por: Chen, Yuanteng, et al.
Publicado: (2025)
por: Chen, Yuanteng, et al.
Publicado: (2025)
Ejemplares similares
-
ProtoNAM: Prototypical Neural Additive Models for Interpretable Deep Tabular Learning
por: Xiong, Guangzhi, et al.
Publicado: (2024) -
CoLiDR: Concept Learning using Aggregated Disentangled Representations
por: Sinha, Sanchit, et al.
Publicado: (2024) -
A Comprehensive Survey on the Risks and Limitations of Concept-based Models
por: Sinha, Sanchit, et al.
Publicado: (2025) -
A Self-explaining Neural Architecture for Generalizable Concept Learning
por: Sinha, Sanchit, et al.
Publicado: (2024) -
Structural Causality-based Generalizable Concept Discovery Models
por: Sinha, Sanchit, et al.
Publicado: (2024)