dMoE: dLLMs with Learnable Block Experts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Feng, Sicheng, Chen, Zigeng, Fang, Gongfan, Ma, Xinyin, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
dVoting: Fast Voting for dLLMs
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
von: Feng, Sicheng, et al.
Veröffentlicht: (2026)
dParallel: Learnable Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
DMax: Aggressive Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs
von: Lu, Haiquan, et al.
Veröffentlicht: (2026)
von: Lu, Haiquan, et al.
Veröffentlicht: (2026)
dKV-Cache: The Cache for Diffusion Language Models
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
Efficient Reasoning Models: A Survey
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
von: Feng, Sicheng, et al.
Veröffentlicht: (2025)
Thinkless: LLM Learns When to Think
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
SlimSAM: 0.1% Data Makes Segment Anything Slim
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Every Step Counts: Decoding Trajectories as Authorship Fingerprints of dLLMs
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
ConciseHint: Boosting Efficient Reasoning via Continuous Concise Hints during Generation
von: Tang, Siao, et al.
Veröffentlicht: (2025)
von: Tang, Siao, et al.
Veröffentlicht: (2025)
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
SparseD: Sparse Attention for Diffusion Language Models
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
MixReasoning: Switching Modes to Think
von: Lu, Haiquan, et al.
Veröffentlicht: (2025)
von: Lu, Haiquan, et al.
Veröffentlicht: (2025)
CoT-Valve: Length-Compressible Chain-of-Thought Tuning
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)
In-Video Instructions: Visual Signals as Generative Control
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
VeriThinker: Learning to Verify Makes Reasoning Model Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
von: Chen, Zigeng, et al.
Veröffentlicht: (2025)
STDec: Spatio-Temporal Stability Guided Decoding for dLLMs
von: Chen, Yuzhe, et al.
Veröffentlicht: (2026)
von: Chen, Yuzhe, et al.
Veröffentlicht: (2026)
TinyFusion: Diffusion Transformers Learned Shallow
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Quantization Meets dLLMs: A Systematic Study of Post-training Quantization for Diffusion LLMs
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
Isomorphic Pruning for Vision Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Q-ARVD: Quantizing Autoregressive Video Diffusion Models
von: Tang, Siao, et al.
Veröffentlicht: (2026)
von: Tang, Siao, et al.
Veröffentlicht: (2026)
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Invisible Safety Threat: Malicious Finetuning for LLM via Steganography
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
LiteFocus: Accelerated Diffusion Inference for Long Audio Synthesis
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
Self-Purification Mitigates Backdoors in Multimodal Diffusion Language Models
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
von: Wan, Guangnian, et al.
Veröffentlicht: (2026)
Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
LD-MoLE: Learnable Dynamic Routing for Mixture of LoRA Experts
von: Zhuang, Yuan, et al.
Veröffentlicht: (2025)
von: Zhuang, Yuan, et al.
Veröffentlicht: (2025)
GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs
von: Deng, Jianing, et al.
Veröffentlicht: (2026)
von: Deng, Jianing, et al.
Veröffentlicht: (2026)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
von: Li, Yunxin, et al.
Veröffentlicht: (2024)
Do Domain-specific Experts exist in MoE-based LLMs?
von: Do, Giang, et al.
Veröffentlicht: (2026)
von: Do, Giang, et al.
Veröffentlicht: (2026)
ROMER: Expert Replacement and Router Calibration for Robust MoE LLMs on Analog Compute-in-Memory Systems
von: Zhou, Wenyong, et al.
Veröffentlicht: (2026)
von: Zhou, Wenyong, et al.
Veröffentlicht: (2026)
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Steering MoE LLMs via Expert (De)Activation
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
MH-MoE: Multi-Head Mixture-of-Experts
von: Huang, Shaohan, et al.
Veröffentlicht: (2024)
von: Huang, Shaohan, et al.
Veröffentlicht: (2024)
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
von: Jing, Linglin, et al.
Veröffentlicht: (2025)
von: Jing, Linglin, et al.
Veröffentlicht: (2025)
MoMoE: Mixture of Moderation Experts Framework for AI-Assisted Online Governance
von: Goyal, Agam, et al.
Veröffentlicht: (2025)
von: Goyal, Agam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
dVoting: Fast Voting for dLLMs
von: Feng, Sicheng, et al.
Veröffentlicht: (2026) -
dParallel: Learnable Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2025) -
DMax: Aggressive Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2026) -
Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs
von: Lu, Haiquan, et al.
Veröffentlicht: (2026) -
dKV-Cache: The Cache for Diffusion Language Models
von: Ma, Xinyin, et al.
Veröffentlicht: (2025)