Pangu Ultra: Pushing the Limits of Dense Large Language Models on Ascend NPUs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yin, Yichun, Huang, Wenyong, Song, Kaikai, Tang, Yehui, Wu, Xueyu, Guo, Wei, Guo, Peng, Wang, Yaoyuan, Meng, Xiaojun, Wang, Yasheng, Li, Dong, Chen, Can, Tu, Dandan, Li, Yin, Yu, Fisher, Tang, Ruiming, Wang, Yunhe, Wang, Baojun, Wang, Bin, Wang, Bo, Liu, Boxiao, Zhang, Changzheng, Tang, Duyu, Mi, Fei, Jin, Hui, Wei, Jiansheng, Qin, Jiarui, Li, Jinpeng, Zhao, Jun, Deng, Liqun, Li, Lin, Xu, Minghui, Zhang, Naifu, Zheng, Nianzu, Li, Qiang, Ruan, Rongju, Cheng, Shengjun, Guo, Tianyu, He, Wei, Li, Wei, Liu, Weiwen, Liu, Wulong, Dai, Xinyi, Dong, Yonghan, Pan, Yu, Li, Yue, Wang, Yufei, Li, Yujun, Ni, Yunsheng, Liu, Zhe, Zhang, Zhenhe, Liu, Zhicheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pangu Ultra MoE: How to Train Your Big MoE on Ascend NPUs
von: Tang, Yehui, et al.
Veröffentlicht: (2025)
von: Tang, Yehui, et al.
Veröffentlicht: (2025)
Pangu Light: Weight Re-Initialization for Pruning and Accelerating LLMs
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
EAGLE-Pangu: Accelerator-Safe Tree Speculative Decoding on Ascend NPUs
von: Han, Chang, et al.
Veröffentlicht: (2026)
von: Han, Chang, et al.
Veröffentlicht: (2026)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
von: Tang, Yehui, et al.
Veröffentlicht: (2025)
von: Tang, Yehui, et al.
Veröffentlicht: (2025)
HiFloat4 Format for Language Model Pre-training on Ascend NPUs
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
OSUM-Pangu: An Open-Source Multidimension Speech Understanding Foundation Model Built upon OpenPangu on Ascend NPUs
von: Liao, Yujie, et al.
Veröffentlicht: (2026)
von: Liao, Yujie, et al.
Veröffentlicht: (2026)
Humanity's Last Code Exam: Can Advanced LLMs Conquer Human's Hardest Code Competition?
von: Li, Xiangyang, et al.
Veröffentlicht: (2025)
von: Li, Xiangyang, et al.
Veröffentlicht: (2025)
A novel porous titanium with engineered surface for bone defect repair in load‐bearing position
von: Wei Liu, et al.
Veröffentlicht: (2024)
von: Wei Liu, et al.
Veröffentlicht: (2024)
Pangu Embedded: An Efficient Dual-system LLM Reasoner with Metacognition
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
A-IO: Adaptive Inference Orchestration for Memory-Bound NPUs
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
von: Zhang, Chen, et al.
Veröffentlicht: (2026)
Adaptive Tool Use in Large Language Models with Meta-Cognition Trigger
von: Li, Wenjun, et al.
Veröffentlicht: (2025)
von: Li, Wenjun, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models in Visual Speech Recognition: Model Scaling, Context-Aware Decoding, and Iterative Polishing
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
Clinical factors predicting the rate of cognitive decline in a US memory clinic: An electronic health record study
von: Yuchan Wang, et al.
Veröffentlicht: (2025)
von: Yuchan Wang, et al.
Veröffentlicht: (2025)
MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models
von: Wang, Dianyi, et al.
Veröffentlicht: (2025)
von: Wang, Dianyi, et al.
Veröffentlicht: (2025)
Mixture of Lookup Experts
von: Jie, Shibo, et al.
Veröffentlicht: (2025)
von: Jie, Shibo, et al.
Veröffentlicht: (2025)
Deep Lookup Network
von: Guo, Yulan, et al.
Veröffentlicht: (2025)
von: Guo, Yulan, et al.
Veröffentlicht: (2025)
EPD-Serve: A Flexible Multimodal EPD Disaggregation Inference Serving System On Ascend
von: Bai, Fan, et al.
Veröffentlicht: (2026)
von: Bai, Fan, et al.
Veröffentlicht: (2026)
Not All Models Suit Expert Offloading: On Local Routing Consistency of Mixture-of-Expert Models
von: Liang, Jingcong, et al.
Veröffentlicht: (2025)
von: Liang, Jingcong, et al.
Veröffentlicht: (2025)
Entropy Law: The Story Behind Data Compression and LLM Performance
von: Yin, Mingjia, et al.
Veröffentlicht: (2024)
von: Yin, Mingjia, et al.
Veröffentlicht: (2024)
Topological Sequence Analysis of Genomes: Delta Complex approaches
von: Liu, Jian, et al.
Veröffentlicht: (2025)
von: Liu, Jian, et al.
Veröffentlicht: (2025)
Detection, Attribution and Projection of Precipitation Structure Changes Over Northwest China
von: Shigen Wang, et al.
Veröffentlicht: (2026)
von: Shigen Wang, et al.
Veröffentlicht: (2026)
Integrated Analysis of the Immune Infiltration Pattern and Novel Diagnostic Biomarkers in Septic Cardiomyopathy
von: Wei Liu, et al.
Veröffentlicht: (2026)
von: Wei Liu, et al.
Veröffentlicht: (2026)
Soil Carbonate Dominates Calcium‐Bound Organic Carbon Storage at the Continental Scale
von: Li Tang, et al.
Veröffentlicht: (2025)
von: Li Tang, et al.
Veröffentlicht: (2025)
Yttrium‐Induced Modulation of Solute Partition and σ‐Phase Suppression in the Solidification of 7Mo Super‐Austenitic Stainless Steel
von: Wenqiang Liu, et al.
Veröffentlicht: (2025)
von: Wenqiang Liu, et al.
Veröffentlicht: (2025)
Large‐Scale, Stretchable, Self‐Protective, and Multifunctional Perovskite Luminescent Filament with Ultra‐High Stability
von: Liyan Yang, et al.
Veröffentlicht: (2024)
von: Liyan Yang, et al.
Veröffentlicht: (2024)
LLM-Based Scientific Equation Discovery via Physics-Informed Token-Regularized Policy Optimization
von: Wang, Boxiao, et al.
Veröffentlicht: (2026)
von: Wang, Boxiao, et al.
Veröffentlicht: (2026)
Advances and Frontiers of LLM-based Issue Resolution in Software Engineering: A Comprehensive Survey
von: Li, Caihua, et al.
Veröffentlicht: (2026)
von: Li, Caihua, et al.
Veröffentlicht: (2026)
PICC Maintenance Challenges Among Esophageal Cancer Patients: A Qualitative Study
von: Guo Huanfei, et al.
Veröffentlicht: (2025)
von: Guo Huanfei, et al.
Veröffentlicht: (2025)
Rethinking Video Tokenization: A Conditioned Diffusion-based Approach
von: Yang, Nianzu, et al.
Veröffentlicht: (2025)
von: Yang, Nianzu, et al.
Veröffentlicht: (2025)
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
von: Chen, Xinxi, et al.
Veröffentlicht: (2024)
von: Chen, Xinxi, et al.
Veröffentlicht: (2024)
The Exploration of Error Bounds in Classification with Noisy Labels
von: Liu, Haixia, et al.
Veröffentlicht: (2025)
von: Liu, Haixia, et al.
Veröffentlicht: (2025)
Generalized spectral characterization of signed bipartite graphs
von: Guo, Songlin, et al.
Veröffentlicht: (2025)
von: Guo, Songlin, et al.
Veröffentlicht: (2025)
Balanced Multi-view Clustering
von: Li, Zhenglai, et al.
Veröffentlicht: (2025)
von: Li, Zhenglai, et al.
Veröffentlicht: (2025)
The Lepton Flavor Changing Decays and One-loop Muon Anomalous Magnetic Moment in the Extended Mirror Twin Higgs Models
von: Liu, Guo-Li, et al.
Veröffentlicht: (2020)
von: Liu, Guo-Li, et al.
Veröffentlicht: (2020)
ChatGPT for Computational Topology
von: Liu, Jian, et al.
Veröffentlicht: (2023)
von: Liu, Jian, et al.
Veröffentlicht: (2023)
Computing Khovanov homology of tangles
von: Shen, Li, et al.
Veröffentlicht: (2025)
von: Shen, Li, et al.
Veröffentlicht: (2025)
Khovanov homology of tangles: algorithm and computation
von: Shen, Li, et al.
Veröffentlicht: (2025)
von: Shen, Li, et al.
Veröffentlicht: (2025)
Persistent Khovanov homology of tangles
von: Liu, Jian, et al.
Veröffentlicht: (2024)
von: Liu, Jian, et al.
Veröffentlicht: (2024)
Evolutionary Khovanov homology
von: Shen, Li, et al.
Veröffentlicht: (2024)
von: Shen, Li, et al.
Veröffentlicht: (2024)
Multi-Granularity Semantic Revision for Large Language Model Distillation
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Pangu Ultra MoE: How to Train Your Big MoE on Ascend NPUs
von: Tang, Yehui, et al.
Veröffentlicht: (2025) -
Pangu Light: Weight Re-Initialization for Pruning and Accelerating LLMs
von: Chen, Hanting, et al.
Veröffentlicht: (2025) -
EAGLE-Pangu: Accelerator-Safe Tree Speculative Decoding on Ascend NPUs
von: Han, Chang, et al.
Veröffentlicht: (2026) -
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
von: Tang, Yehui, et al.
Veröffentlicht: (2025) -
HiFloat4 Format for Language Model Pre-training on Ascend NPUs
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)