AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Jiehao, Huang, Zixiao, Li, Wenhao, Shen, Chuyun, Sheng, Junjie, Wang, Xiangfeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Case Study of Selected PTQ Baselines for Reasoning LLMs on Ascend NPU
by: Luo, Yuchen, et al.
Published: (2026)
by: Luo, Yuchen, et al.
Published: (2026)
TextAtari: 100K Frames Game Playing with Language Agents
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster
by: Feng, Laingjun, et al.
Published: (2025)
by: Feng, Laingjun, et al.
Published: (2025)
Agentic Episodic Control
by: Yang, Xidong, et al.
Published: (2025)
by: Yang, Xidong, et al.
Published: (2025)
GraphThought: Graph Combinatorial Optimization with Thought Generation
by: Huang, Zixiao, et al.
Published: (2025)
by: Huang, Zixiao, et al.
Published: (2025)
Ascend HiFloat8 Format for Deep Learning
by: Luo, Yuanyong, et al.
Published: (2024)
by: Luo, Yuanyong, et al.
Published: (2024)
AscendCraft: Automatic Ascend NPU Kernel Generation via DSL-Guided Transcompilation
by: Wen, Zhongzhen, et al.
Published: (2026)
by: Wen, Zhongzhen, et al.
Published: (2026)
A Survey of Automatic Prompt Engineering: An Optimization Perspective
by: Li, Wenwu, et al.
Published: (2025)
by: Li, Wenwu, et al.
Published: (2025)
HiFloat4 Format for Language Model Pre-training on Ascend NPUs
by: Taghian, Mehran, et al.
Published: (2026)
by: Taghian, Mehran, et al.
Published: (2026)
AscendKernelGen: A Systematic Study of LLM-Based Kernel Generation for Neural Processing Units
by: Cao, Xinzi, et al.
Published: (2026)
by: Cao, Xinzi, et al.
Published: (2026)
Unleashing Low-Bit Inference on Ascend NPUs: A Comprehensive Evaluation of HiFloat Formats
by: Zhao, Pengxiang, et al.
Published: (2026)
by: Zhao, Pengxiang, et al.
Published: (2026)
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
ShadowNPU: System and Algorithm Co-design for NPU-Centric On-Device LLM Inference
by: Yin, Wangsong, et al.
Published: (2025)
by: Yin, Wangsong, et al.
Published: (2025)
Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel Synthesis
by: Zheng, Yujie, et al.
Published: (2026)
by: Zheng, Yujie, et al.
Published: (2026)
OLion: Approaching the Hadamard Ideal by Intersecting Spectral and $\ell_{\infty}$ Implicit Biases
by: Wang, Zixiao, et al.
Published: (2026)
by: Wang, Zixiao, et al.
Published: (2026)
Precision-Scalable Microscaling Datapaths with Optimized Reduction Tree for Efficient NPU Integration
by: Cuyckens, Stef, et al.
Published: (2025)
by: Cuyckens, Stef, et al.
Published: (2025)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
by: Li, Sijia, et al.
Published: (2025)
by: Li, Sijia, et al.
Published: (2025)
Quant.npu: Enabling Efficient Mobile NPU Inference for on-device LLMs via Fully Static Quantization
by: Zhang, Jinghe, et al.
Published: (2026)
by: Zhang, Jinghe, et al.
Published: (2026)
Multiobjective Hydropower Reservoir Operation Optimization with Transformer-Based Deep Reinforcement Learning
by: Wu, Rixin, et al.
Published: (2023)
by: Wu, Rixin, et al.
Published: (2023)
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation
by: Yang, Haochen, et al.
Published: (2026)
by: Yang, Haochen, et al.
Published: (2026)
Where Paths Collide: A Comprehensive Survey of Classic and Learning-Based Multi-Agent Pathfinding
by: Wang, Shiyue, et al.
Published: (2025)
by: Wang, Shiyue, et al.
Published: (2025)
Graph Classification via Reference Distribution Learning: Theory and Practice
by: Wang, Zixiao, et al.
Published: (2024)
by: Wang, Zixiao, et al.
Published: (2024)
BOTS: Batch Bayesian Optimization of Extended Thompson Sampling for Severely Episode-Limited RL Settings
by: Karine, Karine, et al.
Published: (2024)
by: Karine, Karine, et al.
Published: (2024)
MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing
by: Yao, Yinsheng, et al.
Published: (2026)
by: Yao, Yinsheng, et al.
Published: (2026)
SkyRover: A Modular Simulator for Cross-Domain Pathfinding
by: Ma, Wenhui, et al.
Published: (2025)
by: Ma, Wenhui, et al.
Published: (2025)
Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents
by: Wang, Haochen, et al.
Published: (2026)
by: Wang, Haochen, et al.
Published: (2026)
MAESTRO: Multi-Agent Environment Shaping through Task and Reward Optimization
by: Wu, Boyuan
Published: (2025)
by: Wu, Boyuan
Published: (2025)
Training Long-Context LLMs Efficiently via Chunk-wise Optimization
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Automated Generation of Diverse Courses of Actions for Multi-Agent Operations using Binary Optimization and Graph Learning
by: Poddar, Prithvi, et al.
Published: (2025)
by: Poddar, Prithvi, et al.
Published: (2025)
ESPO: Entropy Importance Sampling Policy Optimization
by: Sheng, Yuepeng, et al.
Published: (2025)
by: Sheng, Yuepeng, et al.
Published: (2025)
Adaptive Social Learning via Mode Policy Optimization for Language Agents
by: Wang, Minzheng, et al.
Published: (2025)
by: Wang, Minzheng, et al.
Published: (2025)
Reparameterization Flow Policy Optimization
by: Zhong, Hai, et al.
Published: (2026)
by: Zhong, Hai, et al.
Published: (2026)
Spectral Clustering for Discrete Distributions
by: Wang, Zixiao, et al.
Published: (2024)
by: Wang, Zixiao, et al.
Published: (2024)
Reparameterization Proximal Policy Optimization
by: Zhong, Hai, et al.
Published: (2025)
by: Zhong, Hai, et al.
Published: (2025)
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
by: Winston, Caleb, et al.
Published: (2026)
by: Winston, Caleb, et al.
Published: (2026)
Large Language Model Agent for Hyper-Parameter Optimization
by: Liu, Siyi, et al.
Published: (2024)
by: Liu, Siyi, et al.
Published: (2024)
Learning to Compress Graphs via Dual Agents for Consistent Topological Robustness Evaluation
by: Chai, Qisen, et al.
Published: (2025)
by: Chai, Qisen, et al.
Published: (2025)
Collaborative Multi-Agent Reinforcement Learning for Automated Feature Transformation with Graph-Driven Path Optimization
by: Huang, Xiaohan, et al.
Published: (2025)
by: Huang, Xiaohan, et al.
Published: (2025)
Decision Flow Policy Optimization
by: Hu, Jifeng, et al.
Published: (2025)
by: Hu, Jifeng, et al.
Published: (2025)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
by: Liu, Zongkai, et al.
Published: (2024)
by: Liu, Zongkai, et al.
Published: (2024)
Similar Items
-
A Case Study of Selected PTQ Baselines for Reasoning LLMs on Ascend NPU
by: Luo, Yuchen, et al.
Published: (2026) -
TextAtari: 100K Frames Game Playing with Language Agents
by: Li, Wenhao, et al.
Published: (2025) -
MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster
by: Feng, Laingjun, et al.
Published: (2025) -
Agentic Episodic Control
by: Yang, Xidong, et al.
Published: (2025) -
GraphThought: Graph Combinatorial Optimization with Thought Generation
by: Huang, Zixiao, et al.
Published: (2025)