A Case Study of Selected PTQ Baselines for Reasoning LLMs on Ascend NPU
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Yuchen, Zhu, Fangyue, Zhou, Ruining, Huang, Mingzhe, Zhu, Jian, Fan, Fanyu, Shao, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization
von: Wu, Jiehao, et al.
Veröffentlicht: (2026)
von: Wu, Jiehao, et al.
Veröffentlicht: (2026)
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs
von: Luo, Hang, et al.
Veröffentlicht: (2025)
von: Luo, Hang, et al.
Veröffentlicht: (2025)
SLAM: Towards Efficient Multilingual Reasoning via Selective Language Alignment
von: Fan, Yuchun, et al.
Veröffentlicht: (2025)
von: Fan, Yuchun, et al.
Veröffentlicht: (2025)
From Answers to Questions: EQGBench for Evaluating LLMs' Educational Question Generation
von: Zhou, Chengliang, et al.
Veröffentlicht: (2025)
von: Zhou, Chengliang, et al.
Veröffentlicht: (2025)
Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
von: Chen, Rubing, et al.
Veröffentlicht: (2025)
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task
von: Yan, Yuchen, et al.
Veröffentlicht: (2025)
von: Yan, Yuchen, et al.
Veröffentlicht: (2025)
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
von: Chen, Mingyang, et al.
Veröffentlicht: (2025)
MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster
von: Feng, Laingjun, et al.
Veröffentlicht: (2025)
von: Feng, Laingjun, et al.
Veröffentlicht: (2025)
Simulated Annealing Enhances Theory-of-Mind Reasoning in Autoregressive Language Models
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
Dissecting Failure Dynamics in Large Language Model Reasoning
von: Zhu, Wei, et al.
Veröffentlicht: (2026)
von: Zhu, Wei, et al.
Veröffentlicht: (2026)
InftyThink: Breaking the Length Limits of Long-Context Reasoning in Large Language Models
von: Yan, Yuchen, et al.
Veröffentlicht: (2025)
von: Yan, Yuchen, et al.
Veröffentlicht: (2025)
Benchmarking Contextual and Paralinguistic Reasoning in Speech-LLMs: A Case Study with In-the-Wild Data
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2025)
von: Wang, Qiongqiong, et al.
Veröffentlicht: (2025)
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions
von: Hong, Zijin, et al.
Veröffentlicht: (2025)
von: Hong, Zijin, et al.
Veröffentlicht: (2025)
Pangu Ultra: Pushing the Limits of Dense Large Language Models on Ascend NPUs
von: Yin, Yichun, et al.
Veröffentlicht: (2025)
von: Yin, Yichun, et al.
Veröffentlicht: (2025)
CCR-Bench: A Comprehensive Benchmark for Evaluating LLMs on Complex Constraints, Control Flows, and Real-World Cases
von: Xue, Xiaona, et al.
Veröffentlicht: (2026)
von: Xue, Xiaona, et al.
Veröffentlicht: (2026)
Reasoning Topology Matters: Network-of-Thought for Complex Reasoning Tasks
von: Huang, Fan
Veröffentlicht: (2026)
von: Huang, Fan
Veröffentlicht: (2026)
InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning
von: Yan, Yuchen, et al.
Veröffentlicht: (2026)
von: Yan, Yuchen, et al.
Veröffentlicht: (2026)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
von: Chen, Zui, et al.
Veröffentlicht: (2024)
von: Chen, Zui, et al.
Veröffentlicht: (2024)
Tokenization Constraints in LLMs: A Study of Symbolic and Arithmetic Reasoning Limits
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Optimizing Large Language Models through Quantization: A Comparative Analysis of PTQ and QAT Techniques
von: Hasan, Jahid
Veröffentlicht: (2024)
von: Hasan, Jahid
Veröffentlicht: (2024)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
von: Banerjee, Tanushree, et al.
Veröffentlicht: (2024)
von: Banerjee, Tanushree, et al.
Veröffentlicht: (2024)
Select2Reason: Efficient Instruction-Tuning Data Selection for Long-CoT Reasoning
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
von: Yang, Cehao, et al.
Veröffentlicht: (2025)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
von: Li, Yangning, et al.
Veröffentlicht: (2025)
von: Li, Yangning, et al.
Veröffentlicht: (2025)
Is Depth All You Need? An Exploration of Iterative Reasoning in LLMs
von: Wu, Zongqian, et al.
Veröffentlicht: (2025)
von: Wu, Zongqian, et al.
Veröffentlicht: (2025)
S^3cMath: Spontaneous Step-level Self-correction Makes Large Language Models Better Mathematical Reasoners
von: Yan, Yuchen, et al.
Veröffentlicht: (2024)
von: Yan, Yuchen, et al.
Veröffentlicht: (2024)
SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese
von: Xu, Liang, et al.
Veröffentlicht: (2024)
von: Xu, Liang, et al.
Veröffentlicht: (2024)
A$^2$FM: An Adaptive Agent Foundation Model for Tool-Aware Hybrid Reasoning
von: Chen, Qianben, et al.
Veröffentlicht: (2025)
von: Chen, Qianben, et al.
Veröffentlicht: (2025)
How Do LLMs Perform Two-Hop Reasoning in Context?
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
Key-Point-Driven Mathematical Reasoning Distillation of Large Language Model
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xunyu, et al.
Veröffentlicht: (2024)
TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenarios
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
von: Wei, Shaohang, et al.
Veröffentlicht: (2025)
HiFloat4 Format for Language Model Pre-training on Ascend NPUs
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
von: Taghian, Mehran, et al.
Veröffentlicht: (2026)
Capabilities of GPT-5 on Multimodal Medical Reasoning
von: Wang, Shansong, et al.
Veröffentlicht: (2025)
von: Wang, Shansong, et al.
Veröffentlicht: (2025)
PathCoT: Chain-of-Thought Prompting for Zero-shot Pathology Visual Reasoning
von: Zhou, Junjie, et al.
Veröffentlicht: (2025)
von: Zhou, Junjie, et al.
Veröffentlicht: (2025)
Towards Foundation Models for Knowledge Graph Reasoning
von: Galkin, Mikhail, et al.
Veröffentlicht: (2023)
von: Galkin, Mikhail, et al.
Veröffentlicht: (2023)
LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization
von: Wu, Xingyu, et al.
Veröffentlicht: (2025)
von: Wu, Xingyu, et al.
Veröffentlicht: (2025)
Discerning minds or generic tutors? Evaluating instructional guidance capabilities in Socratic LLMs
von: Liu, Ying, et al.
Veröffentlicht: (2025)
von: Liu, Ying, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization
von: Wu, Jiehao, et al.
Veröffentlicht: (2026) -
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025) -
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026) -
MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer
von: Lin, Honglin, et al.
Veröffentlicht: (2025) -
Causal Graphs Meet Thoughts: Enhancing Complex Reasoning in Graph-Augmented LLMs
von: Luo, Hang, et al.
Veröffentlicht: (2025)