Gespeichert in:
| 1. Verfasser: | Xu, Bruce Changlong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.11663 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Astro: Activation-guided Structured Regularization for Outlier-Robust LLM Post-Training Quantization
von: Chen, Xi, et al.
Veröffentlicht: (2026)
von: Chen, Xi, et al.
Veröffentlicht: (2026)
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
Post Training Quantization of Large Language Models with Microscaling Formats
von: Sharify, Sayeh, et al.
Veröffentlicht: (2024)
von: Sharify, Sayeh, et al.
Veröffentlicht: (2024)
Achieving binary weight and activation for LLMs using Post-Training Quantization
von: Song, Siqing, et al.
Veröffentlicht: (2025)
von: Song, Siqing, et al.
Veröffentlicht: (2025)
BWLA: Breaking the Barrier of W1AX Post-Training Quantization for LLMs
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
Pushing the Limits of Block Rotations in Post-Training Quantization
von: Sanjeet, Sai, et al.
Veröffentlicht: (2026)
von: Sanjeet, Sai, et al.
Veröffentlicht: (2026)
Beacon: Post-Training Quantization with Integrated Grid Selection
von: Zhang, Shihao, et al.
Veröffentlicht: (2025)
von: Zhang, Shihao, et al.
Veröffentlicht: (2025)
PTQTP: Post-Training Quantization to Trit-Planes for Large Language Models
von: Xiao, He, et al.
Veröffentlicht: (2025)
von: Xiao, He, et al.
Veröffentlicht: (2025)
A Quantized VAE-MLP Botnet Detection Model: A Systematic Evaluation of Quantization-Aware Training and Post-Training Quantization Strategies
von: Wasswa, Hassan, et al.
Veröffentlicht: (2025)
von: Wasswa, Hassan, et al.
Veröffentlicht: (2025)
MoBiE: Efficient Inference of Mixture of Binary Experts under Post-Training Quantization
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
Improving Quantization with Post-Training Model Expansion
von: Franco, Giuseppe, et al.
Veröffentlicht: (2025)
von: Franco, Giuseppe, et al.
Veröffentlicht: (2025)
Post-Training Statistical Calibration for Higher Activation Sparsity
von: Chua, Vui Seng, et al.
Veröffentlicht: (2024)
von: Chua, Vui Seng, et al.
Veröffentlicht: (2024)
Exploring Layer-wise Information Effectiveness for Post-Training Quantization in Small Language Models
von: Xiao, He, et al.
Veröffentlicht: (2025)
von: Xiao, He, et al.
Veröffentlicht: (2025)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
von: Yu, Xiaoming, et al.
Veröffentlicht: (2026)
von: Yu, Xiaoming, et al.
Veröffentlicht: (2026)
From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks
von: Xu, Bruce Changlong, et al.
Veröffentlicht: (2026)
von: Xu, Bruce Changlong, et al.
Veröffentlicht: (2026)
DAQ: Density-Aware Post-Training Weight-Only Quantization For LLMs
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
MagR: Weight Magnitude Reduction for Enhancing Post-Training Quantization
von: Zhang, Aozhong, et al.
Veröffentlicht: (2024)
von: Zhang, Aozhong, et al.
Veröffentlicht: (2024)
Towards Next-Level Post-Training Quantization of Hyper-Scale Transformers
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
von: Kim, Junhan, et al.
Veröffentlicht: (2024)
Quamba: A Post-Training Quantization Recipe for Selective State Space Models
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2024)
von: Chiang, Hung-Yueh, et al.
Veröffentlicht: (2024)
Beyond GRPO and On-Policy Distillation: An Empirical Sparse-to-Dense Reward Principle for Language-Model Post-Training
von: Xu, Yuanda, et al.
Veröffentlicht: (2026)
von: Xu, Yuanda, et al.
Veröffentlicht: (2026)
CrossQuant: A Post-Training Quantization Method with Smaller Quantization Kernel for Precise Large Language Model Compression
von: Liu, Wenyuan, et al.
Veröffentlicht: (2024)
von: Liu, Wenyuan, et al.
Veröffentlicht: (2024)
Accumulator-Aware Post-Training Quantization for Large Language Models
von: Colbert, Ian, et al.
Veröffentlicht: (2024)
von: Colbert, Ian, et al.
Veröffentlicht: (2024)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
Qronos: Correcting the Past by Shaping the Future... in Post-Training Quantization
von: Zhang, Shihao, et al.
Veröffentlicht: (2025)
von: Zhang, Shihao, et al.
Veröffentlicht: (2025)
D$^2$Quant: Accurate Low-bit Post-Training Weight Quantization for LLMs
von: Yan, Xianglong, et al.
Veröffentlicht: (2026)
von: Yan, Xianglong, et al.
Veröffentlicht: (2026)
FlexRound: Learnable Rounding based on Element-wise Division for Post-Training Quantization
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2023)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2023)
RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm
von: Yang, Yongyi, et al.
Veröffentlicht: (2025)
von: Yang, Yongyi, et al.
Veröffentlicht: (2025)
Post-Training Quantization of OpenPangu Models for Efficient Deployment on Atlas A2
von: Luo, Yilun, et al.
Veröffentlicht: (2025)
von: Luo, Yilun, et al.
Veröffentlicht: (2025)
The Coverage Principle: How Pre-Training Enables Post-Training
von: Chen, Fan, et al.
Veröffentlicht: (2025)
von: Chen, Fan, et al.
Veröffentlicht: (2025)
DNN Memory Footprint Reduction via Post-Training Intra-Layer Multi-Precision Quantization
von: Ghavami, Behnam, et al.
Veröffentlicht: (2024)
von: Ghavami, Behnam, et al.
Veröffentlicht: (2024)
Interactions Across Blocks in Post-Training Quantization of Large Language Models
von: Shabanovi, Khasmamad, et al.
Veröffentlicht: (2024)
von: Shabanovi, Khasmamad, et al.
Veröffentlicht: (2024)
Quant-dLLM: Post-Training Extreme Low-Bit Quantization for Diffusion Large Language Models
von: Zhang, Tianao, et al.
Veröffentlicht: (2025)
von: Zhang, Tianao, et al.
Veröffentlicht: (2025)
KDRL: Post-Training Reasoning LLMs via Unified Knowledge Distillation and Reinforcement Learning
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
Towards a Unified View of Large Language Model Post-Training
von: Lv, Xingtai, et al.
Veröffentlicht: (2025)
von: Lv, Xingtai, et al.
Veröffentlicht: (2025)
QuantMoE-Bench: Examining Post-Training Quantization for Mixture-of-Experts
von: Li, Pingzhi, et al.
Veröffentlicht: (2024)
von: Li, Pingzhi, et al.
Veröffentlicht: (2024)
SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models
von: Xiao, Guangxuan, et al.
Veröffentlicht: (2022)
von: Xiao, Guangxuan, et al.
Veröffentlicht: (2022)
Post-Training Quantization of Generative and Discriminative LSTM Text Classifiers: A Study of Calibration, Class Balance, and Robustness
von: Rahaman, Md Mushfiqur, et al.
Veröffentlicht: (2025)
von: Rahaman, Md Mushfiqur, et al.
Veröffentlicht: (2025)
LRQ: Optimizing Post-Training Quantization for Large Language Models by Learning Low-Rank Weight-Scaling Matrices
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
How to Parameterize Asymmetric Quantization Ranges for Quantization-Aware Training
von: You, Jaeseong, et al.
Veröffentlicht: (2024)
von: You, Jaeseong, et al.
Veröffentlicht: (2024)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Astro: Activation-guided Structured Regularization for Outlier-Robust LLM Post-Training Quantization
von: Chen, Xi, et al.
Veröffentlicht: (2026) -
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025) -
Post Training Quantization of Large Language Models with Microscaling Formats
von: Sharify, Sayeh, et al.
Veröffentlicht: (2024) -
Achieving binary weight and activation for LLMs using Post-Training Quantization
von: Song, Siqing, et al.
Veröffentlicht: (2025) -
BWLA: Breaking the Barrier of W1AX Post-Training Quantization for LLMs
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)