PTQ1.61: Push the Real Limit of Extremely Low-Bit Post-Training Quantization Methods for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Jiaqi, Zhang, Miao, Wang, Ming, Shang, Yuzhang, Zhang, Kaihao, Guan, Weili, Wang, Yaowei, Zhang, Min |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
PTQ4DiT: Post-training Quantization for Diffusion Transformers
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
von: Wu, Junyi, et al.
Veröffentlicht: (2024)
CLAQ: Pushing the Limits of Low-Bit Post-Training Quantization for LLMs
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
Boost Post-Training Quantization via Null Space Optimization for Large Language Models
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025)
1-Bit FQT: Pushing the Limit of Fully Quantized Training to 1-bit
von: Gao, Chang, et al.
Veröffentlicht: (2024)
von: Gao, Chang, et al.
Veröffentlicht: (2024)
LiDAR-PTQ: Post-Training Quantization for Point Cloud 3D Object Detection
von: Zhou, Sifan, et al.
Veröffentlicht: (2024)
von: Zhou, Sifan, et al.
Veröffentlicht: (2024)
Quant-dLLM: Post-Training Extreme Low-Bit Quantization for Diffusion Large Language Models
von: Zhang, Tianao, et al.
Veröffentlicht: (2025)
von: Zhang, Tianao, et al.
Veröffentlicht: (2025)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
PTQ4SAM: Post-Training Quantization for Segment Anything
von: Lv, Chengtao, et al.
Veröffentlicht: (2024)
von: Lv, Chengtao, et al.
Veröffentlicht: (2024)
PTQ4VM: Post-Training Quantization for Visual Mamba
von: Cho, Younghyun, et al.
Veröffentlicht: (2024)
von: Cho, Younghyun, et al.
Veröffentlicht: (2024)
PTQ4ARVG: Post-Training Quantization for AutoRegressive Visual Generation Models
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
von: Liu, Xuewen, et al.
Veröffentlicht: (2026)
Pushing the Limits of Low-Bit Optimizers: A Focus on EMA Dynamics
von: Xu, Cong, et al.
Veröffentlicht: (2025)
von: Xu, Cong, et al.
Veröffentlicht: (2025)
PTQ4RIS: Post-Training Quantization for Referring Image Segmentation
von: Jiang, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Jiang, Xiaoyan, et al.
Veröffentlicht: (2024)
Pushing the Limits of Block Rotations in Post-Training Quantization
von: Sanjeet, Sai, et al.
Veröffentlicht: (2026)
von: Sanjeet, Sai, et al.
Veröffentlicht: (2026)
QuantVSR: Low-Bit Post-Training Quantization for Real-World Video Super-Resolution
von: Chai, Bowen, et al.
Veröffentlicht: (2025)
von: Chai, Bowen, et al.
Veröffentlicht: (2025)
SKIM: Any-bit Quantization Pushing The Limits of Post-Training Quantization
von: Bai, Runsheng, et al.
Veröffentlicht: (2024)
von: Bai, Runsheng, et al.
Veröffentlicht: (2024)
MEC-Quant: Maximum Entropy Coding for Extremely Low Bit Quantization-Aware Training
von: Pang, Junbiao, et al.
Veröffentlicht: (2025)
von: Pang, Junbiao, et al.
Veröffentlicht: (2025)
SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs
von: Cheng, Wenhua, et al.
Veröffentlicht: (2025)
von: Cheng, Wenhua, et al.
Veröffentlicht: (2025)
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models
von: Xu, Siyuan, et al.
Veröffentlicht: (2026)
von: Xu, Siyuan, et al.
Veröffentlicht: (2026)
HBVLA: Pushing 1-Bit Post-Training Quantization for Vision-Language-Action Models
von: Yan, Xin, et al.
Veröffentlicht: (2026)
von: Yan, Xin, et al.
Veröffentlicht: (2026)
Bi-VLM: Pushing Ultra-Low Precision Post-Training Quantization Boundaries in Vision-Language Models
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Pack-PTQ: Advancing Post-training Quantization of Neural Networks by Pack-wise Reconstruction
von: Li, Changjun, et al.
Veröffentlicht: (2025)
von: Li, Changjun, et al.
Veröffentlicht: (2025)
Bits for Privacy: Evaluating Post-Training Quantization via Membership Inference
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenxiang, et al.
Veröffentlicht: (2025)
PTQ4ADM: Post-Training Quantization for Efficient Text Conditional Audio Diffusion Models
von: Vora, Jayneel, et al.
Veröffentlicht: (2024)
von: Vora, Jayneel, et al.
Veröffentlicht: (2024)
Extreme Limit Theory of Competing Risks under Power Normalization
von: Hu, Kaihao, et al.
Veröffentlicht: (2023)
von: Hu, Kaihao, et al.
Veröffentlicht: (2023)
HESTIA: A Hessian-Guided Differentiable Quantization-Aware Training Framework for Extremely Low-Bit LLMs
von: Wang, Guoan, et al.
Veröffentlicht: (2026)
von: Wang, Guoan, et al.
Veröffentlicht: (2026)
VPTQ: Extreme Low-bit Vector Post-Training Quantization for Large Language Models
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
MARR: Module-Adaptive Residual Reconstruction for Low-Bit Post-Training Quantization
von: Su, Le, et al.
Veröffentlicht: (2026)
von: Su, Le, et al.
Veröffentlicht: (2026)
BiDM: Pushing the Limit of Quantization for Diffusion Models
von: Zheng, Xingyu, et al.
Veröffentlicht: (2024)
von: Zheng, Xingyu, et al.
Veröffentlicht: (2024)
TesseraQ: Ultra Low-Bit LLM Post-Training Quantization with Block Reconstruction
von: Li, Yuhang, et al.
Veröffentlicht: (2024)
von: Li, Yuhang, et al.
Veröffentlicht: (2024)
Shedding the Bits: Pushing the Boundaries of Quantization with Minifloats on FPGAs
von: Aggarwal, Shivam, et al.
Veröffentlicht: (2023)
von: Aggarwal, Shivam, et al.
Veröffentlicht: (2023)
MPQ-DM: Mixed Precision Quantization for Extremely Low Bit Diffusion Models
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
von: Feng, Weilun, et al.
Veröffentlicht: (2024)
SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2025)
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2025)
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
LLDif: Diffusion Models for Low-light Emotion Recognition
von: Wang, Zhifeng, et al.
Veröffentlicht: (2024)
von: Wang, Zhifeng, et al.
Veröffentlicht: (2024)
Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization
von: Pavlov, Gorgi
Veröffentlicht: (2026)
von: Pavlov, Gorgi
Veröffentlicht: (2026)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
von: Zhang, Xi, et al.
Veröffentlicht: (2025)
QVGen: Pushing the Limit of Quantized Video Generative Models
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
von: Huang, Yushi, et al.
Veröffentlicht: (2025)
Robust Ultra Low-Bit Post-Training Quantization via Stable Diagonal Curvature Estimate
von: Kim, Jaemin, et al.
Veröffentlicht: (2026)
von: Kim, Jaemin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Benchmarking Post-Training Quantization in LLMs: Comprehensive Taxonomy, Unified Evaluation, and Comparative Analysis
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025) -
PTQ4DiT: Post-training Quantization for Diffusion Transformers
von: Wu, Junyi, et al.
Veröffentlicht: (2024) -
CLAQ: Pushing the Limits of Low-Bit Post-Training Quantization for LLMs
von: Wang, Haoyu, et al.
Veröffentlicht: (2024) -
Boost Post-Training Quantization via Null Space Optimization for Large Language Models
von: Zhao, Jiaqi, et al.
Veröffentlicht: (2025) -
1-Bit FQT: Pushing the Limit of Fully Quantized Training to 1-bit
von: Gao, Chang, et al.
Veröffentlicht: (2024)