Unlocking the Pre-Trained Model as a Dual-Alignment Calibrator for Post-Trained LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Beier, Wang, Cheng, Wei, Hongxin, Li, Sharon, Du, Xuefeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
von: Luo, Beier, et al.
Veröffentlicht: (2025)
von: Luo, Beier, et al.
Veröffentlicht: (2025)
Rethinking Post-Training Quantization: Introducing a Statistical Pre-Calibration Approach
von: Ghaffari, Alireza, et al.
Veröffentlicht: (2025)
von: Ghaffari, Alireza, et al.
Veröffentlicht: (2025)
Training with Fewer Bits: Unlocking Edge LLMs Training with Stochastic Rounding
von: Liu, Taowen, et al.
Veröffentlicht: (2025)
von: Liu, Taowen, et al.
Veröffentlicht: (2025)
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
von: Zhang, Jingxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Jingxuan, et al.
Veröffentlicht: (2026)
Provable Training Data Identification for Large Language Models
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
Unlocking [CLS] Features for Continual Post-Training
von: Yildirim, Murat Onur, et al.
Veröffentlicht: (2025)
von: Yildirim, Murat Onur, et al.
Veröffentlicht: (2025)
Heterogeneous Low-Bandwidth Pre-Training of LLMs
von: Obeidi, Yazan, et al.
Veröffentlicht: (2026)
von: Obeidi, Yazan, et al.
Veröffentlicht: (2026)
Graph Pre-Training Models Are Strong Anomaly Detectors
von: Cheng, Jiashun, et al.
Veröffentlicht: (2024)
von: Cheng, Jiashun, et al.
Veröffentlicht: (2024)
DAQ: Density-Aware Post-Training Weight-Only Quantization For LLMs
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
von: Luo, Yingsong, et al.
Veröffentlicht: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Dual Prototypes for Adaptive Pre-Trained Model in Class-Incremental Learning
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
Post-Training Statistical Calibration for Higher Activation Sparsity
von: Chua, Vui Seng, et al.
Veröffentlicht: (2024)
von: Chua, Vui Seng, et al.
Veröffentlicht: (2024)
Gradient-Aligned Calibration for Post-Training Quantization of Diffusion Models
von: Hoang, Dung Anh, et al.
Veröffentlicht: (2026)
von: Hoang, Dung Anh, et al.
Veröffentlicht: (2026)
Understanding the Difficulty of Low-Precision Post-Training Quantization for LLMs
von: Xu, Zifei, et al.
Veröffentlicht: (2024)
von: Xu, Zifei, et al.
Veröffentlicht: (2024)
Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training
von: Zhou, Wenjie, et al.
Veröffentlicht: (2026)
von: Zhou, Wenjie, et al.
Veröffentlicht: (2026)
DBellQuant: Breaking the Bell with Double-Bell Transformation for LLMs Post Training Binarization
von: Ye, Zijian, et al.
Veröffentlicht: (2025)
von: Ye, Zijian, et al.
Veröffentlicht: (2025)
Dual Consolidation for Pre-Trained Model-Based Domain-Incremental Learning
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
Self-Supervised Pre-Training for Precipitation Post-Processor
von: An, Sojung, et al.
Veröffentlicht: (2023)
von: An, Sojung, et al.
Veröffentlicht: (2023)
LLM4TS: Aligning Pre-Trained LLMs as Data-Efficient Time-Series Forecasters
von: Chang, Ching, et al.
Veröffentlicht: (2023)
von: Chang, Ching, et al.
Veröffentlicht: (2023)
The Coverage Principle: How Pre-Training Enables Post-Training
von: Chen, Fan, et al.
Veröffentlicht: (2025)
von: Chen, Fan, et al.
Veröffentlicht: (2025)
Layer Importance for Mathematical Reasoning is Forged in Pre-Training and Invariant after Post-Training
von: Nepal, Aadim, et al.
Veröffentlicht: (2025)
von: Nepal, Aadim, et al.
Veröffentlicht: (2025)
Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models
von: Hoang, Dung Anh, et al.
Veröffentlicht: (2025)
von: Hoang, Dung Anh, et al.
Veröffentlicht: (2025)
EDA-DM: Enhanced Distribution Alignment for Post-Training Quantization of Diffusion Models
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
von: Liu, Xuewen, et al.
Veröffentlicht: (2024)
Unlocking the Potential of Model Calibration in Federated Learning
von: Chu, Yun-Wei, et al.
Veröffentlicht: (2024)
von: Chu, Yun-Wei, et al.
Veröffentlicht: (2024)
CLAQ: Pushing the Limits of Low-Bit Post-Training Quantization for LLMs
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
Mixture-of-Channels: Exploiting Sparse FFNs for Efficient LLMs Pre-Training and Inference
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
AlignTune: Modular Toolkit for Post-Training Alignment of Large Language Models
von: Lyngkhoi, R E Zera Marveen, et al.
Veröffentlicht: (2026)
von: Lyngkhoi, R E Zera Marveen, et al.
Veröffentlicht: (2026)
Make Some Noise: Unlocking Language Model Parallel Inference Capability through Noisy Training
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
Better Representations via Adversarial Training in Pre-Training: A Theoretical Perspective
von: Xing, Yue, et al.
Veröffentlicht: (2024)
von: Xing, Yue, et al.
Veröffentlicht: (2024)
Achieving binary weight and activation for LLMs using Post-Training Quantization
von: Song, Siqing, et al.
Veröffentlicht: (2025)
von: Song, Siqing, et al.
Veröffentlicht: (2025)
Efficient Post-Training Pruning of Large Language Models with Statistical Correction
von: Yu, Peiqi, et al.
Veröffentlicht: (2026)
von: Yu, Peiqi, et al.
Veröffentlicht: (2026)
HG-Adapter: Improving Pre-Trained Heterogeneous Graph Neural Networks with Dual Adapters
von: Mo, Yujie, et al.
Veröffentlicht: (2024)
von: Mo, Yujie, et al.
Veröffentlicht: (2024)
Open-Vocabulary Calibration for Fine-tuned CLIP
von: Wang, Shuoyuan, et al.
Veröffentlicht: (2024)
von: Wang, Shuoyuan, et al.
Veröffentlicht: (2024)
D$^2$Quant: Accurate Low-bit Post-Training Weight Quantization for LLMs
von: Yan, Xianglong, et al.
Veröffentlicht: (2026)
von: Yan, Xianglong, et al.
Veröffentlicht: (2026)
Leveraging RAG for Training-Free Alignment of LLMs
von: Halloran, John T.
Veröffentlicht: (2026)
von: Halloran, John T.
Veröffentlicht: (2026)
Convex Dataset Valuation for Post-Training
von: Zeng, Siqi, et al.
Veröffentlicht: (2026)
von: Zeng, Siqi, et al.
Veröffentlicht: (2026)
Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning Models
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
Mitigating Spurious Correlations in LLMs via Causality-Aware Post-Training
von: Gui, Shurui, et al.
Veröffentlicht: (2025)
von: Gui, Shurui, et al.
Veröffentlicht: (2025)
Any-Depth Alignment: Unlocking Innate Safety Alignment of LLMs to Any-Depth
von: Zhang, Jiawei, et al.
Veröffentlicht: (2025)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
von: Luo, Beier, et al.
Veröffentlicht: (2025) -
Rethinking Post-Training Quantization: Introducing a Statistical Pre-Calibration Approach
von: Ghaffari, Alireza, et al.
Veröffentlicht: (2025) -
Training with Fewer Bits: Unlocking Edge LLMs Training with Stochastic Rounding
von: Liu, Taowen, et al.
Veröffentlicht: (2025) -
QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models
von: Zhang, Jingxuan, et al.
Veröffentlicht: (2026) -
Provable Training Data Identification for Large Language Models
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)