Gradient Based Method for the Fusion of Lattice Quantizers
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Liyuan, Cao, Hanzhong, Li, Jiaheng, Yu, Minyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Numerion: A Multi-Hypercomplex Model for Time Series Forecasting
di: Cao, Hanzhong, et al.
Pubblicazione: (2025)
di: Cao, Hanzhong, et al.
Pubblicazione: (2025)
Repetitive Contrastive Learning Enhances Mamba's Selectivity in Time Series Prediction
di: Yan, Wenbo, et al.
Pubblicazione: (2025)
di: Yan, Wenbo, et al.
Pubblicazione: (2025)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
di: Zhang, Xi, et al.
Pubblicazione: (2025)
di: Zhang, Xi, et al.
Pubblicazione: (2025)
Gradient-Free Training of Quantized Neural Networks
di: Cohen, Noa, et al.
Pubblicazione: (2024)
di: Cohen, Noa, et al.
Pubblicazione: (2024)
An Integrated Fusion Framework for Ensemble Learning Leveraging Gradient Boosting and Fuzzy Rule-Based Models
di: Li, Jinbo, et al.
Pubblicazione: (2025)
di: Li, Jinbo, et al.
Pubblicazione: (2025)
Logit Dynamics in Softmax Policy Gradient Methods
di: Li, Yingru
Pubblicazione: (2025)
di: Li, Yingru
Pubblicazione: (2025)
Theory-optimal Quantization Based on Flatness
di: Huang, Xiusheng, et al.
Pubblicazione: (2026)
di: Huang, Xiusheng, et al.
Pubblicazione: (2026)
CrossQuant: A Post-Training Quantization Method with Smaller Quantization Kernel for Precise Large Language Model Compression
di: Liu, Wenyuan, et al.
Pubblicazione: (2024)
di: Liu, Wenyuan, et al.
Pubblicazione: (2024)
Beyond Importance Sampling: Rejection-Gated Policy Optimization
di: Sun, Ziwu, et al.
Pubblicazione: (2026)
di: Sun, Ziwu, et al.
Pubblicazione: (2026)
NestQuant: Nested Lattice Quantization for Matrix Products and LLMs
di: Savkin, Semyon, et al.
Pubblicazione: (2025)
di: Savkin, Semyon, et al.
Pubblicazione: (2025)
Collaborative Management for Chronic Diseases and Depression: A Double Heterogeneity-based Multi-Task Learning Method
di: Chai, Yidong, et al.
Pubblicazione: (2025)
di: Chai, Yidong, et al.
Pubblicazione: (2025)
FedDRL: A Trustworthy Federated Learning Model Fusion Method Based on Staged Reinforcement Learning
di: Chen, Leiming, et al.
Pubblicazione: (2023)
di: Chen, Leiming, et al.
Pubblicazione: (2023)
A Bayesian Hybrid Parameter-Efficient Fine-Tuning Method for Large Language Models
di: Chai, Yidong, et al.
Pubblicazione: (2025)
di: Chai, Yidong, et al.
Pubblicazione: (2025)
Quantifying Multimodal Capabilities: Formal Generalization Guarantees in Pairwise Metric Learning
di: Zhou, Richeng, et al.
Pubblicazione: (2026)
di: Zhou, Richeng, et al.
Pubblicazione: (2026)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
di: Barakat, Anas, et al.
Pubblicazione: (2024)
di: Barakat, Anas, et al.
Pubblicazione: (2024)
GWQ: Gradient-Aware Weight Quantization for Large Language Models
di: Shao, Yihua, et al.
Pubblicazione: (2024)
di: Shao, Yihua, et al.
Pubblicazione: (2024)
Optimize Weight Rounding via Signed Gradient Descent for the Quantization of LLMs
di: Cheng, Wenhua, et al.
Pubblicazione: (2023)
di: Cheng, Wenhua, et al.
Pubblicazione: (2023)
FAQ: Mitigating Quantization Error via Regenerating Calibration Data with Family-Aware Quantization
di: Xiao, Haiyang, et al.
Pubblicazione: (2026)
di: Xiao, Haiyang, et al.
Pubblicazione: (2026)
Gradient Flow Drifting: Generative Modeling via Wasserstein Gradient Flows of KDE-Approximated Divergences
di: Cao, Jiarui, et al.
Pubblicazione: (2026)
di: Cao, Jiarui, et al.
Pubblicazione: (2026)
Learning General Policies with Policy Gradient Methods
di: Ståhlberg, Simon, et al.
Pubblicazione: (2025)
di: Ståhlberg, Simon, et al.
Pubblicazione: (2025)
Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models
di: Chen, Kejia, et al.
Pubblicazione: (2025)
di: Chen, Kejia, et al.
Pubblicazione: (2025)
PoGO: A Scalable Proof of Useful Work via Quantized Gradient Descent and Merkle Proofs
di: Orlicki, José I.
Pubblicazione: (2025)
di: Orlicki, José I.
Pubblicazione: (2025)
Understanding the Generalization of Stochastic Gradient Adam in Learning Neural Networks
di: Tang, Xuan, et al.
Pubblicazione: (2025)
di: Tang, Xuan, et al.
Pubblicazione: (2025)
GradientStabilizer:Fix the Norm, Not the Gradient
di: Huang, Tianjin, et al.
Pubblicazione: (2025)
di: Huang, Tianjin, et al.
Pubblicazione: (2025)
CN-Buzz2Portfolio: A Chinese-Market Dataset and Benchmark for LLM-Based Macro and Sector Asset Allocation from Daily Trending Financial News
di: Chen, Liyuan, et al.
Pubblicazione: (2026)
di: Chen, Liyuan, et al.
Pubblicazione: (2026)
AsymKV: Enabling 1-Bit Quantization of KV Cache with Layer-Wise Asymmetric Quantization Configurations
di: Tao, Qian, et al.
Pubblicazione: (2024)
di: Tao, Qian, et al.
Pubblicazione: (2024)
QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks
di: Tseng, Albert, et al.
Pubblicazione: (2024)
di: Tseng, Albert, et al.
Pubblicazione: (2024)
Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems
di: Hu, Liyuan
Pubblicazione: (2025)
di: Hu, Liyuan
Pubblicazione: (2025)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
di: Kar, Avik, et al.
Pubblicazione: (2026)
di: Kar, Avik, et al.
Pubblicazione: (2026)
Matrix Low-Rank Approximation For Policy Gradient Methods
di: Rozada, Sergio, et al.
Pubblicazione: (2024)
di: Rozada, Sergio, et al.
Pubblicazione: (2024)
Policy Gradient Methods in the Presence of Symmetries and State Abstractions
di: Panangaden, Prakash, et al.
Pubblicazione: (2023)
di: Panangaden, Prakash, et al.
Pubblicazione: (2023)
An Effective Dynamic Gradient Calibration Method for Continual Learning
di: Lin, Weichen, et al.
Pubblicazione: (2024)
di: Lin, Weichen, et al.
Pubblicazione: (2024)
XGrad: Boosting Gradient-Based Optimizers With Weight Prediction
di: Guan, Lei, et al.
Pubblicazione: (2023)
di: Guan, Lei, et al.
Pubblicazione: (2023)
PTQ1.61: Push the Real Limit of Extremely Low-Bit Post-Training Quantization Methods for Large Language Models
di: Zhao, Jiaqi, et al.
Pubblicazione: (2025)
di: Zhao, Jiaqi, et al.
Pubblicazione: (2025)
RWKVQuant: Quantizing the RWKV Family with Proxy Guided Hybrid of Scalar and Vector Quantization
di: Xu, Chen, et al.
Pubblicazione: (2025)
di: Xu, Chen, et al.
Pubblicazione: (2025)
A Depression Detection Method Based on Multi-Modal Feature Fusion Using Cross-Attention
di: Li, Shengjie, et al.
Pubblicazione: (2024)
di: Li, Shengjie, et al.
Pubblicazione: (2024)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
di: Yu, Xiaoming, et al.
Pubblicazione: (2026)
di: Yu, Xiaoming, et al.
Pubblicazione: (2026)
Mollification Effects of Policy Gradient Methods
di: Wang, Tao, et al.
Pubblicazione: (2024)
di: Wang, Tao, et al.
Pubblicazione: (2024)
ODICE: Revealing the Mystery of Distribution Correction Estimation via Orthogonal-gradient Update
di: Mao, Liyuan, et al.
Pubblicazione: (2024)
di: Mao, Liyuan, et al.
Pubblicazione: (2024)
The Curse and Blessing of Mean Bias in FP4-Quantized LLM Training
di: Cao, Hengjie, et al.
Pubblicazione: (2026)
di: Cao, Hengjie, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Numerion: A Multi-Hypercomplex Model for Time Series Forecasting
di: Cao, Hanzhong, et al.
Pubblicazione: (2025) -
Repetitive Contrastive Learning Enhances Mamba's Selectivity in Time Series Prediction
di: Yan, Wenbo, et al.
Pubblicazione: (2025) -
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
di: Zhang, Xi, et al.
Pubblicazione: (2025) -
Gradient-Free Training of Quantized Neural Networks
di: Cohen, Noa, et al.
Pubblicazione: (2024) -
An Integrated Fusion Framework for Ensemble Learning Leveraging Gradient Boosting and Fuzzy Rule-Based Models
di: Li, Jinbo, et al.
Pubblicazione: (2025)