Compensate Quantization Errors+: Quantized Models Are Inquisitive Learners
Fuente:
arXiv
Salvato in:
| Autori principali: | Gao, Yifei, Ou, Jie, Wang, Lei, Cheng, Jun, Zhou, Mengchu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Frequency Matters: Fast Model-Agnostic Data Curation for Pruning and Quantization
di: Monaco, Francesco Pio, et al.
Pubblicazione: (2026)
di: Monaco, Francesco Pio, et al.
Pubblicazione: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024)
Detecting AI-Generated Texts in Cross-Domains
di: Zhou, You, et al.
Pubblicazione: (2024)
di: Zhou, You, et al.
Pubblicazione: (2024)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
di: Khanna, Danush, et al.
Pubblicazione: (2025)
di: Khanna, Danush, et al.
Pubblicazione: (2025)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
di: Liu, Han, et al.
Pubblicazione: (2026)
di: Liu, Han, et al.
Pubblicazione: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
Xinyu: An Efficient LLM-based System for Commentary Generation
di: Wu, Yiquan, et al.
Pubblicazione: (2024)
di: Wu, Yiquan, et al.
Pubblicazione: (2024)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
di: Liu, Han, et al.
Pubblicazione: (2026)
di: Liu, Han, et al.
Pubblicazione: (2026)
Constructing Cloze Questions Generatively
di: Sun, Yicheng, et al.
Pubblicazione: (2024)
di: Sun, Yicheng, et al.
Pubblicazione: (2024)
APIO: Automatic Prompt Induction and Optimization for Grammatical Error Correction and Text Simplification
di: Chernodub, Artem, et al.
Pubblicazione: (2025)
di: Chernodub, Artem, et al.
Pubblicazione: (2025)
PatentGPT: A Large Language Model for Intellectual Property
di: Bai, Zilong, et al.
Pubblicazione: (2024)
di: Bai, Zilong, et al.
Pubblicazione: (2024)
Search-R3: Unifying Reasoning and Embedding in Large Language Models
di: Gui, Yuntao, et al.
Pubblicazione: (2025)
di: Gui, Yuntao, et al.
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
di: Putra, Rizky Ramadhana, et al.
Pubblicazione: (2026)
di: Putra, Rizky Ramadhana, et al.
Pubblicazione: (2026)
Super Tiny Language Models
di: Hillier, Dylan, et al.
Pubblicazione: (2024)
di: Hillier, Dylan, et al.
Pubblicazione: (2024)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
di: Wen, Yuqiao, et al.
Pubblicazione: (2025)
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
di: Buszydlik, Aleksander, et al.
Pubblicazione: (2023)
A Practical Review of Mechanistic Interpretability for Transformer-Based Language Models
di: Rai, Daking, et al.
Pubblicazione: (2024)
di: Rai, Daking, et al.
Pubblicazione: (2024)
Demystifying Instruction Mixing for Fine-tuning Large Language Models
di: Wang, Renxi, et al.
Pubblicazione: (2023)
di: Wang, Renxi, et al.
Pubblicazione: (2023)
A Hierarchical Error Framework for Reliable Automated Coding in Communication Research: Applications to Health and Political Communication
di: Zhao, Zhilong, et al.
Pubblicazione: (2025)
di: Zhao, Zhilong, et al.
Pubblicazione: (2025)
Can Large Language Models Grasp Legal Theories? Enhance Legal Reasoning with Insights from Multi-Agent Collaboration
di: Yuan, Weikang, et al.
Pubblicazione: (2024)
di: Yuan, Weikang, et al.
Pubblicazione: (2024)
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
Partially Recentralization Softmax Loss for Vision-Language Models Robustness
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework
di: Chen, Jie, et al.
Pubblicazione: (2025)
di: Chen, Jie, et al.
Pubblicazione: (2025)
Streamlining Redundant Layers to Compress Large Language Models
di: Chen, Xiaodong, et al.
Pubblicazione: (2024)
di: Chen, Xiaodong, et al.
Pubblicazione: (2024)
Quantized Side Tuning: Fast and Memory-Efficient Tuning of Quantized Large Language Models
di: Zhang, Zhengxin, et al.
Pubblicazione: (2024)
di: Zhang, Zhengxin, et al.
Pubblicazione: (2024)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
di: Zhou, Yangyang, et al.
Pubblicazione: (2026)
di: Zhou, Yangyang, et al.
Pubblicazione: (2026)
Detecting Subtle Differences between Human and Model Languages Using Spectrum of Relative Likelihood
di: Xu, Yang, et al.
Pubblicazione: (2024)
di: Xu, Yang, et al.
Pubblicazione: (2024)
Low-Resource Court Judgment Summarization for Common Law Systems
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
di: Liu, Shuaiqi, et al.
Pubblicazione: (2024)
Knowledge-Augmented Multimodal Clinical Rationale Generation for Disease Diagnosis with Small Language Models
di: Niu, Shuai, et al.
Pubblicazione: (2024)
di: Niu, Shuai, et al.
Pubblicazione: (2024)
Beyond Token Length: Step Pruner for Efficient and Accurate Reasoning in Large Language Models
di: Wu, Canhui, et al.
Pubblicazione: (2025)
di: Wu, Canhui, et al.
Pubblicazione: (2025)
Robustness of Large Language Models to Perturbations in Text
di: Singh, Ayush, et al.
Pubblicazione: (2024)
di: Singh, Ayush, et al.
Pubblicazione: (2024)
Distributed In-Context Learning under Non-IID Among Clients
di: Liang, Siqi, et al.
Pubblicazione: (2024)
di: Liang, Siqi, et al.
Pubblicazione: (2024)
Integrating Emotional and Linguistic Models for Ethical Compliance in Large Language Models
di: Chang, Edward Y.
Pubblicazione: (2024)
di: Chang, Edward Y.
Pubblicazione: (2024)
Language Models are Crossword Solvers
di: Saha, Soumadeep, et al.
Pubblicazione: (2024)
di: Saha, Soumadeep, et al.
Pubblicazione: (2024)
The Dual-Route Model of Induction
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
di: Feucht, Sheridan, et al.
Pubblicazione: (2025)
UNO-Bench: A Unified Benchmark for Exploring the Compositional Law Between Uni-modal and Omni-modal in Omni Models
di: Chen, Chen, et al.
Pubblicazione: (2025)
di: Chen, Chen, et al.
Pubblicazione: (2025)
Dual Debiasing for Noisy In-Context Learning for Text Generation
di: Liang, Siqi, et al.
Pubblicazione: (2025)
di: Liang, Siqi, et al.
Pubblicazione: (2025)
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
di: Lu, Yaxi, et al.
Pubblicazione: (2024)
di: Lu, Yaxi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Frequency Matters: Fast Model-Agnostic Data Curation for Pruning and Quantization
di: Monaco, Francesco Pio, et al.
Pubblicazione: (2026) -
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025) -
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
di: Dumitru, Razvan-Gabriel, et al.
Pubblicazione: (2024) -
Detecting AI-Generated Texts in Cross-Domains
di: Zhou, You, et al.
Pubblicazione: (2024) -
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
di: Khanna, Danush, et al.
Pubblicazione: (2025)