K-Quantization and its Impact on Output Performance
Fuente:
arXiv
Salvato in:
| Autori principali: | Davidsson, Robin Baki, Nugues, Pierre |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Linking Named Entities in Diderot's \textit{Encyclopédie} to Wikidata
di: Nugues, Pierre
Pubblicazione: (2024)
di: Nugues, Pierre
Pubblicazione: (2024)
Matching and Linking Entries in Historical Swedish Encyclopedias
di: Börjesson, Simon, et al.
Pubblicazione: (2025)
di: Börjesson, Simon, et al.
Pubblicazione: (2025)
Mapping the Past: Geographically Linking an Early 20th Century Swedish Encyclopedia with Wikidata
di: Ahlin, Axel, et al.
Pubblicazione: (2024)
di: Ahlin, Axel, et al.
Pubblicazione: (2024)
ATLAS: Article Tracking, Linking, and Analysis of Swedish Encyclopedias
di: Andersson, Albin, et al.
Pubblicazione: (2026)
di: Andersson, Albin, et al.
Pubblicazione: (2026)
EDDA-Coordinata: An Annotated Dataset of Historical Geographic Coordinates
di: Moncla, Ludovic, et al.
Pubblicazione: (2026)
di: Moncla, Ludovic, et al.
Pubblicazione: (2026)
OAC: Output-adaptive Calibration for Accurate Post-training Quantization
di: Edalati, Ali, et al.
Pubblicazione: (2024)
di: Edalati, Ali, et al.
Pubblicazione: (2024)
English K_Quantization of LLMs Does Not Disproportionately Diminish Multilingual Performance
di: Borgersen, Karl Audun, et al.
Pubblicazione: (2025)
di: Borgersen, Karl Audun, et al.
Pubblicazione: (2025)
VQ-Logits: Compressing the Output Bottleneck of Large Language Models via Vector Quantized Logits
di: Shao, Jintian, et al.
Pubblicazione: (2025)
di: Shao, Jintian, et al.
Pubblicazione: (2025)
LLM as a Scorer: The Impact of Output Order on Dialogue Evaluation
di: Chen, Yi-Pei, et al.
Pubblicazione: (2024)
di: Chen, Yi-Pei, et al.
Pubblicazione: (2024)
LLM Output Detectability and Task Performance Can be Jointly Optimized
di: Saito, Koshiro, et al.
Pubblicazione: (2026)
di: Saito, Koshiro, et al.
Pubblicazione: (2026)
The Uneven Impact of Post-Training Quantization in Machine Translation
di: Marie, Benjamin, et al.
Pubblicazione: (2025)
di: Marie, Benjamin, et al.
Pubblicazione: (2025)
On the Impact of Calibration Data in Post-training Quantization and Pruning
di: Williams, Miles, et al.
Pubblicazione: (2023)
di: Williams, Miles, et al.
Pubblicazione: (2023)
A Better LLM Evaluator for Text Generation: The Impact of Prompt Output Sequencing and Optimization
di: Chu, KuanChao, et al.
Pubblicazione: (2024)
di: Chu, KuanChao, et al.
Pubblicazione: (2024)
Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost
di: Nayab, Sania, et al.
Pubblicazione: (2024)
di: Nayab, Sania, et al.
Pubblicazione: (2024)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
di: Neshaei, Seyed Parsa, et al.
Pubblicazione: (2024)
Sustainable LLM Inference for Edge AI: Evaluating Quantized LLMs for Energy Efficiency, Output Accuracy, and Inference Latency
di: Husom, Erik Johannes, et al.
Pubblicazione: (2025)
di: Husom, Erik Johannes, et al.
Pubblicazione: (2025)
LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs
di: Long, Do Xuan, et al.
Pubblicazione: (2024)
di: Long, Do Xuan, et al.
Pubblicazione: (2024)
The Hidden Costs of Translation Accuracy: Distillation, Quantization, and Environmental Impact
di: Vijay, Dhaathri, et al.
Pubblicazione: (2025)
di: Vijay, Dhaathri, et al.
Pubblicazione: (2025)
Beyond Real Weights: Hypercomplex Representations for Stable Quantization
di: Ahad, Jawad Ibn, et al.
Pubblicazione: (2025)
di: Ahad, Jawad Ibn, et al.
Pubblicazione: (2025)
Quantifying the Impact of Structured Output Format on Large Language Models through Causal Inference
di: Yuan, Han, et al.
Pubblicazione: (2025)
di: Yuan, Han, et al.
Pubblicazione: (2025)
The Impact of Item-Writing Flaws on Difficulty and Discrimination in Item Response Theory
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
di: Schmucker, Robin, et al.
Pubblicazione: (2025)
Mitigating the Impact of Outlier Channels for Language Model Quantization with Activation Regularization
di: Nrusimha, Aniruddha, et al.
Pubblicazione: (2024)
di: Nrusimha, Aniruddha, et al.
Pubblicazione: (2024)
DRS: Deep Question Reformulation With Structured Output
di: Li, Zhecheng, et al.
Pubblicazione: (2024)
di: Li, Zhecheng, et al.
Pubblicazione: (2024)
Feature-Aware Malicious Output Detection and Mitigation
di: Dong, Weilong, et al.
Pubblicazione: (2025)
di: Dong, Weilong, et al.
Pubblicazione: (2025)
Training Data Size Sensitivity in Unsupervised Rhyme Recognition
di: Plecháč, Petr, et al.
Pubblicazione: (2026)
di: Plecháč, Petr, et al.
Pubblicazione: (2026)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
di: Wang, Qianli, et al.
Pubblicazione: (2025)
di: Wang, Qianli, et al.
Pubblicazione: (2025)
SESGO: Spanish Evaluation of Stereotypical Generative Outputs
di: Robles, Melissa, et al.
Pubblicazione: (2025)
di: Robles, Melissa, et al.
Pubblicazione: (2025)
Beat-Based Rhythm Quantization of MIDI Performances
di: Wachter, Maximilian, et al.
Pubblicazione: (2025)
di: Wachter, Maximilian, et al.
Pubblicazione: (2025)
FMBench: Adaptive Large Language Model Output Formatting
di: Wang, Yaoting, et al.
Pubblicazione: (2026)
di: Wang, Yaoting, et al.
Pubblicazione: (2026)
Enhancing Automated Interpretability with Output-Centric Feature Descriptions
di: Gur-Arieh, Yoav, et al.
Pubblicazione: (2025)
di: Gur-Arieh, Yoav, et al.
Pubblicazione: (2025)
Advancing Academic Chatbots: Evaluation of Non Traditional Outputs
di: Favero, Nicole, et al.
Pubblicazione: (2025)
di: Favero, Nicole, et al.
Pubblicazione: (2025)
Beemo: Benchmark of Expert-edited Machine-generated Outputs
di: Artemova, Ekaterina, et al.
Pubblicazione: (2024)
di: Artemova, Ekaterina, et al.
Pubblicazione: (2024)
Lost in Space: Finding the Right Tokens for Structured Output
di: Hamilton, Sil, et al.
Pubblicazione: (2025)
di: Hamilton, Sil, et al.
Pubblicazione: (2025)
Decoupling Task-Solving and Output Formatting in LLM Generation
di: Deng, Haikang, et al.
Pubblicazione: (2025)
di: Deng, Haikang, et al.
Pubblicazione: (2025)
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs
di: Yazan, Mert, et al.
Pubblicazione: (2024)
di: Yazan, Mert, et al.
Pubblicazione: (2024)
HarmLevelBench: Evaluating Harm-Level Compliance and the Impact of Quantization on Model Alignment
di: Belkhiter, Yannis, et al.
Pubblicazione: (2024)
di: Belkhiter, Yannis, et al.
Pubblicazione: (2024)
On the Diversity of Synthetic Data and its Impact on Training Large Language Models
di: Chen, Hao, et al.
Pubblicazione: (2024)
di: Chen, Hao, et al.
Pubblicazione: (2024)
Output-Space Search: Targeting LLM Generations in a Frozen Encoder-Defined Output Space
di: Materzok, Tobias
Pubblicazione: (2026)
di: Materzok, Tobias
Pubblicazione: (2026)
OLA: Output Language Alignment in Code-Switched LLM Interactions
di: Oh, Juhyun, et al.
Pubblicazione: (2026)
di: Oh, Juhyun, et al.
Pubblicazione: (2026)
Weight Tying Biases Token Embeddings Towards the Output Space
di: Lopardo, Antonio, et al.
Pubblicazione: (2026)
di: Lopardo, Antonio, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Linking Named Entities in Diderot's \textit{Encyclopédie} to Wikidata
di: Nugues, Pierre
Pubblicazione: (2024) -
Matching and Linking Entries in Historical Swedish Encyclopedias
di: Börjesson, Simon, et al.
Pubblicazione: (2025) -
Mapping the Past: Geographically Linking an Early 20th Century Swedish Encyclopedia with Wikidata
di: Ahlin, Axel, et al.
Pubblicazione: (2024) -
ATLAS: Article Tracking, Linking, and Analysis of Swedish Encyclopedias
di: Andersson, Albin, et al.
Pubblicazione: (2026) -
EDDA-Coordinata: An Annotated Dataset of Historical Geographic Coordinates
di: Moncla, Ludovic, et al.
Pubblicazione: (2026)