Enregistré dans:
| Auteur principal: | Prejzner, Jakub |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2603.04162 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
par: Ociepa, Krzysztof, et autres
Publié: (2026)
par: Ociepa, Krzysztof, et autres
Publié: (2026)
Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation
par: Ociepa, Krzysztof, et autres
Publié: (2024)
par: Ociepa, Krzysztof, et autres
Publié: (2024)
Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation
par: Wróbel, Krzysztof, et autres
Publié: (2026)
par: Wróbel, Krzysztof, et autres
Publié: (2026)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
par: Kinas, Remigiusz, et autres
Publié: (2026)
par: Kinas, Remigiusz, et autres
Publié: (2026)
Bielik 11B v2 Technical Report
par: Ociepa, Krzysztof, et autres
Publié: (2025)
par: Ociepa, Krzysztof, et autres
Publié: (2025)
Bielik 11B v3: Multilingual Large Language Model for European Languages
par: Ociepa, Krzysztof, et autres
Publié: (2025)
par: Ociepa, Krzysztof, et autres
Publié: (2025)
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
par: Liu, Zechun, et autres
Publié: (2025)
par: Liu, Zechun, et autres
Publié: (2025)
Passage Retrieval of Polish Texts Using OKAPI BM25 and an Ensemble of Cross Encoders
par: Pokrywka, Jakub
Publié: (2024)
par: Pokrywka, Jakub
Publié: (2024)
R2Q: Towards Robust 2-Bit Large Language Models via Residual Refinement Quantization
par: Chen, Jiayi, et autres
Publié: (2025)
par: Chen, Jiayi, et autres
Publié: (2025)
Predicting Emotion Intensity in Polish Political Texts: Comparing Supervised Models and Large Language Models in a Resource-Poor Language
par: Plisiecki, Hubert, et autres
Publié: (2024)
par: Plisiecki, Hubert, et autres
Publié: (2024)
GPT-4 passes most of the 297 written Polish Board Certification Examinations
par: Pokrywka, Jakub, et autres
Publié: (2024)
par: Pokrywka, Jakub, et autres
Publié: (2024)
Polish-English medical knowledge transfer: A new benchmark and results
par: Grzybowski, Łukasz, et autres
Publié: (2024)
par: Grzybowski, Łukasz, et autres
Publié: (2024)
PLLuM: A Family of Polish Large Language Models
par: Kocoń, Jan, et autres
Publié: (2025)
par: Kocoń, Jan, et autres
Publié: (2025)
LLMzSzŁ: a comprehensive LLM benchmark for Polish
par: Jassem, Krzysztof, et autres
Publié: (2025)
par: Jassem, Krzysztof, et autres
Publié: (2025)
Bielik v3 Small: Technical Report
par: Ociepa, Krzysztof, et autres
Publié: (2025)
par: Ociepa, Krzysztof, et autres
Publié: (2025)
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
par: Xu, Yuzhuang, et autres
Publié: (2026)
par: Xu, Yuzhuang, et autres
Publié: (2026)
CEAID: Benchmark of Multilingual Machine-Generated Text Detection Methods for Central European Languages
par: Macko, Dominik, et autres
Publié: (2025)
par: Macko, Dominik, et autres
Publié: (2025)
Self-Polish: Enhance Reasoning in Large Language Models via Problem Refinement
par: Xi, Zhiheng, et autres
Publié: (2023)
par: Xi, Zhiheng, et autres
Publié: (2023)
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
par: Statkiewicz, Grzegorz, et autres
Publié: (2026)
par: Statkiewicz, Grzegorz, et autres
Publié: (2026)
Efficient Language Adaptive Pre-training: Extending State-of-the-Art Large Language Models for Polish
par: Ruciński, Szymon
Publié: (2024)
par: Ruciński, Szymon
Publié: (2024)
BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models
par: He, Liulu, et autres
Publié: (2025)
par: He, Liulu, et autres
Publié: (2025)
Polish-ASTE: Aspect-Sentiment Triplet Extraction Datasets for Polish
par: Lango, Marta, et autres
Publié: (2025)
par: Lango, Marta, et autres
Publié: (2025)
Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
par: Liu, Ruikang, et autres
Publié: (2025)
par: Liu, Ruikang, et autres
Publié: (2025)
E^2-LLM: Efficient and Extreme Length Extension of Large Language Models
par: Liu, Jiaheng, et autres
Publié: (2024)
par: Liu, Jiaheng, et autres
Publié: (2024)
AMAQ: Adaptive Mixed-bit Activation Quantization for Collaborative Parameter Efficient Fine-tuning
par: Song, Yurun, et autres
Publié: (2025)
par: Song, Yurun, et autres
Publié: (2025)
A Comprehensive Study on Quantization Techniques for Large Language Models
par: Lang, Jiedong, et autres
Publié: (2024)
par: Lang, Jiedong, et autres
Publié: (2024)
Evaluating Quantized Large Language Models
par: Li, Shiyao, et autres
Publié: (2024)
par: Li, Shiyao, et autres
Publié: (2024)
SignRoundV2: Toward Closing the Performance Gap in Extremely Low-Bit Post-Training Quantization for LLMs
par: Cheng, Wenhua, et autres
Publié: (2025)
par: Cheng, Wenhua, et autres
Publié: (2025)
A Survey of Low-bit Large Language Models: Basics, Systems, and Algorithms
par: Gong, Ruihao, et autres
Publié: (2024)
par: Gong, Ruihao, et autres
Publié: (2024)
ButterflyQuant: Ultra-low-bit LLM Quantization through Learnable Orthogonal Butterfly Transforms
par: Xu, Bingxin, et autres
Publié: (2025)
par: Xu, Bingxin, et autres
Publié: (2025)
Optimizing Large Language Models through Quantization: A Comparative Analysis of PTQ and QAT Techniques
par: Hasan, Jahid
Publié: (2024)
par: Hasan, Jahid
Publié: (2024)
Advancing Beyond Identification: Multi-bit Watermark for Large Language Models
par: Yoo, KiYoon, et autres
Publié: (2023)
par: Yoo, KiYoon, et autres
Publié: (2023)
Comparing Uncertainty Measurement and Mitigation Methods for Large Language Models: A Systematic Review
par: Abbasli, Toghrul, et autres
Publié: (2025)
par: Abbasli, Toghrul, et autres
Publié: (2025)
Cross-Family Speculative Decoding for Polish Language Models on Apple~Silicon: An Empirical Evaluation of Bielik~11B with UAG-Extended MLX-LM
par: Fonal, Krzysztof
Publié: (2026)
par: Fonal, Krzysztof
Publié: (2026)
A Comparative Study of Large Language Models and Human Personality Traits
par: Jiaqi, Wang, et autres
Publié: (2025)
par: Jiaqi, Wang, et autres
Publié: (2025)
Exploring the Trade-Offs: Quantization Methods, Task Difficulty, and Model Size in Large Language Models From Edge to Giant
par: Lee, Jemin, et autres
Publié: (2024)
par: Lee, Jemin, et autres
Publié: (2024)
A Comprehensive Evaluation of Quantization Strategies for Large Language Models
par: Jin, Renren, et autres
Publié: (2024)
par: Jin, Renren, et autres
Publié: (2024)
A Comparative Study of Quality Evaluation Methods for Text Summarization
par: Nguyen, Huyen, et autres
Publié: (2024)
par: Nguyen, Huyen, et autres
Publié: (2024)
When Quantization Affects Confidence of Large Language Models?
par: Proskurina, Irina, et autres
Publié: (2024)
par: Proskurina, Irina, et autres
Publié: (2024)
DLLMQuant: Quantizing Diffusion-based Large Language Models
par: Xu, Chen, et autres
Publié: (2025)
par: Xu, Chen, et autres
Publié: (2025)
Documents similaires
-
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
par: Ociepa, Krzysztof, et autres
Publié: (2026) -
Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation
par: Ociepa, Krzysztof, et autres
Publié: (2024) -
Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation
par: Wróbel, Krzysztof, et autres
Publié: (2026) -
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
par: Kinas, Remigiusz, et autres
Publié: (2026) -
Bielik 11B v2 Technical Report
par: Ociepa, Krzysztof, et autres
Publié: (2025)