Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ociepa, Krzysztof, Flis, Łukasz, Wróbel, Krzysztof, Gwoździej, Adrian, Kinas, Remigiusz |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2026)
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2026)
Bielik 11B v3: Multilingual Large Language Model for European Languages
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
Bielik 11B v2 Technical Report
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
Bielik v3 Small: Technical Report
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
von: Kinas, Remigiusz, et al.
Veröffentlicht: (2026)
von: Kinas, Remigiusz, et al.
Veröffentlicht: (2026)
Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation
von: Wróbel, Krzysztof, et al.
Veröffentlicht: (2026)
von: Wróbel, Krzysztof, et al.
Veröffentlicht: (2026)
PL-Guard: Benchmarking Language Model Safety for Polish
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
FIN-bench-v2: A Unified and Robust Benchmark Suite for Evaluating Finnish Large Language Models
von: Kytöniemi, Joona, et al.
Veröffentlicht: (2025)
von: Kytöniemi, Joona, et al.
Veröffentlicht: (2025)
OPOR-Bench: Evaluating Large Language Models on Online Public Opinion Report Generation
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
von: Yu, Jinzheng, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
The Curious Case of Visual Grounding: Different Effects for Speech- and Text-based Language Encoders
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
von: Sauter, Adrian, et al.
Veröffentlicht: (2025)
Big City Bias: Evaluating the Impact of Metropolitan Size on Computational Job Market Abilities of Language Models
von: Campanella, Charlie, et al.
Veröffentlicht: (2024)
von: Campanella, Charlie, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for Zero-Shot Disease Labeling in CT Radiology Reports Across Organ Systems
von: Garcia-Alcoser, Michael E., et al.
Veröffentlicht: (2025)
von: Garcia-Alcoser, Michael E., et al.
Veröffentlicht: (2025)
Lisbon Computational Linguists at SemEval-2024 Task 2: Using A Mistral 7B Model and Data Augmentation
von: Guimarães, Artur, et al.
Veröffentlicht: (2024)
von: Guimarães, Artur, et al.
Veröffentlicht: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
von: Ovcharov, Volodymyr
Veröffentlicht: (2026)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Precise Length Control in Large Language Models
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
What Drives Performance in Multilingual Language Models?
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
Large Language Models for Biomedical Article Classification
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
Evaluating Pixel Language Models on Non-Standardized Languages
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2024)
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
Socially Responsible Data for Large Multilingual Language Models
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
von: Gao, Ge, et al.
Veröffentlicht: (2024)
von: Gao, Ge, et al.
Veröffentlicht: (2024)
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
Qomhra: A Bilingual Irish and English Large Language Model
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
Dialect Normalization using Large Language Models and Morphological Rules
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
Towards Human Understanding of Paraphrase Types in Large Language Models
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
von: Liu, Han, et al.
Veröffentlicht: (2026)
von: Liu, Han, et al.
Veröffentlicht: (2026)
Task Contamination: Language Models May Not Be Few-Shot Anymore
von: Li, Changmao, et al.
Veröffentlicht: (2023)
von: Li, Changmao, et al.
Veröffentlicht: (2023)
Quality Estimation with $k$-nearest Neighbors and Automatic Evaluation for Model-specific Quality Estimation
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
LinkNER: Linking Local Named Entity Recognition Models to Large Language Models using Uncertainty
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
von: Zhang, Zhen, et al.
Veröffentlicht: (2024)
A Domain-Based Taxonomy of Jailbreak Vulnerabilities in Large Language Models
von: Peláez-González, Carlos, et al.
Veröffentlicht: (2025)
von: Peláez-González, Carlos, et al.
Veröffentlicht: (2025)
Linguistic Interpretability of Transformer-based Language Models: a systematic review
von: López-Otal, Miguel, et al.
Veröffentlicht: (2025)
von: López-Otal, Miguel, et al.
Veröffentlicht: (2025)
Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis
von: Nagao, Moe, et al.
Veröffentlicht: (2026)
von: Nagao, Moe, et al.
Veröffentlicht: (2026)
Refining Packing and Shuffling Strategies for Enhanced Performance in Generative Language Models
von: Chen, Yanbing, et al.
Veröffentlicht: (2024)
von: Chen, Yanbing, et al.
Veröffentlicht: (2024)
KyrgyzBERT: A Compact, Efficient Language Model for Kyrgyz NLP
von: Metinov, Adilet, et al.
Veröffentlicht: (2025)
von: Metinov, Adilet, et al.
Veröffentlicht: (2025)
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
von: Song, Chenyang, et al.
Veröffentlicht: (2023)
von: Song, Chenyang, et al.
Veröffentlicht: (2023)
Aligning Large Language Models for Faithful Integrity Against Opposing Argument
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
von: Zhao, Yong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2026) -
Bielik 11B v3: Multilingual Large Language Model for European Languages
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025) -
Bielik 11B v2 Technical Report
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025) -
Bielik v3 Small: Technical Report
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2025) -
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
von: Kinas, Remigiusz, et al.
Veröffentlicht: (2026)