4 citations · 4 across the 7 of their papers we have counts for
7 papers
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
Krzysztof Ociepa, Łukasz Flis, Remigiusz Kinas +2
The development of the Bielik v3 PL series, encompassing both the 7B and 11B parameter variants, represents a significant milestone in the field of language-specific large language…
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
Remigiusz Kinas, Paweł Kiszczak, Sergio P. Perez +4
This report details the creation of Bielik-Minitron-7B, a compressed 7.35B parameter version of the Bielik-11B-v3.0 model, specifically optimized for European languages. By leverag…
Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation
Krzysztof Wróbel, Jan Maria Kowalski, Jerzy Surma +2
As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety classifiers has become paramount. W…
Bielik 11B v3: Multilingual Large Language Model for European Languages
Krzysztof Ociepa, Łukasz Flis, Remigiusz Kinas +2
We present Bielik 11B v3, a state-of-the-art language model highly optimized for the Polish language, while also maintaining strong capabilities in other European languages. This m…
Bielik v3 Small: Technical Report
Krzysztof Ociepa, Łukasz Flis, Remigiusz Kinas +2
We introduce Bielik v3, a series of parameter-efficient generative text models (1.5B and 4.5B) optimized for Polish language processing. These models demonstrate that smaller, well…
Bielik 11B v2 Technical Report
Krzysztof Ociepa, Łukasz Flis, Krzysztof Wróbel +2
We present Bielik 11B v2, a state-of-the-art language model optimized for Polish text processing. Built on the Mistral 7B v0.2 architecture and scaled to 11B parameters using depth…