7 papers
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
Krzysztof Ociepa, Åukasz Flis, Remigiusz Kinas +2
The development of the Bielik v3 PL series, encompassing both the 7B and 11B parameter variants, represents a significant milestone in the field of language-specific large language…
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
Remigiusz Kinas, PaweÅ Kiszczak, Sergio P. Perez +4
This report details the creation of Bielik-Minitron-7B, a compressed 7.35B parameter version of the Bielik-11B-v3.0 model, specifically optimized for European languages. By leverag…
Bielik 11B v3: Multilingual Large Language Model for European Languages
Krzysztof Ociepa, Åukasz Flis, Remigiusz Kinas +2
We present Bielik 11B v3, a state-of-the-art language model highly optimized for the Polish language, while also maintaining strong capabilities in other European languages. This m…
Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation
Krzysztof Ociepa, Åukasz Flis, Krzysztof Wróbel +2
We introduce Bielik 7B v0.1, a 7-billion-parameter generative text model for Polish language processing. Trained on curated Polish corpora, this model addresses key challenges in l…
Reasoning Language Models: A Blueprint
Maciej Besta, Julia Barth, Eric Schreiber +16
Reasoning language models (RLMs), also known as Large Reasoning Models (LRMs), such as OpenAI's o1 and o3, DeepSeek-R1, and Alibaba's QwQ, have redefined AI's problem-solving capab…
Bielik v3 Small: Technical Report
Krzysztof Ociepa, Åukasz Flis, Remigiusz Kinas +2
We introduce Bielik v3, a series of parameter-efficient generative text models (1.5B and 4.5B) optimized for Polish language processing. These models demonstrate that smaller, well…