Small Language Models
Training Polish language models from scratch. We study scaling laws, tokenizers, data mixtures and distillation to understand how much capability a compact model can learn.
Explore programme →
FABRYKA AI RESEARCH
INDEPENDENT ARTIFICIAL INTELLIGENCE RESEARCH LABORATORY · WARSAW, POLAND
We study how language models learn, how to make them more efficient, and how to evaluate them reliably — with a particular focus on Polish-language AI.
Research → Experiments → Evidence → Applications
01 / RESEARCH PROGRAMMES
How much quality can we achieve with a given budget of data, training and inference — and how can we measure it reliably? These three programmes connect our current work and planned studies.
Training Polish language models from scratch. We study scaling laws, tokenizers, data mixtures and distillation to understand how much capability a compact model can learn.
Explore programme →Making language models practical on commodity GPUs. We study quantization, speculative decoding, including DFlash, routing and the trade-offs between throughput, latency and quality.
Explore programme →Building the data and measurements that Polish AI needs. This programme connects Polish DynaWord, CodeSOTA benchmarks, OCR evaluation and methodologies for comparing Polish LLMs.
Explore programme →02 / SELECTED RESEARCH OUTPUTS
Explore public model weights, data and benchmark evidence. Each linked artifact carries its own documentation and scope.
A compact language-model release. Model card, weights and usage information are available in the public repository.
A public Polish-language dataset supporting our work on language-model training and data composition.
Model comparisons and benchmark evidence. Individual entries identify their sources and evaluation context.
03 / OPEN RESEARCH
Our research connects questions, baselines, experiments and artifacts. We use Fabryka Track to record training runs, compare configurations and investigate results.
Dated experiment write-ups are preserved in our research notes. Track provides the tools for recording experiments; access to individual runs depends on the project.
04 / RESEARCH INFRASTRUCTURE
Training, measurement and serving are part of the same experimental workflow. We build the infrastructure that helps us investigate each step.
Public infrastructure for model comparisons, benchmark results and evaluation evidence.
Explore CodeSOTA ↗Experiment tracking for training runs, metrics, logs, checkpoints and configuration comparisons.
Explore Track ↗GPU systems for training and inference experiments, with hardware and workload conditions documented alongside results.
Inference research →05 / PEOPLE
Meet the Fabryka AI team: researchers, engineers and contributors working across models, data, evaluation, agents and infrastructure.
Meet the Fabryka AI team ↓
Language models, training infrastructure, model optimization and evaluation.
Datasets, evaluations, training infrastructure and systems engineering.

Founder of Xfaang. Builds autonomous AI agents and LLM products, connecting data, evaluation and infrastructure.

Agent teams, web and mobile applications, recruitment and onboarding.

Theoretical physics, computer vision for oncology, and generative and agent systems.

Bots, microservices and production tools, from idea to deployment.

Long-running agent workflows and ML strategies informed by data exploration and visualization.

Data pipelines, systems engineering and technical planning.

Evaluation tooling, reproducible experiments and Polish byte-level BPE tokenizers.

Research tooling, experiment documentation and presentations.

Training infrastructure, developer tools and the Fabryka Track interface.

Small-model training experiments, model-based evaluation and benchmark methodology.

Software engineering and contributions to the Fabryka Track interface.

Polish tokenizers, technical reviews and research planning.

Contributions to research discussions, planning and review.
RESEARCH & PRACTICE
Supporting the team with research methodology, evaluation and domain expertise.
06 / TECHNOLOGY TRANSFER
Selected technologies developed or evaluated at Fabryka become production systems. Our inference platform makes open models available through an API, with tools for chat, routing and token-level analysis.
The platform also serves models developed by other teams. Availability through our API is separate from authorship of a model.
Explore Fabryka API →07 / ABOUT FABRYKA AI
Fabryka AI is an independent artificial intelligence research laboratory in Warsaw. We build datasets, train models from scratch and develop methods for efficient inference and reliable evaluation.