BitSharp · AI engineering studio · Rzeszów

2026-03-05 arXiv 2603.04162 · v2 published
2026-02 9 models + 3 datasets on HF
2026-07-24 PolAgentBench: draft complete · 67 tasks
2026-07-23 3 note theses pending approval
openhaft.pl · production · live

Two papers of our own, an agentic benchmark we wrote, nine public Bielik models and deployments you can click. Leading rather than promising — every claim on this page points to its proof.

daysminutes

time to import a 4,000+ product catalog · openhaft.pl · → deployments

Work, not declarations. Pick a registry.

thumbnails: real screenshots from this project · duotone

First-rate AI engineering — on evidence, not slogans.

BitSharp · AI engineering studio · Rzeszów

Areas of work

RAG and knowledge search

Question answering over company knowledge that cites its source down to the paragraph — hybrid retrieval, reranking, relevance evaluation on gold sets.

proof → notes/retrieval
Agentic systems

Tool orchestration, failure handling, the point where a human steps in — measured with our own benchmark of 67 agentic tasks in Polish.

proof → PolAgentBench
Document data extraction

Invoices, contracts, correspondence: schema enforced at decoding, deterministic validation, a human approves only the exceptions.

proof → deployments
LLM production and optimization

Serving with full observability, model routing, quantization — from 22 GB down to 3.26 GB without losing the Polish.

proof → workshop · research

Working standard

  • Deterministic: fixed seed, temperature 0, results reproducible bit for bit.
  • Commit-stamped: every result carries the hash of the code that produced it.
  • Tests before GPU: 204 unit tests before a single run starts.
  • Caveats stated plainly: methodological artifacts go in the body text, not a footnote.
  • Open artifacts: models, data and evaluation logs are public.
  • Deterministic code where it wins — an LLM only where the rules cannot be written down.

Founder

Jakub Prejzner — AI Engineer. Builds RAG systems, process automation and document data extraction; follows the field and ships from it as it moves. The quantization papers are proof he can take a subject down to research level.

own papers2 (arXiv + draft)
public Bielik models9 · HF
agentic benchmark67 tasks
production deploymentopenhaft.pl
based inRzeszów / PL