Prices,
in writing.
Every line below maps to a step on the AI Model Training page. Pilots are fixed price. Programs are quoted after the scoping week. Retainers are monthly. All prices in EUR, excluding VAT, valid until 31 March 2027.
Fixed scope, fixed price, one number at the end.
Both pilots end with a base-vs-tuned scoreboard on a frozen eval and a written recommendation. Pick by how much data you already have.
Pilot S
Supervised fine-tuning on data you already have. One task, one model, one frozen eval, one report.
- Scoping call + data and licence audit
- Frozen eval set built from your examples
- One fine-tuned open-weight model (4B–12B class)
- Base-vs-tuned scoreboard, GGUF build, model card
- Written go/no-go recommendation
- 2–3 weeks · prepaid
Pilot M
Everything in Pilot S, plus the dataset engineering: building, translating or distilling the training data, and an eval set from scratch.
- Scoping week + data audit on site or remote
- Dataset build: cleaning, dedup, quality scoring, licence gating
- Translation or teacher distillation where data is missing
- Base vs tuned vs the API or RAG setup you were considering
- Deployment recipe (GGUF / vLLM), hand-over call
- 4–5 weeks · 50 % on signing, 50 % on delivery
Starting from one of our published domain models (Slovenian, biomedical, cybersecurity) takes €2,000 off Pilot M.
€20,000 – €90,000, quoted after scoping.
Fine-tuning, evaluation and deployment for a production model, with continued pre-training where the domain needs it. Six weeks at the lower end, four months at the upper. The scoping week is billed as Pilot S and credited against the program.
| What moves the price | Lower end | Upper end |
|---|---|---|
| Duration | 6–8 weeks | 3–4 months |
| Training stages | SFT only, one or two re-tunes | CPT on hundreds of millions of tokens, then SFT, then re-tunes |
| Data readiness | Clean, licensed, in one place | Scattered, needs translation, distillation or OCR |
| Model size | 4B–12B, single GPU serving | 27B–35B MoE, multi-GPU or on-prem cluster |
| Evaluation | One objective track | Objective + multi-judge tracks, several use cases |
| Deployment | GGUF hand-over | vLLM on-prem install, cascade router, air-gapped delivery |
| Payment | 40 % on signing · 40 % at the mid-program eval gate · 20 % on delivery. GPU time for CPT billed separately, estimated in the quote. | |
The model stays current. The scoreboard is the contract.
New data in, retrain, re-score, ship. Base-model upgrades when a better open model wins on your eval. Every version of the weights is yours.
| Lite | Core | Scale | Enterprise | |
|---|---|---|---|---|
| Per month | €1,200 | €3,500 | €7,500 | from €15,000 |
| Models under retainer | 1 | 1 | up to 3 | as scoped |
| Retrain cadence | quarterly | monthly | monthly, or on data threshold | weekly possible |
| Base-model upgrade | once a year | next cycle after release | within days of release | within days, with rollback plan |
| Evaluation | frozen set, re-scored each retrain | frozen set + regression alerts | + new eval slices as use cases grow | + multi-judge track, compliance re-run |
| Cascade router | — | — | maintained (small model first, frontier fallback) | maintained + cost reporting |
| Deployment support | e-mail, 5 business days | e-mail, 2 business days | e-mail + chat, next business day | on-prem / air-gapped, response SLA |
| Documentation refresh | model card per version | model card per version | model + data card per version | + AI Act record pack per version |
| Term | 12 months, billed monthly. 10 % off when the year is prepaid. Included GPU time covers the retrain cadence above; CPT-scale runs are quoted separately. | |||
Priced separately, so the retainer stays about the model.
Most customers run the model on their own hardware. For those who don't, we host it on capacity we control inside the EU. Sized by traffic, never by hardware.
Hosted endpoint · EU
- Light€600 / month — internal tools, a few users
- Standard€1,200 / month — customer-facing, business hours
- High€2,000 / month — 24/7, sustained traffic
- IncludesvLLM or llama.cpp serving, TLS, monitoring, monthly usage report
- ExcludesFrontier-API fallback calls in a cascade (passed through at cost)
GPU time & add-ons
- GPU timeOn request, blocks of 100 GPU-hours, for teams with their own pipeline
- Extra eval set€2,500 — a new frozen benchmark for an additional use case
- On-site day€1,500 + travel — install, workshop or air-gapped hand-over
- Dataset onlyQuoted per corpus — build, translate or distil without training
- SpeechASR / TTS fine-tunes priced as Pilot M or program, by hours of audio
Terms
- All prices in EUR, excluding VAT. Valid until 31 March 2027; this page carries a version, quotes reference it.
- NDA and data-processing agreement signed before any data is transferred. Data stays inside the EU.
- You own the weights, adapters, datasets, eval sets and provenance record of everything delivered, as described on the training page.
- Pilots are prepaid (S) or 50/50 (M). Programs 40/40/20 on milestones. Retainers monthly in advance, 12-month term.
- Open-weight bases only; we flag any base or data licence that is not commercially usable before work starts.
How to read this page
- Not sure which pilot? Send fifty real examples of the task. We answer within two business days with S or M and why.
- Already paying for a frontier API? Bring the monthly bill. The retainer should sit well under what the small model saves.
- Regulated data? Ask about the air-gapped rung. Same prices, plus on-site days.
- Slovenian version of this page: slovensko.
Pick a pilot. Get a number.
Thirty minutes to scope it. A dated schedule after that. Bring fifty examples and, if you have one, your current API bill.