ILLATE
Pricing

Fixed fees, written guarantees

The check is free. The migration is a fixed fee, refunded in full if the model doesn't match. Hosting is optional and priced against what you pay today.

01 Plans

Three steps, priced separately

Parity check Free
  • Cost check of every option on your volumes
  • Fine-tune on your examples, test on a held-out slice
  • PASS / FAIL / INCONCLUSIVE report
  • We tell you if a cheap API already does the job
Migration $1,500–3,000 fixed
  • Fee set by data size and task, agreed before work starts
  • Half up front, half on delivery
  • Full refund if the final model misses the agreed margin
  • You keep the weights, eval set and scripts
Hosting Per call optional
  • OpenAI-compatible endpoint, monitoring and retraining runs
  • Flat price per call, fixed for 12 months
  • Priced below your cheapest option that passes your eval
  • Leave any time and take the model with you

Prices are in US dollars. Invoices go out through Wise or Payoneer. Larger or multi-model engagements are quoted separately.

02 Compare

Your options when a fine-tuned model is retired

You can stay until the deadline, rebuild in-house, or have us migrate it. Here is how they compare.

ILLATE migrationRebuild in-houseStay on the OpenAI fine-tune
Keeps working after OpenAI's datesYesYesNo, on their schedule
Retrain when labels changeYesYesNo new jobs after 6 Jan 2027
You own the weightsYesYesNo
Parity proven before you switchWritten margin, intervals, refundIf you build the evaln/a
Your engineers' timeAbout 2 h a weekWeeks of ML and infra workNone, until it stops
Serving and on-callUs, or your cloudYour team runs GPUsOpenAI
Upfront cost$1,500–3,000, refundableEngineering salaries$0
03 Questions

About the money

What counts as "doesn't match"?

Before the final run we agree in writing on a test set and a margin, usually 2 or 3 accuracy points against your current model. If the final parity report isn't a PASS against that margin, we refund the whole fee.

Why is the check free?

It costs us a few dollars of GPU time and tells both sides whether the work is worth doing. Most of what we learn in the check carries straight into the migration.

Will hosting save us money?

It depends on volume. Short classification calls are cheap on OpenAI, so at low volume the case is control and continuity, not savings. High-volume or long-output workloads usually save. The cost check gives you the number before you commit.

Can we host it ourselves?

Yes. We deliver the weights and a vLLM deployment for your cloud. You pay only the migration fee.

Get your number first

The cost check and parity check are free.

Book a parity check