ImprovePrivate Models

Your best runs, trained into a model that’s yours.

Runs your team already approved train a private open-weight model for the work you repeat most, hosted privately behind your gateway, at up to 60% lower cost at scale. Before anything changes, you see how it scores against your current model on the same cases and when it pays for itself.

invoice-extractionFinance · 120,000 invoices a month
Cost per invoice, current model$0.02596.8% correct on your checks
Passed your barOn 500 invoices it never trained on
Private modelTrained on 4,200 invoices
Correct on your checks96.8% today97.1%
Cost per invoice$0.025 today$0.010
Monthly at 120,000$3,000 today$1,200
TrainingOne time$1,800
Pays back its training in a month$1,800 a month saved at today’s volume
What you get
Your expertise stays yours

Reviewed runs and approved examples train a model private to your workspace, never a shared one.

Proof before you switch

Accuracy and cost per successful task against your current model, on cases it never saw in training.

Up to 60% lower cost at scale

High-volume tasks move to your private model once it clears your bar, with break-even shown before you spend.

Start from what works

Teach it with runs your team already approved.

Pick a task you repeat at volume, like routing tickets or extracting invoices. Reviewed agent runs and examples from your own datasets become the training set.

Training set · ticket-routingSupport
Reviewed agent runsMarked correct by Support5,600
Help-desk exportsTickets with approved routing2,400
Customer names and emailsremoved
Ready to train8,000
You choose what it learnsOnly examples your team marked correct
Reviewed examples

You choose exactly what it learns.

Every example is one your team marked correct, with personal and sensitive details removed before training.

Starting modelOpen-weight families
NemotronNVIDIA
GemmaGoogle
LlamaMeta
GLMZ.ai
DeepSeekDeepSeek

Your workspace shows which models you can train today.

Starting model

Start from an open-weight model that fits the job.

Choose from families by NVIDIA, Google, and Meta, as well as GLM and DeepSeek. The result is your private version, trained on your examples.

Proof first

Know it beats your current model before you switch.

Record scores the trained model and the one you run today on the same held-out cases, so the decision rests on your numbers.

Validation · ticket-routing600 held-out tickets
Routed correctly93.4% today95.0%
Urgent tickets misrouted4 today1
Cost per successful task$0.018 today$0.008
Median response time1.8 s today0.6 s
Ready to adoptMet every bar your team set
Side by side

Accuracy, cost, and speed against today’s model.

Each run ends in a clear verdict: ready to adopt, short of your bar, or too close to call.

Estimate · contract-taggingLegal
TrainingOne time$1,200
HostingScales to zero when idle$2.90 an hour
Break-even volumeAgainst your current model31,000 a month
Your volume14,000 a month
Below break-even todayStay on your current model for now
Costs up front

Training, hosting, and break-even before you spend.

See the one-time training estimate, the hourly hosting price, and the monthly volume where the model pays for itself. When the math says wait, the estimate says so.

Put it to work

Your model, in the Gateway you already use.

Switch the agents you choose. Every other agent stays on the model it uses today.

Gateway · ticket-routing-v1Private to your workspace
support-triageAgentswitched
13 other Support agentsunchanged
Previous modelOne step to restorekept
Before and afterLive conversationsin Monitoring
Only the agents you chooseSame access controls and spend tracking
Your call, agent by agent

Adopt it where it wins. Restore in one step.

It sits beside your other approved models with the same access controls and spend tracking, and Monitoring compares each agent before and after on live conversations.

Private Models

Bring one task you repeat at volume.

We’ll size the training set, estimate training and hosting, and show where a private model pays for itself. Need a specific region? We’ll confirm availability first.