Skip to content
The unique bit

One API. The right model every time.

Model quality is not one number. A model that leads on agentic coding can sit mid-table on maths; one costing a fiftieth as much can be statistically level on the task actually in front of you. Progressive Labs works out which kind of work each request is, then picks the cheapest model that still clears the quality bar you set.

How a request is decided

01

Work out what kind of task this is

Some of it is certain rather than guessed: an attached image means vision; declared tools mean tool use; a 300,000-token prompt means long context. The rest comes from the shape and wording of the request. It costs nothing and takes under a millisecond, because anything that adds latency to every request has to justify itself against a saving measured in thousandths of a penny.

02

Throw away everything that cannot do the job

Context window too small, no tool support, no vision, currently rate-limited or erroring. These are hard facts, not preferences, and they are applied before price is considered at all.

03

Apply your quality floor

Each remaining model has a published benchmark score for this task type, normalised against the best model available for it. Anything below your floor is removed. We test against the lower end of the confidence interval, so a model never sneaks onto the shortlist on the strength of one thin, unverified number.

04

Only now, look at price

Among the survivors — and only among the survivors — the dial decides how hard to push for value. It cannot reach below the floor, because everything below the floor was already gone.

05

Tell you what happened

Which model ran, what it scored, how many alternatives cleared the bar, and what your baseline would have cost. If routing cost you more on a request than pinning would have, we show that too.

The fourteen task types

A task type earns its place only if model rankings genuinely reorder on it. Below: the best model for each, the cheapest that still clears the default floor, and the gap between them. Generated from the live catalogue of 24 models.

Task typeDefault floorBest modelCheapest above the floorPrice gap
Simple chat70%Claude Fable 5DeepSeek V4 Flash75%−99%
General chat82%Claude Fable 5DeepSeek V4 Pro89%−97%
Summarisation85%Claude Fable 5DeepSeek V4 Pro91%−97%
Extraction & classification85%Claude Fable 5DeepSeek V4 Flash88%−99%
Translation85%Claude Fable 5DeepSeek V4 Pro89%−97%
Code generation90%Claude Fable 5Claude Sonnet 591%−70%
Code review & debugging92%Claude Fable 5Claude Opus 593%−50%
Agentic coding96%Claude Fable 5Claude Fable 5100%same model
Maths & logic90%Claude Opus 5Qwen3.7 Plus99%−94%
Science & research92%Claude Fable 5DeepSeek V4 Flash97%−99%
Long-context work90%Claude Fable 5DeepSeek V4 Pro91%−97%
Tool use & agents94%Claude Fable 5Claude Fable 5100%same model
Creative writing85%Claude Fable 5Qwen3.8 Max96%−84%
Vision88%Claude Fable 5DeepSeek V4 Pro90%−97%

“Price gap” compares blended per-token rates, weighted three-to-one toward input because that is how real traffic runs. It is the headroom available on that task type, not a promise about your particular workload — what you actually save depends on what you actually send.

What it will never do

Silently downgrade you to save money
The floor is a hard constraint. Price is only ever a tie-break between models that already cleared it.
Override a model you pinned
Pin a model and that is exactly what runs. We only report afterwards what routing would have chosen.
Guess confidently when it is unsure
Low classifier confidence raises the floor automatically. An uncertain router should be an expensive one.
Fail your request because routing broke
If classification is unavailable we fall back to a conservative default and carry on. A routing problem must never become an outage.
Show you a saving we cannot substantiate
Every figure is measured against the baseline you declared, before the request ran, and labelled as an estimate where it is one.
Throw away a warm prompt cache to chase a cheaper rate
Cache state is part of the cost model. Within a session we keep you on one model so the discount survives.

See it decide on your own work

Put Auto against the model you use today. If we do not beat it on your real prompts, you have lost the price of a few tokens finding out.

Compare against your model