Skip to content

Model catalogue

Every model you can run, grouped by publisher. Prices are per million tokens in british pounds, and are what you actually pay — there is no separate platform fee on top.

You do not have to choose from this list. Pass auto and we pick the best-value model for each request — here is how that works.

Anthropic

5 models

Claude Fable 5

Featured
anthropic/claude-fable-5

Anthropic's most capable model. Top of SWE-bench Verified and top of the creative-writing arena — the one to reach for when being right matters more than what it costs.

reasoningagentic1M context
Input /M£9.18
Output /M£45.90
Context1m
Try now

Claude Opus 5

Featured
anthropic/claude-opus-5

Complex agentic coding and long-horizon work at half the price of Fable 5. Leads Terminal-Bench 2.1 and the maths arena.

codingagentic1M context
Input /M£4.59
Output /M£22.95
Context1m
Try now

Claude Sonnet 5

Featured
anthropic/claude-sonnet-5

Near-Opus coding quality at Sonnet money — 0.852 on SWE-bench Verified for a fifth of Fable 5. One of the best value-per-point models in the catalogue.

balancedcoding1M context
Input /M£2.87
Output /M£14.34
Context1m
Try now

Claude Opus 4.8

anthropic/claude-opus-4-8

Previous-generation Opus. Still highly autonomous on long-horizon agentic work.

agentic1M context
Input /M£4.59
Output /M£22.95
Context1m
Try now

Claude Haiku 4.5

anthropic/claude-haiku-4-5

Fast and cheap for simple, high-volume work.

fastcheap
Input /M£1.07
Output /M£5.36
Context200k
Try now

Alibaba

4 models

Qwen3.8 Max

Featured
alibaba/qwen3.8-max

Alibaba's newest flagship — 2.4T-parameter MoE, natively multimodal, third in the creative-writing arena. Flat pricing across the whole 1M window, with no long-context surcharge.

newflat 1M pricingmultimodal
Input /M£1.91
Output /M£5.74
Context1m
Try now

Qwen3.7 Plus

alibaba/qwen3.7-plus

Outstanding knowledge-task quality for the money — 0.885 MMLU-Pro at a fraction of frontier pricing.

cheapknowledge
Input /M£0.343
Output /M£1.37
Context1m
Try now

Qwen3.7 Flash

alibaba/qwen3.7-flash

The floor of the catalogue. For classification, routing and high-volume trivial chat.

cheapestfast
Input /M£0.037
Output /M£0.159
Context1m
Try now

Qwen3 Coder Next

alibaba/qwen3-coder-next

Open-weight coding specialist with a very large output budget.

codingopen weights
Input /M£0.147
Output /M£0.979
Context262.1k
Try now

OpenAI

3 models

GPT-5.6 Sol

Featured
openai/gpt-5.6-sol

OpenAI's flagship. Leads Terminal-Bench 2 and tops the GPQA science leaderboard.

reasoningagentic
Input /M£4.59
Output /M£27.54
Context1.1m
Try now

GPT-5.6 Terra

openai/gpt-5.6-terra

The balanced GPT-5.6 tier. Strong general and agentic performance at mid-range pricing.

balanced
Input /M£1.91
Output /M£11.48
Context1.1m
Try now

GPT-5.6 Luna

openai/gpt-5.6-luna

Small, quick and very cheap, with the full million-token window.

cheapfast
Input /M£0.245
Output /M£1.47
Context1.1m
Try now

Google

3 models

Gemini 3.6 Flash

Featured
google/gemini-3.6-flash

Punches far above its price on maths — fourth in the arena, level with models costing several times more.

mathsfast
Input /M£1.43
Output /M£7.17
Context1m
Try now

Gemini 3.1 Pro

google/gemini-3.1-pro

Google’s reasoning flagship, strong across science and long-context work.

reasoningscience
Input /M£1.91
Output /M£11.48
Context1m
Try now

Gemini 3.5 Flash Lite

google/gemini-3.5-flash-lite

Cheap, fast, million-token context. A good default for summarisation and extraction.

cheaplong context
Input /M£0.321
Output /M£2.68
Context1m
Try now

DeepSeek

2 models

DeepSeek V4 Pro

Featured
deepseek/deepseek-v4-pro

Frontier-adjacent quality at a tenth of frontier prices, and by far the most aggressive prompt-cache rate on the market — cached input costs under 1% of fresh.

best valueopen weightscache king
Input /M£0.466
Output /M£0.932
Context1m
Try now

DeepSeek V4 Flash

deepseek/deepseek-v4-flash

Remarkable coding quality for the price — 0.79 SWE-bench at pennies per million tokens.

cheapcodingopen weights
Input /M£0.171
Output /M£0.343
Context1m
Try now

Z.ai

2 models

GLM-5.2

zai/glm-5.2

Open-weight flagship that effectively saturates scripted tool-use benchmarks (99.1% on τ²-bench).

tool useopen weights
Input /M£1.50
Output /M£4.71
Context1m
Try now

GLM-4.7 Flash

zai/glm-4.7-flash

Very cheap, and near-perfect on scripted tool workflows (98.8% τ²-bench).

cheaptool useopen weights
Input /M£0.073
Output /M£0.49
Context202.8k
Try now

MiniMax

1 model

MiniMax M3

Featured
minimax/minimax-m3

Statistically level with Gemini 3.1 Pro on SWE-bench Verified at a fraction of the price. The clearest example of why routing pays.

best valuecodingopen weights
Input /M£0.643
Output /M£2.57
Context1m
Try now

Moonshot

2 models

Kimi K3

moonshot/kimi-k3

Top-tier agentic coding, and unusually high prompt-cache hit rates in production.

agenticcoding
Input /M£2.87
Output /M£14.34
Context1m
Try now

Kimi K2.7 Code

moonshot/kimi-k2.7-code

Open-weight coding specialist tuned for long edit sessions.

codingopen weights
Input /M£0.75
Output /M£3.75
Context262.1k
Try now

xAI

1 model

Grok 4.5

xai/grok-4.5

Strong general-purpose reasoning with competitive coding scores.

reasoning
Input /M£1.91
Output /M£5.74
Context500k
Try now

Mistral

1 model

Mistral Medium 3.5

mistral/mistral-medium-3-5

European-hosted general model with solid coding performance.

EU hosting
Input /M£1.43
Output /M£7.17
Context262.1k
Try now