Claude Fable 5
Featuredanthropic/claude-fable-5Anthropic's most capable model. Top of SWE-bench Verified and top of the creative-writing arena — the one to reach for when being right matters more than what it costs.
Every model you can run, grouped by publisher. Prices are per million tokens in british pounds, and are what you actually pay — there is no separate platform fee on top.
You do not have to choose from this list. Pass auto and we pick the best-value model for each request — here is how that works.
anthropic/claude-fable-5Anthropic's most capable model. Top of SWE-bench Verified and top of the creative-writing arena — the one to reach for when being right matters more than what it costs.
anthropic/claude-opus-5Complex agentic coding and long-horizon work at half the price of Fable 5. Leads Terminal-Bench 2.1 and the maths arena.
anthropic/claude-sonnet-5Near-Opus coding quality at Sonnet money — 0.852 on SWE-bench Verified for a fifth of Fable 5. One of the best value-per-point models in the catalogue.
anthropic/claude-opus-4-8Previous-generation Opus. Still highly autonomous on long-horizon agentic work.
anthropic/claude-haiku-4-5Fast and cheap for simple, high-volume work.
alibaba/qwen3.8-maxAlibaba's newest flagship — 2.4T-parameter MoE, natively multimodal, third in the creative-writing arena. Flat pricing across the whole 1M window, with no long-context surcharge.
alibaba/qwen3.7-plusOutstanding knowledge-task quality for the money — 0.885 MMLU-Pro at a fraction of frontier pricing.
alibaba/qwen3.7-flashThe floor of the catalogue. For classification, routing and high-volume trivial chat.
alibaba/qwen3-coder-nextOpen-weight coding specialist with a very large output budget.
openai/gpt-5.6-solOpenAI's flagship. Leads Terminal-Bench 2 and tops the GPQA science leaderboard.
openai/gpt-5.6-terraThe balanced GPT-5.6 tier. Strong general and agentic performance at mid-range pricing.
openai/gpt-5.6-lunaSmall, quick and very cheap, with the full million-token window.
google/gemini-3.6-flashPunches far above its price on maths — fourth in the arena, level with models costing several times more.
google/gemini-3.1-proGoogle’s reasoning flagship, strong across science and long-context work.
google/gemini-3.5-flash-liteCheap, fast, million-token context. A good default for summarisation and extraction.
deepseek/deepseek-v4-proFrontier-adjacent quality at a tenth of frontier prices, and by far the most aggressive prompt-cache rate on the market — cached input costs under 1% of fresh.
deepseek/deepseek-v4-flashRemarkable coding quality for the price — 0.79 SWE-bench at pennies per million tokens.
zai/glm-5.2Open-weight flagship that effectively saturates scripted tool-use benchmarks (99.1% on τ²-bench).
zai/glm-4.7-flashVery cheap, and near-perfect on scripted tool workflows (98.8% τ²-bench).
minimax/minimax-m3Statistically level with Gemini 3.1 Pro on SWE-bench Verified at a fraction of the price. The clearest example of why routing pays.
moonshot/kimi-k3Top-tier agentic coding, and unusually high prompt-cache hit rates in production.
moonshot/kimi-k2.7-codeOpen-weight coding specialist tuned for long edit sessions.
xai/grok-4.5Strong general-purpose reasoning with competitive coding scores.
mistral/mistral-medium-3-5European-hosted general model with solid coding performance.