Price a workload, not a token

Headline input and output prices answer the wrong question. What you actually pay depends on the shape of your traffic — and for agentic coding tools, most of the tokens are cache reads, which many price lists barely mention. Put your own numbers in and see what every model in the catalog would cost, at live KoRouter rates and at each provider's official prices.

Your token mix

Every number is yours to change. If you don't know your mix yet, start from one of these shapes — they are illustrative starting points, not measurements.

Two focused hours in Claude Code — context re-read on every request

M tokens
K tokens
K tokens
K tokens

5.70M tokens in total · on the cheapest model, deepseek-v4-flash, cache reads and writes are 41% of the bill

What it costs, cheapest first

ModelKoRouterOfficialCache share
deepseek-v4-flashDeepSeek$0.03$0.2241%
deepseek-v4-flash-0731DeepSeek$0.03$0.2241%
deepseek-v4-proDeepSeek$0.22$0.3348%
deepseek-v4-pro-0813DeepSeek$0.22$0.3348%
glm-5.2Zhipu GLM$0.90$2.0969%
claude-haiku-4-5Anthropic$1.10$1.6956%
gpt-5.4OpenAI$1.25$3.3837%
gpt-5.6-terraOpenAI$1.32$3.5852%
gpt-5.6-solOpenAI$1.35$6.7556%
glm-5.3Zhipu GLM$2.06$2.0969%
claude-sonnet-5Anthropic$2.19$3.3856%
gpt-5.5OpenAI$2.50$6.7537%
kimi-k3Kimi$3.36$3.7553%
claude-opus-5Anthropic$5.48$8.4456%
claude-opus-4-6Anthropic$5.48$8.4456%
claude-opus-4-7Anthropic$5.48$8.4456%
claude-opus-4-8Anthropic$5.48$8.4456%
gpt-6-astraOpenAI$6.24$16.8856%
claude-fable-5.1Anthropic$8.53$13.1343%
claude-fable-5Anthropic$10.97$16.8856%

Prices come from the live billing catalog — the same one that bills your requests. “Official” is the provider's published list price where they publish one; a blank cell means they don't. Models billed per request or per image are not in this table because a token mix cannot price them.

Why the cache lines matter

Agentic coding tools re-send the system prompt, the project context and the conversation history on every request. Providers cache that context, so it bills at the cache-read rate rather than the input rate — which is why a session can be 80–90% cache reads by token count while the headline input price barely touches the bill. A discount that applies only to input and output, and not to the cache lines, is worth far less than it looks. The discount badge analysis works one example through end to end, and what a real Claude Code session costs breaks a single session down line by line.

On KoRouter one multiplier applies to every price component of a model, cache reads and cache writes included, and the provider's official price sits next to ours in the model marketplace so you can check the arithmetic yourself.

Free to sign up, no monthly fee — pay as you go with credits.

Browse models