All models

deepseek-v4-pro pricing

Per million tokens, in US dollars. Synced from the public price feeds on 2026-08-07, and these are the same numbers KostLens bills from.

What it costs

Provider Input Output Cached input Cache write Cache write, 1h
deepseek $0.435 $0.870 $0.004 $0.000 -
tencent $0.435 $0.870 $0.004 $0.000 -
azure $1.74 $3.48 - - -
fireworks $1.74 $3.48 $0.145 - -

A dash means the provider does not publish that rate, not that it is free.

The same model is not the same price

deepseek-v4-pro is sold by 4 providers here, and the dearest charges 300% more per input token than the cheapest: $1.74 on fireworks against $0.435 on deepseek. That gap is real money on a workload of any size, and it is invisible in a provider console, which only ever shows you its own bill.

The rate most price tables leave out

Reading from the prompt cache is cheap, and everybody publishes that number. WRITING to it is a separate, disjoint set of tokens billed above the input rate, and almost nobody publishes it. On deepseek, deepseek-v4-pro charges $0.435 for fresh input and $0.000 to write into the cache.

It matters because a cost tool that folds cache writes into plain input understates the bill of every heavy cache user, and the people who reuse context are exactly the ones caching the most. KostLens counts them apart.

Know what this costs you, per customer

These are list prices. What they turn into on your bill depends on your token mix, and which of your customers is generating it. KostLens wraps your existing client in three lines, reads the usage the provider already returns, and gives you cost per customer and per feature. Your prompts and API keys never touch our servers.

Start free Price your own volume