Version

Kimi K3 Max

Kimi K3 scores 1674 on the WebDev Arena leaderboard, eighteen points below Opus 5, and at $3 in and $15 out it costs about forty percent less. The thing to know first: K3 Max is not a model in Moonshot's own documentation.

  • 1674 on WebDev Arena
  • 1,048,576 token window
  • cached input at $0.30

current Maker: Moonshot AI

Last checked: This is a reference page. It is re-checked against the vendor sources and updated when a new version ships. What changed

Kimi K3 Max is not the name of a model

The row you see on the Arena leaderboard is called kimi-k3-max, but Moonshot's model catalogue lists no model by that name. It lists kimi-k3. The max suffix is a setting, the amount of effort the model spends reasoning, not a separate product with a separate price.

The same pattern repeats on GLM-5.2 and on DeepSeek, and most comparison lists treat each row as its own model. We kept the slug people actually search for, and attached the price and the context window to the base model, because the source that publishes those two numbers is talking about the base model.

The price, and a second rate that only applies to repeated prompts

Three dollars per million input tokens and fifteen per million output. Against Opus 5 at $5 and $25, that is forty percent less. Moonshot publishes a second input rate too: when the prompt is already cached, input bills at $0.30, a tenth of the standard rate.

We did not put that $0.30 in the price column, for a plain reason. No real workload is one hundred percent cache. If your architecture has a long fixed prefix repeated a thousand times a day, the cached rate is real for you and you should work it out yourself. If every request is fresh, three dollars is what you pay.

Why it comes third rather than second in our table

Kimi trails Fable 5 by two points in the coding table, and all two of them come out of one column: tooling maturity. Fable 5 takes full marks and the whole fifteen points of weight there; Kimi takes zero, because Moonshot's platform documentation lists neither an official command line tool nor an editor extension. On the leaderboard itself, Kimi is ahead of Fable 5.

So if your team writes its own tooling layer and does not need a ready-made agent SDK, those fifteen points mean nothing to you, and Kimi effectively sits higher than our table shows. The table cannot know that about you. You can.

The column Kimi wins outright, and why we distrust it

Kimi's context window is 1,048,576 tokens. That is two to the twentieth, not a rounded marketing million. Every other candidate with a published window sits at exactly one million, so a difference of under five percent hands Kimi all ten points of the context column and everyone else zero.

Min-max normalisation does that, and we do not bend it, because the moment one column gets adjusted by hand the rest of the table stops meaning anything. But a reader should know those ten points are not ten points of real advantage. In practice the two windows are the same size.

Specifications

API model ID kimi-k3 Source
Context window 1 million tokens Source
Input price $3 per million tokens Source
Output price $15 per million tokens Source

Using it from Iran

This column carries a date and says how it was checked, because most listicles guess it. Where we have only read a vendor policy page, the note below says exactly that.

Reachable
not checked
Payment
not checked
Free tier
no

How we checked: Moonshot publishes no supported-countries list, so we can write neither open nor blocked. Unchecked is a real answer here, and it scores nothing in the ranking.

What it is good at

  • 1674 on WebDev Arena, above Fable 5 and below Opus 5
  • The cheapest of the top three in our coding table at $3 and $15
  • The largest context window in the table, 1,048,576 tokens
  • A cached input rate of $0.30, for architectures with a long fixed prefix

Where it falls short

  • Moonshot documents no model called Kimi K3 Max, so what the leaderboard measured is a setting rather than a product.
  • It scores zero in the tooling column: no official command line tool and no editor extension in the documentation.
  • Moonshot publishes no supported-countries list, so the Iran column stays unchecked and takes no points.
  • On the general text leaderboard it scores 1487, below both Opus 5 and Fable 5.
  • The $0.30 rate is only real when the prompt genuinely hits the cache, which you have to work out on your own workload.

Our take

If you write your own tooling layer and your token spend is high, Kimi K3 is the most sensible trade in this table: eighteen points behind Opus 5 at forty percent less. If you want to pick a model up on day one and work with the tooling that comes with it, this is not the one.

Questions people actually ask

Is Kimi K3 Max different from Kimi K3

Moonshot's catalogue only has kimi-k3. Max is the reasoning effort level, set on that same model, so the price and the context window are identical.

What does Kimi K3 cost

Three dollars per million input tokens and fifteen per million output tokens. When the prompt is cached, input bills at $0.30.

Does Kimi work from Iran

We do not know, and that is what the page says. Moonshot publishes no supported-countries list and we have not tested it from an Iranian connection ourselves. Any answer other than unchecked would be a guess today.

For code, Kimi or Opus 5

On the web-dev leaderboard Opus 5 leads at 1692 against Kimi's 1674. The score gap is small and the price gap is not, so the answer turns on how much ready-made tooling is worth to you.

Sources

  1. Kimi K3 pricing vendor source 12 August 2026
  2. Kimi platform models vendor source 12 August 2026