On July 16, Moonshot AI released Kimi K3. It performs nearly as well as the best models from OpenAI and Anthropic, and it costs about a third of the price.

The old story was that Chinese models were cheap but a step behind. That story is over. K3 sits close to the frontier on performance, and the low price is a straight advantage for whoever is buying.

Here are the numbers. K3 charges $15 per million output tokens. Claude Fable 5, one of the top models, charges $50 for the same work (Fortune). On published benchmarks, K3 beats Claude Opus 4.8 and GPT 5.5 on coding and agent tasks. Only two models in the world still rank above it: Fable 5 and GPT 5.6 Sol. It is also the largest open-weight model ever built, at 2.8 trillion parameters and a one million token context window.

What this means for your costs

Every AI feature you run has a price per token. Product descriptions, support replies, ticket routing, content generation. Each one is a model call with a cost attached.

When a model this capable arrives at this price, it pulls the whole market down with it. You do not have to switch to Kimi to benefit. More competition at the top forces every provider, including the one you already use, to lower prices or lose customers.

The market saw it right away. When K3 shipped, shares of its Chinese rival z.ai fell 28 percent in a single day (CNBC). That is what price pressure at the top looks like: it does not just squeeze the labs you have heard of, it squeezes every competitor trying to hold a price line.

What this means for your decisions

You will probably never run K3 yourself. It needs at least 64 accelerators to serve, which no store or small company is going to set up. Open weights here means the inference market runs it for you and undercuts the closed labs on price.

The practical takeaway is about lock-in. The best model keeps changing every few months. It was DeepSeek, then GLM, now K3, and there will be another soon.

If switching providers means rebuilding your systems, you are stuck paying whatever your current vendor charges. K3 ships with an OpenAI-compatible API so you can move without a rewrite. Build so you can always buy the cheapest model that does the job.

A couple of things worth flagging. Moonshot ran its own benchmarks, but independent tests from Artificial Analysis have largely confirmed the picture, placing K3 next to Opus 4.8 and GPT 5.5, behind Fable 5 and GPT 5.6 Sol. On the other side, K3 shows a higher hallucination rate than the previous version, so check accuracy on your own tasks before trusting it blindly. The downloadable weights only land on July 27, with the license still unconfirmed, so self-hosting is not a production plan yet.

None of that changes the direction. Capable AI keeps getting cheaper, and it keeps coming from more places.

We made this point back in February in Intelligence is Everywhere. Advantage is Not. Once intelligence becomes cheap and easy to buy, it stops being the advantage. K3 is the clearest proof of that yet. Six months ago the frontier belonged to a handful of labs charging accordingly. Now a near-frontier model ships at a third of the price, from a company most of your customers have never heard of. The input got commoditized exactly like we said it would.

Which means the same conclusion applies here too: access to the smartest model is not a strategy. How you architect around it is. The company that treats its choice of model as permanent is the one that will overpay for it.