AI DESK · HONG KONG · WEEKLY

Kimi K3 Matches Claude Coding At A Third Of The Price

Kimi K3 matched frontier coding performance at a third of Anthropic's price the same week Nvidia halved its Asia buyer list, and the second event did not stop the first.
AN

Nvidia's Buyer List Halves

Nvidia's Asia buyer whitelist is the paperwork that lets Blackwell-class GPUs (Nvidia's newest AI training chips, the generation after the H100 and H200 that Washington already restricts) leave the country legally. In mid-July Nvidia cut that whitelist, which had covered approved buyers in Singapore, Malaysia and Japan, by more than half. The trigger was a compliance sweep after the US Commerce Department flagged Blackwell-class chips reaching Chinese-linked entities by routing through Malaysia. The export-control framework this whitelist enforces is not a wall around a country, but a chain of approved buyers each vouching for where the chip goes next. Malaysia was the weak link in that chain, and Nvidia responded by cutting the compromised route out of it. A compliance list built to prove the chips were staying put had to be cut in half just to keep working. For a procurement lead sourcing GPU capacity anywhere in Southeast Asia this quarter, the change is immediate: fewer approved counterparties, slower diligence on end-use certificates, and every buyer relationship that touches Malaysia getting re-underwritten from scratch.

Cheaper, Faster, Unverified

Moonshot AI, the Alibaba- and Tencent-backed lab in Beijing, released Kimi K3 on July 17: 2.8 trillion total parameters, a one-million-token context window, and a mixture-of-experts architecture that activates only 16 of 896 possible expert subnetworks per query, the trick that keeps running cost down even at that size. In blind coding-benchmark testing it beat both Anthropic's Claude Fable 5 and OpenAI's GPT-5.6 Sol, and Moonshot is charging $3 per million input tokens and $15 per million output tokens against roughly $50 per million output tokens for Fable 5, about a third of the price for equal or better work. Anthropic CEO Dario Amodei had reportedly not expected a Chinese lab to near the US frontier for another six months; Elon Musk guessed Q1 2027. Both were wrong on timing: Kimi K3 beat both models on the coding benchmark outright. Whether they were wrong on substance is what the July 27 weight release will show: Moonshot is withholding the full model weights until that date, and only then can anyone outside the company confirm what chips actually trained it. TSMC fell 7 percent despite a 77 percent jump in quarterly profit that same week; Nvidia dropped 1.2 percent, briefly losing its title as the world's most valuable company to Apple. For a CTO deciding whether to route production coding workloads to Kimi K3's API today, the instruction is to wait for July 27, when the weight release settles what actually built it.

July 27 is the date that actually matters here. That's when Moonshot releases the weights researchers need to check what chips actually trained Kimi K3, and it's the only way to know whether Nvidia's halved whitelist closed a leak or missed one. Until then, every read on this, mine included, is a bet on which failure mode the export-control regime has: too slow, meaning the leak stayed open long enough for a restricted chip to train a frontier model, or too late, meaning the fix arrives only after the compromised route has already done its work.

Sources

PREVIOUS COLUMNS, AI DESK DESK
The Wang Report's columns are produced by AI under human editorial oversight. See our Editorial Standards.