China's 2.8-trillion-parameter Kimi K3 beats Claude Fable 5 in Frontend Code Arena benchmark— Moonshot AI delivers largest open-weight AI model ever, as China works around U.S. compute limits
Moonshot's own disclosures point to export-grade Nvidia silicon and an unnamed alternative GPU vendor.
Beijing-based Moonshot AI has released Kimi K3, a 2.8 trillion parameter model that the company describes in its technical blog as the world's first open 3T-class system and the largest open-weight AI model to date. Moonshot said K3 still sits behind Anthropic's Claude Fable 5 and OpenAI's GPT 5.6 Sol on overall performance, but it outperformed every other model in the company's evaluation suite, including Claude Opus 4.8 and GPT 5.5, across coding and agentic benchmarks. The model has a 1 million token context window, native vision, and activates just 16 of its 896 experts per token, roughly 1.8% of the pool. Full weights are due by July 27.
Arena ranked K3 first in its Frontend Code evaluation at 1,679 points, ahead of Fable 5, in blind developer testing. API pricing is $0.30 per million cache-hit input tokens, $3 per million on cache misses, and $15 per million output tokens. Kimi K2 launched a year ago at $0.60 per million input tokens, so uncached K3 input costs five times as much.
Big news: Kimi-K3 by @Kimi_Moonshot is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5.This is a 17-place jump from Kimi-k2.6 (#18 -> #1).In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics,… https://t.co/YDN3BufGkC pic.twitter.com/Oa6teaQnWpJuly 16, 2026
Moonshot claims roughly a 2.5x improvement in scaling efficiency over Kimi K2, attributed to two architectural changes: Kimi Delta Attention, a hybrid linear attention scheme, and Attention Residuals, which change how information moves between layers. Quantization-aware training starts at the supervised fine-tuning stage, using MXFP4 weights and MXFP8 activations, a combination Moonshot says it chose for broad hardware compatibility. Bank of America analysts led by Alex Liu said in a note cited by CNBC that K3 shows large-scale pre-training plus architectural work can still deliver step-change gains for flagship Chinese models despite compute constraints.
Moonshot's kernel optimization benchmark ran on Nvidia's H200, and what the blog identifies only as a "GPGPU from an alternative vendor," which the company didn't name. MiniTriton, a Triton-like compiler K3 built from scratch, is charted against Triton on an Nvidia L20, the cut-down Ada-based card sold into China under U.S. export rules. Moonshot recommends serving K3 on supernodes of 64 or more accelerators, keeping expert-parallel traffic inside one high-bandwidth domain. The blog doesn't say where the H200 hardware is; Congress passed a bill in January to close the offshore cloud rental loophole that gave Chinese firms remote access to restricted accelerators.
In one case study, K3 spent a single 48-hour autonomous run designing a simulated inference chip for a nano model built on its own architecture, using open-source EDA tools and the Nangate 45nm library. The design closed timing at 100 MHz within 4mm squared, packed 1.46 million standard cells and an INT4 MAC array, and sustained more than 8,700 tokens per second of simulated decode.
At the moment, every published K3 number is a claim made by Moonshot-reported or drawn from API access and can’t be verified until the weights are made public on July 27. Anthropic accused Moonshot in February of using 3.4 million Claude exchanges to train its models through distillation, and K3 now benchmarks within a few points of the models named in that complaint.
Follow Tom's Hardware on Google News, or add us as a preferred source, to get our latest news, analysis, & reviews in your feeds.
Get Tom's Hardware's best news and in-depth reviews, straight to your inbox.

Luke James is a freelance writer and journalist. Although his background is in legal, he has a personal interest in all things tech, especially hardware and microelectronics, and anything regulatory.
-
heffeque Truth hurts sometimes.Reply
You'll see hurt people lash out statements of disbelief when presented with proof.
Clearly things are changing, but some people prefer to think that the US doesn't have a serious contender.
Sadly Europe is nowhere to be found. -
jp7189 Reply
The problem is the proof is sometimes not as awesome as it seems. Benchmarking AI can be a cat and mouse game as its possible to train AI on a particular benchmark without noticeably improving real world results. Qwen36 looks like it should beat Gemma4, and for a while I was regularly switching between them, but now I find myself sticking with Gemma4 because it produces usable results more often. YMMV.heffeque said:Truth hurts sometimes.
You'll see hurt people lash out statements of disbelief when presented with proof.
Clearly things are changing, but some people prefer to think that the US doesn't have a serious contender.
Sadly Europe is nowhere to be found.
That's not to take anything away from Chinese models as they are producing some good things. Just cautioning against getting over excited about 1st party benchmarks that have yet to be verified. -
PEnns Replyheffeque said:Truth hurts sometimes.
You'll see hurt people lash out statements of disbelief when presented with proof.
Clearly things are changing, but some people prefer to think that the US doesn't have a serious contender.
Sadly Europe is nowhere to be found.
This is another "Made in Japan" moment of denial where it was used as a derogatory statement in the US (and elsewhere, but not to the same extent or vehemence). Then suddenly in the 1970s, EVERYBODY was buying Japanese cars, electronics, etc...!!
And now we have the "Made in China" denial movement whose disciples refuse to accept the obvious and seem to never have studied the word "Hubris" or know its meaning... -
johnnycanadian Reply
This, exactly. I lived through the (pardon the term) "Jap Crap" era of the 1970s when every domestic manufacturer believed that Japan was a joke, capable of only exporting rubber dog doo or cheap tin toys. Then the 1980s happened. Same thing with the Korean vehicle manufacturers: the mid-1980s Hyundai Pony was terrible in every measurable way save for price, but now, Hyundai/Kia/Genesis is firing on all cylinders.PEnns said:And now we have the "Made in China" denial movement whose disciples refuse to accept the obvious and seem to never have studied the word "Hubris" or know its meaning...
We're about to start receiving Chinese EVs in Canada, and a couple of the BYD Denza models look spectacular enough that I'm going to happily trade in my Tesla once dealerships are established in the lower mainland. I had an opportunity to check them out last time I was in Mexico and I was very impressed! -
nookoool Replyjohnnycanadian said:This, exactly. I lived through the (pardon the term) "Jap Crap" era of the 1970s when every domestic manufacturer believed that Japan was a joke, capable of only exporting rubber dog doo or cheap tin toys. Then the 1980s happened. Same thing with the Korean vehicle manufacturers: the mid-1980s Hyundai Pony was terrible in every measurable way save for price, but now, Hyundai/Kia/Genesis is firing on all cylinders.
We're about to start receiving Chinese EVs in Canada, and a couple of the BYD Denza models look spectacular enough that I'm going to happily trade in my Tesla once dealerships are established in the lower mainland. I had an opportunity to check them out last time I was in Mexico and I was very impressed!
Just adding the lesser known was "Taiwan Junk , Bootlegs, pc clones" era -
PEnns Replyjohnnycanadian said:This, exactly. I lived through the (pardon the term) "Jap Crap" era of the 1970s when every domestic manufacturer believed that Japan was a joke, capable of only exporting rubber dog doo or cheap tin toys. Then the 1980s happened. Same thing with the Korean vehicle manufacturers: the mid-1980s Hyundai Pony was terrible in every measurable way save for price, but now, Hyundai/Kia/Genesis is firing on all cylinders.
We're about to start receiving Chinese EVs in Canada, and a couple of the BYD Denza models look spectacular enough that I'm going to happily trade in my Tesla once dealerships are established in the lower mainland. I had an opportunity to check them out last time I was in Mexico and I was very impressed!
And to add insult to injury, the same people who are decrying the "Made in China" stuff....are posting or disseminating their ignorance on phones, tablets, PCs, you name it, the vast majority of which is.....Made in China!! -
russell_john People seem to be missing what American Big AI is scared of happening. This will be Open Source meaning if you have enough memory, storage, CPU and GPU power you can download it and run it yourself.Reply
But I have complete confidence that the Federal Government will find an excuse to keep it out of the hands of Americans and force us all to pay an exorbitant price to the American AI Cartel that are desperate to get their investment back. -
usertests Reply
That will work as well as the war on piracy.russell_john said:But I have complete confidence that the Federal Government will find an excuse to keep it out of the hands of Americans and force us all an exorbitant price to the American AI Cartel that are desperate to get their investment back.
The main obstacle to running frontier open source models locally is the cost of admission.