Kimi K3: The new Chinese AI model that challenges US giants

  • Moonshot AI releases Kimi K3, a 2,8 trillion parameter model, the largest open source model to date.
  • The system outperforms Claude Opus 4.8 and GPT-5.5 in multiple tests, and leads in web interface creation.
  • Their full weights will be published on July 27, and their price per API is much lower than their rivals.
  • The launch reignites the debate about China's technological advantage in the face of chip restrictions.

Kimi K3 artificial intelligence model

Chinese artificial intelligence makes another statement. Moonshot AI has presented Kimi K3A language model with 2,8 trillion parameters, making it the largest open weights system ever created. The company, backed by Alibaba, claims its new model matches or surpasses leading systems from Anthropic and OpenAI in several key tasks, particularly in interface programming and complex reasoning.

The announcement comes just months after DeepSeek proved that China could compete at the forefront of AI. Now, with Kimi K3, the feeling of déjà vu returns: a model that appears without much fanfare sneaks into the top spots of benchmarks and forces Silicon Valley to question whether its leadership is still solid. The developer community has already put it to the test, and the results speak for themselves.

Kimi K3
Related article:
Kimi K3: the new Chinese AI model that challenges OpenAI and Anthropic

A leap in parameters and performance

Kimi K3 programming interface

Kimi K3 boasts 2,8 trillion total parameters, a sparse architecture that activates 16 of its 896 experts , and a context window of one million tokens. This allows it to process lengthy documents, hold long conversations, and execute multi-step programming tasks. According to Moonshot, the model is 75% larger than DeepSeek V4 Pro (1,6 trillion) and three times the size of its predecessor, Kimi K2.

In internal tests, Kimi K3 has outperformed Claude Opus 4.8 and GPT-5.5 in areas such as GPU kernel optimization, long-term coding, and agent-based reasoning . The company acknowledges that it still lags behind the more advanced Claude Fable 5 and GPT-5.6 Sol, but the gap has narrowed considerably. On the Arena.ai platform, which conducts blind evaluations with real developers, Kimi K3 achieved first place in six out of seven categories, including web design, data analysis, and content creation.

Results in benchmarks and comparisons

Kimi K3 benchmark comparison

The Frontend Code Arena ranking, which measures models' ability to build functional web interfaces, places Kimi K3 in first place with 1.679 points, ahead of Claude Fable 5 (1.631) and GPT-5.6 Sol (1.618). Vals AI ranks it second globally with 74,70%, only behind Fable 5 , while Artificial Analysis puts it on par with GPT-5.5 and Claude Opus 4.8 in overall performance. These evaluations, although preliminary, indicate that the Chinese model not only competes but leads in practical and user-visible tasks.

In addition to programming, Kimi K3 excels at tasks that combine vision, reasoning, and agent execution. Moonshot has incorporated a technique called "vision in the loop" that allows the model to review screenshots of its own work and correct errors on the fly. This explains why it's able to generate interactive applications, such as a browser-based recreation of macOS 27, from a simple description and a reference image.

Opening strategy and pricing

Kimi K3 open source

Unlike its US rivals, who keep their models under wraps, Moonshot has confirmed that the full Kimi K3 database will be released on July 27. Any company or researcher will be able to download it, run it on their own servers, and adapt it without paying for licenses. This strategy, the same one popularized by DeepSeek, makes the model reusable infrastructure and accelerates its adoption.

The API pricing is also aggressive: $3 per million input tokens and $15 per million output tokens , significantly lower than Claude Opus 4.8's $5 and $25, respectively, and nearly half of the latter's $1,80 per task. While Kimi K3's average task cost is $0,95, it's still cheaper than GPT-5.6 Sol ($1,04) and much more affordable than Opus 4.8 ($1,80). This makes it an attractive option for companies seeking processing power without breaking the bank.

Geopolitical implications

The launch of Kimi K3 comes amid escalating trade tensions between the United States and China. Export restrictions on advanced chips haven't stopped Chinese labs from innovating , and this model demonstrates that efficient hardware utilization can compensate for limited access to the most powerful processors. Scarcity, when managed effectively, has become an incentive to optimize algorithms and architectures.

Furthermore, the political reaction was swift. David Sacks, former White House AI official, used Kimi K3's result to criticize domestic restrictions in the US, stating that "this is how you lose the AI ​​race." Meanwhile, in China, Premier Xi Jinping championed open source as a way for all sectors to benefit from artificial intelligence at the opening of the Shanghai AI World Conference.

The race won't necessarily be won by the one who builds the biggest model, but by the one who manages to spread their innovations the fastest. Kimi K3 confirms that the Chinese ecosystem, with its focus on open models, low prices, and an active community, has ceased to be an anomaly and has become a structural competitor . DeepSeek was the warning; Kimi K3, the confirmation that AI leadership no longer has a clear owner.


Add as preferred source in Google