ChengRang

Kimi K3

AI Coding Open Source

Moonshot new flagship released 2026/7/16, 2.8T parameters, largest open-weight model globally; 1M context + native vision; tops Frontend Code Arena

Open SourceCodingMoEMoonshotFrontend Code
Visit Kimi K3

Disclaimer: Review content represents our editorial team's views and experience, not commercial recommendation or investment advice. Product info and pricing may change; refer to official sources.

Overview

Kimi K3 is the new flagship large model officially released by Moonshot AI on the evening of July 16, 2026, and is the company's most capable model to date. K3 has a total of **2.8 trillion parameters** (MoE architecture, dynamically activating 16 out of 896 experts), surpassing DeepSeek V4 Pro's 1.6 trillion, making it the **world's largest open-weight model**. Technically, it adopts the self-developed **KDA hybrid linear attention mechanism** (Kimi Delta Attention) + Attention Residuals technology, natively supports visual understanding, and has a **1 million token context window**.

**Performance highlights**: Topped the Frontend Code Arena with **1679 points**, surpassing Claude Fable 5. It has been simultaneously launched on Kimi App, Kimi Work, Kimi Code, and API, with **full model weights open-sourced on 2026/7/27**.

**Rare honesty**: The official blog homepage explicitly states that "overall capabilities still lag behind Claude Fable 5 and GPT-5.6 Sol," especially in the HLE (Humanity's Last Exam) reasoning test, without only showcasing winning benchmarks like most vendors. This attitude of "daring to display the loss table" has gained considerable recognition in the industry.

The day after release (7/19), reports emerged that Moonshot AI could complete a Hong Kong IPO **within 6 months**, with ARR already reaching $300 million, and K3's release is expected to drive several-fold growth.

Key Features

Use Cases

Pros

Pricing

API pricing: Input $3.00/M token (reduced to $0.30/M with cache hit), output $15.00/M token, significantly higher than K2.6 (input $0.95/M, output $4.00/M), corresponding to the 1 million token ultra-long context and stronger reasoning capability. Available directly in Kimi App/Work/Code (free and paid quotas depend on specific product lines). Full open-source weights will be available for free download and self-deployment after 2026/7/27, but due to the massive total parameter size, local personal deployment is essentially impractical; feasible paths are community quantized versions + cloud-hosted inference.

Summary

Kimi K3 is one of the landmark events in the 2026 domestic large model competition—2.8 trillion parameters making it the world's largest open-source model, top frontend code capability globally, and the rare official admission on the release homepage that overall capabilities still lag behind Claude Fable 5 and GPT-5.6 Sol. This honest stance is more commendable than simply stacking parameters. If your core need is frontend code generation or a super-large open-source model for self-controlled deployment, K3 is currently the most noteworthy option; but if you need top-tier general reasoning capability (HLE, etc.), you should still prioritize Fable 5 / GPT-5.6 Sol. After the full weights are open-sourced on 7/27, it is recommended to follow community quantization solutions to lower the deployment threshold.

Version History

Category
AI Coding
Pricing
Open Source
Tags
Open Source · Coding · MoE
Website

Related Tools