China AI Hub China AI Hub

Models / Kimi K3

Kimi K3

Kimi K-series family · Provider: moonshot-ai · Status: active · Released: 2026-07-16

Kimi K3 is a Kimi K-series model developed by moonshot-ai, a Chinese AI company headquartered in 13F, Building 1, JD Technology Building, 76 Zhichun Road, Haidian District, Beijing, China, released 2026-07-16 with a 1,048,576-token context window and open weights.

Kimi K3
Image: AI-generated illustration (Seedream)
Last verified: · Data status: Current · Next review:
Key facts for Kimi K3
Key factValue
Model IDkimi-k3
ArchitectureMoE: 2.8T total / 104B activated; 93 layers (1 dense); 896 experts (16 selected + 2 shared per token); 69 KDA + 24 Gated MLA layers; hidden dim 7168; SiTU-GLU; MoonViT-V2 vision encoder (401M); MXFP4 weights / MXFP8 activations
Context window 1,048,576 tokens
Max output1,048,576 tokens
Open weightsYes
LicenseKimi K3 License (permissive MIT-style, but Model-as-a-Service operators with >$20M aggregate revenue over any 12 months must sign a separate agreement; products with >100M MAU or >$20M monthly revenue must display 'Kimi K3' in the UI)
Self-hostingYes
API availableYes
API pricing $3 input / $15 output per 1M tokens (USD) · provider pricing page
Cloud providersMoonshot AI Platform

Capabilities

Capabilities of Kimi K3
CapabilitySupported
ReasoningYes
CodingYes
MathUnknown
ChineseUnknown
EnglishUnknown
MultilingualUnknown
VisionYes
AudioUnknown
VideoYes
Tool callingYes
Function callingUnknown
Structured outputYes
Agent capabilityYes
RAGUnknown
Computer useUnknown

Benchmark results

Benchmark results for Kimi K3
BenchmarkVersionScoreMetricDateSource typeSource
GPQA Diamond 93.5 accuracy 2026-07 vendor_reported link
HLE-Full 43.5 (56.0 with tools) accuracy 2026-07 vendor_reported link
DeepSWE 67.5 accuracy 2026-07 vendor_reported link
Terminal-Bench 2.1 88.3 accuracy 2026-07 vendor_reported link
MMMU-Pro 81.6 accuracy 2026-07 vendor_reported link
Video-MME (with subtitles) 90 accuracy 2026-07 vendor_reported link

Benchmark scores are single data points, not universal rankings. Vendor-reported scores are labeled as such.

Known limitations

Kimi K3 is Moonshot AI’s flagship open-weight model (“Open Frontier Weights”, released 2026-07-16): a 2.8T-parameter MoE with 104B activated parameters, a 1M-token context window, native visual understanding, and always-on thinking with reasoning_effort low/high/max (default max). It targets software engineering, knowledge work and deep reasoning.

The API model kimi-k3 (min $1 top-up) prices at $3.00 input / $15.00 output per 1M tokens with automatic prefix caching ($0.30 cached input); max_completion_tokens defaults to 131,072 and can be set up to 1,048,576. Temperature (1.0) and top_p (0.95) are fixed. Full weights are on Hugging Face and ModelScope under the Kimi K3 License.

Provider

Pricing

Benchmarks with results for this model

Agents built on this model

API

Data interpretation

Evidence types for Kimi K3
FieldValueEvidence type
Context window1,048,576 tokensOfficial
ArchitectureMoE: 2.8T total / 104B activated; 93 layers (1 dense); 896 experts (16 selected + 2 shared per token); 69 KDA + 24 Gated MLA layers; hidden dim 7168; SiTU-GLU; MoonViT-V2 vision encoder (401M); MXFP4 weights / MXFP8 activationsOfficial
Open weightsYesOfficial
LicenseKimi K3 License (permissive MIT-style, but Model-as-a-Service operators with >$20M aggregate revenue over any 12 months must sign a separate agreement; products with >100M MAU or >$20M monthly revenue must display 'Kimi K3' in the UI)Official
API pricing $3 / $15 per 1M tokens (USD) Official
GPQA Diamond 93.5 accuracy Vendor-reported
HLE-Full 43.5 (56.0 with tools) accuracy Vendor-reported
DeepSWE 67.5 accuracy Vendor-reported
Terminal-Bench 2.1 88.3 accuracy Vendor-reported
MMMU-Pro 81.6 accuracy Vendor-reported
Video-MME (with subtitles) 90 accuracy Vendor-reported
Family position1 model in the Kimi K-series familyChina AI Hub analysis

Evidence types: Official = vendor documentation, pricing pages or model cards. Vendor-reported = benchmark scores published by the vendor. China AI Hub analysis = derived from the database itself. See the sourcing policy.

Sources

Confidence and source hierarchy per the sourcing policy. Facts change; check the source before relying on this page.

What is the context window of Kimi K3?

Kimi K3 has a 1,048,576-token context window and a maximum output of 1,048,576 tokens.

Is Kimi K3 open weight?

Yes — Kimi K3 weights are openly available under the Kimi K3 License (permissive MIT-style, but Model-as-a-Service operators with >$20M aggregate revenue over any 12 months must sign a separate agreement; products with >100M MAU or >$20M monthly revenue must display 'Kimi K3' in the UI) license.

How much does Kimi K3 cost through the API?

$3 per 1M input tokens and $15 per 1M output tokens (USD).

Where does China AI Hub get its Kimi K3 data?

From 4 sources (official pages first), last verified 2026-09-20. Benchmark scores are labeled by source type; see the sourcing policy for details.