Moonshot AI logo

Moonshot AI: Kimi K3

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

Source-linked kimi-k3 Open weights Released 2026-07-16
API record Report
InputT
OutputT
Input price$3/M
Output price$15/M
Context1.04858M
Max output131.072K
Providers63
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Kimi K3
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Kenari kimi-k3 1.04858M 131.072K ReasoningToolsJSON Docs ↗
Nvidia moonshotai/kimi-k3 1.04858M 131.072K ReasoningToolsJSON Docs ↗
SCNet Token Plan Kimi-K3 1.04858M 131.072K ReasoningToolsJSON Docs ↗
Kimi For Coding k3 1.04858M 131.072K ReasoningToolsJSON Docs ↗
Umans AI Coding Plan umans-kimi-k3 1.04858M 131.072K ReasoningToolsJSON Docs ↗
SenseNova (China) kimi-k3 1.04858M 65.536K ReasoningToolsJSON Docs ↗
NanoGPT moonshotai/kimi-k3 1.04858M 943.718K $2 $10 $0.2 ReasoningToolsJSON Docs ↗
CrofAI kimi-k3 1M 262.144K $2 $8 $0.25 ReasoningToolsJSON Docs ↗
OpenRouter moonshotai/kimi-k3 1.04858M 943.718K $2.34 $11.7 $0.261 ReasoningToolsJSON Docs ↗
Vancine kimi-k3 1.04858M 131.072K $2.4 $12 $0.24 ReasoningToolsJSON Docs ↗
DigitalOcean kimi-k3 1.04858M 131.072K $2.55 $12.95 $0.285 ReasoningToolsJSON Docs ↗
Deep Infra moonshotai/Kimi-K3 1.04858M 131.072K $2.85 $14.25 $0.285 ReasoningToolsJSON Docs ↗
Merge Gateway moonshot/kimi-k3 1.04858M 1.04858M $2.9 $14 $0.3 ReasoningToolsJSON Docs ↗
OpenCode Zen kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Cortecs kimi-k3 1.04858M 1.04858M $3 $14.999 ReasoningToolsJSON Docs ↗
NovitaAI moonshotai/kimi-k3 1.04858M 1.04858M $3 $15 $0.3 ReasoningToolsJSON Docs ↗
AIHubMix kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Impossibl moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Fireworks AI accounts/fireworks/models/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Weights & Biases moonshotai/Kimi-K3 1.04858M 1.04858M $3 $15 $0.3 ReasoningToolsJSON Docs ↗
CrossModel moonshot/kimi-k3 1.04858M 1.04858M $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Neon kimi-k3 1.04858M 65.536K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Baseten moonshotai/Kimi-K3 1.04858M 262.144K $3 $15 ReasoningToolsJSON Docs ↗
Nebius Token Factory moonshotai/Kimi-K3 1.04858M 8K $3 $15 $3 ReasoningToolsJSON Docs ↗
ClinePass cline-pass/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
ai& moonshotai/kimi-k3 1.04858M 131.072K $3 $12.5 $0.5 ReasoningToolsJSON Docs ↗
ZenMux moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
EmpirioLabs AI kimi-k3 1M 131.072K $3 $15 $3 ReasoningToolsJSON Docs ↗
Together AI moonshotai/Kimi-K3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Moonshot AI (China) kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Synthetic hf:moonshotai/Kimi-K3 524.288K 65.536K $3 $15 $0.45 ReasoningToolsJSON Docs ↗
Requesty kimi-k3 1.04858M 262.144K $3 $15 $0.45 ReasoningToolsJSON Docs ↗
Kilo Gateway moonshotai/kimi-k3 1.04858M 943.718K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Vercel AI Gateway moonshotai/kimi-k3 1M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Hugging Face moonshotai/Kimi-K3 1M 131.072K $3 $15 ReasoningToolsJSON Docs ↗
Perplexity Agent moonshot-ai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Moonshot AI kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Eden AI moonshot/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Modal moonshotai/Kimi-K3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Neuralwatt kimi-k3 1.04856M 1.04856M $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Berget.AI moonshotai/Kimi-K3 327.68K 32.768K $3 $15 ReasoningToolsJSON Docs ↗
TokenGo moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Abacus moonshotai/Kimi-K3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
CoralBricks kimi-k3 1.04858M 131.072K $3 $15 ReasoningToolsJSON Docs ↗
302.AI kimi-k3 1.04858M 131.072K $3 $15 ReasoningToolsJSON Docs ↗
GitHub Copilot kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Jalapeno Cloud Kimi-K3 1.04858M 131.072K $3 $15 ReasoningToolsJSON Docs ↗
TensorX moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.75 ReasoningToolsJSON Docs ↗
Vivgrid kimi-k3 1M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Pioneer moonshotai/Kimi-K3 1M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Cloudflare AI Gateway moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Umans AI umans-kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Arcee moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Ofox moonshotai/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Opper moonshot/kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
OpenCode Go kimi-k3 1.04858M 131.072K $3 $15 $0.3 ReasoningToolsJSON Docs ↗
Charm Hyper kimi-k3 1.04858M 16K $3.266 $16.332 $0.327 ReasoningToolsJSON Docs ↗
OrcaRouter kimi/kimi-k3 1.04858M 131.072K $3.3 $16.5 $0.33 ReasoningToolsJSON Docs ↗
Venice AI kimi-k3 1M 131.072K $3.75 $18.75 $0.375 ReasoningToolsJSON Docs ↗
GreenPT kimi-k3 1.04858M 131.072K $3.762 $18.81 $0.941 ReasoningToolsJSON Docs ↗
Tinfoil kimi-k3 262.144K 131.072K $4 $20 $0.8 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) kimi-k3-fast 1.04038M 131.072K $4.5 $22.5 $0.45 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $2/M

Across 57 priced providers

Median input $3/M

Midpoint of listed rates

Highest input $4.5/M

2.2× the lowest listed rate

Output range $8 – $22.5

Per million output tokens

NanoGPT $2/M
CrofAI $2/M
OpenRouter $2.34/M
Vancine $2.4/M
DigitalOcean $2.55/M
Deep Infra $2.85/M
Merge Gateway $2.9/M
OpenCode Zen $3/M
Cortecs $3/M
NovitaAI $3/M
AIHubMix $3/M
Impossibl $3/M
Fireworks AI $3/M
Weights & Biases $3/M
CrossModel $3/M
Neon $3/M
Baseten $3/M
Nebius Token Factory $3/M
ClinePass $3/M
ai& $3/M
ZenMux $3/M
EmpirioLabs AI $3/M
Together AI $3/M
Moonshot AI (China) $3/M
Synthetic $3/M
Requesty $3/M
Kilo Gateway $3/M
Vercel AI Gateway $3/M
Hugging Face $3/M
Perplexity Agent $3/M
Moonshot AI $3/M
Eden AI $3/M
Modal $3/M
Neuralwatt $3/M
Berget.AI $3/M
TokenGo $3/M
Abacus $3/M
CoralBricks $3/M
302.AI $3/M
GitHub Copilot $3/M
Jalapeno Cloud $3/M
TensorX $3/M
Vivgrid $3/M
DevPass (LLM Gateway) $3/M
Pioneer $3/M
Cloudflare AI Gateway $3/M
Umans AI $3/M
Arcee $3/M
Ofox $3/M
Opper $3/M
OpenCode Go $3/M
Charm Hyper $3.266/M
OrcaRouter $3.3/M
Venice AI $3.75/M
GreenPT $3.762/M
Tinfoil $4/M
DevPass (LLM Gateway) $4.5/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

2 providers list the identical $2 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Context limits also differ by provider, from 262.144K to 1.04858M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
Attachments Yes
Vision input Yes
Open weights Yes
Model family
kimi-k3
Knowledge cutoff
Not documented
License
Not documented
Release date
2026-07-16
Model ID
moonshotai/kimi-k3
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
Coding IndexCoding Index Agentic IndexAgentic Index IntelligenceIntelligence Index Design Arena: w…Design Arena: website Design Arena: c…Design Arena: codecategories Design Arena: g…Design Arena: gamedev Design Arena: d…Design Arena: dataviz Design Arena: u…Design Arena: uicomponent This model — Coding Index: 76.2 index (92.2th percentile) This model — Agentic Index: 50.6 index (92.7th percentile) This model — Intelligence Index: 43.8 index (88.0th percentile) This model — Design Arena: website: 1354 elo (98.9th percentile) This model — Design Arena: codecategories: 1389 elo (99.4th percentile) This model — Design Arena: gamedev: 1402 elo (98.2th percentile) This model — Design Arena: dataviz: 1365 elo (98.8th percentile) This model — Design Arena: uicomponent: 1369 elo (98.8th percentile)

Each axis is this model's percentile among the 8 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

Coding Index index Source ↗
Coding Index #16 of 205
Agentic Index index Source ↗
Agentic Index #14 of 193
Intelligence Index index Source ↗
Intelligence Index #23 of 192
Design Arena: website elo Source ↗
Design Arena: website #2 of 175
Design Arena: codecategories elo Source ↗
Design Arena: codecategories #1 of 167
Design Arena: gamedev elo Source ↗
Design Arena: gamedev #3 of 166
Design Arena: dataviz elo Source ↗
Design Arena: dataviz #2 of 165
Design Arena: uicomponent elo Source ↗
Design Arena: uicomponent #2 of 161
Design Arena: 3d elo Source ↗
Design Arena: 3d #1 of 157
Design Arena: svg elo Source ↗
Design Arena: svg #5 of 116
Design Arena: fullstack elo Source ↗
Design Arena: fullstack #5 of 67
Design Arena: webapps elo Source ↗
Design Arena: webapps #2 of 67
Design Arena: mobileapps elo Source ↗
Design Arena: mobileapps #5 of 66
Design Arena: androidnative elo Source ↗
Design Arena: androidnative #9 of 59
Design Arena: python-pptxslides elo Source ↗
Design Arena: python-pptxslides #9 of 47
Design Arena: godotgamedev elo Source ↗
Design Arena: godotgamedev #18 of 45
Design Arena: agenticgamedev elo Source ↗
Design Arena: agenticgamedev #5 of 44
Design Arena: htmlslides elo Source ↗
Design Arena: htmlslides #3 of 40
BrowseComp accuracy · 2026-07-16 Source ↗
BrowseComp #2 of 15
CharXiv Reasoning accuracy · 2026-07-16 Source ↗
CharXiv Reasoning #1 of 7
AutomationBench success rate · 2026-07-16 Source ↗
AutomationBench #2 of 5
GDPval-AA — v2 Elo · 2026-07-16 Source ↗
GDPval-AA #2 of 4
AA-Briefcase Elo · 2026-07-16 Source ↗
1548.0 2 models scored — too few for a distribution
DeepSWE — 1.1 resolve rate · Kimi Code · 2026-07-16 Source ↗
67.5 1 model scored — too few for a distribution
FrontierSWE dominance score · Kimi Code · 2026-07-16 Source ↗
81.2 1 model scored — too few for a distribution
JobBench score · 2026-07-16 Source ↗
52.9 1 model scored — too few for a distribution
Program Bench score · Kimi Code · 2026-07-16 Source ↗
77.8 1 model scored — too few for a distribution
SWE Marathon — 1.1 resolve rate · Claude Code · 2026-07-16 Source ↗
42.0 1 model scored — too few for a distribution
SpreadsheetBench — 2 score · Claude Code · 2026-07-16 Source ↗
34.8 1 model scored — too few for a distribution
Terminal-Bench — 2.1 accuracy · Kimi Code · 2026-07-16 Source ↗
88.3 1 model scored — too few for a distribution
ZeroBench pass@5 · 2026-07-16 Source ↗
41.0 1 model scored — too few for a distribution

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Coding Index, the benchmark with the widest published coverage that this model appears in.

0 43 86 $0.01 $0.1 $1 $10 Claude Opus 5 — $5/M, score 78.0 Claude Opus 5 (batch) — $2.5/M, score 78.0 SpaceXAI: Grok 4.5 — $2/M, score 72.4 Qwen: Qwen3.6 Plus — $0.325/M, score 54.5 Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 9.7 Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 20.1 Qwen: Qwen3 Next 80B A3B Thinking — $0.15/M, score 17.4 OpenAI: GPT-4 Turbo (batch) — $5/M, score 21.5 OpenAI: o3 Mini High (batch) — $0.55/M, score 16.3 Qwen: Qwen3.8 Max (0902) — $2/M, score 71.8 Gemini 3.5 Flash — $1.5/M, score 70.1 Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 71.5 Gemma 3 12B IT — $0.05/M, score 5.8 GPT-4.1 mini — $0.4/M, score 20.2 DeepSeek V4 Pro — $0.435/M, score 59.4 Claude Sonnet 5 — $2/M, score 71.5 Mistral: Ministral 3 8B 2512 — $0.15/M, score 9.7 Kimi K3 — $3/M, score 76.2 OpenAI: o1 (batch) — $7.5/M, score 39.7 Gemini 3.7 Flash — $0.75/M, score 76.1 LongCat-2.0 — $0.3/M, score 45.3 Qwen: Qwen3.7 Plus — $0.32/M, score 55.9 Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 74.3 Mistral: Mistral Large 3 2512 — $0.5/M, score 20.1 DeepSeek V3.2 — $0.18/M, score 44.2 Anthropic: Claude Fable 5.1 (batch) — $5/M, score 81.6 Anthropic: Claude Opus 4.7 — $5/M, score 73.6 Google: Gemini 3.5 Flash (batch) — $0.75/M, score 70.1 OpenAI: GPT-5.4 Nano (batch) — $0.1/M, score 56.1 Inception: Mercury 2 — $0.25/M, score 31.1 MiMo-V2.5-Pro — $0.435/M, score 60.2 OpenAI: GPT-5.4 (batch) — $1.25/M, score 71.1 Gemini 3.1 Flash Lite Preview — $0.25/M, score 34.7 Z.ai: GLM 4.6 — $0.43/M, score 45.8 OpenAI: GPT-4o-mini (batch) — $0.075/M, score 11.4 Qwen: Qwen3.5-9B (batch) — $0.17/M, score 28.7 OpenAI: GPT-5.1 (batch) — $0.625/M, score 49.4 Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 43.9 Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 52.1 Gemini 3.6 Flash — $0.75/M, score 69.2 GPT OSS 20B — $0.02/M, score 20.7 Gemma 4 26B A4B IT — $0.042/M, score 39.3 Claude Fable 5 — $10/M, score 76.5 Gemma 4 31B IT — $0.09/M, score 43.4 Anthropic: Claude Sonnet 4 — $3/M, score 37.6 Z.ai: GLM 5.3 Flash — $0.15/M, score 71.5 OpenAI: GPT-3.5 Turbo (batch) — $0.25/M, score 10.7 OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 11.1 MiniMax: MiniMax M3 — $0.3/M, score 58.6 DeepSeek: DeepSeek V4 Pro 0813 (batch) — $0.66/M, score 68.8 Thinking Machines: Inkling Small (batch) — $0.5/M, score 52.9 Anthropic: Claude Fable 5.1 — $10/M, score 81.6 Google: Gemma 4 31B (batch) — $0.39/M, score 43.4 Qwen: Qwen3.8 2.4T A95B (batch) — $2/M, score 71.9 Mistral: Devstral 2 2512 — $0.4/M, score 31.3 MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 60.8 Inkling — $1.87/M, score 52.1 OpenAI: gpt-oss-20b (batch) — $0.05/M, score 20.7 Nemotron 3 Super 120B A12B — $0.2/M, score 37.7 Muse Spark 1.1 — $1.25/M, score 71.3 Gemini 3.1 Pro Preview — $2/M, score 68.8 Gemma 3 27B IT — $0.08/M, score 10.1 Muse Spark 1.2 — $1.25/M, score 72.2 MiMo-V2.5 — $0.14/M, score 56.8 Google: Gemini 3.5 Flash Lite (batch) — $0.15/M, score 49.3 Kimi K2 Thinking — $0.4/M, score 21.0 Mistral: Mistral Medium 3.5 (batch) — $0.75/M, score 46.9 Qwen: Qwen3.8 Max (0803) — $2/M, score 68.9 Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 20.5 DeepSeek V4 Flash 0731 — $0.05/M, score 69.1 GPT-5.1 — $1.25/M, score 49.4 GPT-4 Turbo — $10/M, score 21.5 Qwen: Qwen3.8 27B — $0.42/M, score 68.1 GPT-5.6 Luna — $0.2/M, score 71.4 Z.ai: GLM 5.1 — $0.966/M, score 55.8 o1 — $15/M, score 39.7 GPT-5 Mini — $0.25/M, score 15.6 Google: Gemini 3.8 Flash (batch) — $0.375/M, score 76.3 GPT OSS 120B — $0.03/M, score 30.4 Anthropic: Claude Sonnet 4.5 — $3/M, score 52.1 Kimi K2.6 — $0.95/M, score 61.8 Google: Gemini 3.7 Flash (batch) — $0.375/M, score 76.1 Mistral: Mistral Small 4 (batch) — $0.075/M, score 26.6 inclusionAI: Ling 3.0 Flash VL — $0.06/M, score 57.0 Z.ai: GLM 5.3 (batch) — $0.7/M, score 74.8 DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 69.1 GPT-4 — $30/M, score 13.1 OpenAI: GPT-5.5 (batch) — $2.5/M, score 74.9 MoonshotAI: Kimi K3 (batch) — $3/M, score 76.2 Inkling Small — $0.45/M, score 52.9 Anthropic: Claude Sonnet 5 (batch) — $1/M, score 71.5 Anthropic: Claude Fable 5 (batch) — $5/M, score 76.5 IBM: Granite 4.2 8B — $0.06/M, score 22.4 OpenAI: GPT-5 (batch) — $0.625/M, score 37.8 Google: Gemini 2.5 Pro (batch) — $0.625/M, score 33.3 OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 20.2 SpaceXAI: Grok 4.6 — $2/M, score 76.8 DeepSeek-R1 — $0.7/M, score 24.6 DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 43.5 OpenAI: GPT-6 Astra (batch) — $5/M, score 76.9 OpenAI: GPT-5.6 Terra (batch) — $1/M, score 76.7 Gemma 3 4B IT — $0.04/M, score 2.7 DeepSeek V4 Pro 0813 — $0.442/M, score 68.8 Z.ai: GLM 5.2 (batch) — $0.7/M, score 68.8 OpenAI: gpt-oss-120b (batch) — $0.15/M, score 30.4 Google: Gemini 3.6 Flash (batch) — $0.375/M, score 69.2 GPT-6 Astra — $10/M, score 76.9 Nemotron 3 Nano 30B A3B — $0.05/M, score 14.4 GPT-4o (2024-05-13) — $5/M, score 24.2 Z.ai: GLM 5.3 — $1.4/M, score 74.8 Anthropic: Claude Opus 4.8 — $5/M, score 74.3 SpaceXAI: Grok 4.3 — $1.25/M, score 42.2 GPT-5.6 Terra — $2/M, score 76.7 MiniMax: MiniMax M3 (batch) — $0.3/M, score 58.6 Gemini 3.8 Flash — $0.75/M, score 76.3 NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 49.3 MiniMax: MiniMax M2.7 — $0.3/M, score 52.6 Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 73.6 Solar Pro 4 — $0.3/M, score 52.7 Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 68.8 Mistral: Mistral Medium 3.1 — $0.4/M, score 20.5 Thinking Machines: Inkling (batch) — $1/M, score 52.1 GPT-5.5 — $5/M, score 74.9 Anthropic: Claude Sonnet 4.6 — $3/M, score 63.0 Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 63.0 GPT-5 — $1.25/M, score 37.8 OpenAI: GPT-5 Mini (batch) — $0.125/M, score 15.6 Gemini 2.5 Pro — $1.25/M, score 33.3 inclusionAI: Ling 3.0 Flash — $0.021/M, score 50.6 SpaceXAI: Grok 4.3 (batch) — $1/M, score 42.2 Nemotron 3 Ultra 550B A55B — $0.5/M, score 49.3 OpenAI: GPT-5.6 Luna (batch) — $0.1/M, score 71.4 OpenAI: GPT-5.4 Mini (batch) — $0.375/M, score 56.1 Qwen: Qwen3.8 2.4T A95B — $2/M, score 71.9 OpenAI: GPT-5.6 Sol (batch) — $1/M, score 77.4 GPT-5.4 mini — $0.75/M, score 56.1 Kimi K2.7 Code — $0.95/M, score 60.8 Gemini 3.5 Flash Lite — $0.3/M, score 49.3 Hy3 preview — $0.066/M, score 58.8 DeepSeek V4 Flash — $0.15/M, score 52.0 Nemotron 3.5 Lightning 30B A3B — $0.05/M, score 26.8 GPT-5.6 Sol — $4/M, score 77.4 GPT-4.1 nano — $0.1/M, score 11.1 GPT-5.4 — $2.5/M, score 71.1 GPT-4o mini — $0.15/M, score 11.4 GPT-5.4 nano — $0.2/M, score 56.1 GPT-3.5-turbo — $0.5/M, score 10.7 Kimi K2.5 — $0.3/M, score 46.8 Trinity Large Thinking — $0.25/M, score 25.8 Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 12.1 Z.ai: GLM 5.2 — $0.966/M, score 68.8 Qwen: Qwen3.7 Max — $1.475/M, score 66.0 SpaceXAI: Grok Build 0.1 — $1/M, score 51.5 Qwen: Qwen3.6 35B A3B — $0.1/M, score 41.9 Qwen: Qwen3.6 27B — $0.3/M, score 53.7 inclusionAI: Ling-2.6-flash — $0.01/M, score 25.3 Mistral: Mistral Small 4 — $0.15/M, score 26.6 Kwaipilot: KAT-Coder-Pro V2 — $0.3/M, score 59.5 Qwen: Qwen3.5-9B — $0.1/M, score 28.7 Qwen: Qwen3.5-35B-A3B — $0.312/M, score 37.0 Qwen: Qwen3.5-122B-A10B — $0.26/M, score 45.7 Upstage: Solar Pro 3 — $0.15/M, score 16.2 Z.ai: GLM 4.7 — $0.4/M, score 45.3 Amazon: Nova 2 Lite — $0.3/M, score 23.0 Mistral: Ministral 3 3B 2512 — $0.1/M, score 4.8 Google: Gemma 3n 4B — $0.06/M, score 3.2 Meta: Llama 4 Maverick — $0.2/M, score 16.3 OpenAI: o3 Mini High — $1.1/M, score 16.3 Meta: Llama 3.3 70B Instruct — $0.1/M, score 11.9 Meta: Llama 3.1 8B Instruct — $0.05/M, score 5.4 Nex AGI: Nex-N2-Pro — $0.25/M, score 59.1 Mistral: Mistral Medium 3.5 — $1.5/M, score 46.9 inclusionAI: Ring-2.6-1T — $0.075/M, score 42.8 IBM: Granite 4.1 8B — $0.05/M, score 9.5 Qwen: Qwen3.5 397B A17B — $0.55/M, score 48.2 Qwen: Qwen3 Coder Next — $0.12/M, score 36.2 Mistral: Ministral 3 14B 2512 — $0.2/M, score 14.4 Anthropic: Claude Haiku 4.5 — $1/M, score 43.9 Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 22.1 Qwen: Qwen3 8B — $0.117/M, score 9.0 Qwen: Qwen3 14B — $0.227/M, score 13.8 Qwen: Qwen3 32B — $0.08/M, score 15.3 Meta: Llama 4 Scout — $0.1/M, score 8.2 DeepSeek: DeepSeek V3 0324 — $0.25/M, score 21.2 Cohere: Command A — $2.5/M, score 27.8 Step 3.7 Flash — $0.185/M, score 39.6 Kimi K3 Input price per million tokens (log scale) Index

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Max Output Tokens943718 → 131072
Price Completion11.7 → 15.0
Price Prompt2.34 → 3.0
Max Output Tokens131072 → 943718
Price Completion15.0 → 11.7
Price Prompt3.0 → 2.34
Max Output Tokens943718 → 131072
Price Completion11.7 → 15.0
Price Prompt2.34 → 3.0
Max Output Tokens131072 → 943718
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/moonshotai/kimi-k3
curl
curl "https://model.kyssta.lol/api/v1/models/moonshotai/kimi-k3"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Kimi K3?

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work. It is published by Moonshot AI and catalogued here from Models.dev.

How much does Kimi K3 cost?

Listed input pricing starts at $2 per million tokens from NanoGPT, rising to $4.5 across 57 listed providers.

What is the context length of Kimi K3?

Kimi K3 accepts up to 1.04858M tokens of context and returns up to 131.072K output tokens.

Does Kimi K3 support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, reasoning, and image input.

Which providers serve Kimi K3?

62 providers list this model: Kenari, Nvidia, SCNet Token Plan, Kimi For Coding, Umans AI Coding Plan, SenseNova (China) and 56 more.

Are the weights for Kimi K3 open?

Yes. The weights are published.

When was Kimi K3 released?

The catalog records a release date of 2026-07-16, last verified Sep 11, 2026.

More models from Moonshot AI

View all →