Kimi K3
Best for the hardest, highest-stakes tasks.
- Artificial Analysis Intelligence Index
- 44
- #2 of 6 models we serve
- DeepSWE
- 68.5%
- #3 of 5 models we serve
- Terminal-Bench 2.1
- 85%
- #2 of 6 models we serve
- Context window
- 1M
- 128K max output
How it ranks
Against every model we serve, best first. Starred scores are reported by the publisher on its model card; the rest are independent measurements. About the scores
- GLM 5.345
- Kimi K344
- GLM 5.3 Flash42
- DeepSeek V4.1 Flash39
- DeepSeek V4 Flash34
- Umans Flash18
- DeepSeek V4.1 Flash74.2%* reported by the publisher
- GLM 5.369%
- Kimi K368.5%
- GLM 5.3 Flash63.4%
- DeepSeek V4 Flash54.4%* reported by the publisher
Not published: Umans Flash.
- DeepSeek V4.1 Flash90.6%* reported by the publisher
- Kimi K385%
- GLM 5.3 Flash84.3%
- GLM 5.383.9%
- DeepSeek V4 Flash78.7%
- Umans Flash44.9%
Pricing
- Input
- $3.00 per 1M tokens
- Output
- $15.00 per 1M tokens
- Cache read
- $0.30 per 1M tokens
Every request is billed for exactly the tokens it uses. All prices
Specs
- Context window
- 1M tokens
- Max output
- 128K tokens
- Vision
- Yes
- Tool calling
- Yes
- Reasoning
- Adjustable (none, low, high, max)
- Default effort
- max
- Weights
- Full weights on Hugging Face
- In production since
- July 31, 2026
Quickstart
Pick your tool; the command already carries this model's id. More tools (Pi, Copilot, Zed, omp) in the setup guides.
OpenCode
Built inLaunch OpenCode on this model with the umans CLI. Umans AI is also built into OpenCode: run /connect and pick it.
curl -fsSL https://api.code.umans.ai/cli/install.sh | bash # once
umans opencode --model umans-kimi-k3