Kimi K3 Pricing 2026
Complete pricing guide with plans, and cost analysis
Kimi K3 pricing ranges from $0.30 to $15/per million tokens.
Are you Kimi K3? Claim this profile
All Kimi K3 Plans & Pricing
| Plan | Monthly | Annual | Best For |
|---|---|---|---|
| Kimi K3 API (Pay-as-you-go) | Contact Sales | Contact Sales | Long-horizon coding, knowledge work, and advanced reasoning tasks needing a 1M-token context window |
| Verified pricing · last checked July 2026 · 2 sources
Get this price at Kimi K3 →
| |||
| What's included at Kimi K3 API (Pay-as-you-go) Best for: Long-horizon coding, knowledge work, and advanced reasoning tasks needing a 1M-token context window
| |||
View all features by plan (compare side-by-side)
Kimi K3 API (Pay-as-you-go)
- Kimi K3: $3.00 input / $15.00 output per M tokens
- Cached input (cache hit): $0.30 per M tokens
- 1M-token context window
- 2.8T-parameter, natively multimodal reasoning model with always-on “thinking mode”
Kimi K3 costs $0.30 to $15 per per million tokens as of July 2026. Pricing depends on your chosen tier, contract length, and negotiated discounts.
Use the interactive pricing calculator to estimate your exact cost based on team size and requirements.
- Free tier: No free tier available
Kimi K3 offers 1 pricing tiers: Kimi K3 API (Pay-as-you-go). The Kimi K3 API (Pay-as-you-go) plan is long-horizon coding, knowledge work, and advanced reasoning tasks needing a 1m-token context window.
Compared to other llm api providers software, Kimi K3 is positioned at the budget-friendly price point.
How much does Kimi K3 cost?
Kimi K3 Pricing Overview
Kimi K3 has 1 pricing plans ranging from $0.30 to $15/per million tokens. The Kimi K3 API (Pay-as-you-go) plan requires contacting sales for a custom quote and is designed for long-horizon coding, knowledge work, and advanced reasoning tasks needing a 1m-token context window.
This pricing was last verified in July 19, 2026 from 2 independent sources.
Kimi K3 costs $3.00 per million input tokens and $15.00 per million output tokens as of July 2026, with cache-hit input discounted to $0.30 per million. Moonshot AI released Kimi K3 on July 16, 2026 — a 2.8-trillion-parameter, natively multimodal reasoning model with a 1-million-token context window built for long-horizon coding, knowledge work, and deep reasoning. It is designed as an open-weight model, though full public weight release was still pending at launch.
Usage-Based Rates
Per-unit pricing for Kimi K3 API usage.
Kimi K3 API (Pay-as-you-go)
| Model | Input | Output | Cached | Per |
|---|---|---|---|---|
| kimi-k3 1000K ctx | $3.00 | $15.00 | $0.300 | 1M tokens |
- Pricing cross-referenced via OpenRouter (moonshotai/kimi-k3) and multiple third-party trackers citing Moonshot's official platform.kimi.ai rate card
- Full open-weight release was reported pending as of verification date (2026-07-19) — hosted API access confirmed
How Kimi K3 Pricing Compares
| Software | Starting Price | Top Price |
|---|---|---|
| Kimi K3 | $0.3/per million tokens | $15/per million tokens |
| Amazon Bedrock | $0.07/per million tokens | $75/per million tokens |
| Anyscale | $0.15/per million tokens | $5/per million tokens |
| Baidu ERNIE API | $0.1/per million tokens | $10/per million tokens |
| Cerebras Inference API | $0.1/per million tokens | $6/per million tokens |
| Cohere API | $0.037/per million tokens | $10/per million tokens |
Kimi K3 Pricing FAQ
01 How much does Kimi K3 cost per million tokens?
Kimi K3 costs $3.00 per million input tokens and $15.00 per million output tokens, with cache-hit input tokens discounted to $0.30 per million — a 90% discount versus standard input.
02 When was Kimi K3 released?
Moonshot AI released Kimi K3 on July 16, 2026, as a 2.8-trillion-parameter, natively multimodal reasoning model with a 1-million-token context window.
03 Is Kimi K3 open-weight?
Kimi K3 is designed as an open-weight model, though full public weight release was still reported pending as of its July 16, 2026 launch — check Moonshot AI's official channels for the current weight-release status.
Is this pricing incorrect? — we'll verify and update it.