Kimi K3 API 价格与成本估算
Moonshot Kimi K3 的 API 定价为:输入 $3/百万 tokens,输出 $15/百万 tokens。缓存命中的输入按 $0.3/百万 tokens 计费。上下文窗口 1.05M tokens。该模型暂无 Batch 折扣档。价格核实于 2026-09-04,以官方定价页为准。
输入 / 百万 tokens
$3
输出 / 百万 tokens
$15
缓存输入 / 百万
$0.3
上下文窗口
1.05M
三档负载月成本估算
| 负载 | 每月 tokens | 标准价 | 50% 缓存命中 |
|---|---|---|---|
| 轻量(原型/side project) | 5M in + 1M out | $30.00 | $23.25 |
| 中等(生产小流量) | 50M in + 10M out | $300 | $233 |
| 重度(规模化生产) | 500M in + 100M out | $3000 | $2325 |
注意:Always-on reasoning; output includes thinking tokens.
价位相近的替代模型
| 模型 | 输入/百万 | 输出/百万 | 中等负载月成本 |
|---|---|---|---|
| Moonshot Kimi K3 | $3 | $15 | $300 |
| OpenAI GPT-5.6 terra | $2 | $12 | $220 |
| Google Gemini 3.1 Pro Preview | $2 | $12 | $220 |
| Anthropic Claude Sonnet 5 | $2 | $10 | $200 |
| OpenAI GPT-5.6 sol | $4 | $20 | $400 |
中等负载 = 每月 5000 万输入 + 1000 万输出 tokens,标准价、无缓存。价格核实于 2026-09-04。
Kimi K3 vs Moonshot 全系模型
| 模型 | 输入/百万 | 输出/百万 | 缓存输入/百万 | 上下文 | 中等负载月成本 |
|---|---|---|---|---|---|
| Kimi K3 | $3 | $15 | $0.3 | 1.05M | $300 |
| Kimi K2.7 Code | $0.95 | $4 | $0.19 | 262K | $87.50 |
按你的用量计算
大模型 API 价格计算器
对比 Claude、GPT、Gemini、DeepSeek、Kimi、Grok 等大模型的最新 API 价格,按 token 用量估算月账单——包含多数对比表忽略的提示词缓存价与批量折扣。选择你当前在用的模型,还能直接看到迁移到每个备选每月省(或多花)多少钱。 价格核实日期 2026-09-04. 完全在浏览器本地运行——不上传、无需注册。
按缓存价计费的输入 token 占比
选中后表格会多一列,显示按当前用量换到每个模型每月省/多花多少。
| 模型 | 输入 $/M | 输出 $/M | 缓存 $/M | 上下文 | 估算月成本 | 对比当前 |
|---|---|---|---|---|---|---|
GPT-5.6 luna OpenAI | $0.2 | $1.2 | $0.02 | 1.05M | $22.00 | −$278 (−93%) |
Gemini 3.1 Flash-Lite | $0.25 | $1.5 | $0.025 | 1M | $27.50 | −$273 (−91%) |
DeepSeek V4 Flash DeepSeek · 开源权重 | $0.435 | $1.3 | $0.0145 | 1M | $34.75 | −$265 (−88%) |
Gemini 3.6 Flash | $0.75 | $3.75 | $0.075 | 1M | $75.00 | −$225 (−75%) |
GLM-5 Zhipu (z.ai) · 开源权重 | $1 | $3.2 | $0.2 | 200K | $82.00 | −$218 (−73%) |
Grok 4.20 xAI | $1.25 | $2.5 | $0.2 | 1M | $87.50 | −$213 (−71%) |
Kimi K2.7 Code Moonshot · 开源权重 | $0.95 | $4 | $0.19 | 262K | $87.50 | −$213 (−71%) |
Claude Haiku 4.5 Anthropic | $1 | $5 | $0.1 | 200K | $100 | −$200 (−67%) |
DeepSeek V4 Pro DeepSeek · 开源权重 | $1.3 | $3.91 | $0.043 | 1M | $104 | −$196 (−65%) |
GLM-5.2 Zhipu (z.ai) · 开源权重 | $1.4 | $4.4 | $0.26 | 200K | $114 | −$186 (−62%) |
Grok 4.6 xAI | $2 | $6 | $0.5 | 500K | $160 | −$140 (−47%) |
Grok 4.5 xAI | $2 | $6 | $0.3 | 500K | $160 | −$140 (−47%) |
Claude Sonnet 5 Anthropic | $2 | $10 | $0.2 | 1M | $200 | −$100 (−33%) |
GPT-5.6 terra OpenAI | $2 | $12 | $0.2 | 1.05M | $220 | −$80.00 (−27%) |
Gemini 3.1 Pro Preview | $2 | $12 | $0.2 | 1M | $220 | −$80.00 (−27%) |
Kimi K3当前 Moonshot · 开源权重 | $3 | $15 | $0.3 | 1.048576M | $300 | — |
GPT-5.6 sol OpenAI | $4 | $20 | $0.4 | 1.05M | $400 | +$100 (+33%) |
Claude Opus 5 Anthropic | $5 | $25 | $0.5 | 1M | $500 | +$200 (+67%) |
Claude Fable 5.1 Anthropic | $10 | $50 | $0.25 | 1M | $1000 | +$700 (+233%) |
Claude Fable 5 Anthropic | $10 | $50 | $1 | 1M | $1000 | +$700 (+233%) |
价格单位为美元/百万 token,标准(非批量)档,核实日期 2026-09-04. Claude Fable 5.1: Released Sep 1, 2026. Anthropic's Mythos-class flagship tier above Opus 5 — same $10/$50 base rate as Fable 5, but cache reads drop to 2.5% of the input price ($0.25/MTok vs the 10% every other Claude model charges). Shares its underlying model with Claude Mythos 5.1, offered only to approved organizations. Claude Fable 5: Superseded by Claude Fable 5.1 (Sep 1, 2026), which keeps the same base rate but cuts cache reads to $0.25/MTok. Fable 5 remains available as a pinned snapshot with cache reads at the standard 10% ($1.00/MTok). Shares its underlying model with Claude Mythos 5. Claude Sonnet 5: The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled. GPT-5.6 sol: Promotional price, available at least through Nov 21, 2026 (previously $5 / $30). Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 terra: Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 luna: Requests beyond the long-context threshold bill at 2× input / 1.5× output. Gemini 3.1 Pro Preview: Prompts over 200K tokens bill at $4 / $18. Gemini 3.6 Flash: Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027. Grok 4.6: Requests over 200K tokens bill at 2×. Grok 4.5: Requests over 200K tokens bill at 2×. Kimi K3: Always-on reasoning; output includes thinking tokens. DeepSeek V4 Pro: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak. DeepSeek V4 Flash: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.
批量折扣仅对提供异步批量档的厂商生效。缓存价按读取价建模,不含缓存写入加价(Anthropic)。超过厂商阈值的长上下文加价在对应模型旁注明。
FAQ
Kimi K3 的 API 价格是多少?
输入 $3/百万 tokens,输出 $15/百万 tokens,缓存命中输入 $0.3/百万 tokens。注意:Always-on reasoning; output includes thinking tokens.(核实于 2026-09-04)
用 Kimi K3 跑一个月大概花多少钱?
以每月 5000 万输入 + 1000 万输出 tokens 的中等负载估算约 $300/月;若 50% 输入命中缓存约 $233/月。
Kimi K3 的缓存计费怎么算?
命中提示缓存的输入 tokens 按 $0.3/百万计费,约为标准输入价的 10%。系统提示词、few-shot 示例等重复前缀是主要受益场景。
和 GPT-5.6 terra 相比谁更便宜?
同样中等负载下,Kimi K3 约 $300/月,OpenAI GPT-5.6 terra 约 $220/月。实际选型还应考虑质量、时延和上下文窗口(Kimi K3 为 1.05M,GPT-5.6 terra 为 1.05M)。
Kimi K3 和 Moonshot 其他模型的价格差多少?
按输入价对比:Kimi K2.7 Code 为 $0.95/$4(输入/输出,每百万 tokens,约为 Kimi K3 的 32%)。详见上方全系对比表。