工具大全
LLM API 价格计算器(全部模型)

Claude Sonnet 5 API 价格与成本估算

Anthropic Claude Sonnet 5 的 API 定价为:输入 $2/百万 tokens,输出 $10/百万 tokens。缓存命中的输入按 $0.2/百万 tokens 计费。上下文窗口 1M tokens。Batch API 按标准价的 50% 计费。价格核实于 2026-09-04,以官方定价页为准。

输入 / 百万 tokens

$2

输出 / 百万 tokens

$10

缓存输入 / 百万

$0.2

上下文窗口

1M

在役未宣布弃用(核实于 2026-09-05)。 Tentative retirement not sooner than Jun 30, 2027.
查看完整退役时间线 →

三档负载月成本估算

负载每月 tokens标准价50% 缓存命中Batch API
轻量(原型/side project)5M in + 1M out$20.00$15.50$10.00
中等(生产小流量)50M in + 10M out$200$155$100
重度(规模化生产)500M in + 100M out$2000$1550$1000

注意:The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled.

价位相近的替代模型

模型输入/百万输出/百万中等负载月成本
Anthropic Claude Sonnet 5$2$10$200
OpenAI GPT-5.6 terra$2$12$220
Google Gemini 3.1 Pro Preview$2$12$220
xAI Grok 4.6$2$6$160
xAI Grok 4.5$2$6$160

中等负载 = 每月 5000 万输入 + 1000 万输出 tokens,标准价、无缓存。价格核实于 2026-09-04。

Claude Sonnet 5 vs Anthropic 全系模型

模型输入/百万输出/百万缓存输入/百万上下文中等负载月成本
Claude Sonnet 5$2$10$0.21M$200
Claude Fable 5.1$10$50$0.251M$1000
Claude Fable 5$10$50$11M$1000
Claude Opus 5$5$25$0.51M$500
Claude Haiku 4.5$1$5$0.1200K$100

按你的用量计算

返回

大模型 API 价格计算器

对比 Claude、GPT、Gemini、DeepSeek、Kimi、Grok 等大模型的最新 API 价格,按 token 用量估算月账单——包含多数对比表忽略的提示词缓存价与批量折扣。选择你当前在用的模型,还能直接看到迁移到每个备选每月省(或多花)多少钱。 价格核实日期 2026-09-04. 完全在浏览器本地运行——不上传、无需注册。

百万 token
百万 token

按缓存价计费的输入 token 占比

选中后表格会多一列,显示按当前用量换到每个模型每月省/多花多少。

模型输入 $/M输出 $/M缓存 $/M上下文估算月成本对比当前

GPT-5.6 luna

OpenAI

$0.2$1.2$0.021.05M$22.00−$178 (−89%)

Gemini 3.1 Flash-Lite

Google

$0.25$1.5$0.0251M$27.50−$173 (−86%)

DeepSeek V4 Flash

DeepSeek · 开源权重

$0.435$1.3$0.01451M$34.75−$165 (−83%)

Gemini 3.6 Flash

Google

$0.75$3.75$0.0751M$75.00−$125 (−63%)

GLM-5

Zhipu (z.ai) · 开源权重

$1$3.2$0.2200K$82.00−$118 (−59%)

Grok 4.20

xAI

$1.25$2.5$0.21M$87.50−$113 (−56%)

Kimi K2.7 Code

Moonshot · 开源权重

$0.95$4$0.19262K$87.50−$113 (−56%)

Claude Haiku 4.5

Anthropic

$1$5$0.1200K$100−$100 (−50%)

DeepSeek V4 Pro

DeepSeek · 开源权重

$1.3$3.91$0.0431M$104−$95.90 (−48%)

GLM-5.2

Zhipu (z.ai) · 开源权重

$1.4$4.4$0.26200K$114−$86.00 (−43%)

Grok 4.6

xAI

$2$6$0.5500K$160−$40.00 (−20%)

Grok 4.5

xAI

$2$6$0.3500K$160−$40.00 (−20%)

Claude Sonnet 5当前

Anthropic

$2$10$0.21M$200

GPT-5.6 terra

OpenAI

$2$12$0.21.05M$220+$20.00 (+10%)

Gemini 3.1 Pro Preview

Google

$2$12$0.21M$220+$20.00 (+10%)

Kimi K3

Moonshot · 开源权重

$3$15$0.31.048576M$300+$100 (+50%)

GPT-5.6 sol

OpenAI

$4$20$0.41.05M$400+$200 (+100%)

Claude Opus 5

Anthropic

$5$25$0.51M$500+$300 (+150%)

Claude Fable 5.1

Anthropic

$10$50$0.251M$1000+$800 (+400%)

Claude Fable 5

Anthropic

$10$50$11M$1000+$800 (+400%)

价格单位为美元/百万 token,标准(非批量)档,核实日期 2026-09-04. Claude Fable 5.1: Released Sep 1, 2026. Anthropic's Mythos-class flagship tier above Opus 5 — same $10/$50 base rate as Fable 5, but cache reads drop to 2.5% of the input price ($0.25/MTok vs the 10% every other Claude model charges). Shares its underlying model with Claude Mythos 5.1, offered only to approved organizations. Claude Fable 5: Superseded by Claude Fable 5.1 (Sep 1, 2026), which keeps the same base rate but cuts cache reads to $0.25/MTok. Fable 5 remains available as a pinned snapshot with cache reads at the standard 10% ($1.00/MTok). Shares its underlying model with Claude Mythos 5. Claude Sonnet 5: The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled. GPT-5.6 sol: Promotional price, available at least through Nov 21, 2026 (previously $5 / $30). Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 terra: Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 luna: Requests beyond the long-context threshold bill at 2× input / 1.5× output. Gemini 3.1 Pro Preview: Prompts over 200K tokens bill at $4 / $18. Gemini 3.6 Flash: Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027. Grok 4.6: Requests over 200K tokens bill at 2×. Grok 4.5: Requests over 200K tokens bill at 2×. Kimi K3: Always-on reasoning; output includes thinking tokens. DeepSeek V4 Pro: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak. DeepSeek V4 Flash: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.

批量折扣仅对提供异步批量档的厂商生效。缓存价按读取价建模,不含缓存写入加价(Anthropic)。超过厂商阈值的长上下文加价在对应模型旁注明。

FAQ

Claude Sonnet 5 的 API 价格是多少?

输入 $2/百万 tokens,输出 $10/百万 tokens,缓存命中输入 $0.2/百万 tokens。注意:The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled.(核实于 2026-09-04)

用 Claude Sonnet 5 跑一个月大概花多少钱?

以每月 5000 万输入 + 1000 万输出 tokens 的中等负载估算约 $200/月;若 50% 输入命中缓存约 $155/月;离线任务走 Batch API 还可按 50% 计费。

Claude Sonnet 5 的缓存计费怎么算?

命中提示缓存的输入 tokens 按 $0.2/百万计费,约为标准输入价的 10%。系统提示词、few-shot 示例等重复前缀是主要受益场景。

和 GPT-5.6 terra 相比谁更便宜?

同样中等负载下,Claude Sonnet 5 约 $200/月,OpenAI GPT-5.6 terra 约 $220/月。实际选型还应考虑质量、时延和上下文窗口(Claude Sonnet 5 为 1M,GPT-5.6 terra 为 1.05M)。

Claude Sonnet 5 和 Anthropic 其他模型的价格差多少?

按输入价对比:Claude Fable 5.1 为 $10/$50(输入/输出,每百万 tokens,约为 Claude Sonnet 5 的 500%);Claude Fable 5 为 $10/$50(输入/输出,每百万 tokens,约为 Claude Sonnet 5 的 500%);Claude Opus 5 为 $5/$25(输入/输出,每百万 tokens,约为 Claude Sonnet 5 的 250%);Claude Haiku 4.5 为 $1/$5(输入/输出,每百万 tokens,约为 Claude Sonnet 5 的 50%)。详见上方全系对比表。

其他模型的 API 价格