工具大全
LLM API 价格计算器(全部模型)

DeepSeek V4 Flash API 价格与成本估算

DeepSeek DeepSeek V4 Flash 的 API 定价为:输入 $0.435/百万 tokens,输出 $1.3/百万 tokens。缓存命中的输入按 $0.0145/百万 tokens 计费。上下文窗口 1M tokens。该模型暂无 Batch 折扣档。价格核实于 2026-09-04,以官方定价页为准。

输入 / 百万 tokens

$0.435

输出 / 百万 tokens

$1.3

缓存输入 / 百万

$0.0145

上下文窗口

1M

三档负载月成本估算

负载每月 tokens标准价50% 缓存命中
轻量(原型/side project)5M in + 1M out$3.47$2.42
中等(生产小流量)50M in + 10M out$34.75$24.24
重度(规模化生产)500M in + 100M out$348$242

注意:Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.

价位相近的替代模型

模型输入/百万输出/百万中等负载月成本
DeepSeek DeepSeek V4 Flash$0.435$1.3$34.75
Google Gemini 3.1 Flash-Lite$0.25$1.5$27.50
OpenAI GPT-5.6 luna$0.2$1.2$22.00
Google Gemini 3.6 Flash$0.75$3.75$75.00
Zhipu (z.ai) GLM-5$1$3.2$82.00

中等负载 = 每月 5000 万输入 + 1000 万输出 tokens,标准价、无缓存。价格核实于 2026-09-04。

DeepSeek V4 Flash vs DeepSeek 全系模型

模型输入/百万输出/百万缓存输入/百万上下文中等负载月成本
DeepSeek V4 Flash$0.435$1.3$0.01451M$34.75
DeepSeek V4 Pro$1.3$3.91$0.0431M$104

按你的用量计算

返回

大模型 API 价格计算器

对比 Claude、GPT、Gemini、DeepSeek、Kimi、Grok 等大模型的最新 API 价格,按 token 用量估算月账单——包含多数对比表忽略的提示词缓存价与批量折扣。选择你当前在用的模型,还能直接看到迁移到每个备选每月省(或多花)多少钱。 价格核实日期 2026-09-04. 完全在浏览器本地运行——不上传、无需注册。

百万 token
百万 token

按缓存价计费的输入 token 占比

选中后表格会多一列,显示按当前用量换到每个模型每月省/多花多少。

模型输入 $/M输出 $/M缓存 $/M上下文估算月成本对比当前

GPT-5.6 luna

OpenAI

$0.2$1.2$0.021.05M$22.00−$12.75 (−37%)

Gemini 3.1 Flash-Lite

Google

$0.25$1.5$0.0251M$27.50−$7.25 (−21%)

DeepSeek V4 Flash当前

DeepSeek · 开源权重

$0.435$1.3$0.01451M$34.75

Gemini 3.6 Flash

Google

$0.75$3.75$0.0751M$75.00+$40.25 (+116%)

GLM-5

Zhipu (z.ai) · 开源权重

$1$3.2$0.2200K$82.00+$47.25 (+136%)

Grok 4.20

xAI

$1.25$2.5$0.21M$87.50+$52.75 (+152%)

Kimi K2.7 Code

Moonshot · 开源权重

$0.95$4$0.19262K$87.50+$52.75 (+152%)

Claude Haiku 4.5

Anthropic

$1$5$0.1200K$100+$65.25 (+188%)

DeepSeek V4 Pro

DeepSeek · 开源权重

$1.3$3.91$0.0431M$104+$69.35 (+200%)

GLM-5.2

Zhipu (z.ai) · 开源权重

$1.4$4.4$0.26200K$114+$79.25 (+228%)

Grok 4.6

xAI

$2$6$0.5500K$160+$125 (+360%)

Grok 4.5

xAI

$2$6$0.3500K$160+$125 (+360%)

Claude Sonnet 5

Anthropic

$2$10$0.21M$200+$165 (+476%)

GPT-5.6 terra

OpenAI

$2$12$0.21.05M$220+$185 (+533%)

Gemini 3.1 Pro Preview

Google

$2$12$0.21M$220+$185 (+533%)

Kimi K3

Moonshot · 开源权重

$3$15$0.31.048576M$300+$265 (+763%)

GPT-5.6 sol

OpenAI

$4$20$0.41.05M$400+$365 (+1051%)

Claude Opus 5

Anthropic

$5$25$0.51M$500+$465 (+1339%)

Claude Fable 5.1

Anthropic

$10$50$0.251M$1000+$965 (+2778%)

Claude Fable 5

Anthropic

$10$50$11M$1000+$965 (+2778%)

价格单位为美元/百万 token,标准(非批量)档,核实日期 2026-09-04. Claude Fable 5.1: Released Sep 1, 2026. Anthropic's Mythos-class flagship tier above Opus 5 — same $10/$50 base rate as Fable 5, but cache reads drop to 2.5% of the input price ($0.25/MTok vs the 10% every other Claude model charges). Shares its underlying model with Claude Mythos 5.1, offered only to approved organizations. Claude Fable 5: Superseded by Claude Fable 5.1 (Sep 1, 2026), which keeps the same base rate but cuts cache reads to $0.25/MTok. Fable 5 remains available as a pinned snapshot with cache reads at the standard 10% ($1.00/MTok). Shares its underlying model with Claude Mythos 5. Claude Sonnet 5: The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled. GPT-5.6 sol: Promotional price, available at least through Nov 21, 2026 (previously $5 / $30). Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 terra: Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 luna: Requests beyond the long-context threshold bill at 2× input / 1.5× output. Gemini 3.1 Pro Preview: Prompts over 200K tokens bill at $4 / $18. Gemini 3.6 Flash: Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027. Grok 4.6: Requests over 200K tokens bill at 2×. Grok 4.5: Requests over 200K tokens bill at 2×. Kimi K3: Always-on reasoning; output includes thinking tokens. DeepSeek V4 Pro: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak. DeepSeek V4 Flash: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.

批量折扣仅对提供异步批量档的厂商生效。缓存价按读取价建模,不含缓存写入加价(Anthropic)。超过厂商阈值的长上下文加价在对应模型旁注明。

FAQ

DeepSeek V4 Flash 的 API 价格是多少?

输入 $0.435/百万 tokens,输出 $1.3/百万 tokens,缓存命中输入 $0.0145/百万 tokens。注意:Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.(核实于 2026-09-04)

用 DeepSeek V4 Flash 跑一个月大概花多少钱?

以每月 5000 万输入 + 1000 万输出 tokens 的中等负载估算约 $34.75/月;若 50% 输入命中缓存约 $24.24/月。

DeepSeek V4 Flash 的缓存计费怎么算?

命中提示缓存的输入 tokens 按 $0.0145/百万计费,约为标准输入价的 3%。系统提示词、few-shot 示例等重复前缀是主要受益场景。

和 Gemini 3.1 Flash-Lite 相比谁更便宜?

同样中等负载下,DeepSeek V4 Flash 约 $34.75/月,Google Gemini 3.1 Flash-Lite 约 $27.50/月。实际选型还应考虑质量、时延和上下文窗口(DeepSeek V4 Flash 为 1M,Gemini 3.1 Flash-Lite 为 1M)。

DeepSeek V4 Flash 和 DeepSeek 其他模型的价格差多少?

按输入价对比:DeepSeek V4 Pro 为 $1.3/$3.91(输入/输出,每百万 tokens,约为 DeepSeek V4 Flash 的 299%)。详见上方全系对比表。

其他模型的 API 价格