Gemini 3.6 Flash API 价格与成本估算
Google Gemini 3.6 Flash 的 API 定价为:输入 $0.75/百万 tokens,输出 $3.75/百万 tokens。缓存命中的输入按 $0.075/百万 tokens 计费。上下文窗口 1M tokens。Batch API 按标准价的 50% 计费。价格核实于 2026-09-04,以官方定价页为准。
输入 / 百万 tokens
$0.75
输出 / 百万 tokens
$3.75
缓存输入 / 百万
$0.075
上下文窗口
1M
三档负载月成本估算
| 负载 | 每月 tokens | 标准价 | 50% 缓存命中 | Batch API |
|---|---|---|---|---|
| 轻量(原型/side project) | 5M in + 1M out | $7.50 | $5.81 | $3.75 |
| 中等(生产小流量) | 50M in + 10M out | $75.00 | $58.13 | $37.50 |
| 重度(规模化生产) | 500M in + 100M out | $750 | $581 | $375 |
注意:Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027.
价位相近的替代模型
| 模型 | 输入/百万 | 输出/百万 | 中等负载月成本 |
|---|---|---|---|
| Google Gemini 3.6 Flash | $0.75 | $3.75 | $75.00 |
| Zhipu (z.ai) GLM-5 | $1 | $3.2 | $82.00 |
| xAI Grok 4.20 | $1.25 | $2.5 | $87.50 |
| Moonshot Kimi K2.7 Code | $0.95 | $4 | $87.50 |
| Anthropic Claude Haiku 4.5 | $1 | $5 | $100 |
中等负载 = 每月 5000 万输入 + 1000 万输出 tokens,标准价、无缓存。价格核实于 2026-09-04。
Gemini 3.6 Flash vs Google 全系模型
| 模型 | 输入/百万 | 输出/百万 | 缓存输入/百万 | 上下文 | 中等负载月成本 |
|---|---|---|---|---|---|
| Gemini 3.6 Flash | $0.75 | $3.75 | $0.075 | 1M | $75.00 |
| Gemini 3.1 Pro Preview | $2 | $12 | $0.2 | 1M | $220 |
| Gemini 3.1 Flash-Lite | $0.25 | $1.5 | $0.025 | 1M | $27.50 |
按你的用量计算
大模型 API 价格计算器
对比 Claude、GPT、Gemini、DeepSeek、Kimi、Grok 等大模型的最新 API 价格,按 token 用量估算月账单——包含多数对比表忽略的提示词缓存价与批量折扣。选择你当前在用的模型,还能直接看到迁移到每个备选每月省(或多花)多少钱。 价格核实日期 2026-09-04. 完全在浏览器本地运行——不上传、无需注册。
按缓存价计费的输入 token 占比
选中后表格会多一列,显示按当前用量换到每个模型每月省/多花多少。
| 模型 | 输入 $/M | 输出 $/M | 缓存 $/M | 上下文 | 估算月成本 | 对比当前 |
|---|---|---|---|---|---|---|
GPT-5.6 luna OpenAI | $0.2 | $1.2 | $0.02 | 1.05M | $22.00 | −$53.00 (−71%) |
Gemini 3.1 Flash-Lite | $0.25 | $1.5 | $0.025 | 1M | $27.50 | −$47.50 (−63%) |
DeepSeek V4 Flash DeepSeek · 开源权重 | $0.435 | $1.3 | $0.0145 | 1M | $34.75 | −$40.25 (−54%) |
Gemini 3.6 Flash当前 | $0.75 | $3.75 | $0.075 | 1M | $75.00 | — |
GLM-5 Zhipu (z.ai) · 开源权重 | $1 | $3.2 | $0.2 | 200K | $82.00 | +$7.00 (+9%) |
Grok 4.20 xAI | $1.25 | $2.5 | $0.2 | 1M | $87.50 | +$12.50 (+17%) |
Kimi K2.7 Code Moonshot · 开源权重 | $0.95 | $4 | $0.19 | 262K | $87.50 | +$12.50 (+17%) |
Claude Haiku 4.5 Anthropic | $1 | $5 | $0.1 | 200K | $100 | +$25.00 (+33%) |
DeepSeek V4 Pro DeepSeek · 开源权重 | $1.3 | $3.91 | $0.043 | 1M | $104 | +$29.10 (+39%) |
GLM-5.2 Zhipu (z.ai) · 开源权重 | $1.4 | $4.4 | $0.26 | 200K | $114 | +$39.00 (+52%) |
Grok 4.6 xAI | $2 | $6 | $0.5 | 500K | $160 | +$85.00 (+113%) |
Grok 4.5 xAI | $2 | $6 | $0.3 | 500K | $160 | +$85.00 (+113%) |
Claude Sonnet 5 Anthropic | $2 | $10 | $0.2 | 1M | $200 | +$125 (+167%) |
GPT-5.6 terra OpenAI | $2 | $12 | $0.2 | 1.05M | $220 | +$145 (+193%) |
Gemini 3.1 Pro Preview | $2 | $12 | $0.2 | 1M | $220 | +$145 (+193%) |
Kimi K3 Moonshot · 开源权重 | $3 | $15 | $0.3 | 1.048576M | $300 | +$225 (+300%) |
GPT-5.6 sol OpenAI | $4 | $20 | $0.4 | 1.05M | $400 | +$325 (+433%) |
Claude Opus 5 Anthropic | $5 | $25 | $0.5 | 1M | $500 | +$425 (+567%) |
Claude Fable 5.1 Anthropic | $10 | $50 | $0.25 | 1M | $1000 | +$925 (+1233%) |
Claude Fable 5 Anthropic | $10 | $50 | $1 | 1M | $1000 | +$925 (+1233%) |
价格单位为美元/百万 token,标准(非批量)档,核实日期 2026-09-04. Claude Fable 5.1: Released Sep 1, 2026. Anthropic's Mythos-class flagship tier above Opus 5 — same $10/$50 base rate as Fable 5, but cache reads drop to 2.5% of the input price ($0.25/MTok vs the 10% every other Claude model charges). Shares its underlying model with Claude Mythos 5.1, offered only to approved organizations. Claude Fable 5: Superseded by Claude Fable 5.1 (Sep 1, 2026), which keeps the same base rate but cuts cache reads to $0.25/MTok. Fable 5 remains available as a pinned snapshot with cache reads at the standard 10% ($1.00/MTok). Shares its underlying model with Claude Mythos 5. Claude Sonnet 5: The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled. GPT-5.6 sol: Promotional price, available at least through Nov 21, 2026 (previously $5 / $30). Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 terra: Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 luna: Requests beyond the long-context threshold bill at 2× input / 1.5× output. Gemini 3.1 Pro Preview: Prompts over 200K tokens bill at $4 / $18. Gemini 3.6 Flash: Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027. Grok 4.6: Requests over 200K tokens bill at 2×. Grok 4.5: Requests over 200K tokens bill at 2×. Kimi K3: Always-on reasoning; output includes thinking tokens. DeepSeek V4 Pro: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak. DeepSeek V4 Flash: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.
批量折扣仅对提供异步批量档的厂商生效。缓存价按读取价建模,不含缓存写入加价(Anthropic)。超过厂商阈值的长上下文加价在对应模型旁注明。
FAQ
Gemini 3.6 Flash 的 API 价格是多少?
输入 $0.75/百万 tokens,输出 $3.75/百万 tokens,缓存命中输入 $0.075/百万 tokens。注意:Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027.(核实于 2026-09-04)
用 Gemini 3.6 Flash 跑一个月大概花多少钱?
以每月 5000 万输入 + 1000 万输出 tokens 的中等负载估算约 $75.00/月;若 50% 输入命中缓存约 $58.13/月;离线任务走 Batch API 还可按 50% 计费。
Gemini 3.6 Flash 的缓存计费怎么算?
命中提示缓存的输入 tokens 按 $0.075/百万计费,约为标准输入价的 10%。系统提示词、few-shot 示例等重复前缀是主要受益场景。
和 GLM-5 相比谁更便宜?
同样中等负载下,Gemini 3.6 Flash 约 $75.00/月,Zhipu (z.ai) GLM-5 约 $82.00/月。实际选型还应考虑质量、时延和上下文窗口(Gemini 3.6 Flash 为 1M,GLM-5 为 200K)。
Gemini 3.6 Flash 和 Google 其他模型的价格差多少?
按输入价对比:Gemini 3.1 Pro Preview 为 $2/$12(输入/输出,每百万 tokens,约为 Gemini 3.6 Flash 的 267%);Gemini 3.1 Flash-Lite 为 $0.25/$1.5(输入/输出,每百万 tokens,约为 Gemini 3.6 Flash 的 33%)。详见上方全系对比表。