已追踪 25 个模型 · September 10, 2026 核实
LLM API 价格,核实并标注日期

在你基于它构建之前,先比较每个 AI 模型的价格。

10 家服务商的输入与输出价格 — OpenAI、Anthropic、Google、Mistral、DeepSeek、xAI 等 — 均直接对照各服务商官方定价页核实,每一行都标注来源与日期。

追踪模型数 · 平均 $/百万 TOKENS
$2.03
平均输入
$9.73
平均输出
基于全部 25 个追踪模型计算,于 September 10, 2026 首次核实。我们每 48 小时重新核实一次,下次核实后将显示趋势。
最便宜的追踪模型
Mistral Small 4
$0.15 / $0.60
最贵的追踪模型
Claude Fable 5.1 (并列)
$10.00 / $50.00
价格差距
66.7x
最贵与最便宜输入价格之比
追踪的服务商
10
共 25 个模型
输入价格分布 — 所有追踪模型
$/百万 tokens · 对数刻度 · 完整比较 →
Claude Fable 5.1
$10.00
GPT-6 Astra
$10.00
Claude Opus 5
$5.00
GPT-5.6 Sol
$4.00
Command A
$2.50
Claude Sonnet 5
$2.00
GPT-5.6 Terra
$2.00
Gemini 3.1 Pro Preview
$2.00
Grok 4.6
$2.00
Qwen3.8-Max
$2.00
Mistral Medium 3.5
$1.50
Grok 4.3
$1.25
Amazon Nova 2 Pro
$1.25
Claude Haiku 4.5
$1.00
Grok Build 0.1
$1.00
Gemini 3.8 Flash
$0.75
Mistral Large 3
$0.50
Qwen3.7-Plus
$0.40
DeepSeek V4.1 Flash
$0.30
Amazon Nova 2 Lite
$0.30
Llama 4 Maverick
$0.27
Gemini 3.1 Flash-Lite
$0.25
GPT-5.6 Luna
$0.20
Llama 4 Scout
$0.18
Mistral Small 4
$0.15
旗舰 均衡 经济

今日模型价格

来源:各服务商官方定价页 · September 10, 2026 核实
已显示 25 个模型中的 25 个
模型 服务商 输入 /百万 输出 /百万 上下文 档位 来源
Mistral $0.15 $0.60 未公开 经济 mistral.ai — API pricing ↗
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate.
Meta (via Together AI) $0.18 $0.59 100万 tokens 经济 together.ai — Llama 4 Scout ↗
OpenAI $0.20 $1.20 约105万 tokens 经济 platform.openai.com/docs — Pricing ↗
Audio input priced separately at $0.50/1M.
Google $0.25 $1.50 未公开 经济 ai.google.dev — Gemini API pricing ↗
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate, one of several inference providers serving this open-weight model.
Meta (via Together AI) $0.27 $0.85 约105万 tokens 经济 together.ai — Llama 4 Maverick ↗
Peak-hour, cache-miss rate shown. Off-peak (01:00–04:00 & 06:00–10:00 UTC, Mon–Fri) is half price; cache-hit input is far cheaper still ($0.006/1M peak).
DeepSeek $0.30 $1.20 100万 tokens 经济 api-docs.deepseek.com — Pricing ↗
Bedrock Standard tier. Flex tier: $0.15 / $1.25.
Amazon $0.30 $2.50 100万 tokens 经济 aws.amazon.com — Bedrock pricing ↗
Rate shown for 0–256K input tokens; both 256K–1M tier and above bill higher on input only.
Alibaba $0.40 $1.60 100万 tokens 经济 alibabacloud.com — Model Studio pricing ↗
Despite the "Large" name, priced below Medium 3.5 in Mistral's current lineup.
Mistral $0.50 $1.50 未公开 经济 mistral.ai — API pricing ↗
Promotional rate through Dec 31, 2026; steps up to $1.50 / $7.50 on Jan 1, 2027.
Google $0.75 $3.75 未公开 均衡 ai.google.dev — Gemini API pricing ↗
Anthropic $1.00 $5.00 20万 tokens 经济 platform.claude.com/docs — Pricing ↗
xAI $1.00 $2.00 25.6万 tokens 经济 docs.x.ai — Pricing ↗
Long-context tier (>200K input) doubles to $2.50 / $5.
xAI $1.25 $2.50 100万 tokens 均衡 docs.x.ai — Pricing ↗
Preview model, Bedrock Standard tier. A cheaper "Flex" tier and pricier "Priority" tier also exist.
Amazon $1.25 $10.00 未公开(预览版) 均衡 aws.amazon.com — Bedrock pricing ↗
Positioned by Mistral as its top general-purpose / agentic model — priced above "Large 3".
Mistral $1.50 $7.50 未公开 均衡 mistral.ai — API pricing ↗
Launched at introductory pricing through Aug 31, 2026; the scheduled increase to $3/$15 was cancelled — $2/$10 is now the standing price.
Anthropic $2.00 $10.00 100万 tokens 均衡 platform.claude.com/docs — Pricing ↗
OpenAI $2.00 $12.00 约105万 tokens 均衡 platform.openai.com/docs — Pricing ↗
Requests over 200K input tokens bill at $4 / $18 instead.
Google $2.00 $12.00 ≤20万 tokens 档位 旗舰 ai.google.dev — Gemini API pricing ↗
Long-context tier (>200K input) doubles to $4 / $12.
xAI $2.00 $6.00 50万 tokens 旗舰 docs.x.ai — Pricing ↗
Singapore-region rate, flat regardless of prompt length.
Alibaba $2.00 $6.00 100万 tokens 均衡 alibabacloud.com — Model Studio pricing ↗
Current flagship; priced on Cohere's docs site rather than its main pricing page.
Cohere $2.50 $10.00 25.6万 tokens 均衡 docs.cohere.com — Command A ↗
Pricing is promotional, guaranteed only through Nov 21, 2026.
OpenAI $4.00 $20.00 约105万 tokens 均衡 platform.openai.com/docs — Pricing ↗
Optional "fast mode" (research preview) available at $10 / $50, first-party API only.
Anthropic $5.00 $25.00 100万 tokens 旗舰 platform.claude.com/docs — Pricing ↗
Cache hits priced at 0.025x base input (vs 0.1x standard) — the lowest cache rate Anthropic offers.
Anthropic $10.00 $50.00 100万 tokens 旗舰 platform.claude.com/docs — Pricing ↗
Short-context tier shown; long-context requests (very large prompts) bill at $20 / $75.
OpenAI $10.00 $50.00 约105万 tokens 旗舰 platform.openai.com/docs — Pricing ↗

价格对比一览

完整比较页面 →
最便宜的输入价格
Mistral · 每百万 tokens 输入 $0.15 / 输出 $0.60。适合高频、低复杂度任务的起点 — 分类、抽取、简单摘要。
所有追踪模型的平均值
$2.03 / $9.73
涵盖旗舰与经济档位在内的 25 个追踪模型的平均输入/输出价格 — 仅作为判断报价是否合理的粗略参考,并非推荐。
最贵的输入价格
Claude Fable 5.1 与 GPT-6 Astra 并列
Anthropic · 每百万 tokens 输入 $10.00 / 输出 $50.00 — 是最便宜追踪价格的 66.7 倍。我们不追踪基准测试分数,请将此视为价格比较,而非能力排名。

我们如何核实每一个价格

完整方法论 →
01
我们查阅官方来源
每个价格都来自服务商自己的公开定价页 — 绝不使用第三方聚合数据、抓取的 API 响应或估算值。对于开放权重模型(如 Llama),我们会注明具体是哪家托管方的价格,因为不同托管方对同一模型的收费并不相同。
02
我们为每次核实标注日期
这是一个新的追踪项目 — 上表反映的是 September 10, 2026 的单次核实结果。我们每 48 小时重新核实一次价格;任何变化都会标注新日期并被记录,而不会被悄悄覆盖。
03
我们保持免费和透明
浏览无需注册账号。我们不接受此处列出的任何服务商的赞助,也与其无关联 — 每一行都以相同方式核实和处理。展示的是标准的短上下文价格;大多数服务商还提供缓存、批处理或长上下文档位,我们会链接但不会在表格中完全展开。

常见问题

所有常见问题 →

Mistral Small 4(Mistral)是我们追踪中最便宜的模型,每百万输入 tokens $0.15,每百万输出 tokens $0.60。还有几个经济型模型价格紧随其后 — 详见上方完整表格。

差距很大。Claude Fable 5.1(Anthropic)每输入 token 的收费是 Mistral Small 4 的 66.7 倍(与 GPT-6 Astra 并列榜首)。旗舰级定价通常反映的是更大、推理能力更深的模型,而非输出质量的线性提升 — 在假设贵的模型就是对的之前,值得结合自己的实际任务做基准测试。

输入 tokens 是你发送给模型的内容 — 提示词、上下文、文档、对话历史。输出 tokens 是模型生成返回的内容。输出价格几乎总是更高,通常是输入价格的 3 到 5 倍,因为生成文本在计算上比读取文本更昂贵。

Meta 并不销售 Llama 的官方 API 访问 — 它公开发布模型权重,并让其他公司来托管。我们展示的 Llama 4 价格来自 Together AI,是提供该模型的多家推理服务商之一;其他托管方(Groq、Fireworks 等)的定价可能不同。

本页每个价格均于 September 10, 2026 直接对照服务商自己的定价页核实。此后我们每 48 小时重新核实一次 — 这是首次核实,所以目前还没有历史记录,但未来每次核实都会标注日期,任何变化都会被记录,历史记录将从此开始积累。

没有。PerTokens 独立运营,不接受任何赞助。每个服务商都以相同方式列出,使用相同的来源(其自身的公开定价页)和相同的核实日期。