在你基于它构建之前,先比较每个 AI 模型的价格。
10 家服务商的输入与输出价格 — OpenAI、Anthropic、Google、Mistral、DeepSeek、xAI 等 — 均直接对照各服务商官方定价页核实,每一行都标注来源与日期。
今日模型价格
来源:各服务商官方定价页 · September 10, 2026 核实| 模型 | 服务商 | 输入 /百万 | 输出 /百万 | 上下文 | 档位 | 来源 |
|---|---|---|---|---|---|---|
| Mistral | $0.15 | $0.60 | 未公开 | 经济 | mistral.ai — API pricing ↗ | |
|
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate.
|
Meta (via Together AI) | $0.18 | $0.59 | 100万 tokens | 经济 | together.ai — Llama 4 Scout ↗ |
| OpenAI | $0.20 | $1.20 | 约105万 tokens | 经济 | platform.openai.com/docs — Pricing ↗ | |
|
Audio input priced separately at $0.50/1M.
|
$0.25 | $1.50 | 未公开 | 经济 | ai.google.dev — Gemini API pricing ↗ | |
|
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate, one of several inference providers serving this open-weight model.
|
Meta (via Together AI) | $0.27 | $0.85 | 约105万 tokens | 经济 | together.ai — Llama 4 Maverick ↗ |
|
Peak-hour, cache-miss rate shown. Off-peak (01:00–04:00 & 06:00–10:00 UTC, Mon–Fri) is half price; cache-hit input is far cheaper still ($0.006/1M peak).
|
DeepSeek | $0.30 | $1.20 | 100万 tokens | 经济 | api-docs.deepseek.com — Pricing ↗ |
|
Bedrock Standard tier. Flex tier: $0.15 / $1.25.
|
Amazon | $0.30 | $2.50 | 100万 tokens | 经济 | aws.amazon.com — Bedrock pricing ↗ |
|
Rate shown for 0–256K input tokens; both 256K–1M tier and above bill higher on input only.
|
Alibaba | $0.40 | $1.60 | 100万 tokens | 经济 | alibabacloud.com — Model Studio pricing ↗ |
|
Despite the "Large" name, priced below Medium 3.5 in Mistral's current lineup.
|
Mistral | $0.50 | $1.50 | 未公开 | 经济 | mistral.ai — API pricing ↗ |
|
Promotional rate through Dec 31, 2026; steps up to $1.50 / $7.50 on Jan 1, 2027.
|
$0.75 | $3.75 | 未公开 | 均衡 | ai.google.dev — Gemini API pricing ↗ | |
| Anthropic | $1.00 | $5.00 | 20万 tokens | 经济 | platform.claude.com/docs — Pricing ↗ | |
| xAI | $1.00 | $2.00 | 25.6万 tokens | 经济 | docs.x.ai — Pricing ↗ | |
|
Long-context tier (>200K input) doubles to $2.50 / $5.
|
xAI | $1.25 | $2.50 | 100万 tokens | 均衡 | docs.x.ai — Pricing ↗ |
|
Preview model, Bedrock Standard tier. A cheaper "Flex" tier and pricier "Priority" tier also exist.
|
Amazon | $1.25 | $10.00 | 未公开(预览版) | 均衡 | aws.amazon.com — Bedrock pricing ↗ |
|
Positioned by Mistral as its top general-purpose / agentic model — priced above "Large 3".
|
Mistral | $1.50 | $7.50 | 未公开 | 均衡 | mistral.ai — API pricing ↗ |
|
Launched at introductory pricing through Aug 31, 2026; the scheduled increase to $3/$15 was cancelled — $2/$10 is now the standing price.
|
Anthropic | $2.00 | $10.00 | 100万 tokens | 均衡 | platform.claude.com/docs — Pricing ↗ |
| OpenAI | $2.00 | $12.00 | 约105万 tokens | 均衡 | platform.openai.com/docs — Pricing ↗ | |
|
Requests over 200K input tokens bill at $4 / $18 instead.
|
$2.00 | $12.00 | ≤20万 tokens 档位 | 旗舰 | ai.google.dev — Gemini API pricing ↗ | |
|
Long-context tier (>200K input) doubles to $4 / $12.
|
xAI | $2.00 | $6.00 | 50万 tokens | 旗舰 | docs.x.ai — Pricing ↗ |
|
Singapore-region rate, flat regardless of prompt length.
|
Alibaba | $2.00 | $6.00 | 100万 tokens | 均衡 | alibabacloud.com — Model Studio pricing ↗ |
|
Current flagship; priced on Cohere's docs site rather than its main pricing page.
|
Cohere | $2.50 | $10.00 | 25.6万 tokens | 均衡 | docs.cohere.com — Command A ↗ |
|
Pricing is promotional, guaranteed only through Nov 21, 2026.
|
OpenAI | $4.00 | $20.00 | 约105万 tokens | 均衡 | platform.openai.com/docs — Pricing ↗ |
|
Optional "fast mode" (research preview) available at $10 / $50, first-party API only.
|
Anthropic | $5.00 | $25.00 | 100万 tokens | 旗舰 | platform.claude.com/docs — Pricing ↗ |
|
Cache hits priced at 0.025x base input (vs 0.1x standard) — the lowest cache rate Anthropic offers.
|
Anthropic | $10.00 | $50.00 | 100万 tokens | 旗舰 | platform.claude.com/docs — Pricing ↗ |
|
Short-context tier shown; long-context requests (very large prompts) bill at $20 / $75.
|
OpenAI | $10.00 | $50.00 | 约105万 tokens | 旗舰 | platform.openai.com/docs — Pricing ↗ |
价格对比一览
完整比较页面 →我们如何核实每一个价格
完整方法论 →常见问题
所有常见问题 →Mistral Small 4(Mistral)是我们追踪中最便宜的模型,每百万输入 tokens $0.15,每百万输出 tokens $0.60。还有几个经济型模型价格紧随其后 — 详见上方完整表格。
差距很大。Claude Fable 5.1(Anthropic)每输入 token 的收费是 Mistral Small 4 的 66.7 倍(与 GPT-6 Astra 并列榜首)。旗舰级定价通常反映的是更大、推理能力更深的模型,而非输出质量的线性提升 — 在假设贵的模型就是对的之前,值得结合自己的实际任务做基准测试。
输入 tokens 是你发送给模型的内容 — 提示词、上下文、文档、对话历史。输出 tokens 是模型生成返回的内容。输出价格几乎总是更高,通常是输入价格的 3 到 5 倍,因为生成文本在计算上比读取文本更昂贵。
Meta 并不销售 Llama 的官方 API 访问 — 它公开发布模型权重,并让其他公司来托管。我们展示的 Llama 4 价格来自 Together AI,是提供该模型的多家推理服务商之一;其他托管方(Groq、Fireworks 等)的定价可能不同。
本页每个价格均于 September 10, 2026 直接对照服务商自己的定价页核实。此后我们每 48 小时重新核实一次 — 这是首次核实,所以目前还没有历史记录,但未来每次核实都会标注日期,任何变化都会被记录,历史记录将从此开始积累。
没有。PerTokens 独立运营,不接受任何赞助。每个服务商都以相同方式列出,使用相同的来源(其自身的公开定价页)和相同的核实日期。