Vergleichen Sie den Preis jedes KI-Modells — bevor Sie darauf aufbauen.
Input- und Output-Preise für 10 Anbieter — OpenAI, Anthropic, Google, Mistral, DeepSeek, xAI und mehr — direkt auf der Preisseite jedes Anbieters geprüft, mit Quelle und Datum in jeder Zeile.
Modellpreise heute
Quellen: die öffentliche Preisseite jedes Anbieters · verifiziert am September 10, 2026| Modell | Anbieter | Input /1M | Output /1M | Kontext | Stufe | Quelle |
|---|---|---|---|---|---|---|
| Mistral | $0.15 | $0.60 | Nicht veröffentlicht | Budget | mistral.ai — API pricing ↗ | |
|
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate.
|
Meta (via Together AI) | $0.18 | $0.59 | 1M Tokens | Budget | together.ai — Llama 4 Scout ↗ |
| OpenAI | $0.20 | $1.20 | ~1,05M Tokens | Budget | platform.openai.com/docs — Pricing ↗ | |
|
Audio input priced separately at $0.50/1M.
|
$0.25 | $1.50 | Nicht veröffentlicht | Budget | ai.google.dev — Gemini API pricing ↗ | |
|
Meta doesn't sell first-party API access — price shown is Together AI's hosted rate, one of several inference providers serving this open-weight model.
|
Meta (via Together AI) | $0.27 | $0.85 | ~1,05M Tokens | Budget | together.ai — Llama 4 Maverick ↗ |
|
Peak-hour, cache-miss rate shown. Off-peak (01:00–04:00 & 06:00–10:00 UTC, Mon–Fri) is half price; cache-hit input is far cheaper still ($0.006/1M peak).
|
DeepSeek | $0.30 | $1.20 | 1M Tokens | Budget | api-docs.deepseek.com — Pricing ↗ |
|
Bedrock Standard tier. Flex tier: $0.15 / $1.25.
|
Amazon | $0.30 | $2.50 | 1M Tokens | Budget | aws.amazon.com — Bedrock pricing ↗ |
|
Rate shown for 0–256K input tokens; both 256K–1M tier and above bill higher on input only.
|
Alibaba | $0.40 | $1.60 | 1M Tokens | Budget | alibabacloud.com — Model Studio pricing ↗ |
|
Despite the "Large" name, priced below Medium 3.5 in Mistral's current lineup.
|
Mistral | $0.50 | $1.50 | Nicht veröffentlicht | Budget | mistral.ai — API pricing ↗ |
|
Promotional rate through Dec 31, 2026; steps up to $1.50 / $7.50 on Jan 1, 2027.
|
$0.75 | $3.75 | Nicht veröffentlicht | Ausgewogen | ai.google.dev — Gemini API pricing ↗ | |
| Anthropic | $1.00 | $5.00 | 200K Tokens | Budget | platform.claude.com/docs — Pricing ↗ | |
| xAI | $1.00 | $2.00 | 256K Tokens | Budget | docs.x.ai — Pricing ↗ | |
|
Long-context tier (>200K input) doubles to $2.50 / $5.
|
xAI | $1.25 | $2.50 | 1M Tokens | Ausgewogen | docs.x.ai — Pricing ↗ |
|
Preview model, Bedrock Standard tier. A cheaper "Flex" tier and pricier "Priority" tier also exist.
|
Amazon | $1.25 | $10.00 | Nicht veröffentlicht (Vorschau) | Ausgewogen | aws.amazon.com — Bedrock pricing ↗ |
|
Positioned by Mistral as its top general-purpose / agentic model — priced above "Large 3".
|
Mistral | $1.50 | $7.50 | Nicht veröffentlicht | Ausgewogen | mistral.ai — API pricing ↗ |
|
Launched at introductory pricing through Aug 31, 2026; the scheduled increase to $3/$15 was cancelled — $2/$10 is now the standing price.
|
Anthropic | $2.00 | $10.00 | 1M Tokens | Ausgewogen | platform.claude.com/docs — Pricing ↗ |
| OpenAI | $2.00 | $12.00 | ~1,05M Tokens | Ausgewogen | platform.openai.com/docs — Pricing ↗ | |
|
Requests over 200K input tokens bill at $4 / $18 instead.
|
$2.00 | $12.00 | Stufe ≤200K Tokens | Flaggschiff | ai.google.dev — Gemini API pricing ↗ | |
|
Long-context tier (>200K input) doubles to $4 / $12.
|
xAI | $2.00 | $6.00 | 500K Tokens | Flaggschiff | docs.x.ai — Pricing ↗ |
|
Singapore-region rate, flat regardless of prompt length.
|
Alibaba | $2.00 | $6.00 | 1M Tokens | Ausgewogen | alibabacloud.com — Model Studio pricing ↗ |
|
Current flagship; priced on Cohere's docs site rather than its main pricing page.
|
Cohere | $2.50 | $10.00 | 256K Tokens | Ausgewogen | docs.cohere.com — Command A ↗ |
|
Pricing is promotional, guaranteed only through Nov 21, 2026.
|
OpenAI | $4.00 | $20.00 | ~1,05M Tokens | Ausgewogen | platform.openai.com/docs — Pricing ↗ |
|
Optional "fast mode" (research preview) available at $10 / $50, first-party API only.
|
Anthropic | $5.00 | $25.00 | 1M Tokens | Flaggschiff | platform.claude.com/docs — Pricing ↗ |
|
Cache hits priced at 0.025x base input (vs 0.1x standard) — the lowest cache rate Anthropic offers.
|
Anthropic | $10.00 | $50.00 | 1M Tokens | Flaggschiff | platform.claude.com/docs — Pricing ↗ |
|
Short-context tier shown; long-context requests (very large prompts) bill at $20 / $75.
|
OpenAI | $10.00 | $50.00 | ~1,05M Tokens | Flaggschiff | platform.openai.com/docs — Pricing ↗ |
So vergleichen sich die Preise
Vollständige Vergleichsseite →Wie wir jeden Preis verifizieren
Vollständige Methodik →Häufig gestellte Fragen
Alle FAQ →Mistral Small 4 (Mistral) ist das günstigste von uns erfasste Modell, mit $0.15 pro Million Input-Tokens und $0.60 pro Million Output-Tokens. Mehrere andere Budget-Modelle liegen dicht dahinter — siehe die vollständige Tabelle oben.
Groß. Claude Fable 5.1 (Anthropic) verlangt 66.7x mehr pro Input-Token als Mistral Small 4 (gleichauf mit GPT-6 Astra an der Spitze). Flaggschiff-Preise spiegeln in der Regel größere Modelle mit tieferem Reasoning wider, nicht eine lineare Verbesserung der Ausgabequalität — es lohnt sich, dies an der eigenen Aufgabe zu messen, bevor man annimmt, das teurere Modell sei die richtige Wahl.
Input-Tokens sind das, was Sie an das Modell senden — Ihr Prompt, Kontext, Dokumente, Gesprächsverlauf. Output-Tokens sind das, was es zurückgibt. Output ist fast immer teurer, oft 3- bis 5-mal so teuer wie der Input-Preis, weil das Generieren von Text rechnerisch aufwendiger ist als das Lesen.
Meta verkauft keinen eigenen API-Zugang zu Llama — das Unternehmen veröffentlicht die Modellgewichte und lässt andere Firmen es hosten. Die von uns gezeigten Preise für Llama 4 stammen von Together AI, einem von mehreren Inferenz-Anbietern, die es bereitstellen; andere Hoster (Groq, Fireworks und andere) berechnen es möglicherweise anders.
Jeder Preis auf dieser Seite wurde am September 10, 2026 direkt anhand der Preisseite des Anbieters geprüft. Ab sofort prüfen wir alle 48 Stunden erneut — dies war die erste Prüfung, es gibt also noch keinen Verlauf, aber jede künftige Prüfung erhält ein Datum und jede Änderung wird protokolliert, sodass der Verlauf von hier an entsteht.
Nein. PerTokens ist unabhängig und nicht gesponsert. Jeder Anbieter wird auf die gleiche Weise gelistet, mit derselben Quelle (seiner eigenen öffentlichen Preisseite) und demselben Verifizierungsdatum.