Best LLM API for…
Shortlists built from verified pricing and published context windows — not benchmarked capability. Pick a use case below to see which tracked models fit the budget.
Lowest input/output rates for classification, extraction, and bulk processing.
Largest published context windows for long documents and codebases.
Below-average input pricing across every tier — for teams watching burn.
The top-tier models, cheapest-first, when budget allows the flagship rate.
Balanced and flagship models with enough context for real codebases.
Budget-tier models sorted by output rate — the cost driver in chat.
Budget-tier models sorted by input rate — the cost driver for long inputs.
Balanced-tier models sorted by output rate — the cost driver for long-form copy.
Flagship-tier models sorted by the cheapest output rate — reasoning-heavy tasks are output-token-heavy.