Skip to content
AI Dev Toolkit.
Esc
  • AI Token CounterCount tokens for GPT, Claude, Gemini, DeepSeek, Qwen and more.Tool
  • LLM API Cost CalculatorEstimate per-request, daily and monthly API costs.Tool
  • AI Model ComparisonCompare prices, context windows and features across models.Tool
  • AI Model Pricing PagesSpecs, real costs and cheaper alternatives for popular models.Tool
  • Context Window CheckerSee whether your text fits each model's context window.Tool
  • Subscription vs API CalculatorFind out whether a chat plan or the API is cheaper for you.Tool
  • GPU / VRAM CalculatorCheck how much VRAM a local model needs and which GPUs fit.Tool
  • Claude Code Error DatabaseExact Claude Code error messages with tested fixes.Tool

Tokens & Costs

AI API providers and their pricing

Each AI company prices its models differently. Pick a provider to see every model it offers, what each costs per million tokens, and which ones support caching, batch discounts and image input.

ProviderModelsInput price range / 1MLargest contextNewest model
OpenAI60$0.018 – $1501,050,000GPT-6.1 Sol
Qwen53$0.03 – $41,048,576Qwen3.8 Max Prime
Google20$0.05 – $21,048,576Gemini 3.8 Flash
Mistral20$0.029 – $21,048,576Mistral Large 4
Anthropic16$0.10 – $151,000,000Claude Haiku 5.5
Z.ai16$0.0605 – $2.801,048,576GLM 5.3 Prime
Meta14$0.05 – $1.251,048,576Muse Spark 1.3
DeepSeek11$0.259 – $0.95531,048,576DeepSeek V4.1 Flash
MiniMax8$0.21 – $0.551,000,192MiniMax M3
Moonshot AI7$0.45 – $0.531,048,576Kimi K3
xAI7$1 – $22,000,000Grok 4.7
ByteDance Seed6$0.075 – $0.50262,144Seed 2.1 Turbo
Amazon5$0.035 – $2.501,000,000Nova 2 Lite
Cohere5$0.0375 – $2.50256,000Command A+
NVIDIA5$0.0469 – $0.50262,144Nemotron 3.5 Lightning
Perplexity5$1 – $3200,000Sonar Pro Search

Comparing specific models across providers? Use the AI model comparison table.

Prices updated 2026-10-09

About

How providers differ

Each provider here is a company that builds models and sells access to them through an API. They price differently: most offer a range from small, cheap models for simple tasks to large ones for hard reasoning. For 10 of the 16 providers here, the most expensive model costs at least ten times as much per input token as the cheapest.

Features matter as much as price. Prompt caching cuts the cost of repeated input, batch processing trades speed for a discount (8 of the 16 providers here list batch prices for at least one model), and 11 of them publish open-weight models that you can also run on your own hardware; the VRAM calculator tells you what that needs.

The list includes established providers with at least 3 priced text models in the data. Each provider page lists every model with its prices, context window and features, so you can pick within a family before comparing across them.

Further reading: Claude vs GPT vs Gemini pricing, explained with live prices, The cheapest LLM APIs right now, ranked from daily price data, What is a token in AI? A plain-English guide with real examples.

FAQ

Questions people ask

Do I need a separate API key for each provider?

Yes, if you use them directly: each provider issues its own keys and bills separately. Routing services such as OpenRouter offer one key for many providers’ models. Some providers have free tiers; the free Gemini API key guide shows Google’s.

Why do providers count tokens differently?

Each model family has its own tokenizer, so the same text becomes a different number of tokens on each one. A lower price per token doesn’t always mean a lower bill. The token counter shows the counts side by side.

How often is this page updated?

Every day, from the same price data as the rest of the site. These figures are from 2026-10-09.

Which provider is cheapest?

It depends on the model size you need. Compare like for like: small models against small models, flagship against flagship. The model comparison can filter by features and sort by price.