Provider
Cohere API pricing: every model compared
Cohere offers 5 models through its API in our data, priced from $0.0375 to $2.50 per million input tokens. The newest is Command A+, first listed on 2026-09-22.
- Models
- 5
- Cheapest input
- $0.0375
- Priciest input
- $2.50
- Largest context
- 256,000
- With caching
- 1 of 5
- Open weights
- 1 of 5
Cheapest
Command R7B (12-2024)Cohere$0.0375 in · $0.15 outLargest context
Command ACohere256,000 tokensNewest
Command A+Cohere2026-09-22Line-up
Every Cohere model and price
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Command A+ | $0.30 | $1.50 | 192,000 | 2026-09-22 |
| Command A | $2.50 | $10 | 256,000 | 2025-03-13 |
| Command R7B (12-2024) | $0.0375 | $0.15 | 128,000 | 2024-12-14 |
| Command R (08-2024) | $0.15 | $0.60 | 128,000 | 2024-08-30 |
| Command R+ (08-2024) | $2.50 | $10 | 128,000 | 2024-08-30 |
Prices in US dollars per 1M tokens, newest first. Release dates are when our source first listed the model.
Overview
Cohere’s pricing at a glance
The spread between Cohere’s cheapest and most expensive model is large: Command R+ (08-2024) costs 67× more per input token than Command R7B (12-2024). Picking the smallest model that does the job well is usually the biggest saving available, ahead of any discount.
1 of 5 of the line-up has a cached-input price, which helps chatbots and agents that resend the same instructions. No batch prices are published in our data. 1 of 5 accept images as input.
FAQ
Cohere API questions
How much does the Cohere API cost?
Cohere's 5 models range from $0.0375 to $2.50 per million input tokens, and output costs more than input on almost every model. The exact bill depends on your token volumes, which you can estimate in the LLM cost calculator.
What is Cohere's cheapest model?
By blended price (three parts input to one part output) it is Command R7B (12-2024), at $0.0375 input and $0.15 output per million tokens. The most expensive is Command R+ (08-2024).
Which Cohere model has the largest context window?
Command A, with 256,000 tokens per request.
Does Cohere offer prompt caching or batch discounts?
In our data, 1 of 5 of Cohere's models have a published cached-input price and none have a published batch price. Both can cut costs substantially for repeated prompts or work that can wait.