# Command API pricing, every Cohere model per million tokens

On 25 September 2026, Cohere models on this list cost from $0.0375 per million input tokens (Command R7B (12-2024)) to $2.50 (Command A), and from $0.15 per million output tokens (Command R7B (12-2024)) to $10.00 (Command A).

Prices checked 25 September 2026 at 06.00 UTC . Source OpenRouter public model list . US dollars per million tokens

Input tokens per call

Output tokens per call

Calls a day

Search models

| Model | Input | Output | Cached input | Batch in / out | Context | Your month | Listed |
|---|---|---|---|---|---|---|---|
| Command A+ | $0.30 | $1.50 | $0.15 |  | 192K | $40.50 | 2026-09-22 |
| Command A | $2.50 | $10.00 |  |  | 256K | $300.00 | 2025-03-13 |
| Command R7B (12-2024) | $0.0375 | $0.15 |  |  | 128K | $4.50 | 2024-12-14 |
| Command R (08-2024) | $0.15 | $0.60 |  |  | 128K | $18.00 | 2024-08-30 |
| Command R+ (08-2024) | $2.50 | $10.00 |  |  | 128K | $300.00 | 2024-08-30 |

[Compare Cohere with every other provider](https://shurco.ai/tools/llm-api-pricing/)

## Common questions

### How much does the Command API cost?

On 25 September 2026, Cohere models on this list cost from $0.0375 per million input tokens (Command R7B (12-2024)) to $2.50 (Command A), and from $0.15 per million output tokens (Command R7B (12-2024)) to $10.00 (Command A).

### What is the cheapest Command model?

For a typical mix of three input tokens to every output token, the cheapest is Command R7B (12-2024) at $0.0375 per million input tokens and $0.15 per million output tokens.

### What is the newest Command model on the list?

Command A+, listed on 22 September 2026, at $0.30 per million input tokens and $1.50 per million output tokens.

### Which Command model has the largest context window?

Command A, with 256K tokens of context.

## Other providers

[GPT](https://shurco.ai/tools/llm-api-pricing/openai/)[Claude](https://shurco.ai/tools/llm-api-pricing/anthropic/)[Gemini](https://shurco.ai/tools/llm-api-pricing/google/)[DeepSeek](https://shurco.ai/tools/llm-api-pricing/deepseek/)[Mistral](https://shurco.ai/tools/llm-api-pricing/mistral/)[Llama](https://shurco.ai/tools/llm-api-pricing/meta/)[Grok](https://shurco.ai/tools/llm-api-pricing/xai/)[Qwen](https://shurco.ai/tools/llm-api-pricing/qwen/)[Kimi](https://shurco.ai/tools/llm-api-pricing/moonshot/)[GLM](https://shurco.ai/tools/llm-api-pricing/zai/)[Nova](https://shurco.ai/tools/llm-api-pricing/amazon/)[MiniMax](https://shurco.ai/tools/llm-api-pricing/minimax/)

## Choosing a model for a production agent

Price is one input. Quality on your own cases, latency and the cost of retries decide the rest. [Tell me what you are building.](https://shurco.ai/#contact)

Web version https://shurco.ai/tools/llm-api-pricing/cohere/
