Command A pricing
Command A API cost from Cohere — $2.50 per 1M input tokens and $10.00 per 1M output tokens. Flagship, RAG-tuned. Estimate your real monthly bill below, then compare it against every other model.
Flagship 256K context · all prices USD per 1M tokens, list price last verified 2026-06-27.
Command A capabilities & limits
Representative specs for this tier — confirm exact limits in the Cohere docs.
API acting up? Check Cohere status ↗ · all AI API status
Call Command A from your code
The exact model string and endpoint to use. Works with the OpenAI SDK — just point base_url at the URL below.
command-a-03-2025https://api.cohere.com/compatibility/v1CO_API_KEYcurl https://api.cohere.com/compatibility/v1/chat/completions \
-H "Authorization: Bearer $CO_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "command-a-03-2025",
"messages": [{"role": "user", "content": "Hello"}]
}' Model strings follow each provider's naming convention; confirm the current id in the Cohere docs before relying on it.
What a Command A API call actually costs
Blended input + output cost at Command A's list rates — useful yardsticks before you commit to volume.
Command A cost calculator
Enter your monthly volume to estimate the Command A bill. Runs entirely in your browser.
Cohere models — Command A in context
How Command A prices against its siblings. Cheapest reference cost first.
| Model | Input /1M | Output /1M | Cost / call* | Context | Type |
|---|---|---|---|---|---|
| Command R7B Cheapest Cohere tier | $0.037 | $0.150 | $0.0002 | 128K | Fast / cheap |
| Command R | $0.150 | $0.600 | $0.0007 | 128K | Balanced |
| Command A Flagship, RAG-tuned | $2.50 | $10.00 | $0.013 | 256K | Flagship |
*Reference cost of one call with 1,000 input + 1,000 output tokens — a neutral yardstick. Use the calculator for your real usage. Cheapest row highlighted.
How Command A pricing works
Command A is billed per token: a lower input price of $2.50 per million tokens for everything you send, and a higher output price of $10.00 per million for everything the model generates. To turn that into a real Command A bill, multiply by your monthly request volume and average prompt/response length — the calculator above does exactly that.
Command A does not list a separate cached-input rate, so every input token bills at $2.50 per million. If repeated context is a big part of your prompts, a model with prompt caching may end up cheaper in practice.
List prices move fast — always confirm the current numbers on the official Cohere pricing page before relying on them for budgeting.
Command A pricing FAQ
How much does the Command A API cost?
Command A costs $2.50 per 1M input tokens and $10.00 per 1M output tokens. A typical 3K-input / 500-output call works out to about $0.013.
What is Command A's context window?
Command A has a 256K-token context window. Tokens you put in the prompt are billed at the $2.50 per 1M input rate.
Does Command A charge extra for output tokens?
Yes. Output (generated) tokens cost $10.00 per 1M, versus $2.50 per 1M for input — about 4.0× more, which is normal for LLM APIs.
Is there a cheaper alternative to Command A?
In the same flagship tier, MiniMax-M3 (MiniMax) is cheaper at $0.300 in / $1.20 out. The overall cheapest model on the market is Gemini 1.5 Flash-8B.