Command A API pricing & cost
Current token rates, context limits, workload examples and unit-economics shortcuts for Command A.
Pricing page fallback verified September 3, 2026; client checks TokenCOGS live registry on load.
What the token rates mean in dollars
Example workload: 2,000 input tokens + 500 output tokens per request, standard text-token pricing.
| Volume | Estimated API cost | Approx. / request |
|---|---|---|
| 1,000 requests | $10.00 | $0.01 |
| 100,000 requests | $1,000.00 | $0.01 |
| 1,000,000 requests | $10,000.00 | $0.01 |
Where Command A sits on price
Command A sits toward the premium end of this provider’s current model lineup. Price alone is not a quality ranking, but this position is useful when you are screening models for unit-economics fit before benchmarking latency and quality.
Cost comparison is not a quality, latency or reliability ranking. Benchmark the task you actually need to run.
Compare all Cohere models →What can move the invoice
- A cached-input rate is not stored for this model in the current registry, so caching savings should be verified directly with the provider.
Ways this provider may reduce spend
No verified provider-specific credits or promotions are currently stored in the TokenCOGS registry.
Turn pricing into unit economics
Command A pricing questions
How much does Command A cost per 1M tokens?
The current TokenCOGS snapshot lists Command A at $2.5 per 1M input tokens and $10 per 1M output tokens. Live registry data can update these values when the provider changes pricing.
What would 1,000 Command A requests cost?
At 2,000 input tokens and 500 output tokens per request, 1,000 requests are approximately $10.00 before provider-specific extras. Change the workload in the linked calculator for your own usage pattern.
Is Command A the cheapest Cohere model?
Command A sits toward the premium end of this provider’s current model lineup. Price alone is not a quality ranking, but this position is useful when you are screening models for unit-economics fit before benchmarking latency and quality.