Command R7B — API Pricing & Benchmarks
Command R7B is Cohere's small model, released in 2024-12. The API costs $0.04 per 1M input tokens and $0.15 per 1M output tokens ($0.07 blended at a 3:1 ratio). It supports a 128K-token context window.
Explore provider routes and additional benchmark sources for Command R7B
| Provider | Cohere |
| Input price / 1M tokens | $0.04 |
| Output price / 1M tokens | $0.15 |
| Cached input / 1M tokens | — |
| Blended price (3:1) | $0.07 |
| Context window | 128K |
| Max output tokens | 4K |
| Knowledge cutoff | 2024-06 |
| License | open weights |
| Input modalities | T |
| Output speed (tok/s) | — |
| Released | 2024-12 |
Where Command R7B fits
On a blended 3:1 basis Command R7B costs $0.07 per 1M tokens, cheaper than 98% of the 80 priced models in this index. It is built for high-volume, latency-sensitive work — classification, extraction, routing and autocomplete — where per-token price matters more than peak intelligence. The weights are open, so it can be self-hosted for data-residency or cost reasons instead of consumed via API.
Price history
| Period | Input $/1M | Output $/1M | Cached $/1M |
|---|---|---|---|
| 2026-07-18 → current | $0.04 | $0.15 | — |
No list-price changes recorded since tracking began. Full dataset: price-history.json (CC BY 4.0).
Alternatives at a similar price
- Nova Micro (Amazon) — $0.06/1M blended, AA Index 4.4
- GPT-4.1 nano (OpenAI) — $0.18/1M blended, AA Index 9.6
- Qwen3.8-Flash (Alibaba Qwen) — $0.23/1M blended, AA Index 55.8
- GLM-5.3-Flash (Zhipu AI (Z.ai)) — $0.24/1M blended, AA Index 57.5
About Cohere
Cohere targets enterprise deployments with its Command series, emphasizing retrieval-augmented generation, tool use and private-deployment options over leaderboard placement.