← All open-weight models

Cohere · General language

Command R7B

Cohere's cheapest served model by a wide margin ($0.0375/1M in) with a 128K window — right for latency-sensitive chatbots, classification and on-device/consumer-GPU deployment where 4K output is enough.

Open weights · Non-commercial (CC-BY-NC) · Data checked September 6, 2026 · Not a hands-on evaluation

Model ID command-r7b-12-2024

Parameters
8B
Architecture
Not recorded
Context window
128,000 tokensMaximum input
Max output
4,000 tokensPer response
Input price
$0.0375 per 1M tokens
Output price
$0.15 per 1M tokens
Cached input
Not published
Blended price
$0.0656 per 1M tokens3:1 input to output
Knowledge cutoff
Not recorded
Released
2024-12
Status
ga
Size band
4B to 15B

License and openness

Command R7B is released under the CC-BY-NC-4.0. This license does not permit commercial use without a separate agreement with the provider. It is suited to research, evaluation and personal projects. License terms can change between versions; confirm the text published with the weights you download.

Capabilities

Input modalitiestext
Output modalitiestext

What it is for

At a three-to-one input-to-output ratio, Command R7B costs $0.07 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $2.61 per month at list price, before caching or batch discounts.

Cost at list price

Input$0.0375 per 1M tokens
Output$0.15 per 1M tokens
Cached inputNot published
Blended, 3:1$0.0656 per 1M tokens
Estimate a workload across all models ↗

Sources

Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.

Other open-weight models from Cohere

See all 19 from Cohere ↗

Browse all open-weight models ↗