Primary use stated or inferred cautiously from official documentation.
Command R 08‑2024
Previous commercial model optimized for RAG, citations, tools, and high performance on enterprise workloads.
The essential
Maximum input capacity when published by the source.
Documented maximum generation limit.
Verified availability channels.
Primary functional family and verified specialties.
Declared terms for API access, model weights, or self-hosting.
Limits and integration
| API ID | command-r-08-2024 |
|---|---|
| Model type | Text and reasoning |
| Access model | Paid proprietary |
| License | Proprietary |
| Deployment | Hosted API |
| Release | AUG 2024 |
| Knowledge cutoff | Not published |
| Entrance | Text · Image |
| Exit | Text |
| Context window | 128,000 tokens |
| maximum output | 4,000 tokens |
| Reasoning | Not applicable or not published |
| Published tools | Citations · RAG · Features |
| Structured Outings | Supported |
| Batch processing | Not published |
| Prompt cache | Not published |
| Fine-tuning | Not published |
| Verified platforms | Cohere API · Amazon Bedrock · Oracle OCI |
RAG with citations and high-volume text workloads.
Cohere recommends Command A for most new uses; retain R for compatibility and cost.
What it can do
Orientation
RAG with citations and high-volume text workloads.
Context
128,000 tokens · 4,000 tokens
Tools and integration
Citations · RAG · Functions
Access
Cohere API and associated clouds
Documented cost
| Concept | Worth | Unit/condition |
|---|---|---|
| Standard input | 0,15 USD / 1 M tokens | Standard API rate |
| Cached input | Not published | Reading reused prefixes |
| Cache write or storage | Not published | The condition varies by provider |
| Standard output | 0,60 USD / 1 M tokens | May include reasoning tokens |
| Batch input | Not published | Asynchronous processing |
| Batch output | Not published | Asynchronous processing |
Consult the primary source: the billing unit depends on the model type and access channel.
i Prices change and may depend on level, region or context length. Check the source before making a decision.
How to read the results
A public benchmark provides guidance, but does not replace an evaluation with your data, tools, budget, and error tolerance.
| Benchmark | Result | Metric | Source |
|---|---|---|---|
| Evaluaciones Command R | Published | RAG and tool use | View source ↗ |
Inferama only highlights a “best result” when the metric, test set, configuration, and date allow for an equivalent comparison. The supplier's figures are presented as claims from its own source.
Chronology
Command R 08‑2024
Version added to Inferama's verified catalog.
Related analysis
Command R 08-2024 and its safety modes: what they control and what you should test
STRICT, CONTEXTUAL, and NONE—called OFF in Chat V2—change the safety instructions included in a request. They do not certify an application or show how a workflow with documents or tools will behave. This guide explains their scope and proposes a testing protocol.
25 Sep 2026 ↗ COMPARATIVACommand A+ vs. Command R 08-2024: How to Decide on a Migration for a Multilingual RAG Assistant
Changing models does not necessarily improve an enterprise assistant. This comparison proposes a reproducible protocol to determine whether Command A+ delivers a net improvement over Command R 08-2024 when both operate on the same corpus, retriever, output contract, and human review process.
22 Sep 2026 ↗