Command R+ by Cohere costs $2.50 per 1 Million input tokens and $10.00 per 1 Million output tokens. It features a 128k context window and has a blended 3:1 production rate of $4.375/1M tokens.
Command R+
104B parameter model built specifically for accurate retrieval-augmented generation and tool integration.
Verified Specifications & Benchmark Data
| Developer / Provider | Cohere |
| Model Family | Command |
| Blended 3:1 Rate (Production Benchmark) | $4.375 / 1M tokens |
| Batch API Discount (24hr SLA) | None |
| Multimodal Vision | ❌ Text Only |
| Function Calling / Structured Outputs | ✅ Native Tool Calling |
| MMLU Benchmark Score | 75.7% |
| Latency & Throughput Tier | Fast (~40 t/s) |
| Optimal Architecture & Use Cases | Grounded RAG with citation generation, enterprise business tooling |
Frequently Asked Questions about Command R+
How much does Command R+ cost per 1M tokens?
Command R+ pricing is set at $2.50 per 1 million input tokens and $10.00 per 1 million output tokens. For high-volume batch workloads, 24-hour batch queues provide a None.
How much money does Command R+ prompt caching save?
Prompt caching is currently not natively offered for Command R+ on direct serverless endpoints.
What is the context window limit of Command R+?
Command R+ has a maximum context window of 128k (128,000 tokens), supporting up to 4,096 completion tokens per response.
Direct Matchups with Command R+
See how Command R+ compares against other leading frontier and open-weights models in cost and latency.
Compare Cohere vs OpenAI pricing, context limits, and cost per 1M tokens.
Compare Cohere vs Anthropic pricing, context limits, and cost per 1M tokens.
Compare Cohere vs Together AI (Meta) pricing, context limits, and cost per 1M tokens.