Gemini 1.5 Flash is 97% cheaper than Command R+ on blended 3:1 token pricing. Command R+ costs $2.50/1M in and $10.00/1M out, while Gemini 1.5 Flash costs $0.07/1M in and $0.30/1M out.
Command R+ vs Gemini 1.5 Flash
Compare verified input/output token rates, prompt caching discounts, latency tiers, and context limits between Cohere and Google.
| Metric / Feature | Command R+ (Cohere) | Gemini 1.5 Flash (Google) |
|---|---|---|
| Input Price ($ / 1M Tokens) | $2.50 | $0.07 |
| Output Price ($ / 1M Tokens) | $10.00 | $0.30 |
| Prompt Caching Input Rate | Not Available | $0.019/1M |
| Blended 3:1 Rate (Production) | $4.375 / 1M | $0.131 / 1M |
| Max Context Window | 128k | 128k |
| Multimodal Vision | ❌ Text Only | ✅ Supported |
| MMLU Benchmark | 75.7% | 78.9% |
| Throughput & Latency | Fast (~40 t/s) | Ultra-Fast (~100 t/s) |
| Best For | Grounded RAG with citation generation, enterprise business tooling | High-frequency API calls, audio processing, bulk document summarization |
When to Choose Command R+
Choose Command R+ if your workload requires grounded rag with citation generation, enterprise business tooling. Ideal for teams needing Cohere's infrastructure and ecosystem tooling.
When to Choose Gemini 1.5 Flash
Choose Gemini 1.5 Flash if your primary objective is high-frequency api calls, audio processing, bulk document summarization. Ideal for scaling high-throughput pipelines with Google.
Want to compare Command R+ against another model?
Select another target model to view immediate pricing & latency trade-offs.