{T}
TokenMath.net
Verified against Cohere Official API Pricing
CohereflagshipGenerally Available

Command R+ (08-2024) Token Counter & Cost Calculator

Enterprise-grade foundation model optimized for conversational RAG, tool use, and citations

Standard Input / 1M
$2.50
Prompt tokens
Cached Input / 1M
$0.25
Prompt caching rate
Output / 1M
$10.00
Generation completion
Context Window
128k
Max out: 4.1k
API & Developer Integration Details
generally available
API Model Stringcommand-r-plus-08-2024
Provider & FamilyCohere (cohere)
Prompt Caching Rules1,024 min tokens (90% off)
Long-Context TiersStandard flat rate throughout full window
First-party audit timestamp: 2026-09-15 14:30 UTCOfficial Documentation

Interactive Cost & Token Simulator for Command R+ (08-2024)

Live Calculation
Input Tokens2,500
Output Tokens800
Requests / Day5,000
Prompt Caching Ratio (50%)$0.25/M cached rate
Cost / Request$0.0114
Daily Spend (5,000 reqs)$57.1875
Monthly Run-Rate (30d)$1,715.63

Standard Workload Cost Scenarios

ScenarioInput TokensOutput TokensUncached CostWith Prompt Caching
Short Chat Query1,000500$0.00750$0.00525
Document Summarization10,0002,000$0.0450$0.0225
Codebase & Context Analysis100,00020,000$0.4500$0.2250
Batch Corpus Processing1,000,000100,000$3.50$1.25
Technical Architecture & Pricing VerificationVerified: 2026-09-15 14:30 UTC
Model ArchitectureCommand R+ (08-2024)
Provider OrganizationCohere
Tokenizer EncodingCohere BPE (~256k vocabulary)
API ConnectivityCohere Chat API (/v2/chat) with native citations, web search connector, and multi-step tool calling.
Batch API 50% DiscountNot Available
Tier 1 Rate Quota100 RPM · 100k TPM
Blended 3:1 Cost / 1M Tokens$4.38
Prompt caching provides a 90% read discount (.25/1M) on cached documents and prefixes.Cohere Official API Pricing

Verified Pricing History

September 2026: $2.5/M input · $10/M output ($0.25/M cached)
Verified Cohere API pricing.

When to Choose Command R+ (08-2024)

Ideal for production workloads demanding flagship capabilities, deep context depth (128k tokens), and reliability from Cohere. Excellent when predictable tokenomics and prompt caching support are paramount.

When Another Model May Be Better

If your use-case requires sub-second streaming latency or ultra-high frequency classification at micro-cent pricing, consider lighter budget options such as Gemini Flash-Lite or Claude Haiku. For deep formal logic, consider dedicated reasoning models like o3.

Frequently Asked Questions About Command R+ (08-2024)

How much does 1 million tokens cost with Command R+ (08-2024)?

For Command R+ (08-2024), 1 million input tokens costs $2.50, while 1 million output tokens costs $10.00. If using prompt caching, repetitive input prefixes are discounted to $0.25 per million.

What is the context window for Command R+ (08-2024)?

Command R+ (08-2024) features a maximum context window of 128,000 tokens (~96,000 words), with a maximum output limit of 4,096 tokens per completion.

Which tokenizer does Command R+ (08-2024) use?

Command R+ (08-2024) utilizes the Cohere BPE (~256k vocabulary). Token counting on TokenMath runs client-side to ensure maximum privacy.

Does Command R+ (08-2024) support prompt caching discounts?

Yes. Command R+ (08-2024) supports prompt caching with a cached input rate of $0.25/1M (saving up to 90% on repeated input context).