Skip to main content

Provider pricing hub

Cohere API Pricing and Models

Cohere API models for generation, retrieval, reranking, and SaaS cost planning.

Pricing summary

Cohere pricing at a glance

These facts are derived from the canonical provider models currently available to CostRivo.

Models tracked

7

Cached-price models

0

Context window range

128,000 tokens

Latest verification

Sep 9, 2026

Model pricing

Cohere models

Compare provider models with the same pricing unit before opening a model-specific calculator.

Swipe sideways to see all columns.

Canonical model pricing and verified model facts for this provider
ModelInput priceOutput priceCached inputContext windowCapabilitiesVerificationAction

Command

Command

$1.00 / 1M tokens$2.00 / 1M tokensNot availableNot listedtext
CohereCommandexisting customersExisting customers only
Official source

Verified Sep 9, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Command Light

Command

$0.30 / 1M tokens$0.60 / 1M tokensNot availableNot listedtext
CohereCommand-lightexisting customersExisting customers only
Official source

Verified Sep 9, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Command R 03-2024

Command R

$0.50 / 1M tokens$1.50 / 1M tokensNot availableNot listedtext
CohereCommand R 03-2024existing customersExisting customers only
Official source

Verified Sep 9, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Command R+ 04-2024

Command R+

$3.00 / 1M tokens$15.00 / 1M tokensNot availableNot listedtext
CohereCommand R+ 04-2024existing customersExisting customers only
Official source

Verified Sep 9, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Command R+ 08-2024

Command R+

$2.50 / 1M tokens$10.00 / 1M tokensNot availableNot listedtext
CohereCommand R+ 08-2024existing customersExisting customers only
Official source

Verified Sep 9, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Aya Expanse 32B

Aya Expanse

$0.50 / 1M tokens$1.50 / 1M tokensNot available128,000 tokenstext
CohereAya ExpanseStable
Official source

Verified Jul 14, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Command R7B

Command R

$0.0375 / 1M tokens$0.15 / 1M tokensNot available128,000 tokenstext, reasoning, function calling
CohereCommand R7BStable
Official source

Verified Jul 14, 2026. Estimates vary by usage and provider pricing conditions.

Open model

Price-derived model highlights

Each criterion is calculated independently from eligible, verified Cohere token pricing. The same model may lead more than one criterion.

Lowest input price

Command R7B

$0.0375 / 1M

Lowest listed cost for 1M input tokens.

View model pricing

Lowest output price

Command R7B

$0.15 / 1M

Lowest listed cost for 1M output tokens.

View model pricing

Lowest combined token price

Command R7B

$0.1875 total

Cost of 1M input tokens plus 1M output tokens.

View model pricing

Calculator

Estimate Cohere API cost

Cohere is preselected with Aya Expanse 32B as the starting model.

Model selection

Choose the provider and model you want to estimate.
CohereAya Expanse

Input price

$0.50 / 1M tokens

Output price

$1.50 / 1M tokens

CohereAya ExpanseStable
Official source

Verified Jul 14, 2026. Estimates vary by usage and provider pricing conditions.

Usage assumptions

Estimate traffic and token usage for an average request.
Active seats, customers, or internal users.
Average AI calls per user each day.
Prompt, history, and retrieved context per request.
Generated answer length; SaaS founders should test long replies.
Use 30 for always-on products or fewer for batch jobs.

Estimated results

Run the calculator to see projected cost and usage volume.

Enter your usage details, then select Calculate estimate to see your projected cost.

Estimated cost = input usage cost + output usage cost + supported optional charges.

Model catalog

Cohere model catalog

Calculator actions appear only for exact model IDs with compatible verified token pricing.

15 current5 historicalReviewed Jul 14, 2026Official catalog source

15 of 15 models

Filters

Generative

4

Command A+

command-a-plus-05-2026

Command A

GenerativeActiveAPI available

Open-source enterprise agent model with multimodal reasoning, multilingual support, and tool use.

Context
128,000 tokens
Maximum output
64,000 tokens
Input
Text, Image
Output
Text
Knowledge cutoff
2025-04-01
Languages
48
Parameters
218B total, 25B active
License
Apache-2.0

Capabilities

ReasoningMultilingualImage InputsTool UseStructured OutputsCitations

Endpoints

Chat V2Chat V1Chat Completions

Deployment

Cohere ApiModel VaultPrivate DeploymentOpen Weights

Catalog details only

Free rate-limited access is not a production token price.

Official sourceVerified Jul 14, 2026

Command A Translate

command-a-translate-08-2025

Command A

GenerativeActiveAPI available

Context-aware multilingual translation model for enterprise localization workflows.

Context
8,000 tokens
Maximum output
8,000 tokens
Input
Text
Output
Text
Knowledge cutoff
2024-06-01
Languages
23
License
CC-BY-NC-4.0

Capabilities

TranslationMultilingualTool UseStructured Outputs

Endpoints

Chat V2Chat V1Chat Completions

Deployment

Cohere ApiPrivate DeploymentEnterprise Contact

Catalog details only

Free rate-limited access is not a production token price.

Official sourceVerified Jul 14, 2026

Command A Vision

command-a-vision-07-2025

Command A

GenerativeActiveAPI available

Multimodal model for document understanding, OCR, chart interpretation, and visual question answering.

Context
128,000 tokens
Maximum output
8,000 tokens
Input
Text, Image
Output
Text
Knowledge cutoff
2024-06-01
License
Apache-2.0

Capabilities

MultimodalImage UnderstandingDocument AnalysisOcrStructured OutputsReasoning

Limitations

Tool Use Not SupportedDoes Not Generate Images

Endpoints

Chat V2Chat V1Chat Completions

Deployment

Cohere ApiPrivate DeploymentEnterprise Contact

Catalog details only

Free rate-limited access is not a production token price.

Official sourceVerified Jul 14, 2026

Command R7B

command-r7b-12-2024

Command R

GenerativeActiveAPI available

Small, fast enterprise model for RAG, agents, tool use, code assistants, and latency-sensitive workloads.

Context
128,000 tokens
Maximum output
4,000 tokens
Input
Text
Output
Text
Knowledge cutoff
2024-06-01
Parameters
7B
License
CC-BY-NC-4.0

Capabilities

RagReasoningTool UseAgentsMultilingualCitations

Endpoints

Chat V2Chat V1Chat Completions

Deployment

Cohere ApiPrivate DeploymentOpen Weights

Verified token pricing

$0.0375 input / $0.15 output per 1M tokens

Use calculator
Official sourceVerified Jul 14, 2026

Coding

1

North Mini Code

north-mini-code-1-0

North

CodingActiveAPI available

Mixture-of-experts coding model for agentic software engineering, terminal agents, and code generation.

Context
256,000 tokens
Maximum output
64,000 tokens
Input
Text
Output
Text
Parameters
30B total, 3B active
License
Apache-2.0

Capabilities

Agentic CodingReasoningTool UseStructured OutputsCode Generation

Endpoints

Chat V2Chat V1Chat Completions

Deployment

Cohere ApiModel VaultPrivate DeploymentOpen Weights

Catalog details only

Free rate-limited access is not a production token price.

Official sourceVerified Jul 14, 2026

Audio

1

Transcribe

cohere-transcribe-03-2026

Cohere Transcribe

AudioActiveAPI available

Automatic speech recognition model for multilingual enterprise transcription.

Input
Audio
Output
Text
Languages
14
Parameters
2B
License
Apache-2.0

Capabilities

Speech To TextBatch ProcessingMultilingual Transcription

Limitations

No Automatic Language DetectionNo TimestampsNo Speaker Diarization

Endpoints

Audio Transcriptions

Deployment

Cohere ApiModel VaultOpen Weights

Catalog details only

Free rate-limited access is not a production token price.

Official sourceVerified Jul 14, 2026

Embeddings

1

Embed 4

embed-v4.0

Embed

EmbeddingsActiveAPI available

Multilingual multimodal embedding model for search, retrieval, RAG, and mixed text-image documents.

Context
128,000 tokens
Input
Text, Image, Mixed Text Image, Pdf
Output
Embedding
Languages
More than 100
License
Proprietary

Capabilities

Semantic SearchRetrievalRagMultimodal EmbeddingsClassification

Endpoints

Embed

Deployment

Cohere ApiModel VaultPrivate Deployment

Catalog details only

Embedding pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Rerank

2

Rerank 4 Fast

rerank-v4.0-fast

Rerank 4

RerankActiveAPI available

Multilingual low-latency reranker for high-throughput retrieval and semi-structured JSON.

Context
32,768 tokens
Input
Text, Json
Output
Ranking Scores
Languages
More than 100
License
Proprietary

Capabilities

RerankingMultilingualSemi Structured DataLow LatencyHigh Throughput

Endpoints

Rerank

Deployment

Cohere ApiModel VaultPrivate Deployment

Catalog details only

Per-search pricing is not compatible with token generation costs.

Official sourceVerified Jul 14, 2026

Rerank 4 Pro

rerank-v4.0-pro

Rerank 4

RerankActiveAPI available

Multilingual high-accuracy reranker for complex retrieval and enterprise search.

Context
32,768 tokens
Input
Text, Json
Output
Ranking Scores
Languages
More than 100
License
Proprietary

Capabilities

RerankingMultilingualSemi Structured DataHigh Accuracy

Endpoints

Rerank

Deployment

Cohere ApiModel VaultPrivate Deployment

Catalog details only

Per-search pricing is not compatible with token generation costs.

Official sourceVerified Jul 14, 2026

Aya

6

Aya Expanse 32B

c4ai-aya-expanse-32b

Aya Expanse

AyaActiveAPI available

32B multilingual text model for generation, summarization, translation, and enterprise communication.

Context
128,000 tokens
Maximum output
4,000 tokens
Input
Text
Output
Text
Languages
23
Parameters
32B
License
CC-BY-NC-4.0

Capabilities

Multilingual GenerationTranslationSummarizationEnterprise Text

Endpoints

Chat

Deployment

Cohere ApiOpen Weights

Verified token pricing

$0.5 input / $1.5 output per 1M tokens

Use calculator
Official sourceVerified Jul 14, 2026

Aya Vision 32B

c4ai-aya-vision-32b

Aya Vision

AyaActiveAPI available

32B multilingual multimodal model for image captioning, visual QA, generation, and translation.

Context
16,000 tokens
Maximum output
4,000 tokens
Input
Text, Image
Output
Text
Languages
23
Parameters
32B
License
CC-BY-NC-4.0

Capabilities

MultimodalImage CaptioningVisual QaTranslationMultilingual Generation

Endpoints

Chat

Deployment

Cohere ApiOpen Weights

Catalog details only

Unavailable pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Tiny Aya – Global

tiny-aya-global

Tiny Aya

AyaActiveAPI available

Best balance across languages and regions.

Context
8,000 tokens
Maximum output
8,000 tokens
Input
Text
Output
Text
Languages
70
Parameters
3.35B
License
CC-BY-NC-4.0

Capabilities

TranslationMultilingual UnderstandingTarget Language GenerationOn Device

Endpoints

Chat

Deployment

Cohere ApiOpen WeightsLocal Deployment

Catalog details only

Unavailable pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Tiny Aya – Earth

tiny-aya-earth

Tiny Aya

AyaActiveAPI available

Optimized for West Asian and African languages.

Context
8,000 tokens
Maximum output
8,000 tokens
Input
Text
Output
Text
Languages
70
Parameters
3.35B
License
CC-BY-NC-4.0

Capabilities

TranslationMultilingual UnderstandingTarget Language GenerationOn Device

Endpoints

Chat

Deployment

Cohere ApiOpen WeightsLocal Deployment

Catalog details only

Unavailable pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Tiny Aya – Fire

tiny-aya-fire

Tiny Aya

AyaActiveAPI available

Optimized for South Asian languages.

Context
8,000 tokens
Maximum output
8,000 tokens
Input
Text
Output
Text
Languages
70
Parameters
3.35B
License
CC-BY-NC-4.0

Capabilities

TranslationMultilingual UnderstandingTarget Language GenerationOn Device

Endpoints

Chat

Deployment

Cohere ApiOpen WeightsLocal Deployment

Catalog details only

Unavailable pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Tiny Aya – Water

tiny-aya-water

Tiny Aya

AyaActiveAPI available

Optimized for European and Asia-Pacific languages.

Context
8,000 tokens
Maximum output
8,000 tokens
Input
Text
Output
Text
Languages
70
Parameters
3.35B
License
CC-BY-NC-4.0

Capabilities

TranslationMultilingual UnderstandingTarget Language GenerationOn Device

Endpoints

Chat

Deployment

Cohere ApiOpen WeightsLocal Deployment

Catalog details only

Unavailable pricing is not compatible with generation input/output costs.

Official sourceVerified Jul 14, 2026

Pricing sources

Pricing source and update notes

CostRivo shows official source references and verification metadata where available. Review provider pricing pages before making high-volume purchasing decisions.

Official provider pricing

Open Cohere pricing references for current provider terms, tiers, and availability notes.

Open source

Pricing table

Review CostRivo's cross-provider pricing table, verification dates, source links, and lifecycle labels.

Open source

FAQ

Cohere cost planning questions

Short answers for using this provider calculator.

How is Cohere API cost calculated?

The calculator multiplies input, output, and cached input tokens by the selected model pricing, then scales the result by request volume.

Which Cohere models have the lowest tracked token prices?

Lowest input price: Command R7B at $0.0375 per 1M tokens; Lowest output price: Command R7B at $0.15 per 1M tokens; Lowest combined token price: Command R7B at $0.1875 for 1M input plus 1M output tokens. Each criterion is calculated independently from eligible verified token prices.

Does this estimate include cached input pricing?

Cached input pricing is not listed for the current provider models in this data set.

When was Cohere pricing last verified?

The most recent model verification shown by CostRivo is Sep 9, 2026. Individual model rows retain their own source and verification details.

Pricing updates

Get pricing updates

Get notified when AI model prices change, new providers are added, product updates ship, launch notes go out, or Costrivo introduces future premium planning features.

No spam. Pricing and product updates only. We only store your email, this page, and signup time.

Optional and separate from calculator inputs.