Skip to content

Command A

cohere/command-a
Language

The enterprise generation model, built for RAG, citations and private deployment.

Model overview

Context window

256K

tokens

Input price

quoted on request

Output price

quoted on request

Weekly volume

tokens / week

Open weights

No

API only

Capabilities

4

of 11 tags

Description

Command A is a language model from Cohere. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Cohere’s stated focus is enterprise retrieval and language stack. The weights are not published, so it is API-only.

Capabilities

Tool calling
Calls functions you define and returns their arguments as structured data, the basis of agents.
JSON mode
Constrains the response to valid JSON matching a schema you supply.
Streaming
Emits tokens as they are generated, so answers appear progressively rather than all at once.
Fine-tunable
Can be further trained on your own data to specialise it.

Strengths & weaknesses

Strengths

  • Grounded generation with citations
  • Private/on-prem options
  • Enterprise multilingual

Weaknesses

  • General benchmarks trail flagships
  • Premium pricing for the tier

Pricing

Free
Pricing for Command A, per 1M tokens
RatePriceUnit
Input$0.00per 1M tokens
Output not charged

Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.

Context window

256Ktokens

covers prompt and response together, so a long input leaves less room for the answer.

256K tokens against a peer range of 32K to 10M, median 256K. Compared across 24 language models. Logarithmic scale.

Supported features

Feature support for Command A
FeatureSupport
Tool callingSupported
JSON modeSupported
StreamingSupported
Vision inputNot supported
Audio inputNot supported
Long contextNot supported
Open weightsNot supported
ReasoningNot supported
Fine-tunableSupported
BatchNot supported
CachingNot supported

Example use cases

  • Enterprise RAG

    Grounded generation with citations

  • Regulated deployments

    Private/on-prem options

  • Grounded assistants

    Enterprise multilingual

Comparison

Command A compared with Embed v4 and Mistral Large 3
AttributeCommand AEmbed v4Mistral Large 3
Context window256K128K256K
Input price$0.00$0.00$0.00
Output price not charged not charged not charged
Tool callingSupportedNot supportedSupported
JSON modeSupportedNot supportedSupported
Vision inputNot supportedSupportedNot supported
Open weightsNot supportedNot supportedNot supported
Prompt cachingNot supportedNot supportedSupported

Best value in each row is highlighted. Illustrative placeholder data.

Provider information

View Cohere
Cohere3 models

Cohere · Toronto, Canada · est. 2019

Cohere never chased consumer chat. It built the enterprise retrieval stack instead.

Recent releases

    Documentation

    Frequently asked questions

    Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Every figure on this page is an illustrative placeholder.