Skip to content
  • Models
  • Rankings
  • Ori
Sign Up
Sign Up
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for cohere

Cohere: Command A

cohere/command-a

Model weights
Compare

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary and open-weights models Command A delivers maximum performance with minimum hardware costs, excelling on business-critical agentic and multilingual tasks.

Modalities

In / Out Price

$2.50 / $10per 1M

Context

256K

Released

Mar 13, 2025

Knowledge Cutoff

Aug 2024

Compare
ProvidersPricingPerformanceUptimeBenchmarksAppsActivityFAQExplore

Providers

This model is hosted by one provider. OpenRouter forwards every request to it directly — no routing decisions to make.

Pricing

The average price customers actually pay for this model, next to the prices providers post. Caching and discounts mean the price actually paid is often well below the listed one.

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better).

Uptime

Uptime is the percentage of the past 3 days that at least one provider was responding to requests. Availability is the percentage of time that inference was successfully served. OpenRouter continuously monitors and uses the next-best provider when one returns an error.

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows where this model lands among all models on OpenRouter.

Benchmark score summary for Cohere: Command A (Artificial Analysis)
SourceBenchmarkScore
Artificial AnalysisCommand A+ Intelligence Index13.9
Artificial AnalysisCommand A+ Coding Index27.8
Artificial AnalysisCommand A+ Agentic Index3.6
Artificial AnalysisCommand A+ GPQA Diamond76.1%
Artificial AnalysisCommand A+ HLE12.0%
Artificial AnalysisCommand A+ IFBench73.9%
Artificial AnalysisCommand A+ τ²-Bench Telecom80.7%
Artificial AnalysisCommand A+ AA-LCR52.7%
Artificial AnalysisCommand A+ GDPval-AA7.9%
Artificial AnalysisCommand A+ CritPt0.3%
Artificial AnalysisCommand A+ SciCode38.5%
Artificial AnalysisCommand A+ Terminal-Bench Hard25.0%
Artificial AnalysisCommand A+ AA-Omniscience Accuracy8.9%
Artificial AnalysisCommand A+ AA-Omniscience Non-Hallucination Rate85.8%

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

Activity

Token volume and request traffic to this model over time.

Quick Start

Drop-in code to call this model. OpenRouter's API is OpenAI-compatible — most SDKs work by just swapping the base URL. The only thing that changes between models is the model slug below.

Explore more models

AI Model RankingsRanking
$2.50$10.000.44s48 tps
99.92%

Throughput

48tok/s

P50, best across providers

Latency

0.44s

P50, best provider

Uptime (3d)The model was reachable. Request routed to a provider.

100.00%

Availability (3d)The model returned inference from any provider. Errors and empty responses count against it.

99.96%

Availability over the last 3 days

Last 72 hours
Availability 99.96%
3 Days Ago2 Days AgoYesterdayNow

Availability over the last 24 hours

OpenRouter Availability
99.93%

When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.

1.
Favicon for https://github.com/conifer-ai/conifer
Conifer toolkit serve
new
17.4Mtokens
2.
Favicon for https://github.com/persona-base-explore
exp272-s1-clause
new
12.2Mtokens
3.
Favicon for https://github.com/llm-milgram
llm-milgram obedience census
new
9.09Mtokens
4.
Favicon for https://sillytavern.app/
SillyTavern
SillyTavern is the LLM frontend for power users, a chat interface that connects to any model and gives deep control through character creation, roleplay, and prompt customization.
6.78Mtokens
5.
Favicon for https://chathub.gg/
ChatHub
new
4.39Mtokens

Frequently asked questions

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary and open-weights models Command A delivers maximum performance with minimum hardware costs, excelling on business-critical agentic and multilingual tasks.

Command A costs $2.50/M input tokens and $10.00/M output tokens.

Command A has a 256,000 token context window. It supports up to 8,192 completion tokens.

The Command A endpoint shown on this page does not accept tools, so function calling is unavailable there. It also supports structured outputs via a JSON schema in response_format.

North Mini Code (free), Command R7B (12-2024), Command R (08-2024) and 1 more are other text models from Cohere.

Command A was released on March 13, 2025. Its knowledge cutoff is August 31, 2024.

More models from Cohere

North Mini Code

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized for code generation, agentic software engineering, and terminal tasks, and is trained to generalize across agent harnesses such as OpenCode and SWE-Agent.

It offers a 256K-token context window with up to 64K tokens of output, supports interleaved reasoning and tool use via JSON schema, and is released open-weight under the Apache 2.0 license. Its small active-parameter footprint enables low-latency inference, including on local hardware.

Text256K contextFree
Rerank 4 Pro

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing required, and state of the art performance with low latency.

Rerank$0.0025/search
Rerank 4 Fast

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing required, and high performance with lowest latency.

Rerank$0.002/search
Rerank v3.5

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search pipelines.

Rerank$0.001/search
Command R7B

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning and multiple steps.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context$0.0375 / $0.15
Command R

command-r-08-2024 is an update of the Command R with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and is competitive with the previous version of the larger Command R+ model.

Read the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context$0.15 / $0.60
Command R+

command-r-plus-08-2024 is an update of the Command R+ with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint the same.

Read the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context$2.50 / $10
Command R+

Command R+ is a new, 104B-parameter LLM from Cohere. It's useful for roleplay, general consumer usecases, and Retrieval Augmented Generation (RAG).

It offers multilingual support for ten key languages to facilitate global business operations. See benchmarks and the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context
Command R+

Command R+ is a new, 104B-parameter LLM from Cohere. It's useful for roleplay, general consumer usecases, and Retrieval Augmented Generation (RAG).

It offers multilingual support for ten key languages to facilitate global business operations. See benchmarks and the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context
Command R

Command-R is a 35B parameter model that performs conversational language tasks at a higher quality, more reliably, and with a longer context than previous models. It can be used for complex workflows like code generation, retrieval augmented generation (RAG), tool use, and agents.

Read the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context
Command

Command is an instruction-following conversational model that performs language tasks with high quality, more reliably and with a longer context than our base generative models.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text4K context
Command R

Command-R is a 35B parameter model that performs conversational language tasks at a higher quality, more reliably, and with a longer context than previous models. It can be used for complex workflows like code generation, retrieval augmented generation (RAG), tool use, and agents.

Read the launch post here.

Use of this model is subject to Cohere's Usage Policy and SaaS Agreement.

Text128K context