LLMcompare

Search

Search for a command to run...

A · closed · Command

Command

Cohere

B · closed · Command

Command A Reasoning

Cohere

0:2

overall capability wins (0 ties across 2 scored). Command A Reasoning leads overall.

  • Pricing3:0
  • Speed2:0
  • Specs0:2

Verdict

Command A Reasoning offers ~62.5x larger context (256K vs 4K); Command A Reasoning allows ~8.0x more max output (33K vs 4K); Command is ~3.5x cheaper on a blended token basis than Command A Reasoning.

  • Command A Reasoning offers ~62.5x larger context (256K vs 4K)
  • Command A Reasoning allows ~8.0x more max output (33K vs 4K)
  • Command is ~3.5x cheaper on a blended token basis than Command A Reasoning

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Command
$1.50
Command A Reasoning
$5

Command is about 3.3x cheaper on this workload.

Price per 1M tokens

Full comparison

Organization

Cohere

Cohere

Family

Command

Command

License

Proprietary

Proprietary

Open weights

No

No

Release

Mar 1, 2023

Aug 1, 2025

Knowledge cutoff

-

-

API / provider

Cohere

Cohere

Modalities

text → text

text → text

Specs

Context window

4K

256K

B +252K

Max output

4K

33K

B +29K

Parameters

—

—

—

Pricing

Input $/1M

$1

$2.50

A +$1.50

Output $/1M

$2

$10

A +$8

Blended $/1M (3∶1)

$1.25

$4.38

A +$3.13

Speed

tok/s

70

45

A +25

TTFT (s)

0.4

0.8

A +0.40

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

—

—

—

Humanity's Last Exam

—

—

—

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

—

—

—

AA-LCR v1.1

—

—

—

CritPt

—

—

—

MMMU-Pro

—

—

—

IFBench

—

—

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

—

—

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

—

—

BigCodeBench

—

—

—

SciCode

—

—

—

CursorBench

—

—

—

SWE-Rebench

—

—

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

—

—

—

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

—

—

—

Command

Context
4K / 4K out
Parameters
—
Price
$1 / $2
Speed
70 tok/s · 0.4s TTFT
Modalities
text → text
License
Proprietary

Original Cohere Command API model — early enterprise RAG/chat baseline before Command R / R+ / A.

Command A Reasoning

Context
256K / 33K out
Parameters
—
Price
$2.50 / $10
Speed
45 tok/s · 0.8s TTFT
Modalities
text → text
License
Proprietary

Cohere's first explicit reasoning model — thinks before answering for nuanced agent and STEM workflows at 256K context.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

These two models have no overlapping published benchmarks in our dataset. Compare specs and pricing instead.

Pick a different pair · Back to catalog