LLMcompare

Search

Search for a command to run...

A · closed · Composer

Composer 2.5

Cursor

B · closed · GPT-5.x

GPT-5.6 Terra

OpenAI

0:3

overall capability wins (0 ties across 3 scored). GPT-5.6 Terra leads overall.

  • Coding0:1
  • Pricing3:0
  • Speed0:2
  • Specs0:2

Verdict

GPT-5.6 Terra leads coding (widest gap: +8.8 on CursorBench); GPT-5.6 Terra offers ~5.3x larger context (1.1M vs 200K); GPT-5.6 Terra allows ~2.0x more max output (128K vs 64K).

  • GPT-5.6 Terra leads coding (widest gap: +8.8 on CursorBench)
  • GPT-5.6 Terra offers ~5.3x larger context (1.1M vs 200K)
  • GPT-5.6 Terra allows ~2.0x more max output (128K vs 64K)
  • Composer 2.5 is ~4.5x cheaper on a blended token basis than GPT-5.6 Terra

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Composer 2.5
$1.13
GPT-5.6 Terra
$5

Composer 2.5 is about 4.4x cheaper on this workload.

Price per 1M tokens

Full comparison

Organization

Cursor

OpenAI

Family

Composer

GPT-5.x

License

Proprietary

Proprietary

Open weights

No

No

Release

May 18, 2026

Jul 9, 2026

Knowledge cutoff

-

-

API / provider

Cursor

OpenAI

Modalities

text → text

text, image → text

Specs

Context window

200K

1.1M

B +850K

Max output

64K

128K

B +64K

Parameters

—

—

—

Pricing

Input $/1M

$0.50

$2

A +$1.50

Output $/1M

$2.50

$12

A +$9.50

Blended $/1M (3∶1)

$1

$4.50

A +$3.50

Speed

tok/s

90

135

B +45

TTFT (s)

0.35

0.3

B +0.05

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

—

92.5

—

Humanity's Last Exam

—

42.9

—

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

—

46.8

—

AA-LCR v1.1

—

83.0

—

CritPt

—

30.0

—

MMMU-Pro

—

80.7

—

IFBench

—

71.2

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

63.4

—

SWE-bench Multilingual

79.8

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

—

88.0

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

20.8

—

BigCodeBench

—

—

—

SciCode

—

55.0

—

CursorBench

56.1

64.9

B +8.8 pts

SWE-Rebench

—

—

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

—

1,521

—

Terminal-Bench 4.0

—

21.5

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

—

1,466

—

Composer 2.5

Context
200K / 64K out
Parameters
—
Price
$0.50 / $2.50
Speed
90 tok/s · 0.35s TTFT
Modalities
text → text
License
Proprietary

Cursor's own agentic coding model (standard tier) — Cursor reports 69.3% Terminal-Bench 2.0 and 79.8% SWE-bench Multilingual; draws from the Cursor Models usage pool at $0.50/$2.50.

GPT-5.6 Terra

Context
1.1M / 128K out
Parameters
—
Price
$2 / $12
Speed
135 tok/s · 0.3s TTFT
Modalities
text, image → text
License
Proprietary

Balanced GPT-5.6 mid-tier for everyday work; July 30, 2026 API cut brought Standard rates to $2/$12 (20% lower), with strong throughput for the price.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

Coding

CursorBench: B +8.8

Pick a different pair · Back to catalog