LLMcompare

Search

Search for a command to run...

A · open · Other

Ling-3.0 Flash Fin

InclusionAI

B · closed · Other

Step-3.5 Flash

StepFun

1:0

overall capability wins (0 ties across 1 scored). Ling-3.0 Flash Fin leads overall.

  • Specs1:0

Verdict

Ling-3.0 Flash Fin offers larger context (262K vs 256K); Ling-3.0 Flash Fin ships open weights (MIT).

  • Ling-3.0 Flash Fin offers larger context (262K vs 256K)
  • Ling-3.0 Flash Fin ships open weights (MIT)

Cost for 1M in + 250K out

Illustrative chat workload at primary-provider list prices.

Ling-3.0 Flash Fin
-
Step-3.5 Flash
$0.16

Price per 1M tokens

Full comparison

Organization

InclusionAI

StepFun

Family

Other

Other

License

MIT

Proprietary

Open weights

Yes

No

Release

Sep 3, 2026

Feb 18, 2026

Knowledge cutoff

-

-

API / provider

No primary API price listed

StepFun

Modalities

text → text

text, image → text

Specs

Context window

262K

256K

A +6K

Max output

-

16K

—

Parameters

124B (5.1B act.)

—

—

Pricing

Input $/1M

—

$0.09

—

Output $/1M

—

$0.30

—

Blended $/1M (3∶1)

-

$0.14

—

Speed

tok/s

-

120

—

TTFT (s)

-

0.25

—

Reasoning

MMLU-Pro

—

—

—

GPQA Diamond

—

83.1

—

Humanity's Last Exam

—

21.1

—

AIME 2025

—

—

—

MATH-500

—

—

—

Humanity's Last Exam (with tools)

—

—

—

AA-Omniscience Accuracy

—

23.6

—

AA-LCR v1.1

—

50.0

—

CritPt

—

2.5

—

MMMU-Pro

—

—

—

IFBench

—

64.6

—

Chartography

—

—

—

Chartography (With Tools)

—

—

—

Coding

SWE-bench Verified

—

—

—

SWE-bench Pro

—

—

—

SWE-bench Multilingual

—

—

—

LiveCodeBench

—

—

—

Terminal-Bench 2.1

—

—

—

Aider Polyglot

—

—

—

Terminal-Bench 3

—

—

—

BigCodeBench

—

—

—

SciCode

—

40.4

—

CursorBench

—

—

—

SWE-Rebench

—

59.6

—

NL2Repo-Bench

—

—

—

DeepSWE

—

—

—

WebDev Arena

—

—

—

Terminal-Bench 4.0

—

—

—

CursorBench 4.0

—

—

—

FrontierCode v1.1 (Main)

—

—

—

Arena

LMArena Elo

—

1,394

—

Ling-3.0 Flash Fin

Context
262K
Parameters
124B (5.1B act.)
Price
—
Speed
—
Modalities
text → text
License
MIT

InclusionAI / Ant Group's finance-specialized fine-tune of Ling-3.0 Flash (124B total / 5.1B active MoE, MIT, 256K context), built with financial institutions for source-grounded financial research, multi-document analysis, valuation modeling, and spreadsheet workflows. Ant Group's model card reports evaluation on FinFIRST, FinSearchComp Verified, FinCRAFT, Finance Agent, APEX-Agents, SpreadsheetBench, and τ³-Banking without publishing numeric scores, so no benchmark values are recorded here.

Step-3.5 Flash

Context
256K / 16K out
Parameters
—
Price
$0.09 / $0.30
Speed
120 tok/s · 0.25s TTFT
Modalities
text, image → text
License
Proprietary

StepFun's fast multimodal API tier — competitive STEM scores at Flash pricing.

Benchmark charts

Winner bars are emphasized. Per-benchmark deltas sit above each chart.

These two models have no overlapping published benchmarks in our dataset. Compare specs and pricing instead.

Pick a different pair · Back to catalog