Skip to content
Bencher

Models · Cognition

SWE-2

Only from the makerNot on OpenRouterAvailable nowGenerally availableClosed (can't be downloaded)Proprietary

SWE-2 is made by Cognition. We don't have enough test results yet to rank it.

Cognition model used as 'sidekick' in Devin Fusion CLI; not on OpenRouter.

Writing & creativity
—
Research & analysis
—
Coding
—
Price
Price unknown— / —
Price per 1M (blended)Blended / 1M
—
MemoryContext
—
Longest answerMax output
—
Test resultsResults
1 (1 independent1 indep.)
Out sinceReleased
—
Made byVendor
Cognition
UnderstandsInputs
—

Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.

Thinking levels

Reasoning settings

You can't choose how long this one thinks.

No adjustable reasoning setting listed.

Read more

Links

—

Every test result

Every result

All the numbers,with where they came from.

Every result we've found for this model, with who measured it and a link to where we read it.

All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.

All results for this model
NotesNotes
FrontierCode50.0%Maxswe-2_maxIndependent testIndependentFrontierCode (Cognition) via Epoch AI ↗1 Oct 2026devinread from Epoch AI benchmark_data.zip (frontiercode_external.csv); original leaderboard https://cognition.com/frontiercode; harness devin; Mean@5