Skip to content
Bencher

Models · Qwen (Alibaba) · Out sinceReleased 2 Apr 2026

Qwen3.6 Plus

Available on OpenRouterOn OpenRouterAvailable nowGenerally availableClosed (can't be downloaded)Proprietary

Qwen3.6 Plus is made by Qwen (Alibaba). We don't have enough test results yet to rank it. It's cheap to use.

Official list price not re-verified (OpenRouter: $0.325/$1.95). Context/max output from OpenRouter.

Writing & creativity
—
Research & analysis
—
Coding
—
Price
Cheap$0.33 / $1.95
Price per 1M (blended)Blended / 1M
$0.73
MemoryContext
1M
Longest answerMax output
66K
Test resultsResults
22 (3 independent3 indep.)
Out sinceReleased
2 Apr 2026
Made byVendor
Qwen (Alibaba)
UnderstandsInputs
text, image, video

Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.

Using it through OpenRouter

On OpenRouter

OpenRouter is a service that gives access to many AI models in one place. This model is listed there as qwen/qwen3.6-plus.

Route it as qwen/qwen3.6-plus at $0.33 in / $1.95 out per 1M tokens, 1M context. Listed since 2 Apr 2026.

frequency_penaltyinclude_reasoninglogprobsmax_tokenspresence_penaltyreasoningresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Every test result

Every result

All the numbers,with where they came from.

Every result we've found for this model, with who measured it and a link to where we read it.

All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.

All results for this model
NotesNotes
AIME 202695.3%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—Full AIME 2026 I & II
GPQA Diamond90.4%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
HMMT February 202687.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—HMMT Feb 26
Humanity's Last Exam28.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
Humanity's Last Exam (with tools)50.6%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
IMO-AnswerBench83.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
LiveCodeBench86.0%DefaultIndependent testIndependentVals.ai ↗1 Sep 2026—±0.977 stderr; $0.078/test
LiveCodeBench87.1%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—LiveCodeBench v6
LLM Creative Story-Writing Benchmark (Lech Mazur)-2.3DefaultQwen 3.6 PlusIndependent testIndependentLech Mazur (LLM Creative Story-Writing Benchmark) ↗28 Sep 2026—rank 48/56; Thurstone comparison score (centered at 0); est. win chance 19%; 95% bootstrap -2.423 to -2.101
MCP Atlas74.1%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—Public set
MCPMark48.2%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
MMLU-Pro88.5%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
MMMLU89.5%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
NL2Repo-Bench37.9%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026Claude Code
SimpleQA Verified44.1%Defaultqwen3.6-plusIndependent testIndependentEpoch AI ↗27 Aug 2026—Epoch-run (no tools); ±1.57 stderr
SkillsBench45.7%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026OpenCodeAvg5, 78-task subset
SWE-bench Multilingual73.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026Internal scaffold (bash + file-edit)
SWE-Bench Pro (public, v1)56.6%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026Internal scaffold (bash + file-edit)Qwen-corrected ('refined') SWE-bench Pro
SWE-bench Verified78.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026Internal scaffold (bash + file-edit)200K ctx
tau3-bench70.7%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—
Terminal-Bench 2.061.6%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026Terminus 2 (Harbor)3h timeout, avg of 5 runs
Toolathlon39.8%DefaultthinkingMaker's own figureVendor-reportedQwen (Alibaba) ↗2 Apr 2026—Reported as 'Tool Decathlon'