Models · Anthropic · Out sinceReleased 29 Sep 2025
Claude Sonnet 4.5
Claude Sonnet 4.5 is made by Anthropic. We don't have enough test results yet to rank it. It's expensive to use.
Auto-created from OpenRouter catalog; verify details.
- —
- —
- —
- Expensive$3 / $15
- $6
- 1M
- —
- 3 (3 independent3 indep.)
- 29 Sep 2025
- Anthropic
- —
Scores are out of 100 for each area, compared with every model we track; “#” is its rank. Context is how much text it can read at once (1M is roughly 700,000 words). Blended price mixes the cost of what you send and what it writes back.
You can't choose how long this one thinks.
No adjustable reasoning setting listed.
OpenRouter is a service that gives access to many AI models in one place. This model is listed there as anthropic/claude-sonnet-4.5.
Route it as anthropic/claude-sonnet-4.5 at $3 in / $15 out per 1M tokens, 1M context. Listed since 29 Sep 2025.
include_reasoningmax_completion_tokensmax_tokensreasoningresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Every test result
Every result
All the numbers,with where they came from.
Every result we've found for this model, with who measured it and a link to where we read it.
All raw rows (vendor and independent kept separate), with setting label, source, date, harness and notes.
| NotesNotes | ||||||
|---|---|---|---|---|---|---|
| SWE-bench Multilingual | 67.0% | Default | Independent testIndependentSWE-bench ↗ | 13 Feb 2026 | mini-SWE-agent 2.0.0a0 | Official SWE-bench bash-only (mini-SWE-agent) run; data file https://raw.githubusercontent.com/SWE-bench/swe-bench.github.io/master/data/leaderboards.json; leaderboard not updated since 2026-02-26 |
| SWE-bench Verified (bash-only, mini-SWE-agent) | 71.4% | High | Independent testIndependentSWE-bench ↗ | 17 Feb 2026 | mini-SWE-agent 2.0.0 | Official SWE-bench bash-only (mini-SWE-agent) run; data file https://raw.githubusercontent.com/SWE-bench/swe-bench.github.io/master/data/leaderboards.json; leaderboard not updated since 2026-02-26 |
| Vectara Hallucination Leaderboard (HHEM) | 12.0% | Defaultanthropic/claude-sonnet-4-5-20250929 | Independent testIndependentVectara ↗ | 22 Sep 2026 | — | factual consistency 88.0 %; answer rate 95.6 %; avg summary 127.8 words; HHEM-2.3 judge; effort not stated (API default) |