Claude Sonnet 5 benchmarks
Claude Sonnet 5 is a proprietary model from Anthropic, released 30 Jun 2026. It ranks 37th of 51 on our overall leaderboard and 14th of 51 for long context. At A$5.62 per million tokens (blended), it costs 1.6× the median model we track. It writes about 73 tokens a second, and answers start after 1.4 min on average, thinking included. It can run with requests kept in Australia on AWS Bedrock in Melbourne.
- CONTEXT WINDOW
- 1,000,000 tokens
- MAX OUTPUT
- 128,000 tokens
- INPUT
- text, image
- WEIGHTS
- Proprietary
- RELEASED
- 30 Jun 2026
- OUTPUT SPEED
- 73 tok/s
- ANSWER STARTS AFTER
- 1.4 min (thinking included)
Claude Sonnet 5 across our rankings.
Benchmark results.
Independent test scores are for the adaptive reasoning, max effort setting. “Among tracked” ranks Claude Sonnet 5 against the 51 models on this site; LMArena ranks run across its full leaderboard.
| BENCHMARK | RESULT | AMONG TRACKED | SOURCE DETAIL |
|---|---|---|---|
| INDEPENDENT TESTS · ARTIFICIAL ANALYSIS | |||
| AA Intelligence Index | 38.2 | #28 of 51 | Adaptive reasoning, max effort |
| AA Coding Index | 71.5 | #20 of 42 | Adaptive reasoning, max effort |
| GPQA Diamond | 91.1% | #27 of 44 | Adaptive reasoning, max effort |
| Humanity’s Last Exam | 41.3% | #28 of 51 | Adaptive reasoning, max effort |
| AA-LCR | 82.0% | #16 of 51 | Adaptive reasoning, max effort |
| Terminal-Bench 2.1 | 80.5% | #21 of 42 | Adaptive reasoning, max effort |
| τ²-Bench (banking) | 37.3% | #21 of 41 | Adaptive reasoning, max effort |
| SciCode | 54.3% | #25 of 46 | Adaptive reasoning, max effort |
| BLIND HUMAN VOTES · LMARENA | |||
| LMArena Text35,301 votes · ±5 | 1461 | #32 of 42 | #51 of 402 on LMArena · 13 Sep 2026 |
| LMArena WebDev11,511 votes · ±7 | 1538 | #26 of 45 | #31 of 129 on LMArena · 22 Sep 2026 |
| LMArena Vision10,289 votes · ±8 | 1266 | #24 of 33 | #33 of 152 on LMArena · 13 Sep 2026 |
| LMArena Document7,411 votes · ±8 | 1476 | #12 of 26 | #15 of 44 on LMArena · 13 Sep 2026 |
| LMArena Search40,230 votes · ±7 | 1196 | #10 of 12 | #12 of 34 on LMArena · 24 Aug 2026 |
| LMArena Agent30,631 sessions | #9 | #8 of 37 | #9 of 46 on LMArena · 15 Sep 2026 |
What Claude Sonnet 5 costs in Australian dollars.
List prices converted at the snapshot’s RBA rate, excluding GST. Reasoning models are billed for their thinking as output tokens, so heavier settings cost more than the job estimates show.
| PER MILLION TOKENS | AUD | USD LIST |
|---|---|---|
| Input tokens | A$2.81 | US$2.00 |
| Output tokens | A$14.04 | US$10.00 |
| Blended (3 in : 1 out) | A$5.62 | US$4.00 |
| JOB | COST (AUD) | MEDIAN MODEL |
|---|---|---|
| Customer support replyPER 1,000 REPLIES | A$9.83 | ≈ A$6.58 |
| Summarise a 30-page documentPER 100 DOCUMENTS | A$6.74 | ≈ A$4.74 |
| Agentic coding taskPER 10 TASKS | A$5.33 | ≈ A$3.72 |
| SETTING | INTELLIGENCE | A$ / 1M | SPEED |
|---|---|---|---|
| Adaptive reasoning, max effort | 38.2 | A$5.62 | 73 tok/s |
| Adaptive reasoning, xhigh effort | 34.4 | A$5.62 | 72 tok/s |
| Adaptive reasoning, high effort | 31.7 | A$5.62 | 65 tok/s |
| Adaptive reasoning, medium effort | 28.1 | A$5.62 | 66 tok/s |
| Adaptive reasoning, low effort | 24.3 | A$5.62 | 64 tok/s |
| Non-reasoning, high effort | 23.2 | A$5.62 | 61 tok/s |
Running Claude Sonnet 5 in Australia.
It can run with requests kept in Australia on AWS Bedrock in Melbourne. How the platforms compare →
| PLATFORM | AVAILABILITY | DETAIL |
|---|---|---|
| AWS Bedrock · Sydney | GLOBAL | AWS's model card also lists an Australia-only au. profile, but its regional summary shows global only — treat as global until they agree. Provider docs, checked 23 Sep 2026 |
| AWS Bedrock · Melbourne | IN REGION | In-region through the bedrock-mantle endpoint (Anthropic Messages API); the standard runtime endpoint uses the au. profile. Provider docs, checked 23 Sep 2026 |
| Azure · Australia East | NOT IN AU | |
| Google Vertex AI · Sydney | GLOBAL |
Compare it with.
Claude Sonnet 5, answered.
How much does Claude Sonnet 5 cost in Australian dollars?
Claude Sonnet 5 lists at US$2.00 per million input tokens and US$10.00 per million output tokens — A$2.81 and A$14.04 at A$1 = US$0.7123 (RBA, 22 Sep 2026), excluding GST. A typical customer support reply works out at about A$9.83 per 1,000 replies, before any reasoning tokens.
Can I use Claude Sonnet 5 in Australia with data kept onshore?
Claude Sonnet 5 can run with requests kept in Australia on AWS Bedrock in Melbourne. Availability comes from the cloud providers’ own documentation; check your provider’s current terms before relying on it for data-residency obligations.
Is Claude Sonnet 5 good for coding?
It ranks 22nd of 42 on our coding ranking, with an Artificial Analysis Coding Index of 71.5 and an LMArena WebDev rating of 1538 (31st of 129). The current leader is Claude Fable 5.1.
How good is Claude Sonnet 5 at agent and automation work?
It ranks 19th of 41 on our agents ranking, completing 37.3% of τ²-Bench banking customer-service tasks and 80.5% of Terminal-Bench 2.1 tasks.
How fast is Claude Sonnet 5?
Artificial Analysis measures Claude Sonnet 5 at about 73 tok/s of output, with the answer starting after 1.4 min on average once thinking time is included (adaptive reasoning, max effort setting).
What is Claude Sonnet 5’s context window?
Anthropic documents a context window of 1,000,000 tokens (1M), with up to 128,000 output tokens. On AA-LCR, which tests reasoning across ~100,000-token document sets, it scores 82.0%.
Ratings: LMArena leaderboard dataset (CC BY 4.0), rescaled for the lens scores · leaderboards to 22 Sep 2026
Evaluations, prices and speed: Artificial Analysis (artificialanalysis.ai) · fetched 23 Sep 2026
Exchange rate: Reserve Bank of Australia, table F11.1 · A$1 = US$0.7123 on 22 Sep 2026 · prices exclude GST
Snapshot 23 Sep 2026 · updated weekly · How the rankings work →