Claude Sonnet 5
Mid-tierby Anthropic
Claude Sonnet 5 is Anthropic's most agentic Sonnet-class model, released June 30, 2026 as an upgrade to Sonnet 4.6 that narrows the gap to Opus 4.8 on reasoning, tool use, coding, computer use, and knowledge work while staying priced below the flagship. It launched at $2 per million input tokens and $10 per million output as introductory pricing through August 31, 2026, but Anthropic cancelled the scheduled September 1 increase to $3 and $15, so $2 and $10 is now the standard rate, with cache hits at $0.20. It ships a 1 million token context window with context compaction and adaptive thinking with selectable effort levels up to xhigh. Anthropic's own numbers show 85.2 on SWE-bench Verified, 63.2 on SWE-bench Pro, 78.3 on SWE-bench Multilingual, 81.2 on OSWorld-Verified, 84.7 on BrowseComp, and 80.4 on Terminal-Bench 2.1 (beating Opus 4.8 at 74.6 on that specific benchmark). Available on the Claude API, Amazon Bedrock, Google Vertex, and Microsoft Foundry at launch.
Input Price
$2.00
per 1M tokens
Output Price
$10.00
per 1M tokens
Context Window
1M
tokens
Released
2026-06
API access
Capabilities
Key Strengths
- ✓1M token context window with context compaction
- ✓$2/$10 is now standard after Anthropic cancelled the $3/$15 step
- ✓Adaptive thinking up to xhigh effort level
- ✓Beats Opus 4.8 on Terminal-Bench 2.1 (80.4 vs 74.6)
- ✓Day-one availability on Claude API, Bedrock, Vertex, and Foundry
Best For
- ▸Cost-sensitive agentic coding pipelines
- ▸Long-running tool-use workflows
- ▸Production RAG and knowledge work
- ▸Computer-use and browser agents
Benchmark Scores
| Benchmark | Score | Description |
|---|---|---|
| SWE-bench | 85.2 | Real-world software engineering tasks from GitHub issues (SWE-bench Verified) |
| MMLU-Pro | 87.5 | General knowledge and reasoning across 57 subjects |
| GPQA Diamond | 88.9 | Graduate-level science questions verified by domain experts |
| BrowseComp | 84.7 | Agentic web search and browsing over hard-to-find facts |
| FrontierCode v1.1 | 42.7 | Agentic coding on frontier software engineering tasks (Main split) |
| Terminal-Bench 4.0 | 8.1 | Long-horizon agentic work in a terminal across software, science, ML, operations, hardware, security, and media (66 tasks, all-or-nothing verifiers) |
| Humanity's Last Exam (tools) | 57.4 | Multidisciplinary expert-level reasoning with tool access |
Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.
Pricing Details
Input tokens
$2.00
per 1M tokens
Output tokens
$10.00
per 1M tokens
Estimated cost per 1K requests
$7.00
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.
Related Models
Claude Mythos 5.1
FlagshipAnthropic
$10.00 in / $50.00 out
Claude Opus 5.5
FlagshipAnthropic
$4.00 in / $20.00 out
Claude Opus 5
FlagshipAnthropic
$5.00 in / $25.00 out
Claude Fable 5.1
FlagshipAnthropic
$10.00 in / $50.00 out
Claude Fable 5
FlagshipAnthropic
$10.00 in / $50.00 out
Claude Opus 4.8
FlagshipAnthropic
$5.00 in / $25.00 out
Claude Opus 4.7
Mid-tierAnthropic
$5.00 in / $25.00 out
Claude Opus 4.6
FlagshipAnthropic
$5.00 in / $25.00 out
Claude Sonnet 4.6
Mid-tierAnthropic
$3.00 in / $15.00 out
Claude Haiku 4.5
BudgetAnthropic
$1.00 in / $5.00 out
GPT-6 Sol
Mid-tierOpenAI
$2.00 in / $10.00 out
GPT-5.6 Terra
Mid-tierOpenAI
$2.00 in / $12.00 out
o3-mini
Mid-tierOpenAI
$1.10 in / $4.40 out
Gemini 3.8 Flash
Mid-tier$0.75 in / $3.75 out