Claude Sonnet 4.5 is the best coding model in the world. It's the strongest model for building complex agents. It's the best model at using computers. And it shows substantial gains in reasoning and math.
Modalities
Context window
1,000,000
Pricing
input / output per 1M
Reasoning
Streaming
Real-time token-by-token response streaming
Function calling
Connect the model to external tools and systems
Structured outputs
Return responses in JSON schema format
Reasoning
Extended thinking before responding
Memory
Persistent memory across conversations
Context editing
Edit conversation context dynamically
Web search
Search the internet for real-time information
File search
Search and retrieve from uploaded files
Code execution
Execute code in a sandboxed environment
Computer use
Control and interact with computer interfaces
File creation
Create documents, spreadsheets, and slides
| SWE BENCH VERIFIED | 77.2 |
| OSWORLD | 61.4 |
| HIGH COMPUTE SWE BENCH | 82.0 |
| Release date | 2025-09-29 |
| Model ID | claude-sonnet-4.5 |
| Provider | Anthropic |
Get detailed information about Claude Sonnet 4.5, including its context window of 1000000 tokens, pricing per million tokens, supported input and output modalities, and benchmark scores. This model from Anthropic offers specific capabilities for natural language processing, code generation, and complex reasoning tasks that set it apart from alternatives.
Compare input and output token pricing for Claude Sonnet 4.5 against other models in its class. Understanding LLM pricing is essential for budgeting your AI applications at scale. We break down the cost per million tokens for both input and output so you can estimate the total cost of your workloads and compare value across providers.
Review benchmark performance data for Claude Sonnet 4.5 across key evaluation metrics. Compare its reasoning, coding, and language understanding capabilities against competing models to determine if it is the right fit for your specific requirements, whether that involves complex analysis, creative generation, or efficient inference at scale.