Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with improvements across system design, code security, and specification adherence. The model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking.
Sonnet 4.5 also introduces stronger agentic capabilities, including improved tool orchestration, speculative parallel execution, and more efficient context and memory management. With enhanced context tracking and awareness of token usage across tool calls, it is particularly well-suited for multi-context and long-running workflows. Use cases span software engineering, cybersecurity, financial analysis, research agents, and other domains requiring sustained reasoning and tool use.
| $3.00 | $15.00 | $0.30 | 1.40s | 37 tps | ||
| $3.00 | $15.00 | $0.30 | -- | -- | ||
| $3.00 | $15.00 | $0.30 | 1.74s | 38 tps | ||
| $3.00 | $15.00 | $0.30 | 1.07s | 29 tps | ||
| $3.00 | $15.00 | $0.30 | 1.07s | 14 tps | ||
Not used in Standard routing:Why these endpoints are not used | ||||||
| $3.30 | $16.50 | $0.33 | 1.50s | 42 tps | ||
| $3.30 | $16.50 | $0.33 | 0.64s | 21 tps | ||
P50, best across providers
P50, best provider
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.
