Claude Sonnet 4.5
A refined iteration of Claude Sonnet 4 with improved performance on graduate-level reasoning and coding benchmarks. Claude Sonnet 4.5 delivers notably stronger results on GPQA and competitive maths whilst maintaining the same pricing as its predecessor.
You might also consider
OpenAI's most powerful reasoning model, using extended chain-of-thought to tackle the hardest problems in mathematics, science, and coding. o3 sets new standards on GPQA and competitive maths at the cost of higher latency and price.
Anthropic's most powerful and intelligent model, built for the most demanding tasks where quality outweighs cost. Claude Opus 4 leads on complex multi-step reasoning, graduate-level science, and nuanced long-form writing.
xAI's most capable model, trained on a 100,000-GPU cluster and setting new benchmarks in mathematics and scientific reasoning. Grok 3 integrates real-time data from the X platform and leads the Arena ELO leaderboard among commercial models.