← All models

Claude 4.5 Sonnet (Non-reasoning)

deprecated by benchmark source
Anthropic · released 2025-09-29 · 11 offers

Context 1M tokens

Top 5 cheapest providers (Adjusted $/task)

The same list price can give a different adjusted $/task (caching, token efficiency) — click a price for its inputs.

Within the active global provider, residency and confidentiality filters.

#ProviderAdjusted $/task
1Amazon Bedrock
2Google
3Google Vertex AI
4AWS Bedrock
5Azure

Composite

5 of 7 inputs · 4 from the model family6 radar axes: DesignArena's two boards share one

49.5

includes −2.5 for its Benchmaxxing signal (from 52.0; why, switch off in Options)

AA CodingCoding Agent v1.4AA IntelligenceAA AgenticEpoch ECISoftware ECIDesignArenaAA Intelligence: percentile 69Epoch ECI: percentile 44Software ECI: percentile 31DesignArena: percentile 11

AA Coding Coding Agent v1.4 AA Intelligence 19.3AA Agentic Epoch ECI 146.8Software ECI 147.7DesignArena 1075/1060

Radar: percentile among all models measured on each input; a gap means not measured.

Benchmark sheet

12 of 140 registered benchmark versions · bars show the percentile among all models measured on each benchmark.

Composite attachments (used in the score, not counted as exact benchmarks):
  • Epoch ECI · attached
  • Software ECI · attached
  • DesignArena Web Apps (agentic) · attached
  • DesignArena Full-Stack · attached
Compare this model →

Agentic

Instruction-following

Knowledge

Long-context

  • AA-LCR v1.15054.0%

    AA-LCR v1.1 v1.1 · Published board — Tests reasoning across multiple long documents with corrected answer keys and grading.

Math

  • AIME 2025 (AA) v20253837.0%

    AIME 2025 (AA) v2025 · Published board — Advanced mathematical problem solving on AIME I and II 2025.

Reasoning

Safety/Alignment

Science

Tool-use

Vision

Missing coverage · 132 benchmark versions

No result does not mean a zero, or that the model was never tested. Collection failures and disputed versions retain their distinct status.

Variants / reasoning settings

Artificial Analysis snapshot 2026-09-19 · Data: Artificial Analysis · Data: Epoch AI (CC BY)

VariantAA CodingAA Intelligence
Claude 4.5 Sonnet (Reasoning)52.121.2
Claude 4.5 Sonnet (Non-reasoning)19.3
Token offers by platform · 8 offers (Adjusted $/task)

Click any underlined price to see how it is estimated and where each input comes from. How we calculate adjusted cost.

8 offers within the active global filters; “—” means the catalog is active but no public token price is available.

OpenRouter (5)

Amazon Bedrock
Google
Amazon Bedrock
Google
Azure

Google Vertex AI (2)

Google Vertex AI
Google Vertex AI

AWS Bedrock (1)

AWS Bedrock
Claude 4.5 Sonnet (Non-reasoning) — benchmarks & cost | Benchmark Heaven